An Actor–Critic-based adapted Deep Reinforcement Learning model for multi-step traffic state prediction
Selim Reza, Marta Campos Ferreira, José J. M. Machado, João Manuel R. S. Tavares
Universidade do Porto INESC TEC Instituto de Ciência e Inovação em Engenharia Mecânica e Engenharia Industrial
阅读操作
确认中在文库中上传 PDF 后可生成中文音频讲解。
摘要与影响
Traffic state prediction is critical to decision-making in various traffic management applications. Despite significant advancements in Deep Learning (DL) models, such as Long Short-Term Memory (LSTM), Graph Neural Networks (GNN), and attention-based transformer models, multi-step predictions remain challenging. The state-of-the-art models face a common limitation: the predictions’ accuracy decreases as the prediction horizon increases, a phenomenon known as error accumulation. In addition, with the arrival of non-recurrent events and external noise, the models fail to maintain good prediction accuracy. Deep Reinforcement Learning (DRL) has been widely applied to diverse tasks, including optimising intersection traffic signal control. However, its potential to address multi-step traffic prediction challenges remains underexplored. This study introduces an Actor–Critic-based adapted DRL method to explore the solution to the challenges associated with multi-step prediction. The Actor network makes predictions by capturing the temporal correlations of the data sequence, and the Critic network optimises the Actor by evaluating the prediction quality using Q-values. This novel combination of Supervised Learning and Reinforcement Learning (RL) paradigms, along with non-autoregressive modelling, helps the model to mitigate the error accumulation problem and increase its robustness to the arrival of non-recurrent events. It also introduces a Denoising Autoencoder to deal with external noise effectively. The proposed model was trained and evaluated on three benchmark traffic flow and speed datasets. Baseline multi-step prediction models were implemented for comparison based on performance metrics such as Mean Absolute Error (MAE) and Root Mean Squared Error (RMSE). The results reveal that the proposed method outperforms the baselines by achieving average improvements of 0.26 to 21.29% in terms of MAE and RMSE for up to 24 time steps of prediction length on the three used datasets, at the expense of relatively higher computational costs. On top of that, this adapted DRL approach outperforms traditional DRL models, such as Deep Deterministic Policy Gradient (DDPG), in accuracy and computational efficiency. • A novel Actor–Critic-based adapted DRL method for multi-step traffic state prediction is proposed. • The Actor generates predictions, and the Critic assesses them using Q-values. • The model effectively handles non-recurrent events and mitigates external noise using a Denoising Autoencoder. • The proposed model predicts long-term traffic states more accurately than the baselines. • It necessitates approximately 10.47 times more computational resources compared to its alternatives.
逐年被引趋势
关键指标
同类平均 = 1
同领域 · 同年份 · 同类型
Google Scholar 与 OpenAlex 的被引统计范围不同,数值存在差异属正常。
AI 辅助阅读
依据:摘要
可就本文提问;依据不足时会说明。
学术脉络
学科主题
工程Traffic Prediction and Management Techniques
Traffic control and management · Hydrological Forecasting Using AI
参考文献 57
此处列出前 3 条
引用本文 7
按被引量排序,此处列出前 3 条