预印本开放获取
Deep Deterministic Policy Gradient Algorithm: A Systematic Review
Ebrahim Hamid Sumiea, Said Jadid Abdulkadir, Safwan Mahmood Al-Selwi, Alawi Alqushaibi, Mohammed Gamal Ragab, Suliman Mohamed Fati, Hitham Alhussian
Prince Sultan University Universiti Teknologi Petronas
阅读操作
确认中在文库中上传 PDF 后可生成中文音频讲解。
摘要与影响
摘要 · 节选
暂未获取摘要。可打开原文或 PDF,后续可基于全文生成更完整的速读。
逐年被引趋势
740
24
725
26
关键指标
15
被引次数 · OpenAlex
-
领域内被引倍数
同类平均 = 1
同类平均 = 1
-
引用位次
同领域 · 同年份 · 同类型
同领域 · 同年份 · 同类型
123
参考文献
Google Scholar 与 OpenAlex 的被引统计范围不同,数值存在差异属正常。
AI 辅助阅读
依据:文献信息
论文问答
当前基于文献信息回答
可就本文提问;依据不足时会说明。
学术脉络
学科主题
计算机 / AIReinforcement Learning in Robotics
Advanced Neural Network Applications · Autonomous Vehicle Technology and Safety
参考文献 123
Batch Normalization: Accelerating Deep Network Training by Reducing Internal Covariate Shift
被引 24,392Sergey Ioffe, Christian Szegedy · arXiv (Cornell University) · 2015
Advances in neural information processing systems 7
被引 22,314Krzysztof J. Cios, Mark E. Shields · Neurocomputing · 1997
MuJoCo: A physics engine for model-based control
被引 4,629Emanuel Todorov, Tom Erez, Yuval Tassa · 2012
此处列出前 3 条
引用本文 15
The innovation path of VR technology integration into music classroom teaching in colleges and universities
被引 11Yupeng Han, Han X. Lin, Chun Zeng · Scientific Reports · 2025
Deep deterministic policy gradient - model-agnostic meta-learning framework: Efficient adaptation in continuous control tasks
被引 8Ebrahim Hamid Sumiea, Said Jadid Abdulkadir, Hitham Alhussian · Results in Engineering · 2025
A deep reinforcement learning-based method for dynamic quality of service aware energy and occupant comfort management in intelligent buildings
被引 6Amirhossein Azimi, Omid Akbari · e-Prime - Advances in Electrical Engineering Electronics and Energy · 2024
按被引量排序,此处列出前 3 条