Learning Selective Communication for Multi-Agent Path Finding
Ziyuan Ma, Yudong Luo, Jia Pan
Simon Fraser University University of Waterloo University of Hong Kong
阅读操作
确认中在文库中上传 PDF 后可生成中文音频讲解。
摘要与影响
Learning communication via deep reinforcement learning (RL) or imitation learning (IL) has recently been shown to be an effective way to solve Multi-Agent Path Finding (MAPF). However, existing communication based MAPF solvers focus on broadcast communication, where an agent broadcasts its message to all other or predefined agents. It is not only impractical but also leads to redundant information that could even impair the multi-agent cooperation. A succinct communication scheme should learnwhichinformation is relevant and influential to each agent’s decision making process. To address this problem, we consider a request-reply scenario and proposeDecision Causal Communication(DCC), a simple yet efficient model to enable agents to select neighbors to conduct communication during both training and execution. Specifically, a neighbor is determined as relevant and influential only when the presence of this neighbor causes the decision adjustment on the central agent. This judgment is learned only based on agent’s local observation and thus suitable for decentralized execution to handle large scale problems. Empirical evaluation in obstacle-rich environment indicates the high success rate with low communication overhead of our method.
逐年被引趋势
关键指标
同类平均 = 1
同领域 · 同年份 · 同类型
Google Scholar 与 OpenAlex 的被引统计范围不同,数值存在差异属正常。
AI 辅助阅读
依据:摘要
可就本文提问;依据不足时会说明。
学术脉络
学科主题
计算机 / AIReinforcement Learning in Robotics
Robotic Path Planning Algorithms · Multimodal Machine Learning Applications
参考文献 67
此处列出前 3 条
引用本文 79
按被引量排序,此处列出前 3 条