Automating eHMI Action Design with LLMs for Automated Vehicle Communication
Xia Ding, Xinyue Gui, Fan Gao, Dongyuan Li, Mark Colley, Takeo Igarashi
The University of Tokyo University College London
阅读操作
确认中在文库中上传 PDF 后可生成中文音频讲解。
摘要与影响
The absence of explicit communication channels between automated vehicles (AVs) and other road users requires the use of external Human-Machine Interfaces (eHMIs) to convey messages effectively in uncertain scenarios.Currently, most eHMI studies employ predefined text messages and manually designed actions to convey these messages, which limits the real-world deployment of eHMIs, where adaptability in dynamic scenarios is essential.Given the generalizability and versatility of large language models (LLMs), they could potentially serve as automated action designers for the message-action design task.To validate this idea, we make three contributions: (1) We propose a pipeline that integrates LLMs and 3D renderers, using LLMs as action designers to generate executable actions for controlling eHMIs and rendering action clips.(2) We collect a user-rated Action-Design Scoring dataset comprising a total of 320 action sequences for eight intended messages and four representative eHMI modalities.The dataset validates that LLMs can translate intended messages into actions close to a human level, particularly for reasoning-enabled LLMs.(3) We introduce two automated raters, Action Reference Score (ARS) and Vision-Language Models (VLMs), to benchmark 18 LLMs, finding that the VLM aligns with human preferences yet varies across eHMI modalities. 1
逐年被引趋势
暂无年度引用数据
关键指标
同类平均 = 1
同领域 · 同年份 · 同类型
Google Scholar 与 OpenAlex 的被引统计范围不同,数值存在差异属正常。
AI 辅助阅读
依据:摘要
可就本文提问;依据不足时会说明。
学术脉络
学科主题
工程Safety Systems Engineering in Autonomy
Human-Automation Interaction and Safety · Autonomous Vehicle Technology and Safety