研究论文
SVA: Towards speech-Enabled vision-Language-Action model
Lingxiao Li, Jiacheng Fan, Xiaohui Ni, Sujuan Qin, Wenmin Li, Fei Gao
Beijing University of Posts and Telecommunications State Key Laboratory of Cryptology
来源Pattern Recognition
年份2025
阅读操作
确认中在文库中上传 PDF 后可生成中文音频讲解。
摘要与影响
摘要 · 节选
暂未获取摘要。可打开原文或 PDF,后续可基于全文生成更完整的速读。
逐年被引趋势
110
126
关键指标
1
被引次数 · OpenAlex
0.55
领域内被引倍数
同类平均 = 1
同类平均 = 1
前 28%
引用位次
同领域 · 同年份 · 同类型
同领域 · 同年份 · 同类型
12
参考文献
Google Scholar 与 OpenAlex 的被引统计范围不同,数值存在差异属正常。
AI 辅助阅读
依据:文献信息
论文问答
当前基于文献信息回答
可就本文提问;依据不足时会说明。
学术脉络
学科主题
计算机 / AIMultimodal Machine Learning Applications
Robot Manipulation and Learning · Social Robot Interaction and HRI
参考文献 12
CochleaNet: A robust language-independent audio-visual model for real-time speech enhancement
被引 93Mandar Gogate, Kia Dashtipour, Ahsan Adeel · Information Fusion · 2020
Visual SLAM for robot navigation in healthcare facility
被引 123Baofu Fang, Gaofei Mei, Xiaohui Yuan · Pattern Recognition · 2021
Hierarchical Reinforcement Learning With Universal Policies for Multistep Robotic Manipulation
被引 93Xintong Yang, Ze Ji, Jing Wu · IEEE Transactions on Neural Networks and Learning Systems · 2021
此处列出前 3 条
引用本文 1
Vision Language Action Models for Embodied Intelligence A Structured Taxonomy Critical Analysis and Future Research Directions
被引 0Ola Farid, Hamdi A. Mahmoud · Computational Discovery and Intelligent Systems (CDIS) · 2026
按被引量排序,此处列出前 3 条