arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

视觉与机器人

VLA / 视觉-语言-动作模型

视觉-语言-动作模型、机器人基础模型和语言条件机器人控制。

2025-09-10 至 2025-09-10 共收录 4 信号源:cs.RO, cs.CV, cs.AI, cs.LG
2509.07962 2025-09-10 cs.RO 91%

TA-VLA: Elucidating the Design Space of Torque-aware Vision-Language-Action Models

Zongzheng Zhang, Haobo Xu, Zhuo Yang, Chenghao Yue, Zehao Lin, Huan-ang Gao, Ziwei Wang, Hao Zhao

机构 * Beijing Academy of Artificial Intelligence, BAAI(北京人工智能研究院) Institute for AI Industry Research (AIR), Tsinghua Univeristy(清华大学人工智能产业研究院) Nanyang Technological University(南洋理工大学)

专题命中 VLA模型 :VLA(title,abstract);vision-language-action(title,abstract);action model(title);分类 cs.RO

Comments Accepted to CoRL 2025, project page: \url{https://zzongzheng0918.github.io/Torque-Aware-VLA.github.io/}

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.06951 2025-09-10 cs.RO cs.CV 89%

F1: A Vision-Language-Action Model Bridging Understanding and Generation to Actions

Qi Lv, Weijie Kong, Hao Li, Jia Zeng, Zherui Qiu, Delin Qu, Haoming Song, Qizhi Chen, Xiang Deng, Jiangmiao Pang

机构 * Shanghai AI Laboratory(上海人工智能实验室) Harbin Institute of Technology (Shenzhen)(哈尔滨工业大学(深圳))

专题命中 VLA模型 :vision-language-action(title,abstract);action model(title);VLA(abstract,comments);分类 cs.RO、cs.CV

Comments Homepage: https://aopolin-lv.github.io/F1-VLA/

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.07957 2025-09-10 cs.RO 83%

Graph-Fused Vision-Language-Action for Policy Reasoning in Multi-Arm Robotic Manipulation

Shunlei Li, Longsen Gao, Jiuwen Cao, Yingbai Hu

机构 * Machine Learning and I-health International Cooperation Base of Zhejiang Province, Artificial Intelligence Institute, Hangzhou Dianzi University, Zhejiang(浙江省机器学习与健康国际合作基地、人工智能学院、杭州电子大学) Electrical and Computer Engineering Department, The University of New Mexico(新墨西哥大学电气与计算机工程系) School of Computation, Information and Technology, Technical University of Munich(慕尼黑技术大学计算、信息与技术学院)

专题命中 VLA模型 :vision-language-action(title,abstract);VLA(abstract);分类 cs.RO

Comments This paper is submitted to IEEE IROS 2025 Workshop AIR4S

详情

展开后加载摘要…

URL PDF HTML 收藏
2509.07594 2025-09-10 cs.IR 50%

ELEC: Efficient Large Language Model-Empowered Click-Through Rate Prediction

Rui Dong, Wentao Ouyang, Xiangzheng Liu

专题命中 VLA模型 :action model(abstract)

Comments SIGIR 2025

详情

展开后加载摘要…

URL PDF HTML 收藏