arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~

视觉与机器人

VLA / 视觉-语言-动作模型

视觉-语言-动作模型、机器人基础模型和语言条件机器人控制。

2025-10-17 至 2025-10-17 共收录 7 信号源:cs.RO, cs.CV, cs.AI, cs.LG

1. VLA模型 5 篇

2510.14902 2025-10-17 cs.RO 91%

VLA^2: Empowering Vision-Language-Action Models with an Agentic Framework for Unseen Concept Manipulation

Han Zhao, Jiaxuan Zhang, Wenxuan Song, Pengxiang Ding, Donglin Wang

专题命中 VLA模型 :vision-language-action(title,abstract);VLA(title,abstract);action model(title);分类 cs.RO

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14300 2025-10-17 cs.RO cs.AI 84%

Expertise need not monopolize: Action-Specialized Mixture of Experts for Vision-Language-Action Learning

Weijie Shen, Yitian Liu, Yuhao Wu, Zhixuan Liang, Sijia Gu, Dehui Wang, Tian Nian, Lei Xu, Yusen Qin, Jiangmiao Pang, Xinping Guan, Xiaokang Yang, Yao Mu

机构 * MoE key Lab of Artificial Intelligence, AI Institute, Shanghai Jiao Tong University(人工智能混合专家实验室,人工智能研究所,上海交通大学) School of Automation and Intelligent Sensing, Shanghai Jiao Tong University(自动化与智能感知学院,上海交通大学) School of Computer Science, Shanghai Jiao Tong University(计算机科学学院,上海交通大学) Shanghai AI Laboratory(上海人工智能实验室) Tsinghua Shenzhen International Graduate School, Tsinghua University(清华大学深圳国际研究生院,清华大学) The University of Hong Kong(香港大学) Tongji University(同济大学) D-Robotics Key Laboratory of System Control and Information Processing, Ministry of Education of China(系统控制与信息处理重点实验室,中华人民共和国教育部) Shanghai Key Laboratory of Integrated Administration Technologies for Information Security(上海信息安全管理集成技术重点实验室)

专题命中 VLA模型 :vision-language-action(title,abstract);VLA(abstract);分类 cs.RO、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.13921 2025-10-17 cs.RO cs.AI 62%

APEX: Empowering LLMs with Physics-Based Task Planning for Real-time Insight

Wanjing Huang, Weixiang Yan, Zhen Zhang, Ambuj Singh

专题命中 VLA模型 :action model(abstract);分类 cs.RO、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.14880 2025-10-17 cs.IR 50%

Fantastic (small) Retrievers and How to Train Them: mxbai-edge-colbert-v0 Tech Report

Rikiya Takehi, Benjamin Clavié, Sean Lee, Aamir Shakir

专题命中 VLA模型 :action model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.04241 2025-10-17 cs.HC 50%

RunPacer: A Smartwatch-Based Vibrotactile Feedback System for Symmetric Co-Running by Visually Impaired Individuals and Guides

Yichen Yu, Huan-Song Xu, Ming-Yen Lin

专题命中 VLA模型 :action model(abstract)

Comments 6 pages, 1 figure

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 数据集与评测 1 篇

2510.14968 2025-10-17 cs.RO cs.AI cs.CV cs.LG cs.SY eess.SY 70%

RDD: Retrieval-Based Demonstration Decomposer for Planner Alignment in Long-Horizon Tasks

Mingxuan Yan, Yuping Wang, Zechun Liu, Jiachen Li

机构 * University of California, Riverside(加州大学河滨分校) University of Michigan(密歇根大学) Meta AI

专题命中 数据集与评测 :vision-language-action(abstract);分类 cs.RO、cs.CV、cs.AI

Comments 39th Conference on Neural Information Processing Systems (NeurIPS 2025); Project Website: rdd-neurips.github.io

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 部署与泛化 1 篇

2510.14293 2025-10-17 cs.RO cs.AI cs.CV cs.LG 70%

Learning Human-Humanoid Coordination for Collaborative Object Carrying

Yushi Du, Yixuan Li, Baoxiong Jia, Yutang Lin, Pei Zhou, Wei Liang, Yanchao Yang, Siyuan Huang

机构 * Department of Electrical and Electronic Engineering, the University of Hong Kong(香港大学电子与电气工程系) State Key Laboratory of General Artificial Intelligence, BIGAI(通用人工智能国家重点实验室,BIGAI) School of Computer Science and Technology, Beijing Institute of Technology(北京理工大学计算机科学与技术学院) Yuanpei College, Peking University(北京大学元培学院)

专题命中 部署与泛化 :action model(abstract);分类 cs.RO、cs.CV、cs.AI

详情

展开后加载摘要…

URL PDF HTML 收藏