VLA^2: Empowering Vision-Language-Action Models with an Agentic Framework for Unseen Concept Manipulation
专题命中 VLA模型 :vision-language-action(title,abstract);VLA(title,abstract);action model(title);分类 cs.RO
视觉与机器人
视觉-语言-动作模型、机器人基础模型和语言条件机器人控制。
专题命中 VLA模型 :vision-language-action(title,abstract);VLA(title,abstract);action model(title);分类 cs.RO
机构 * MoE key Lab of Artificial Intelligence, AI Institute, Shanghai Jiao Tong University(人工智能混合专家实验室,人工智能研究所,上海交通大学) ; School of Automation and Intelligent Sensing, Shanghai Jiao Tong University(自动化与智能感知学院,上海交通大学) ; School of Computer Science, Shanghai Jiao Tong University(计算机科学学院,上海交通大学) ; Shanghai AI Laboratory(上海人工智能实验室) ; Tsinghua Shenzhen International Graduate School, Tsinghua University(清华大学深圳国际研究生院,清华大学) ; The University of Hong Kong(香港大学) ; Tongji University(同济大学) ; D-Robotics ; Key Laboratory of System Control and Information Processing, Ministry of Education of China(系统控制与信息处理重点实验室,中华人民共和国教育部) ; Shanghai Key Laboratory of Integrated Administration Technologies for Information Security(上海信息安全管理集成技术重点实验室)
专题命中 VLA模型 :vision-language-action(title,abstract);VLA(abstract);分类 cs.RO、cs.AI
专题命中 VLA模型 :action model(abstract);分类 cs.RO、cs.AI
专题命中 VLA模型 :action model(abstract)
专题命中 VLA模型 :action model(abstract)
Comments 6 pages, 1 figure
机构 * University of California, Riverside(加州大学河滨分校) ; University of Michigan(密歇根大学) ; Meta AI
专题命中 数据集与评测 :vision-language-action(abstract);分类 cs.RO、cs.CV、cs.AI
Comments 39th Conference on Neural Information Processing Systems (NeurIPS 2025); Project Website: rdd-neurips.github.io
机构 * Department of Electrical and Electronic Engineering, the University of Hong Kong(香港大学电子与电气工程系) ; State Key Laboratory of General Artificial Intelligence, BIGAI(通用人工智能国家重点实验室,BIGAI) ; School of Computer Science and Technology, Beijing Institute of Technology(北京理工大学计算机科学与技术学院) ; Yuanpei College, Peking University(北京大学元培学院)
专题命中 部署与泛化 :action model(abstract);分类 cs.RO、cs.CV、cs.AI