arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

VLA / 视觉-语言-动作模型

视觉-语言-动作模型、机器人基础模型和语言条件机器人控制。

2025-11-04 至 2025-11-04 共收录 14 信号源:cs.RO, cs.CV, cs.AI, cs.LG

1. VLA模型 12 篇

2511.00091 2025-11-04 cs.CV cs.RO 88%

Self-Improving Vision-Language-Action Models with Data Generation via Residual RL

Wenli Xiao, Haotian Lin, Andy Peng, Haoru Xue, Tairan He, Yuqi Xie, Fengyuan Hu, Jimmy Wu, Zhengyi Luo, Linxi "Jim" Fan, Guanya Shi, Yuke Zhu

机构 * NVIDIA(NVIDIA公司) CMU(卡内基梅隆大学) UC Berkeley(加州大学伯克利分校) UT Austin(德克萨斯大学奥斯汀分校)

专题命中 VLA模型 :vision-language-action(title,abstract);action model(title);VLA(abstract);分类 cs.RO、cs.CV

Comments 26 pages

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.01224 2025-11-04 cs.RO 88%

Embodiment Transfer Learning for Vision-Language-Action Models

Chengmeng Li, Yaxin Peng

机构 * Shanghai University(上海大学)

专题命中 VLA模型 :vision-language-action(title,abstract);action model(title);VLA(abstract);分类 cs.RO

详情

展开后加载摘要…

URL PDF HTML 收藏
2505.06111 2025-11-04 cs.RO cs.AI cs.LG 80%

UniVLA: Learning to Act Anywhere with Task-centric Latent Actions

Qingwen Bu, Yanting Yang, Jisong Cai, Shenyuan Gao, Guanghui Ren, Maoqing Yao, Ping Luo, Hongyang Li

专题命中 VLA模型 :vision-language-action(abstract);VLA(abstract);action model(abstract);分类 cs.RO、cs.AI、cs.LG

Comments Accepted to RSS 2025. Code is available at https://github.com/OpenDriveLab/UniVLA

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.01483 2025-11-04 astro-ph.HE astro-ph.SR 78%

On the nature and Galactic origin of the Be binary MWC 656. New insights from VLA, Gaia, and Fermi-LAT

Sergio A. Dzib, Frederic Jaron

专题命中 VLA模型 :VLA(title,abstract)

Comments Accepted to be published in Astronomy & Astrophysics. 5 pages, 1 figure, and 1 table

详情

展开后加载摘要…

URL PDF HTML 收藏
2507.05227 2025-11-04 cs.RO cs.CV cs.LG cs.MM cs.SY eess.SY 75%

NavigScene: Bridging Local Perception and Global Navigation for Beyond-Visual-Range Autonomous Driving

Qucheng Peng, Chen Bai, Guoxiang Zhang, Bo Xu, Xiaotong Liu, Xiaoyin Zheng, Chen Chen, Cheng Lu

机构 * Center for Research in Computer Vision, University of Central Florida(计算机视觉研究中心,中央佛罗里达大学)

专题命中 VLA模型 :vision-language-action(abstract);action model(abstract);分类 cs.RO、cs.CV、cs.LG

Comments Accepted by ACM Multimedia 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2301.03655 2025-11-04 stat.ML cs.LG 74%

Bayesian Additive Main Effects and Multiplicative Interaction Models using Tensor Regression for Multi-environmental Trials

Antonia A. L. Dos Santos, Danilo A. Sarti, Rafael A. Moral, Andrew C. Parnell

机构 * Hamilton Institute, Department of Mathematics(哈里顿研究所,数学系) Statistics, Maynooth University, Ireland(统计学,梅诺特大学,爱尔兰) School of Mathematics(数学学院) Statistics, Insight Centre for Data Analytics, University College Dublin, Ireland(统计学,洞察数据分析师中心,都柏林大学学院,爱尔兰)

专题命中 VLA模型 :action model(title);分类 cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2510.23763 2025-11-04 cs.RO cs.CL cs.CV 73%

RoboOmni: Proactive Robot Manipulation in Omni-modal Context

Siyin Wang, Jinlan Fu, Feihong Liu, Xinzhe He, Huangxuan Wu, Junhao Shi, Kexin Huang, Zhaoye Fei, Jingjing Gong, Zuxuan Wu, Yu-Gang Jiang, See-Kiong Ng, Tat-Seng Chua, Xipeng Qiu

机构 * Fudan University(复旦大学) Shanghai Innovation Institute(上海创新研究院) National University of Singapore(新加坡国立大学)

专题命中 VLA模型 :vision-language-action(abstract);VLA(abstract);分类 cs.RO、cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.01652 2025-11-04 eess.AS cs.SD 50%

Leveraging Language Information for Target Language Extraction

Mehmet Sinan Yıldırım, Ruijie Tao, Wupeng Wang, Junyi Ao, Haizhou Li

机构 * Department of Electrical and Computer Engineering, National University of Singapore(新加坡国立大学电气与计算机工程系) School of Artificial Intelligence, Shenzhen Research Institute of Big Data, The Chinese University of Hong Kong, Shenzhen, China(香港中文大学深圳研究院人工智能学院)

专题命中 VLA模型 :action model(abstract)

Comments Accepted to APSIPA ASC 2025

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.01264 2025-11-04 physics.flu-dyn 50%

Convective flux analysis on the propagation mechanism of oblique detonation waves

Yunfeng Liu

专题命中 VLA模型 :action model(abstract)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.00531 2025-11-04 cs.LO 50%

Runtime Verification of Interactions Using Automata

Chana Weil-Kennedy, Darine Rammal, Christophe Gaston, Arnault Lapitre

专题命中 VLA模型 :action model(abstract)

Comments Preprint

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.00241 2025-11-04 astro-ph.GA 50%

The Universality of Dark Matter Density Profiles for Milky Way Analog Galaxies

Maria Clara Cavalcante-Siviero, K. Menéndez-Delmestre, P. P. B. Beaklini, T. S. Gonçalves, D. C. Rodrigues, N. G. de Isídio, A. E. Araújo-Carvalho

专题命中 VLA模型 :VLA(abstract)

Comments Submitted to Astrophysical Journal

详情

展开后加载摘要…

URL PDF HTML 收藏
2401.11186 2025-11-04 hep-ph nucl-th 50%

$π$- and $K$-Mesons Properties for Large $N_f$

Aftab Ahmad

专题命中 VLA模型 :action model(abstract)

Comments 25 pages, 7figures

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 动作表示与策略 1 篇

2510.16240 2025-11-04 cs.RO 70%

Cosmos-Surg-dVRK: World Foundation Model-based Automated Online Evaluation of Surgical Robot Policy Learning

Lukas Zbinden, Nigel Nelson, Juo-Tung Chen, Xinhao Chen, Ji Woong Kim, Mahdi Azizian, Axel Krieger, Sean Huver

机构 * NVIDIA Johns Hopkins University(约翰霍普金斯大学) Stanford University(斯坦福大学)

专题命中 动作表示与策略 :vision-language-action(abstract);action model(abstract);分类 cs.RO

Comments minor metadata and notation fixes; +3 citations

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 部署与泛化 1 篇

2505.15685 2025-11-04 cs.RO 70%

From Grounding to Manipulation: Case Studies of Foundation Model Integration in Embodied Robotic Systems

Xiuchao Sui, Daiying Tian, Qi Sun, Ruirui Chen, Dongkyu Choi, Kenneth Kwok, Soujanya Poria

机构 * IHPC, Agency for Science, Technology and Research, Singapore(科技研究局智能技术中心,新加坡) Nanyang Technological University, Singapore(南洋理工大学,新加坡)

专题命中 部署与泛化 :vision-language-action(abstract);VLA(abstract);分类 cs.RO

Comments EMNLP 2025 camera ready

详情

展开后加载摘要…

URL PDF HTML 收藏