arXivDaily arXiv每日学术速递 周一至周五更新

视觉与机器人

VLA / 视觉-语言-动作模型

视觉-语言-动作模型、机器人基础模型和语言条件机器人控制。

2025-11-20 至 2025-11-20 共收录 5 信号源:cs.RO, cs.CV, cs.AI, cs.LG

1. VLA模型 3 篇

2511.14759 2025-11-20 cs.LG cs.RO 84%

$π^{*}_{0.6}$: a VLA That Learns From Experience

Physical Intelligence, Ali Amin, Raichelle Aniceto, Ashwin Balakrishna, Kevin Black, Ken Conley, Grace Connors, James Darpinian, Karan Dhabalia, Jared DiCarlo, Danny Driess, Michael Equi, Adnan Esmail, Yunhao Fang, Chelsea Finn, Catherine Glossop, Thomas Godden, Ivan Goryachev, Lachy Groom, Hunter Hancock, Karol Hausman, Gashon Hussein, Brian Ichter, Szymon Jakubczak, Rowan Jen, Tim Jones, Ben Katz, Liyiming Ke, Chandra Kuchi, Marinda Lamb, Devin LeBlanc, Sergey Levine, Adrian Li-Bell, Yao Lu, Vishnu Mano, Mohith Mothukuri, Suraj Nair, Karl Pertsch, Allen Z. Ren, Charvi Sharma, Lucy Xiaoyang Shi, Laura Smith, Jost Tobias Springenberg, Kyle Stachowicz, Will Stoeckle, Alex Swerdlow, James Tanner, Marcel Torne, Quan Vuong, Anna Walling, Haohuan Wang, Blake Williams, Sukwon Yoo, Lili Yu, Ury Zhilinsky, Zhiyuan Zhou

机构 * Physical Intelligence

专题命中 VLA模型 :VLA(title,abstract);vision-language-action(abstract);分类 cs.RO、cs.LG

详情

展开后加载摘要…

URL PDF HTML 收藏
2506.09553 2025-11-20 cs.CV 74%

GLD-Road:A global-local decoding road network extraction model for remote sensing images

Ligao Deng, Yupeng Deng, Yu Meng, Jingbo Chen, Zhihao Xi, Diyou Liu, Qifeng Chu

机构 * Aerospace Information Research Institute, Chinese Academy of Sciences(中国科学院航空信息研究所) School of Electronic, Electrical and Communication Engineering, University of Chinese Academy of Sciences(中国科学院大学电子电气与通信工程学院) Heilongjiang Geographic Information Engineering Institute(黑龙江地理信息工程研究所)

专题命中 VLA模型 :action model(title);分类 cs.CV

Journal ref ISPRS J. Photogramm. Remote Sens. 228, 741-755 (2025)

详情

展开后加载摘要…

URL PDF HTML 收藏
2511.15368 2025-11-20 cond-mat.soft cond-mat.stat-mech 50%

Shaping the aggregates of discotic particles with directional pair interactions

B. Martínez-Haya, N. Morillo, A. Cuetos

专题命中 VLA模型 :action model(abstract)

Comments Accepted in JCP

详情

展开后加载摘要…

URL PDF HTML 收藏

2. 机器人基础模型 1 篇

2511.00917 2025-11-20 cs.RO cs.AI 62%

Maestro: Orchestrating Robotics Modules with Vision-Language Models for Zero-Shot Generalist Robots

Junyao Shi, Rujia Yang, Kaitian Chao, Selina Bingqing Wan, Yifei Shao, Jiahui Lei, Jianing Qian, Long Le, Pratik Chaudhari, Kostas Daniilidis, Chuan Wen, Dinesh Jayaraman

专题命中 机器人基础模型 :VLA(abstract);分类 cs.RO、cs.AI

Comments Plan to resubmit after significant revisions

详情

展开后加载摘要…

URL PDF HTML 收藏

3. 数据集与评测 1 篇

2511.14161 2025-11-20 cs.RO cs.CV 73%

RoboTidy : A 3D Gaussian Splatting Household Tidying Benchmark for Embodied Navigation and Action

Xiaoquan Sun, Ruijian Zhang, Kang Pang, Bingchen Miao, Yuxiang Tan, Zhen Yang, Ming Li, Jiayu Chen

机构 * Huazhong University of Science and Technology(华中科技大学) The University of Hong Kong(香港大学) INFIFORCE Intelligent Technology Co., Ltd.(INFIFORCE智能技术有限公司) Zhejiang University(浙江大学) Guangming Lab, Shenzhen(深圳光明实验室)

专题命中 数据集与评测 :vision-language-action(abstract);VLA(abstract);分类 cs.RO、cs.CV

详情

展开后加载摘要…

URL PDF HTML 收藏