CommentsIEEE CVF Conference on Computer Vision and Pattern Recognition 2026. Project page with code, models and examples: szymanowiczs.github.io/lagernvs
See, Plan, Rewind: Progress-Aware Vision-Language-Action Models for Robust Robotic Manipulation
看、规划、回退:面向鲁棒机器人操作的进度感知视觉-语言-动作模型
Tingjun Dai, Mingfei Han, Tingwen Du, Zhiheng Liu, Zihao Zhang, Zhihui Li, Salman Khan, Jun Yu, Xiaojun Chang
机构
*
School of Information Science and Technology, University of Science and Technology of China(信息科学与技术学院,中国科学技术大学)
;
University of Technology Sydney(新南威尔士大学)
;
Department of Computer Vision, Mohamed Bin Zayed University of Artificial Intelligence(人工智能与计算机视觉系,Mohamed Bin Zayed人工智能大学)
;
The University of Hong Kong(香港大学)
;
Institute of AI for Industry, Chinese Academy of Sciences(产业人工智能研究所,中国科学院)
;
School of Intelligent Science and Engineering, Harbin Institute of Technology (Shenzhen)(智能科学与工程学院,哈尔滨工业大学(深圳))
机构
*
Washington University in St. Louis(华盛顿大学圣路易斯分校)
;
The Chinese University of Hong Kong(香港中文大学)
;
ShanghaiTech University(上海科技大学)
;
University of California, Los Angeles(加州大学洛杉矶分校)