机构
*
Beijing Academy of Artificial Intelligence, Beijing, China(北京人工智能研究院)
;
School of Computation, Information and Technology, Technical University of Munich, Garching, Germany(慕尼黑技术大学计算与信息学院)
;
Department of Shenyang Institute of Computing Technology, University of Chinese Academy of Sciences, Beijing, China(中国科学院沈阳计算技术研究所部门)
;
State Key Laboratory for Novel Software Technology, Nanjing University, Nanjing, China(新型软件技术国家重点实验室)
;
Department of Computer Science and Technology, Tsinghua University, Beijing, China(清华大学计算机科学与技术系)
Embodied Multimodal Grounding for Open-Vocabulary Mobile Manipulation via Semantic 3D Gaussian Splatting
基于语义三维高斯溅射的开放词汇移动操作具身多模态定位
Huosen Ou, Dongni Song, Yuncong Wang, Tao Zhou, Yiding Ji
机构
*
The Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州))
;
Midea Group(美的集团)
;
The Hong Kong University of Science and Technology(香港科技大学)
机构
*
The Hong Kong University of Science and Technology(香港科技大学)
;
National University of Singapore(新加坡国立大学)
;
Nanyang Technological University(南洋理工大学)
;
Wuhan University(武汉大学)
;
Sun Yat-sen University(中山大学)
;
Xidian University(西安电子科技大学)
;
Southeast University(东南大学)
机构
*
School of Computer Science and Engineering, Nanjing University of Science and Technology(南京理工大学计算机科学与工程学院)
;
Shenzhen Institutes of Advanced Technology, Chinese Academy of Sciences(中国科学院深圳先进技术研究院)
;
School of Artificial Intelligence, Beijing Normal University(北京师范大学人工智能学院)
;
Department of Automation, Tsinghua University(清华大学自动化系)
;
College of Artificial Intelligence, Nanjing Forestry University(南京林业大学人工智能学院)
ST-WAM: Semantic-Temporal World Action Model for Robust Manipulation under Visual Distribution Shifts
ST-WAM:面向视觉分布偏移下鲁棒操作的语义-时间世界动作模型
Mingxin Wang, Bin Hu, Bin Qian, Kaitao Jiang, Haoning Wu, Feng Yan, Bowen Jing, Ruiyang Hao, Enyi Wang, Kangning Niu, Yandan Yang, Mu Xu, Yan Wang, Houde Liu, Tianlun Li
CommentsAccepted at the ICRA 2026 Workshop on Reinforcement Learning for Imitation Learning (RL4IL), Vienna. 5 pages, 2 figures. v2: corrects the unified-critic run's curriculum level (10 of 40) and per-run environment counts, adds a Confounding Factors section, and softens the causal framing; measurements unchanged. https://mturan33.github.io/critic-architecture-matters/
Comments8 pages, 6 figures. Accepted manuscript. Published in the 2025 IEEE-RAS 24th International Conference on Humanoid Robots (Humanoids), pp. 1233-1240
Journal ref2025 IEEE-RAS 24th International Conference on Humanoid Robots (Humanoids), pp. 1233-1240 (2025)
GenVid2Robot: From Video Generation to Robot Manipulation via Rigid-Geometric Consistency
GenVid2Robot:通过刚性几何一致性从视频生成到机器人操作
Haohui Huang, Xi Yuan, Panpan Liao, Tao Teng, Chenguang Yang, Jing Guo, Yi Guo
机构
*
School of Automation, Guangdong University of Technology(广东工业大学自动化学院)
;
University of Liverpool(利物浦大学)
;
Department of Computing, The Hong Kong Polytechnic University(香港理工大学计算学系)
;
State Key Laboratory of Submarine Geoscience, School of Automation and Intelligent Sensing, Shanghai Jiao Tong University(上海交通大学海洋地球科学国家重点实验室,自动化与智能感知学院)
Language-Guided Grasping under Partial Observation for Mobile Manipulation in Field Inspection and Maintenance
用于现场检查和维护中移动操作的部分观察下语言引导抓取
Dilermando Almeida, Juliano Negri, Guilherme Lazzarini, Thiago H. Segreto, Ranulfo Bezerra, Gustavo J. G. Lahr, Ricardo V. Godoy, Marcelo Becker
机构
*
Department of Mechanical Engineering, Federal University of Uberlândia(联邦大学伯南迪利亚机械工程系)
;
Department of Mechanical Engineering, University of São Paulo(圣保罗大学机械工程系)
;
Graduate School of Information Sciences, Tohoku University(东北大学信息科学研究生院)
;
Faculdade Israelita de Ensino e Pesquisa Albert Einstein, Hospital Israelita Albert Einstein(艾伯特·爱因斯坦以色列教学与研究学院,艾伯特·爱因斯坦医院)
机构
*
College of Computer Science and Technology, National University of Defense Technology(国防科技大学计算机科学与技术学院)
;
School of Mathematical Sciences, Peking University(北京大学数学学院)
;
Institute for Theoretical Computer Science, Shanghai University of Finance and Economics(上海财经大学理论计算机科学研究所)
;
Information Technology Development, Aetos Capital Group, Sydney(悉尼Aetos资本集团信息技术部)
;
Faculty of Computing, Harbin Institute of Technology(哈尔滨工业大学计算机学院)
;
Department of Computer Science and Technology, Tsinghua University(清华大学计算机科学与技术系)
NoContactNoWorries: Estimating Contact through Vision and Proprioception for In-Hand Dexterous Manipulation
NoContactNoWorries: 通过视觉和本体感觉估计手内灵巧操作的接触
Soham Patil, Avirup Das, Sourabh Bhosale, Spandan Roy
机构
*
Robotics Research Center (RRC), International Institute of Information Technology (IIIT), Hyderabad, India(印度海得拉巴国际信息技术研究所机器人研究中心)
;
Department of Computer Science, The University of Manchester(曼彻斯特大学计算机科学系)