LaViRA: Language-Vision-Robot Actions Translation for Zero-Shot Vision Language Navigation in Continuous Environments
LaViRA: 语言-视觉-机器人动作翻译用于连续环境中的零样本视觉语言导航
Hongyu Ding, Ziming Xu, Yudong Fang, You Wu, Zixuan Chen, Jieqi Shi, Jing Huo, Yifan Zhang, Yang Gao
机构
*
School of Computer Science, Nanjing University(南京大学计算机科学学院)
;
School of Intelligence Science and Technology, Nanjing University(南京大学智能科学与技术学院)
;
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
University of Chinese Academy of Sciences(中国科学院大学)
MLLM-4D: Towards Visual-based Spatial-Temporal Intelligence
MLLM-4D: 向基于视觉的空间-时间智能迈进
Xingyilang Yin, Chengzhengxu Li, Jiahao Chang, Chi-Man Pun, Xiaodong Cun
机构
*
University of Macau(澳门大学)
;
Xi'an Jiaotong University(西安交通大学)
;
The Chinese University of Hong Kong, Shenzhen(香港中文大学(深圳))
;
GVC Lab, Great Bay University(大湾大学GVC实验室)
机构
*
Australian Institute for Machine Learning, Adelaide University(澳大利亚机器学习研究所,阿德莱德大学)
;
The University of Manchester(曼彻斯特大学)
;
Zhejiang University(浙江大学)
;
Agency for Science, Technology and Research (A*STAR)(科技研究局(A*STAR))
机构
*
Southeast University(东南大学)
;
Purple Mountain Labs(紫金山实验室)
;
The Chinese University of Hong Kong(香港中文大学)
;
The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州))
;
Fudan University(复旦大学)
;
Shanghai Jiao Tong University(上海交通大学)
机构
*
Department of XXX, University of YYY, Location, Country(XXX系,YYY大学,地点,国家)
;
School of ZZZ, Institute of WWW, Location, Country(ZZZ学院,WWW研究所,地点,国家)
;
Xpeng Motors Technology Co Ltd(小鹏汽车科技有限公司)
DV-VLN: Dual Verification for Reliable LLM-Based Vision-and-Language Navigation
DV-VLN:基于可靠大语言模型的视觉-语言导航的双重验证
Zijun Li, Shijie Li, Zhenxi Zhang, Bin Li, Shoujun Zhou
机构
*
Robotics Engineering Program, College of Engineering, Zhejiang Normal University(浙江师范大学工程学院机器人工程专业)
;
Shenzhen Institutes of Advanced Technology (SIAT), Chinese Academy of Sciences(中国科学院深圳先进技术研究院)
;
Department of Health Technology and Informatics, The Hong Kong Polytechnic University(香港理工大学健康科技与信息技术系)
CoINS: Counterfactual Interactive Navigation via Skill-Aware VLM
基于技能感知的反事实交互导航:通过技能感知视觉语言模型
Kangjie Zhou, Zhejia Wen, Zhiyong Zhuo, Zike Yan, Pengying Wu, Ieng Hou U, Shuaiyang Li, Han Gao, Kang Ding, Wenhan Cao, Wei Pan, Chang Liu
机构
*
School of Advanced Manufacturing and Robotics, Peking University(北京大学先进制造与机器人学院)
;
Department of Mechanical and Automation Engineering, The Chinese University of Hong Kong(香港中文大学机械与自动化工程系)
;
College of Design and Engineering, National University Of Singapore(新加坡国立大学设计与工程学院)
;
Department of Computer Science, The University of Manchester(曼彻斯特大学计算机科学系)
SpatialReasoner: Active Perception for Large-Scale 3D Scene Understanding
SpatialReasoner: 大规模3D场景理解中的主动感知
Hongpei Zheng, Shijie Li, Yanran Li, Hujun Yin
机构
*
University of Manchester(曼彻斯特大学)
;
Institute for Infocomm Research (I2R), A*STAR, Singapore(信息与通信研究 institute(I2R),A*STAR,新加坡)
;
University of Bedfordshire(贝德福德郡大学)
机构
*
Australian Institute for Machine Learning, the University of Adelaide(澳大利亚机器学习研究所、阿德莱德大学)
;
CREATE Lab, Swiss Federal Institute of Technology Lausanne (EPFL)(洛桑联邦理工学院CREATE实验室)
PathFormer: A Transformer with 3D Grid Constraints for Digital Twin Robot-Arm Trajectory Generation
Ahmed Alanazi, Duy Ho, Yugyung Lee
机构
*
Department of Computer Science, University of Missouri–Kansas City (UMKC)(密苏里大学哥伦比亚分校计算机科学系)
;
Department of Computer Science, California State University, Fullerton(加州州立大学富尔顿分校计算机科学系)
When LLMs step into the 3D World: A Survey and Meta-Analysis of 3D Tasks via Multi-modal Large Language Models
Xianzheng Ma, Brandon Smart, Yash Bhalgat, Shuai Chen, Xinghui Li, Jian Ding, Jindong Gu, Dave Zhenyu Chen, Songyou Peng, Jia-Wang Bian, Philip H Torr, Marc Pollefeys, Matthias Nießner, Ian D Reid, Angel X. Chang, Iro Laina, Victor Adrian Prisacariu
机构
*
University of Oxford(牛津大学)
;
King Abdullah University of Science and Technology(国王 Abdullah 科学与技术大学)
;
Technical University of Munich(慕尼黑技术大学)
;
Mohamed bin Zayed University of Artificial Intelligence(穆罕默德·本·扎耶德人工智能大学)
;
Simon Fraser University(西蒙·弗雷泽大学)
;
ETH Zurich(苏黎世联邦理工学院)
CommentsThe paper is currently under investigation regarding concerns of potential academic misconduct. While the investigation is ongoing, the authors have voluntarily requested to withdraw the manuscript