MAG-Nav: Language-Driven Object Navigation Leveraging Memory-Reserved Active Grounding
机构 * Tencent Robotics X(腾讯机器人X) ; Harbin Institute of Technology(哈尔滨工业大学)
专题命中 视觉定位与Grounding :grounding(title,abstract);visual language model(abstract)
AI 大模型
视觉语言模型、视觉推理、视觉问答、图文理解和视觉 grounding。
机构 * Tencent Robotics X(腾讯机器人X) ; Harbin Institute of Technology(哈尔滨工业大学)
专题命中 视觉定位与Grounding :grounding(title,abstract);visual language model(abstract)
机构 * TNO(荷兰技术院) ; Intelligent Imaging(智能成像)
专题命中 视觉定位与Grounding :vision language model(abstract);VLM(abstract);分类 cs.CV
机构 * RMIT University, Australia(皇家墨尔本理工大学)
专题命中 视觉定位与Grounding :vision language model(abstract);分类 cs.CV
Comments 19 Pages, 24 Figures
机构 * LITIV, Polytechnique Montréal(Polytechnique Montréal 的 LITIV) ; Data Science Laboratory, Université du Québec (TELUQ)(Université du Québec (TELUQ) 的 Data Science Laboratory)
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV
Comments 10 pages
机构 * School of Computation, Information, and Technology, Technical University of Munich(计算、信息与技术学院,慕尼黑技术大学) ; Graduate Institute of International and Development Studies, Geneva(国际与发展研究研究生院,日内瓦) ; JPMorgan AI Research(摩根大通人工智能研究)
专题命中 视觉定位与Grounding :grounding(abstract)
Comments Accepted to ACL 2025-Main Conference
专题命中 视觉定位与Grounding :grounding(abstract)
机构 * School of Computing, National University of Singapore(新加坡国立大学计算机学院) ; CSAIL, Massachusetts Institute of Technology(麻省理工学院计算机科学与人工智能实验室)
专题命中 视觉定位与Grounding :grounding(abstract)
Comments This is the journal version accepted to the International Journal of Robotics Research (IJRR). It extends our prior work presented at Robotics: Science and Systems (RSS) 2024, with a new compositional program induction pipeline from natural language, and expanded evaluations on personalized bookshelf and bedroom furniture layout tasks
专题命中 视觉定位与Grounding :grounding(abstract)
Comments Accespted at the European Conference on Artificial Intelligence 2025