Move to Understand a 3D Scene: Bridging Visual Grounding and Exploration for Efficient and Versatile Embodied Navigation
机构 * Tsinghua University(清华大学) ; Beijing Institute of Technology(北京理工大学) ; Beihang University(北航) ; State Key Laboratory of General Artificial Intelligence, BIGAI, China(国家一般人工智能重点实验室, BIGAI, 中国)
专题命中 三维重建 :3D vision(abstract,comments);3D reconstruction(abstract);point cloud(abstract);分类 cs.CV
Comments Embodied AI; 3D Vision Language Understanding; ICCV 2025 Highlight; https://mtu3d.github.io; Spatial intelligence