SeqVLM: Proposal-Guided Multi-View Sequences Reasoning via VLM for Zero-Shot 3D Visual Grounding
机构 * School of Informatics, Xiamen University(厦门大学信息学院) ; School of Computer Science, Nanjing University(南京大学计算机科学学院) ; School of Computer Science and Technology, East China Normal University(华东师范大学计算机科学与技术学院) ; Key Laboratory of Multimedia Trusted Perception and Efficient Computing, Ministry of Education of China, Xiamen University(教育部多媒体可信感知与高效计算重点实验室,厦门大学)
专题命中 视觉空间推理 :reasoning(title,abstract);分类 cs.AI