DEGround: An Effective Baseline for Ego-centric 3D Visual Grounding with a Homogeneous Framework
DEGround:一种有效的基于自身视角的3D视觉定位基线框架
机构 * School of Computer Science, Wuhan University(武汉大学计算机学院) ; MMLab, The Chinese University of Hong Kong(香港中文大学MMLab) ; Tsinghua University(清华大学)
专题命中 视觉定位与Grounding :grounding(title,abstract);分类 cs.CV
AI总结 本文提出DEGround框架,通过统一的框架实现检测与定位的物体级共享,引入两个任务特定模块提升细粒度指令定位性能,实验表明在多个基准上表现最佳,尤其在EmbodiedScan数据集上精度提升显著。
Comments 1st place on EmbodiedScan visual grounding