机构
*
South China University of Technology(华南理工大学)
;
Institute for Infocomm Research (I 2 R), A*STAR(信息通信研究所(I 2 R),A*STAR)
;
WeChat Vision, Tencent Inc.(微信视觉,腾讯公司)
;
Foshan University(佛山大学)
;
Nanyang Technological University(南洋理工大学)
;
National University of Singapore(新加坡国立大学)
专题命中
视觉定位与Grounding
:grounding(abstract);multimodal large language model(abstract);MLLM(abstract);分类 cs.CV
PlaceIt3D: Language-Guided Object Placement in Real 3D Scenes
Ahmed Abdelreheem, Filippo Aleotti, Jamie Watson, Zawar Qureshi, Abdelrahman Eldesokey, Peter Wonka, Gabriel Brostow, Sara Vicente, Guillermo Garcia-Hernando