Vision-Language Models Are Not Pragmatically Competent in Referring Expression Generation
专题命中 视觉定位与Grounding :vision-language model(title,abstract);grounding(abstract);VLM(comments)
Comments COLM 2025 & CVinW @ CVPR 2025 (Spotlight). Homepage: https://vlm-reg.github.io/