Open-Vocabulary Semantic Segmentation with Uncertainty Alignment for Robotic Scene Understanding in Indoor Building Environments
专题命中 视觉定位与Grounding :vision language model(abstract);分类 cs.CV
Comments 32 pages, 7 figures
AI 大模型
视觉语言模型、视觉推理、视觉问答、图文理解和视觉 grounding。
专题命中 视觉定位与Grounding :vision language model(abstract);分类 cs.CV
Comments 32 pages, 7 figures
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV
Comments Accepted to CVPR 2025
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV
Comments 5 pages, 4 figures
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.LG
Comments The 29th Pacific-Asia Conference on Knowledge Discovery and Data Mining (PAKDD2025)
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV
Comments CVPR 2025. Project website at https://comca-attributes.github.io/
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV
Comments CVPR 2025. Project page: https://relationfield.github.io
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV
Comments CVPR 2025
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV
Comments De-Factify 4.0 workshop at the 39th Annual AAAI Conference on Artificial Intelligence (AAAI 2025)
专题命中 视觉定位与Grounding :multimodal large language model(abstract);分类 cs.AI
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.LG
专题命中 视觉定位与Grounding :multimodal large language model(abstract);分类 cs.CV
Comments Accepted by IJCV
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI
Comments 32 pages, 10 figures, 5 tables
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.LG
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV
Comments MICCAI 2024
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI
Comments 7 pages, 4 figures
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI
Journal ref (2005) in Athanassios Raftopoulos (ed.), Cognitive penetrability of perception: Attention, action, attention and bottom-up constraints (Huntington: Nova Science), 157-70
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV
Comments CVPR 2025
专题命中 视觉定位与Grounding :multimodal large language model(abstract);分类 cs.CV
Comments This is the journal version. For conference version (T2I-CompBench): arXiv:2307.06350v2. Project page: https://karine-h.github.io/T2I-CompBench-new/
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV
Comments Accepted to CVPR'25
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV
Comments Accepted to ICLR 2025. Project page available at https://ponimatkin.github.io/wildpose/
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV
Comments Accepted to TMM 2025
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV
Comments Accepted at WACV 2025 (Oral)
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV
Comments Accepted to CVPR 2024
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI
Comments 10 pages, 1 figure, HICSS 57: Hawaii International Conference on System Sciences, Honolulu, HI, published January 2024
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV
Comments Revised version according to comments from reviewers of ICLR2025
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV
Comments Accepted at Neurips 2024