SuperCap: Multi-resolution Superpixel-based Image Captioning
专题命中 视觉定位与Grounding :vision language model(abstract);分类 cs.CV
Comments 12 pages, 4 figures
AI 大模型
视觉语言模型、视觉推理、视觉问答、图文理解和视觉 grounding。
专题命中 视觉定位与Grounding :vision language model(abstract);分类 cs.CV
Comments 12 pages, 4 figures
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV
Comments 16 pages, 10 figures
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV
Comments Code & models will be released at https://github.com/sming256/TimeLoc. The first 4 authors contributes equally
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV
Comments Project page: https://pz0826.github.io/GAGS-Webpage/
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV
Comments Accepted by CVPR 2025; Code & models: https://github.com/hustvl/MaskAdapter
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.AI
Comments Accepted to ICRA 2025
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV
Comments Preprint. Update: (1) better performance and (2) versatile segmentation. Code and models are available at: https://github.com/hustvl/EVF-SAM
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI
Comments Under review
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV
Comments Project Page: https://irvlutd.github.io/NIDSNet/
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI
Comments 9 pages, 10-12 refs
专题命中 视觉定位与Grounding :visual language model(abstract);分类 cs.CV
Comments 8 pages, 6 figures
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV
Comments ICRA 2025
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.LG
Comments Accepted by ICLR 2025; Code available at https://github.com/Graph-COM/SubgraphRAG
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI
Journal ref IEEE Journal of Selected Topics in Signal Processing, vol. 18, no. 8, pp. 1427-1440, Dec. 2024
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV
Comments 10 pages, 3 figures
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV
Comments 3DV 2025
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV
Comments Code is released at https://github.com/AFeng-x/PixWizard
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV
Comments Accepted by ICLR 2025, Project page: https://liuxuannan.github.io/MMFakeBench.github.io/
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI
Comments Preprint submitted to the 18th International Conference on Metadata and Semantics Research 2024 and published as a full, revised article
Journal ref Sfakakis, M., Garoufallou, E., Damigos, M., Salaba, A., Papatheodorou, C. (eds) Metadata and Semantic Research. MTSR 2024. Communications in Computer and Information Science, vol 2331. Springer, Cham
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI
Comments 18 pages, 10 figures. To appear in the Proceedings of the 2025 ACM CHI Conference on Human Factors in Computing Systems, Yokohama, Japan. https://hugoromat.github.io/ai_instruments/
专题命中 视觉定位与Grounding :multimodal large language model(abstract);分类 cs.AI
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI
Comments Project page: zaidkhan.me/MutaGReP
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.LG
专题命中 视觉定位与Grounding :vision language model(abstract);分类 cs.CV
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.AI
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV
Comments ICLR 2025
专题命中 视觉定位与Grounding :vision-language model(abstract);分类 cs.CV
Comments Submitted to: Pattern Recognition Letters, Klara Reichard and Giulia Rizzoli equally contributed to this work
专题命中 视觉定位与Grounding :grounding(abstract);分类 cs.CV