ChatTracker: Enhancing Visual Tracking Performance via Chatting with Multimodal Large Language Model
专题命中 视觉定位与Grounding :multimodal large language model(title,abstract);MLLM(abstract);分类 cs.CV
AI 大模型
视觉语言模型、视觉推理、视觉问答、图文理解和视觉 grounding。
专题命中 视觉定位与Grounding :multimodal large language model(title,abstract);MLLM(abstract);分类 cs.CV
专题命中 视觉定位与Grounding :vision-language model(title,abstract);VLM(abstract);分类 cs.CV
Comments 13 pages, 10 figures
专题命中 视觉定位与Grounding :multimodal large language model(title,abstract);MLLM(abstract);分类 cs.CV
Comments Accepted as Proceedings Paper at ML4H 2024
专题命中 视觉定位与Grounding :VLM(title,abstract);vision-language model(abstract);分类 cs.CV
Comments WACV 2025 Accepted
专题命中 视觉定位与Grounding :vision language model(title,abstract);LLaVA(abstract);分类 cs.CV
Comments 52 pages, 13 figures
专题命中 视觉定位与Grounding :grounding(title,abstract);visual language model(abstract);分类 cs.CV
Comments Accepted by NeurIPS 2024. The project page: https://github.com/linhuixiao/OneRef
专题命中 视觉定位与Grounding :grounding(title,abstract);multimodal large language model(abstract);分类 cs.CV
专题命中 视觉定位与Grounding :vision-language model(title,abstract);grounding(abstract);分类 cs.CV
Comments 8th Conference on Robot Learning (CoRL 2024), Munich, Germany
专题命中 视觉定位与Grounding :vision language model(title,abstract);VLM(abstract);分类 cs.CV
Comments IEEE Robotics and Automation Letters
专题命中 视觉定位与Grounding :grounding(title,abstract);vision language model(abstract);分类 cs.AI
Comments Published in IROS 2024
专题命中 视觉定位与Grounding :grounding(title,abstract);vision-language model(abstract);分类 cs.CV
专题命中 视觉定位与Grounding :vision-language model(title,abstract);VLM(abstract);分类 cs.CV
专题命中 视觉定位与Grounding :visual language model(title,abstract);grounding(abstract);分类 cs.CV
专题命中 视觉定位与Grounding :grounding(title,abstract);visual language model(abstract);分类 cs.CV
Comments Accepted by ECCV 2024
专题命中 视觉定位与Grounding :vision language model(title,abstract);VLM(abstract);分类 cs.CV
专题命中 视觉定位与Grounding :grounding(title,abstract);VLM(abstract);分类 cs.CV
专题命中 视觉定位与Grounding :vision-language model(title,abstract);VLM(abstract);分类 cs.CV
Comments 29 pages, Accepted for publication in ECCV 2024
专题命中 视觉定位与Grounding :multimodal large language model(title,abstract);MLLM(abstract);分类 cs.CV
专题命中 视觉定位与Grounding :vision-language model(title,abstract);VLM(abstract);分类 cs.CV
Comments CVPR 2024 Accepted
专题命中 视觉定位与Grounding :grounding(title,abstract);visual question answering(abstract);分类 cs.CV
专题命中 视觉定位与Grounding :grounding(title,abstract);multimodal large language model(abstract);分类 cs.AI
专题命中 视觉定位与Grounding :vision-language model(title,abstract);VLM(abstract);分类 cs.CV
Comments Accepted to NAACL 2024 Main Conference
专题命中 视觉定位与Grounding :grounding(title,abstract);multimodal large language model(abstract);分类 cs.AI
专题命中 视觉定位与Grounding :grounding(title,abstract);visual question answering(abstract);分类 cs.CV
专题命中 视觉定位与Grounding :grounding(title,abstract);vision-language model(abstract);分类 cs.CV
Comments 17 pages, 6 tables, 9 figurs. Code, data, and model are available at: https://github.com/sunsmarterjie/ChatterBox
专题命中 视觉定位与Grounding :grounding(title,abstract);visual question answering(abstract);分类 cs.CV
Journal ref EMNLP 2023 Findings
专题命中 视觉定位与Grounding :vision-language model(title,abstract);LLaVA(abstract);分类 cs.CV
Comments Accepted by AAAI 2024; Project page: https://anomalygpt.github.io
专题命中 视觉定位与Grounding :grounding(title,abstract);multimodal large language model(abstract);分类 cs.CV
专题命中 视觉定位与Grounding :visual language model(title,abstract);LLaVA(abstract);分类 cs.AI
专题命中 视觉定位与Grounding :vision-language model(title,abstract);grounding(abstract);分类 cs.CV
Comments Code, demo and models are available at https://github.com/QwenLM/Qwen-VL