Improving Medical Large Vision-Language Models with Abnormal-Aware Feedback
机构 * SKL-IOTSC, CIS, University of Macau(SKL-IOTSC、CIS、澳门大学)
专题命中 视觉定位与Grounding :vision-language model(title,abstract);分类 cs.CV、cs.AI、cs.LG
Comments 16 pages
AI 大模型
视觉语言模型、视觉推理、视觉问答、图文理解和视觉 grounding。
机构 * SKL-IOTSC, CIS, University of Macau(SKL-IOTSC、CIS、澳门大学)
专题命中 视觉定位与Grounding :vision-language model(title,abstract);分类 cs.CV、cs.AI、cs.LG
Comments 16 pages
机构 * Queen Mary University of London(伦敦玛丽女王大学) ; Institute of Information Engineering, Chinese Academy of Sciences(中国科学院信息工程研究所) ; Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所) ; The Hong Kong University of Science and Technology (GZ)(香港科学与技术大学)
专题命中 视觉定位与Grounding :MLLM(title,abstract);multimodal large language model(abstract)
机构 * School of Communication Engineering, Jilin University(吉林大学通信工程学院) ; College of Software Engineering, Jilin University(吉林大学软件工程学院) ; School of Mathematics, Jilin University(吉林大学数学学院) ; College of Electronic Science and Engineering, Jilin University(吉林大学电子科学工程学院)
专题命中 视觉定位与Grounding :VLM(title,abstract);vision-language model(abstract)
Comments 8 pages, 5 figures
机构 * Nanyang Technological University(南洋理工大学) ; Nanjing University of Science and Technology(南京理工大学)
专题命中 视觉定位与Grounding :grounding(title,abstract);分类 cs.CV、cs.AI、cs.LG
专题命中 视觉定位与Grounding :vision language model(title,abstract);分类 cs.CV、cs.AI、cs.LG
专题命中 视觉定位与Grounding :vision-language model(title,abstract);VLM(abstract)
专题命中 视觉定位与Grounding :vision-language model(title,abstract);分类 cs.CV、cs.AI、cs.LG
Comments Accepted to CVPR 2025. Project page: https://plan-lab.github.io/calico/
专题命中 视觉定位与Grounding :vision-language model(title,abstract);分类 cs.CV、cs.AI、cs.LG
Comments Accepted in CVPR 2025
专题命中 视觉定位与Grounding :grounding(title,abstract);分类 cs.CV、cs.AI、cs.LG
专题命中 视觉定位与Grounding :grounding(title,abstract);分类 cs.CV、cs.AI、cs.LG
专题命中 视觉定位与Grounding :visual language model(title,abstract);VLM(abstract)
Journal ref Computer Graphics forum Volume 44 (2025), Number 3
专题命中 视觉定位与Grounding :grounding(title,abstract);分类 cs.CV、cs.AI、cs.LG
Comments CVPR 2025. Project website: https://3d-grand.github.io
专题命中 视觉定位与Grounding :grounding(title,abstract);分类 cs.CV、cs.AI、cs.LG
专题命中 视觉定位与Grounding :vision-language model(title,abstract);分类 cs.CV、cs.AI、cs.LG
专题命中 视觉定位与Grounding :grounding(title,abstract);分类 cs.CV、cs.AI、cs.LG
Comments To be published in the Proceedings of AAAI 2025. The first three authors contributed equally. Project: https://github.com/DirtyHarryLYL/HAKE-AVA
专题命中 视觉定位与Grounding :grounding(title,abstract);visual question answering(abstract)
Comments Accepted to NAACL 2025 Main
专题命中 视觉定位与Grounding :grounding(title,abstract);multimodal large language model(abstract)
Comments Accepted for ICRA 2025. Project page: https://sites.google.com/umn.edu/etog-etrg/home
专题命中 视觉定位与Grounding :vision language model(title,abstract);VLM(abstract)
Comments Accepted to COLING2025
专题命中 视觉定位与Grounding :vision-language model(title,abstract);分类 cs.CV、cs.AI、cs.LG
专题命中 视觉定位与Grounding :grounding(title,abstract);分类 cs.CV、cs.AI、cs.LG
Comments NeurIPS 2024 Datasets and Benchmarks Track
专题命中 视觉定位与Grounding :vision-language model(title,abstract);分类 cs.CV、cs.AI、cs.LG
Comments NeurIPS 2024
专题命中 视觉定位与Grounding :vision-language model(title,abstract);VLM(abstract)
Comments 17 pages, 9 figures
专题命中 视觉定位与Grounding :VLM(title,abstract);vision-language model(abstract)
专题命中 视觉定位与Grounding :vision-language model(title,abstract);分类 cs.CV、cs.AI、cs.LG
Comments ECCV2024; Codes and Supp. are available at: https://github.com/YBZh/LAPT
专题命中 视觉定位与Grounding :grounding(title,abstract);分类 cs.CV、cs.AI、cs.LG
Comments Accepted at ECCV 2024
专题命中 视觉定位与Grounding :grounding(title,abstract);分类 cs.CV、cs.AI、cs.LG
专题命中 视觉定位与Grounding :vision-language model(title,abstract);分类 cs.CV、cs.AI、cs.LG
Comments Accept to ACL 2024; 19 Pages, 15 Figures, 6 Tables
专题命中 视觉定位与Grounding :vision-language model(title,abstract);分类 cs.CV、cs.AI、cs.LG
Comments Medical Imaging with Deep Learning (MIDL) 2024 (Oral)
专题命中 视觉定位与Grounding :VLM(title,abstract);visual language model(abstract)
专题命中 视觉定位与Grounding :grounding(title,abstract);multimodal large language model(abstract)