DeCafNet: Delegate and Conquer for Efficient Temporal Grounding in Long Videos
机构 * Microsoft(微软公司) ; Northeastern University(东北大学)
专题命中 视觉定位与Grounding :grounding(title,abstract);分类 cs.CV、cs.AI
Comments Accepted by CVPR 2025
AI 大模型
视觉语言模型、视觉推理、视觉问答、图文理解和视觉 grounding。
机构 * Microsoft(微软公司) ; Northeastern University(东北大学)
专题命中 视觉定位与Grounding :grounding(title,abstract);分类 cs.CV、cs.AI
Comments Accepted by CVPR 2025
机构 * Gaoling School of Artificial Intelligence, Renmin University of China(中国人民大学 Gallagher人工智能学院) ; Huawei Noah’s Ark Lab(华为诺亚实验室)
专题命中 视觉定位与Grounding :grounding(title,abstract);分类 cs.CV、cs.AI
专题命中 视觉定位与Grounding :grounding(title,abstract);分类 cs.CV、cs.AI
机构 * Department of Electrical and Computer Engineering(电气与计算机工程系) ; Princeton University(普林斯顿大学)
专题命中 视觉定位与Grounding :grounding(title,abstract);分类 cs.CV、cs.AI
机构 * Harbin Institute of Technology (Shenzhen)(哈尔滨工业大学(深圳)) ; Pengcheng Laboratory(鹏城实验室) ; Shandong Jianzhu University(山东建筑大学)
专题命中 视觉定位与Grounding :grounding(title,abstract);分类 cs.CV、cs.AI
Comments Accepted by CVPR 2025
机构 * Hong Kong University of Science and Technology(香港科技大学)
专题命中 视觉定位与Grounding :grounding(title,abstract);分类 cs.CV、cs.AI
Comments 10 pages
专题命中 视觉定位与Grounding :multimodal large language model(title,abstract);分类 cs.CV、cs.AI
Comments 12 pages, 11 figures, 13IHMMSec2025
专题命中 视觉定位与Grounding :grounding(title,abstract);分类 cs.CV、cs.AI
Comments Accepted to the IEEE/CVF Conference on Computer Vision and Pattern Recognition (CVPR) 2025
专题命中 视觉定位与Grounding :grounding(title,abstract);分类 cs.CV、cs.AI
Comments CVPR 2025. Project page: https://yawen-shao.github.io/GREAT/ Code: https://github.com/yawen-shao/GREAT_code
专题命中 视觉定位与Grounding :vision language model(title,abstract);分类 cs.CV、cs.LG
专题命中 视觉定位与Grounding :grounding(title,abstract);分类 cs.CV、cs.LG
Comments Accepted at CVPR2025 with a top score
专题命中 视觉定位与Grounding :grounding(title,abstract);分类 cs.CV、cs.AI
专题命中 视觉定位与Grounding :grounding(title,abstract);分类 cs.CV、cs.AI
专题命中 视觉定位与Grounding :grounding(title,abstract);分类 cs.CV、cs.LG
Comments Accepted in 2025 IEEE International Symposium on Biomedical Imaging (ISBI 2025)
专题命中 视觉定位与Grounding :vision-language model(title,abstract);分类 cs.CV、cs.AI
专题命中 视觉定位与Grounding :multimodal large language model(title,abstract);分类 cs.CV、cs.AI
Comments Accepted by ICLR 2025. The code and data are available at https://github.com/jam-cc/MMAD
专题命中 视觉定位与Grounding :grounding(title,abstract);分类 cs.AI、cs.LG
Comments Accepted at ACL 2024 (main conference)
专题命中 视觉定位与Grounding :grounding(title,abstract);分类 cs.CV、cs.AI
专题命中 视觉定位与Grounding :vision-language model(title,abstract);分类 cs.CV、cs.AI
Comments Accepted at ICLR 2025. The first two authors contributed equally
专题命中 视觉定位与Grounding :grounding(title,abstract);分类 cs.CV、cs.LG
Comments 10 pages, 3 figures, IEEE J-BHI Special Issue on Foundation Models in Medical Imaging
专题命中 视觉定位与Grounding :grounding(title,abstract);分类 cs.CV、cs.AI
Comments AAAI2025
专题命中 视觉定位与Grounding :grounding(title,abstract);分类 cs.CV、cs.AI
Comments 12 pages, 4 figures, 8 tables
专题命中 视觉定位与Grounding :vision-language model(title,abstract);分类 cs.CV、cs.AI
Comments 4 pages,4 figures,www source track
专题命中 视觉定位与Grounding :multimodal large language model(title,abstract);分类 cs.CV、cs.AI
Comments Accepted by CVPPA @ECCV2024. Dataset: https://github.com/Yutong-Zhou-cv/AgriBench
专题命中 视觉定位与Grounding :grounding(title,abstract);分类 cs.CV、cs.AI
Comments Accepted to WACV 2025
专题命中 视觉定位与Grounding :grounding(title,abstract);分类 cs.CV、cs.AI
专题命中 视觉定位与Grounding :grounding(title,abstract);分类 cs.CV、cs.LG
Comments Project Page: https://catherine-r-he.github.io/SynGround/
专题命中 视觉定位与Grounding :grounding(title,abstract);分类 cs.AI、cs.LG
Comments ICLR 2024 Spotlight
Journal ref In International Conference on Learning Representations (ICLR) (2024)
专题命中 视觉定位与Grounding :vision-language model(title,abstract);分类 cs.CV、cs.AI
专题命中 视觉定位与Grounding :VLM(title);vision-language model(abstract);分类 cs.CV、cs.AI