Towards Agentic AI for Multimodal-Guided Video Object Segmentation
机构 * Applied Artificial Intelligence Institute, Deakin University, Australia(应用人工智能研究所,德金大学,澳大利亚)
专题命中 视频多模态 :multimodal(title,abstract);multi-modal(abstract);分类 cs.CV
AI 大模型
跨文本、图像、视频、音频等模态的大模型与学习方法。
机构 * Applied Artificial Intelligence Institute, Deakin University, Australia(应用人工智能研究所,德金大学,澳大利亚)
专题命中 视频多模态 :multimodal(title,abstract);multi-modal(abstract);分类 cs.CV
机构 * Amazon(亚马逊) ; School of Psychology and Neuroscience, University of Glasgow(心理学与神经科学学院,格拉斯哥大学) ; Ben-Gurion University of the Negev(内盖夫本·古里安大学) ; University of Cambridge(剑桥大学) ; ETH Zurich(苏黎世联邦理工学院)
专题命中 视频多模态 :multimodal(title,abstract);分类 cs.AI
Comments Accepted at 2025 IEEE/RSJ International Conference on Intelligent Robots and Systems (IROS)
机构 * Medical AI Research Center ( MedARC )(医学人工智能研究中心(MedARC)) ; Baylor College of Medicine(贝勒医学院)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
Comments Perspective piece on Algonauts 2025 Challenge conclusion
机构 * Shanghai Artificial Intelligence Laboratory(上海人工智能实验室) ; The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州)) ; The Hong Kong University of Science and Technology(香港科学与技术大学) ; Northwestern Polytechnical University(西北工业大学)
专题命中 视频多模态 :cross-modal(abstract);分类 cs.CV
机构 * University of Central Florida(中央佛罗里达大学)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
Comments Accepted to VISION'25 - ICCV 2025 workshop
机构 * University of Central Florida(中央佛罗里达大学)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV