VideoAds for Fast-Paced Video Understanding
机构 * Northwestern University(西北大学) ; Boston University(波士顿大学)
专题命中 视频多模态 :multi-modal(abstract);MLLM(abstract);分类 cs.CV
Comments ICCV2025
AI 大模型
跨文本、图像、视频、音频等模态的大模型与学习方法。
机构 * Northwestern University(西北大学) ; Boston University(波士顿大学)
专题命中 视频多模态 :multi-modal(abstract);MLLM(abstract);分类 cs.CV
Comments ICCV2025
专题命中 视频多模态 :MLLM(abstract);分类 cs.CV
Comments 16 pages, 9 figures
机构 * ByteDance(字节跳动) ; NUS(国立大学新加坡) ; HKUST(GZ)(香港科技大学(珠海)) ; THU(清华大学)
专题命中 视频多模态 :multimodal(abstract);分类 cs.AI
机构 * The University of Texas at San Antonio(德克萨斯大学圣安东尼奥分校)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
机构 * Nagoya University(名古屋大学)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
Comments ACM Multimedia Asia 2025
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
机构 * Department of Mechanical Engineering, University of Delaware(机械工程系,德雷克塞尔大学) ; Departments of Animal & Food Sciences, Biological Sciences, and Medical & Molecular Sciences, University of Delaware(动物与食品科学系、生物科学系和医学与分子科学系,德雷克塞尔大学)
专题命中 视频多模态 :multi-modal(abstract)
机构 * University of Toronto(多伦多大学) ; McMaster University(麦马斯特大学) ; McGill University(麦吉尔大学) ; University of Manitoba(曼尼托巴大学) ; University of California, Los Angeles(加州大学洛杉矶分校) ; University of Montreal(蒙特利尔大学) ; Mila ; CUHK(香港中文大学) ; HKUST(GZ)(香港理工大学(广州)) ; Nanyang Technological University(南洋理工大学) ; Stanford University(斯坦福大学) ; CG Matrix Technology Limited(CG矩阵科技有限公司)
专题命中 跨模态检索 :multi-modal(title,abstract);分类 cs.AI
机构 * Department of Computer Science Bar-Ilan University(巴伊兰大学计算机科学系) ; Electrical and Computer Engineering Technion(技术学院电子与计算机工程系)
专题命中 跨模态检索 :multimodal(abstract);cross-modal(abstract);分类 cs.CV
Comments Accepted to NeurIPS 2025
机构 * Emory University(埃默里大学) ; University of Electronic Science and Technology of China(电子科技大学) ; University of Illinois Chicago(伊利诺伊大学香槟分校) ; The Hong Kong Polytechnic University(香港理工大学)
专题命中 跨模态检索 :multimodal(abstract);分类 cs.CV、cs.AI
Comments 12 pages, 5 figures
机构 * Center for Research in Computer Vision, University of Central Florida(计算机视觉研究中心,中央佛罗里达大学) ; SRI International(SRI国际)
专题命中 跨模态检索 :multimodal(abstract);分类 cs.CV
Comments Published at CVPR 2023
机构 * Inflection AI ; Georgia Institute of Technology(佐治亚理工学院) ; ChatAlpha AI
专题命中 跨模态检索 :multi-modal(abstract);分类 cs.CL
机构 * Beijing University of Posts and Telecommunications(北京邮电大学)
专题命中 跨模态检索 :multimodal(abstract);分类 cs.CV
Comments This paper was originally submitted to ACM MM 2025 on April 12, 2025
专题命中 跨模态检索 :multi-modal(abstract,comments)
Comments Published in the Proceedings of the ICML 2025 Workshop on Multi-modal Foun- dation Models and Large Language Models for Life Sciences, Vancouver, Canada. 2025
专题命中 多模态生成 :cross-modal(title,abstract);multimodal(abstract);分类 cs.CV
专题命中 多模态生成 :multimodal(title,abstract);cross-modal(abstract)
Comments Accepted by IEEE 25th BIBE
机构 * Nanjing University of Information Science \& Technology Nanjing China ; East China Normal University Shanghai China ; The Hong Kong University of Science ; Brown University Providence America ; Southwest Jiaotong University Chengdu China ; Nanjing University of Information Science \& Technology ; East China Normal University ; Brown University ; Southwest Jiaotong University
专题命中 多模态生成 :multimodal(title,abstract);分类 cs.CL、cs.AI
Comments 10 pages, 5 figures, accepted to appear in the Proceedings of the 33rd ACM International Conference on Multimedia (MM '25)
专题命中 多模态生成 :multi-modal(title,abstract)
Comments 31 pages, 9 figures
机构 * Minzu University of China(民族大学) ; City University of Macau(澳门城市大学) ; Key Laboratory of Computing Power Network and Information Security, Ministry of Education, Shandong Computer Science Center (National Supercomputer Center in Jinan), Qilu University of Technology (Shandong Academy of Sciences)(计算能力网络与信息安全重点实验室,教育部,山东计算机科学中心(济南国家超级计算机中心),齐鲁工业大学(山东省科学院)) ; The University of Adelaide(阿德莱德大学)
专题命中 多模态生成 :multimodal(abstract);cross-modal(abstract);分类 cs.AI
Comments 16 pages, 13 figures
机构 * UMass Amherst(马萨诸塞大学阿姆赫斯特分校) ; Sony AI(索尼人工智能) ; UC San Diego(加州大学圣地亚哥分校)
专题命中 多模态生成 :multimodal(abstract);multi-modal(abstract);分类 cs.CV
Comments Project page: https://talkcuts.github.io/
机构 * Wuhan University(武汉大学) ; DAMO Academy, Alibaba Group(达摩院,阿里巴巴集团) ; Hupan Lab(虎扑实验室) ; The Chinese University of Hong Kong(香港中文大学) ; Tsinghua University(清华大学) ; Huazhong University of Science and Technology(华中科技大学) ; Zhejiang University(浙江大学)
专题命中 多模态生成 :multi-modal(abstract);MLLM(abstract)
Comments 13 pages, 6 figures
专题命中 多模态生成 :multimodal(abstract);multi-modal(abstract)
Comments 8 pages, 2 images
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV、cs.AI
Comments NeurIPS 2025; Also Oral at ICML 2025 FM4LS workshop
机构 * University of Tübingen, Tübingen AI Center(图宾根大学,图宾根人工智能中心) ; University of Tübingen, Tübingen AI Center, MPI for Informatics, SIC(图宾根大学,图宾根人工智能中心,马克斯·普朗克信息研究所,SIC) ; University of Tübingen(图宾根大学)
专题命中 多模态生成 :multi-modal(abstract);分类 cs.CV
Comments Accepted to ACM SIGGRAPH Asia 2025. Project website: https://yuxuan-xue.com/infini-human
机构 * Department of Electronic Engineering, Beijing National Research Center for Information Science and Technology (BNRist), Tsinghua University(电子工程系,信息科学与技术国家研究中心(BNRist),清华大学) ; International School, Beijing University of Posts and Telecommunications(国际学院,北京邮电大学)
专题命中 多模态生成 :cross-modal(abstract);分类 cs.AI
Comments 9 pages, 4 figures. Code: https://github.com/tsinghua-fib-lab/MSTDiff
机构 * Tencent GYLab(腾讯GY实验室) ; Fudan University(复旦大学)
专题命中 多模态生成 :multimodal(abstract);分类 cs.CV
机构 * German Aerospace Center (DLR)(德国航空航天中心) ; University of Lübeck(吕贝克大学) ; German Research Center for Artificial Intelligence (DFKI)(德国人工智能研究中心)
专题命中 多模态生成 :分类 cs.CV、cs.AI;multimodal(comments);multimodal foundation model(comments)
Comments Accepted for presentation at ICCV Workshops 2025, "The 4th Workshop on What is Next in Multimodal Foundation Models?" (MMFM)
专题命中 多模态生成 :multi-modal(abstract)
机构 * University of Science and Technology of China(中国科学技术大学) ; National University of Singapore(新加坡国立大学) ; DP Technology(DP技术)
专题命中 多模态生成 :multi-modal(abstract)
Comments NeurIPS 2025
机构 * Frontier AI Research Centre, Macquarie University(前沿人工智能研究中心,麦考瑞大学)
专题命中 多模态生成 :multimodal(abstract)
Comments 35 pages, 3 figures