STARN-GAT: A Multi-Modal Spatio-Temporal Graph Attention Network for Accident Severity Prediction
专题命中 视频多模态 :multi-modal(title,abstract);分类 cs.AI
Comments 10 pages
AI 大模型
跨文本、图像、视频、音频等模态的大模型与学习方法。
专题命中 视频多模态 :multi-modal(title,abstract);分类 cs.AI
Comments 10 pages
机构 * Beijing University of Technology(北京理工大学) ; Shanghai Jiao Tong University(上海交通大学) ; Chinese Academy of Sciences(中国科学院) ; The Hong Kong Polytechnic University(香港理工大学)
专题命中 视频多模态 :multimodal(title,abstract);分类 cs.CV
Comments Accepted by ICCV 2025 (Poster)
专题命中 视频多模态 :multimodal(abstract);cross-modal(abstract);multimodal foundation model(abstract);分类 cs.CV、cs.CL
Comments 10 pages, 10 figures
机构 * Peking University(北京大学) ; Guangdong University of Technology(广东工业大学) ; The University of Sheffield(谢菲尔德大学) ; University of Science and Technology Beijing(北京科技大学) ; Tsinghua University(清华大学) ; The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州)) ; Nanjing University(南京大学) ; University of Trento(特伦特大学) ; Queen's University Belfast(贝尔法斯特女王大学)
专题命中 视频多模态 :multimodal(abstract);MLLM(abstract);分类 cs.CV
Comments Paper was accepted by ACM MM 2025; Code: https://github.com/YihuaJerry/EventVAD
机构 * Harbin Institute of Technology, Shenzhen(哈尔滨工业大学(深圳)) ; Tsinghua Shenzhen International Graduate School, Tsinghua University(清华大学深圳国际研究生院) ; Research Center of Artificial Intelligence, Peng Cheng Laboratory(鹏城实验室人工智能研究中心) ; The Hong Kong University of Science and Technology(香港科学与技术大学)
专题命中 视频多模态 :cross-modal(abstract);分类 cs.CV、cs.MM
Comments Accepted by ICCV'25. 13 pages, 6 figures, 4 tables
机构 * School of Cyberspace Security,Gansu University of Political Science and Law(网络安全学院、政治学科学校)
专题命中 视频多模态 :cross-modal(abstract);分类 cs.CV
Comments 13 pages,7 figures
机构 * ARC Lab, Tencent PCG(腾讯PCG ARC实验室) ; Search Application Department, Tencent CSIG(腾讯CSIG搜索应用部门) ; Tencent Hunyuan(腾讯文生视频) ; Big Data Platform Department, Tencent PCG(腾讯PCG大数据平台部门)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
Comments Project Page: https://tencentarc.github.io/posts/arc-video-announcement/
机构 * Nagoya University(名古屋大学)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
Comments 10 pages, 3 figures
机构 * School of Computer Science, Peking University(北京大学计算机科学系) ; Intelligent Game and Decision Lab (IGDL)(智能游戏与决策实验室) ; College of Computer, National University of Defense Technology(国防科技大学计算机学院) ; School of Computer Science, Shanghai Jiao Tong University(上海交通大学计算机科学系)
专题命中 视频多模态 :cross-modal(abstract);分类 cs.CV
Comments Accepted by IJCAI 2025 (International Joint Conference on Artificial Intelligence)
专题命中 视频多模态 :multi-modal(abstract);分类 cs.MM
Comments Accepted by ACM Multimedia 2025
机构 * Graduate School of Information Science and Technology, The University of Osaka(信息科学与技术研究生学校,大阪大学) ; D3 Center, The University of Osaka(大阪大学D3中心) ; Graduate School of Maritime Sciences, Kobe University(海洋科学研究生学校, Kobe大学)
专题命中 视频多模态 :multimodal(abstract);分类 cs.AI
机构 * University of Amsterdam(阿姆斯特丹大学) ; SAI, Shanghai Jiao Tong University(上海交通大学SAI研究所) ; Xiaohongshu Inc(小红书公司)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
机构 * GenGenAI ; Yonsei University(延世大学)
专题命中 视频多模态 :multi-modal(abstract);分类 cs.CV
Comments ICCVW 2025
专题命中 视频多模态 :multimodal(abstract)
Comments Preliminary version; a revised version will be uploaded later
专题命中 视频多模态 :multi-modal(abstract)
Comments 19 pages, 11 figures, submitted to Advanced Intelligent Systems
专题命中 视频多模态 :multi-modal(abstract)