LatentFlow: Cross-Frequency Experimental Flow Reconstruction from Sparse Pressure via Latent Mapping
专题命中 视频多模态 :cross-modal(abstract);分类 cs.AI
Comments The paper is submitted to IAAI26. Total 9 pages with 8 figures
AI 大模型
跨文本、图像、视频、音频等模态的大模型与学习方法。
专题命中 视频多模态 :cross-modal(abstract);分类 cs.AI
Comments The paper is submitted to IAAI26. Total 9 pages with 8 figures
机构 * AMAP, Alibaba Group(阿里集团AMAP)
专题命中 视频多模态 :multimodal(abstract);分类 cs.AI
机构 * Xidian University(西电大学)
专题命中 视频多模态 :multi-modal(abstract);分类 cs.CV
机构 * Universidad de Buenos Aires, Facultad de Ciencias Exactas y Naturales(布宜诺斯艾利斯大学,精确科学与自然学院) ; Institute of Engineering Sciences, Universidad de O’Higgins(工程科学研究所,奥希金斯大学) ; CONICET-UBA, Instituto de Ciencias de la Computacion (ICC)(CONICET-UBA,计算科学研究所) ; L3S Research Center, Leibniz Universität Hannover(L3S研究中心,汉诺威莱布尼茨大学)
专题命中 视频多模态 :multi-modal(abstract);分类 cs.CV
Journal ref 2025 IEEE/CVF International Conference on Computer Vision Workshops (ICCVW); 2nd Workshop on Neuromorphic Vision (NeVi)
机构 * Louisiana State University(路易斯安那州立大学) ; Northeastern University(东北大学) ; Yale University(耶鲁大学)
专题命中 视频多模态 :multimodal(abstract);分类 cs.AI
Comments This paper will appear in CCS 2025
机构 * MoE Key Lab of Artificial Intelligence, AI Institute Shanghai Jiao Tong University Shanghai China(人工智能联合实验室,人工智能研究院,上海交通大学)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
Comments ACM MM 2025; Code is released at https://github.com/VISION-SJTU/AnchorSync
机构 * University of California, Berkeley(加州大学伯克利分校) ; MIT Media Lab(MIT媒体实验室) ; Princeton University(普林斯顿大学)
专题命中 视频多模态 :multimodal(abstract);分类 cs.AI
机构 * Visual Geometry Group, University of Oxford(视觉几何组,牛津大学) ; Naver Labs Europe(Naver欧洲实验室)
专题命中 视频多模态 :multi-modal(abstract);分类 cs.CV
Comments 17 pages, 6 figures, ICCV 2025 Highlight, Project page: https://geo4d.github.io/
机构 * Department of Automation, University of Science and Technology of China(自动化系,中国科学技术大学) ; Department of Computer Science, City University of Hong Kong(计算机科学系,香港城市大学) ; Department of Mechanical Engineering, City University of Hong Kong(机械工程系,香港城市大学) ; College of Computer Science, Beijing University of Technology(计算机科学学院,北京理工大学)
专题命中 视频多模态 :cross-modal(abstract);分类 cs.CV
Comments 11 pages, 8 figures, under review
专题命中 视频多模态 :multi-modal(abstract);分类 cs.CV
机构 * Lawrence Berkeley National Laboratory(伯克利国家实验室) ; University of California, Irvine(加州大学尔湾分校) ; University of California, Berkeley(加州大学伯克利分校) ; Covalent Metrology(协力计量)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
Comments This paper has been accepted for presentation at the 59th International Conference on Parallel Processing (ICPP 2025), DRAI workshop
专题命中 视频多模态 :MLLM(abstract);分类 cs.CV
专题命中 视频多模态 :multi-modal(abstract);分类 cs.MM
Comments in submission
机构 * Medical AI Research Center ( MedARC )(医学人工智能研究中心(MedARC)) ; Baylor College of Medicine(贝勒医学院)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
Comments Perspective piece on Algonauts 2025 Challenge conclusion
机构 * Shanghai Artificial Intelligence Laboratory(上海人工智能实验室) ; The Hong Kong University of Science and Technology (Guangzhou)(香港科学与技术大学(广州)) ; The Hong Kong University of Science and Technology(香港科学与技术大学) ; Northwestern Polytechnical University(西北工业大学)
专题命中 视频多模态 :cross-modal(abstract);分类 cs.CV
机构 * University of Central Florida(中央佛罗里达大学)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
Comments Accepted to VISION'25 - ICCV 2025 workshop
机构 * University of Central Florida(中央佛罗里达大学)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
机构 * VCIP, School of Computer Science, Nankai University(VCIP,计算机科学学院,南开大学) ; Shanghai AI Laboratory(上海人工智能实验室) ; The Chinese University of Hong Kong(香港中文大学) ; Peking University(北京大学)
专题命中 视频多模态 :multi-modal(abstract);分类 cs.CV
专题命中 视频多模态 :multi-modal(abstract);分类 cs.CV
机构 * Oregon State University(俄勒冈州立大学)
专题命中 视频多模态 :multi-modal(abstract);分类 cs.CV
Comments Accepted to ICCV 2025 [Poster]
机构 * School of Data Science, Fudan University(复旦大学数据科学学院)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
专题命中 视频多模态 :cross-modal(abstract);分类 cs.CV
机构 * Beihang University(北航大学)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
Comments 9 pages, 3 figures, 2 tables. Accepted at CV4A11y, ICCV 2025
机构 * Wangxuan Institute of Computer Technology, Peking University(北京大学计算机技术研究院) ; State Key Laboratory of General Artificial Intelligence, Peking University(北京大学通用人工智能国家重点实验室)
专题命中 视频多模态 :multi-modal(abstract);分类 cs.CV
Comments Accepted by ICCV2025
机构 * State Key Laboratory for Novel Software Technology, Nanjing University(新型软件技术国家重点实验室,南京大学) ; MyBank, Ant Group(蚂蚁集团MyBank) ; Shanghai AI Lab(上海AI实验室)
专题命中 视频多模态 :image-text(abstract);分类 cs.CV
Comments Accepted by ICCV2025
机构 * McGill University(麦吉尔大学) ; Mila – Quebec AI Institute(魁北克人工智能研究所)
专题命中 视频多模态 :multi-modal(abstract);分类 cs.CV
Comments Accepted to MICCAI 2025 (LMID Workshop)
机构 * School of Computing(计算学院) ; Information Systems, Singapore Management University, Singapore(信息系统,新加坡管理大学,新加坡)
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
Comments Accepted at 4DMR@IJCAI25: International IJCAI Workshop on 1st Challenge and Workshop for 4D Micro-Expression Recognition for Mind Reading, August 29, 2025, Guangzhou, China
机构 * Southwestern University of Finance and Economics(西南财经大学) ; Kyoto Institute of Technology(京都技术大学)
专题命中 视频多模态 :multi-modal(abstract);分类 cs.AI
Comments Accepted in SIGKDD 2025
机构 * National Engineering Research Center for Multimedia Software, School of Computer Science, Wuhan University(国家多媒体软件工程研究中心,计算机学院,武汉大学) ; Hong Kong University of Science and Technology(香港科学与技术大学) ; Horizon Robotics
专题命中 视频多模态 :multimodal(abstract);分类 cs.CV
Comments ICCV 2025
专题命中 视频多模态 :multimodal(abstract);分类 cs.AI
Comments Accepted to IROS 2025. Source code available