机构
*
Faculty of Computer Science and Technology, Qilu University of Technology (Shandong Academy of Sciences)(计算机科学与技术学院,齐鲁工业大学(山东科学院))
;
School of Computing, National University of Singapore(computing 学院,新加坡国立大学)
;
School of Computing and Information Technology, Great Bay University(computing 与信息学院,大湾大学)
;
National Engineering Laboratory for Big Data System Computing Technology, Shenzhen University(大数据系统计算技术国家工程实验室,深圳大学)
专题命中
多模态生成
:multi-modal(title,abstract);分类 cs.CV
CommentsAccepted by IEEE Transactions on Information Forensics and Security 2025
机构
*
Department of Civil and Environmental Engineering, University of California, Berkeley(加州大学伯克利分校土木与环境工程系)
;
The Singapore-MIT Alliance for Research and Technology(新加坡-麻省理工联盟研究技术中心)
;
Department of Urban and Regional Planning, University of Florida(佛罗里达大学城市与区域规划系)
;
Department of Civil and Environmental Engineering, Massachusetts Institute of Technology(麻省理工学院土木与环境工程系)
;
Department of Urban Planning, Tsinghua University(清华大学城市规划系)
;
Department of Urban Studies and Planning, Massachusetts Institute of Technology(麻省理工学院城市研究与规划系)
Foundation Molecular Grammar: Multi-Modal Foundation Models Induce Interpretable Molecular Graph Languages
Michael Sun, Weize Yuan, Gang Liu, Wojciech Matusik, Jie Chen
机构
*
MIT CSAIL(麻省理工学院计算机科学与人工智能实验室)
;
MIT Chemistry(麻省理工学院化学系)
;
MIT-IBM Watson AI Lab, IBM Research(麻省理工-IBM Watson人工智能实验室,IBM研究院)
;
University of Notre Dame(诺埃伯大学)
Cross-Modal Causal Intervention for Medical Report Generation
Weixing Chen, Yang Liu, Ce Wang, Jiarui Zhu, Guanbin Li, Cheng-Lin Liu, Liang Lin
机构
*
School of Computer Science and Engineering, Sun Yat-sen University(中山大学计算机科学与工程学院)
;
Guangdong Key Laboratory of Big Data Analysis and Processing(广东大数据分析与处理重点实验室)
;
School of Science, Sun Yat-sen University(中山大学理学院)
;
Hong Kong Polytechnic University(香港理工大学)
;
State Key Laboratory of Multimodal Artificial Intelligence Systems, Institute of Automation, Chinese Academy of Sciences(多模态人工智能系统国家重点实验室,中国科学院自动化研究所)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
专题命中
多模态生成
:cross-modal(title,abstract);分类 cs.CV
CommentsAccepted by IEEE TIP 2025, 16 pages, 11 figures, 7 tables
Journal refIEEE Transactions on Image Processing 34 (2025) 2970-2985
Collaborative Multi-LoRA Experts with Achievement-based Multi-Tasks Loss for Unified Multimodal Information Extraction
Li Yuan, Yi Cai, Xudong Shen, Qing Li, Qingbao Huang, Zikun Deng, Tao Wang
机构
*
School of Software Engineering, South China University of Technology, Guangzhou, China(华南理工大学软件学院)
;
Key Laboratory of Big Data and Intelligent Robot (SCUT), MOE of China(大数据与智能机器人重点实验室)
;
Department of Computing, The Hong Kong Polytechnic University, Hong Kong, China(香港理工大学计算机系)
;
School of Electrical Engineering, Guangxi University, Nanning, China(广西大学电气工程学院)
;
Department of Biostatistics & Health Informatics, King’s College London, London, United Kingdom(伦敦国王学院生物统计与健康信息学系)
VividListener: Expressive and Controllable Listener Dynamics Modeling for Multi-Modal Responsive Interaction
Shiying Li, Xingqun Qi, Bingkun Yang, Chen Weile, Zezhao Tian, Muyi Sun, Qifeng Liu, Man Zhang, Zhenan Sun
机构
*
Beijing University of Posts and Telecommunications(北京邮电大学)
;
Hong Kong University of Science and Technology(香港科技大学)
;
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
机构
*
S-Lab, Nanyang Technological University(南洋理工大学S实验室)
;
Shanghai AI Laboratory Research(上海人工智能实验室)
;
The Chinese University of Hong Kong(香港中文大学)
;
SenseTime Research and Tetras.AI(SenseTime研究部和Tetras.AI)
;
SenseTime Research(商汤科技研究院)
Overcoming False Illusions in Real-World Face Restoration with Multi-Modal Guided Diffusion Model
Keda Tao, Jinjin Gu, Yulun Zhang, Xiucheng Wang, Nan Cheng
机构
*
Xidian University(西安电子科技大学)
;
The University of Sydney(悉尼大学)
;
Shanghai Jiao Tong University(上海交通大学)
;
State Key Laboratory of Integrated Services Networks(集成服务网络国家重点实验室)
机构
*
Zhejiang University(浙江大学)
;
Department of Computer Science, University of Manchester(曼彻斯特大学计算机科学系)
;
School of Informatics, The University of Edinburgh(爱丁堡大学信息学院)