机构
*
State Key Lab of CAD&CG, Zhejiang University(浙江大学CAD与CG国家重点实验室)
;
UESTC
;
University of Virginia(弗吉尼亚大学)
;
HKUST(GZ)(香港科技大学(广州))
;
Cornell University(康奈尔大学)
;
Zhejiang University(浙江大学)
;
National University of Singapore(新加坡国立大学)
机构
*
Shanghai Jiao Tong University(上海交通大学)
;
S-Lab, Nanyang Technological University(南洋理工大学S实验室)
;
Shanghai AI Laboratory(上海人工智能实验室)
;
University of Science and Technology of China(中国科学技术大学)
;
Stanford University(斯坦福大学)
;
Shanghai Innovation Institute(上海创新研究院)
;
The Chinese University of Hong Kong(香港中文大学)
;
Fudan University(复旦大学)
;
CPII under InnoHK(InnoHK下的CPII)
;
Adobe Research(Adobe研究)
CommentsPeer reviewed and accepted at CVPR 2026 at the GRAIL-V (Grounded Retrieval and Agentic Intelligence for Vision-Language) workshop (non-archival track)
MoE-dqINR: A Unified Mixture-of-Experts Implicit Neural Representation Framework for Scan-Specific Dynamic and Quantitative MRI Reconstruction
MoE-dqINR:用于特定扫描动态和定量MRI重建的统一混合专家隐式神经表示框架
Yinzhe Wu, Fanwen Wang, Zhenxuan Zhang, Zi Wang, Chengyan Wang, Guang Yang
机构
*
Department of Bioengineering and I-X, Imperial College London(生物工程系和I-X,帝国理工学院伦敦分校)
;
Cardiovascular Research Centre, Royal Brompton Hospital(心脏血管研究中心,皇家布隆特医院)
;
National Heart and Lung Institute, Imperial College London(国家心脏和肺研究所,帝国理工学院伦敦分校)
;
School of Biomedical Engineering & Imaging Sciences, King’s College London(生物医学工程与成像科学学院,伦敦国王学院)
;
Shanghai Pudong Hospital and Human Phenome Institute, Fudan University(上海浦东医院和人类表型研究所,复旦大学)
;
International Human Phenome Institute (Shanghai), Shanghai, China(国际人类表型研究所(上海),上海,中国)
MAIL++: Multi-Modal Bi-directional Agent Layer for Vision-Language Models
MAIL++: 视觉语言模型的多模态双向智能体层
Kaixiang Chen, Pengfei Fang, Hui Xue
机构
*
School of Computer Science and Engineering, Southeast University(东南大学计算机科学与工程学院)
;
Key Laboratory of New Generation Artificial Intelligence Technology and Its Interdisciplinary Applications (Southeast University), Ministry of Education, China(新一代人工智能技术及其交叉应用国家重点实验室(东南大学),中华人民共和国教育部,中国)
机构
*
Nanjing University of Aeronautics and Astronautics(南京航空航天大学)
;
School of Software Technology, Zhejiang University, Ningbo, China(浙江大学宁波校区软件学院)
;
Ningbo Global Innovation Center, Zhejiang University, Ningbo, China(浙江大学宁波全球创新中心)
;
Collaborative Innovation Center of Novel Software Technology and Industrialization(新型软件技术与产业化协同创新中心)
Comments9 pages, 16 figures. Accepted at the ICLR 2026 Workshop on Principled Design for Trustworthy AI: Interpretability, Robustness, and Safety across Modalities
CapCLIP: A Vision-Language Representation Alignment Approach for Wireless Capsule Endoscopy Analysis
CapCLIP:一种用于无线胶囊内镜分析的视觉-语言表示对齐方法
Haroon Wahab, Irfan Mehmood, Hassan Ugail
机构
*
School of Computer Science, AI and Electronics Faculty of Engineering and Digital Technologies(计算机科学与电子工程学院,工程与数字技术学院)
;
School of Management Faculty of Mgmt, Law & Social Sciences(管理学院,管理、法律与社会科学学院)
;
Centre for Visual Computing and Intelligent Systems(视觉计算与智能系统中心)
CommentsThis article is withdrawn because the experimental results and analysis require substantial revision. The current version should not be cited as a reliable representation of the work
CROC: Evaluating and Training T2I Metrics with Pseudo- and Human-Labeled Contrastive Robustness Checks
CROC:通过伪标签和人工标注的对比鲁棒性检查评估和训练T2I度量
Christoph Leiter, Yuki M. Asano, Margret Keuper, Steffen Eger
机构
*
University of Mannheim(曼海姆大学)
;
University of Technology Nuremberg(纽伦堡技术大学)
;
Max Planck Institute for Informatics, Saarland Informatics Campus(马克斯·普朗克信息研究所,萨尔兰信息校园)
机构
*
University of Trento(特伦托大学)
;
Fondazione Bruno Kessler(布鲁诺·凯斯勒基金会)
;
ENSTA Paris, Institut Polytechnique de Paris(巴黎高等先进技术学院,巴黎综合理工学院)
;
Hefei University of Technology(合肥工业大学)
;
University of Bergamo(贝加莫大学)
;
Technical University of Munich(慕尼黑工业大学)