XAI-CLIP: ROI-Guided Perturbation Framework for Explainable Medical Image Segmentation in Multimodal Vision-Language Models
XAI-CLIP: 通过区域感兴趣引导扰动框架实现多模态视觉-语言模型中可解释的医学图像分割
Thuraya Alzubaidi, Sana Ammar, Maryam Alsharqi, Islem Rekik, Muzammil Behzad
机构
*
King Fahd University of Petroleum and Minerals(国王法赫德石油和矿物大学)
;
Massachusetts Institute of Technology(麻省理工学院)
;
Imperial College London(伦敦帝国学院)
;
KFUPM-SDAIA Joint Research Centre for Artificial Intelligence(KFUPM-SDAIA联合人工智能研究中心)
CoTZero: Annotation-Free Human-Like Vision Reasoning via Hierarchical Synthetic CoT
CoTZero:通过分层合成CoT实现无标注的人类级视觉推理
Chengyi Du, Yazhe Niu, Dazhong Shen, Luxin Xu
机构
*
University of Electronic Science and Technology of China(电子科技大学)
;
Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)
;
The Chinese University of Hong Kong MMLab(香港中文大学 MMLab)
;
The College of Computer Science and Technology(计算机科学与技术学院)
;
Nanjing University of Aeronautics and Astronautics(南京航空航天大学)
EAGLE: Elevating Geometric Reasoning through LLM-empowered Visual Instruction Tuning
通过LLM赋能的视觉指令微调提升几何推理能力
Zhihao Li, Yao Du, Yang Liu, Yan Zhang, Yufang Liu, Mengdi Zhang, Xunliang Cai, Charles Ling, Boyu Wang
机构
*
Department of Computer Science, Western University(计算机科学系,西部大学)
;
Meituan Inc.(美团公司)
;
Department of Automation, Tsinghua University(自动化系,清华大学)
;
School of Computer Science and Technology, East China Normal University(计算机科学与技术学院,东华大学)
Reducing Aleatoric and Epistemic Uncertainty through Multi-modal Data Acquisition
通过多模态数据采集减少偶然性和epistemic不确定性
Arthur Hoarau, Benjamin Quost, Sébastien Destercke, Willem Waegeman
机构
*
Université de Lorraine, CentraleSupélec CNRS, LORIA, Metz, France(洛林大学、中央超导电子学学院 CNRS、LORIA、法国梅茨)
;
Université de technologie de Compiègne UMR CNRS 7253 Heudiasyc, France(图卢兹理工学院 UMR CNRS 7253 Heudiasyc、法国)
;
CNRS, Université de technologie de Compiègne UMR CNRS 7253 Heudiasyc, France(CNRS、图卢兹理工学院 UMR CNRS 7253 Heudiasyc、法国)
;
University of Ghent Ghent, Belgium(根特大学、比利时根特)
CommentsAccepted for publication at the 10th ACM International Conference on Intelligent Systems, Metaheuristics & Swarm Intelligence (ISMSI 2026), April 24-26, Cebu City, Phillipines
机构
*
The University of New South Wales, Sydney, New South Wales, Australia(新南威尔士大学)
;
The University of Sydney, Sydney, New South Wales, Australia(悉尼大学)
;
School of Software and Microelectronics, Peking University, Zibo, Shandong, China(北京大学软件与微电子学院)
;
Shandong University of Technology, Beijing, China(山东理工大学)
CommentsThis version corrects the author affiliation to reflect the accurate institutional information at the time of publication. No technical content of the paper has been changed
VideoVeritas: AI-Generated Video Detection via Perception Pretext Reinforcement Learning
VideoVeritas:通过感知预设强化学习实现AI生成视频检测
Hao Tan, Jun Lan, Senyuan Shi, Zichang Tan, Zijian Yu, Huijia Zhu, Weiqiang Wang, Jun Wan, Zhen Lei
机构
*
MAIS, Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所MAIS部)
;
School of Artificial Intelligence, University of Chinese Academy of Sciences(中国科学院大学人工智能学院)
;
Shenzhen Institute of Advanced Technology (SIAT), Chinese Academy of Sciences(中国科学院深圳先进技术研究所)