AI Wizards at EXIST 2026: Hierarchical Soft-Label Learning for Multimodal Sexism Identification in Memes
2026年EXIST竞赛中的AI奇才:用于表情包中多模态性别歧视识别的分层软标签学习
Matteo Fasulo, Antonio Gravina, Luca Tedeschini, Luca Babboni
机构
*
EXIST Lab(EXIST实验室)
;
CLEF(信息检索评价论坛)
;
Swiss Data Science Center, ETH Zürich(瑞士苏黎世联邦理工学院瑞士数据科学中心)
;
Everest Systems GmbH(珠穆朗玛峰系统有限公司)
;
Villanova.ai S.P.A(维拉诺瓦人工智能股份公司)
机构
*
National Key Laboratory of Human-Machine Hybrid Augmented Intelligence(人机混合增强智能国家级重点实验室)
;
Institute of Artificial Intelligence and Robotics(人工智能与机器人研究院)
;
MiLM Plus
;
Xiaomi Inc(小米公司)
;
Zhongguancun Academy(中关村学院)
;
Beijing, China(北京市)
Can Retrieval Heads See Images? Multimodal Retrieval Heads in Long-Context Vision-Language Models
检索头能看见图像吗?长上下文视觉语言模型中的多模态检索头
Aaron Branson Cigres Li, Zhaowei Wang, Yu Zhao, Yiming Du, Haobo Li, Xiyu Ren, Ginny Wong, Simon See, Lishu Luo, Haodong Duan, Pasquale Minervini, Yangqiu Song
机构
*
HKUST(香港科技大学)
;
University of Edinburgh(爱丁堡大学)
;
CUHK(香港中文大学)
;
NVAITC, NVIDIA, Santa Clara, USA(NVIDIA Santa Clara 分公司)
;
Tsinghua University(清华大学)
DiCE-CIR: Direct Composition Learning for Efficient Zero-Shot Composed Image Retrieval
DiCE-CIR:用于高效零样本合成图像检索的直接合成学习
Gwang-Ho Na, Ho-Joong Kim, Seong-Whan Lee
机构
*
Institute of Information & Communications Technology Planning & Evaluation (IITP)(信息通信技术规划与评估研究所(IITP))
;
Department of Artificial Intelligence, Korea University(韩国大学人工智能系)
High-Fidelity One-Step Generative Visuomotor Policy via Recursive Correction, Frequency Consistency, and Contrastive Flow Matching
通过递归校正、频率一致性和对比流匹配实现高保真一步生成视觉运动策略
Yuran Chen, Xinye Cai, Zhonglin Gong, Yang Huang
机构
*
School of Safety Science and Engineering, Anhui University of Science and Technology(安徽理工大学安全科学与工程学院)
;
State Key Laboratory of Digital Intelligent Technology for Unmanned Coal Mining, Anhui University of Science and Technology(安徽理工大学煤矿智能开采技术与装备国家重点实验室)
;
School of Artificial Intelligence, Anhui University of Science and Technology(安徽理工大学人工智能学院)
Language-guided Medical Image Segmentation with Target-informed Multi-level Contrastive Alignments
基于目标感知多级对比对齐的语言引导医学图像分割
Mingjian Li, Mingyuan Meng, Shuchang Ye, Mingye Zou, Michael Fulham, Lei Bi, Jinman Kim
机构
*
School of Computer Science, The University of Sydney(悉尼大学计算机科学学院)
;
Institute of Translational Medicine, Shanghai Jiao Tong University(上海交通大学转化医学研究院)
;
Zhongguancun Academy & Zhongguancun Institute of Artificial Intelligence, Beijing, China(北京中关村学院及中关村人工智能研究院)
;
Department of Molecular Imaging, Royal Prince Alfred Hospital(皇家珀斯阿尔弗雷德医院分子成像部)
CommentsAccepted to ICML 2026. This version updates the ICML submission with an optimized model checkpoint. Project page: https://omni-diffusion.github.io
GeoSearcher: Anchor-Guided Progressive Reasoning for Remote Sensing Visual Grounding with Process Supervision
GeoSearcher: 基于锚点引导的渐进推理遥感视觉定位与过程监督
Dianyu Wang, Peirong Zhang, Xuyang Li, Xiaoxuan Liu, Lei Wang
机构
*
Key Laboratory of Target Cognition and Application Technology (TCAT), Chinese Academy of Sciences(中国科学院目标认知与应用技术重点实验室)
;
School of Electronic, Electrical and Communication Engineering, University of Chinese Academy of Sciences(中国科学院大学电子电气与通信工程学院)
Region-Aware Multimodal Large Language Model via SlowFast Tokenization and Pseudo-Mask Guidance for 3D CT Report Generation
区域感知多模态大语言模型:基于慢快标记化与伪掩码引导的3D CT报告生成
Sunggu Kyung, Jinyoung Seo, Hyunseok Lim, Dongyeong Kim, Hyungbin Park, Jimin Sung, Jihyun Kim, Wooyoung Jo, Yoojin Nam, Namkug Kim
机构
*
Department of Convergence Medicine, University of Ulsan College of Medicine, Asan Medical Center, Seoul, Republic of Korea(韩国首尔峨山医疗中心蔚山大学医学院融合医学系)
;
University of Ulsan College of Medicine, Seoul, Republic of Korea(韩国首尔蔚山大学医学院)
;
Department of Radiology and Research Institute of Radiology, University of Ulsan College of Medicine, Asan Medical Center, Seoul, Republic of Korea(韩国首尔峨山医疗中心蔚山大学医学院放射科与放射学研究所)
See the Emotion: A Facial Emoji Proxy Modeling for EEG Emotion Recognition
看见情感:用于脑电图情感识别的面部表情符号代理建模
Jingjing Hu, Guo Dan, Haofan Cheng, Ying Zeng, Zhan Si, Jinxing Zhou, Meng Wang
机构
*
Hefei University of Technology(合肥工业大学)
;
PLA Information Engineering University(中国人民解放军信息工程大学)
;
University of Science and Technology of China(中国科学技术大学)
;
MBZUAI
Zhiyue Xu, Fandi Meng, Kaijie Xu, Clark Verbrugge, Simon Lucas, Jian Zhao
机构
*
Institute of Automation, Chinese Academy of Sciences(中国科学院自动化研究所)
;
Zhongguancun Academy(中关村学院)
;
Zhongguancun Institute of Artificial Intelligence(中关村人工智能研究所)
;
School of Computer Science, McGill University, Montreal, Quebec, Canada(麦吉尔大学计算机科学学院)
;
Game AI Research Group, Queen Mary University of London, London, United Kingdom(伦敦女王学院游戏人工智能研究组)