机构
*
School of Information Engineering, Minzu University of China(中国民族大学信息工程学院)
;
School of Artificial Intelligence, Beijing University of Posts and Telecommunications(北京邮电大学人工智能学院)
;
National Library of China, Beijing, China(中国国家图书馆)
;
School of Cyberspace Security, Beijing University of Posts and Telecommunications(北京邮电大学网络安全学院)
Dyslexify: A Mechanistic Defense Against Typographic Attacks in CLIP
Dyslexify: 一种针对CLIP中印刷攻击的机制性防御
Lorenz Hufe, Constantin Venhoff, Erblina Purelku, Maximilian Dreyer, Sebastian Lapuschkin, Wojciech Samek
机构
*
Fraunhofer Heinrich Hertz Institute(弗劳恩霍夫海因里希·赫兹研究所)
;
University of Oxford(牛津大学)
;
Technological University Dublin(都柏林技术大学)
;
Technische Universität Berlin(柏林技术大学)
Journal refManh Luong. (2025). Unbiased Sliced Wasserstein Kernels for High-Quality Audio Captioning. In Advances in Neural Information Processing Systems 38 (NeurIPS 2025)
Annotation-Efficient Vision-Language Model Adaptation to the Polish Language Using the LLaVA Framework
利用LLaVA框架实现波兰语视觉-语言模型的高效标注
Grzegorz Statkiewicz, Alicja Dobrzeniecka, Karolina Seweryn, Aleksandra Krasnodębska, Karolina Piosek, Katarzyna Bogusz, Sebastian Cygert, Wojciech Kusa
Adapting Vision-Language Models for E-commerce Understanding at Scale
大规模电商理解中的视觉-语言模型适应
Matteo Nulli, Vladimir Orshulevich, Tala Bazazo, Christian Herold, Michael Kozielski, Marcin Mazur, Szymon Tuzel, Cees G. M. Snoek, Seyyed Hadi Hashemi, Omar Javed, Yannick Versley, Shahram Khadivi
机构
*
eBay Inc.(eBay公司)
;
University of Amsterdam(阿姆斯特丹大学)
机构
*
University of Science and Technology of China(中国科学技术大学)
;
Shanghai AI Laboratory(上海人工智能实验室)
;
East China Normal University(华东师范大学)
;
The Chinese University of Hong Kong(香港中文大学)
A Vision-Language Foundation Model for Zero-shot Clinical Collaboration and Automated Concept Discovery in Dermatology
一种用于零样本临床协作和皮肤病自动概念发现的视觉-语言基础模型
Siyuan Yan, Xieji Li, Dan Mo, Philipp Tschandl, Yiwen Jiang, Zhonghua Wang, Ming Hu, Lie Ju, Cristina Vico-Alonso, Yizhen Zheng, Jiahe Liu, Juexiao Zhou, Camilla Chello, Jen G. Cheung, Julien Anriot, Luc Thomas, Clare Primiero, Gin Tan, Aik Beng Ng, Simon See, Xiaoying Tang, Albert Ip, Xiaoyang Liao, Adrian Bowling, Martin Haskett, Shuang Zhao, Monika Janda, H. Peter Soyer, Victoria Mar, Harald Kittler, Zongyuan Ge
机构
*
AIM for Health Lab(AIM for Health实验室)
;
Faculty of Information Technology, Monash University(信息技术学院,墨尔本大学)
;
Department of Dermatology, Medical University of Vienna(皮肤科,维也纳医学大学)
;
Faculty of Engineering, Monash University(工程学院,墨尔本大学)
;
Institute of Ophthalmology, University College London(眼科研究所,伦敦大学学院)
;
Dermatology Department. Fundacion Hospital 12 de Octubre(皮肤科部门,12月12日基金会医院)
;
School of Data Science, The Chinese University of Hong Kong, Shenzhen(数据科学学院,香港中文大学(深圳))
;
Frazer Institute, The University of Queensland(弗拉泽研究所,昆士兰大学)
;
Dermatology Research Centre, Brisbane, Australia(皮肤科研究中心,澳大利亚布里斯班)
;
Department of Medical and Cardiovascular Sciences, Sapienza University of Rome(医学与心血管科学系,罗马萨皮恩扎大学)
;
Victorian Melanoma Service, Alfred Care Group, Bayside Health, Melbourne, Australia(维多利亚黑色素瘤服务,阿尔弗雷德护理集团,墨尔本布里斯班健康中心,澳大利亚)
;
SkIIN Discovery Program, Monash University(SkIIN发现计划,墨尔本大学)
;
Claude Bernard Lyon-1 University(克劳德·伯纳德里昂-1大学)
;
Centre Léon Berard, Lyon, France(莱昂·贝拉德中心,法国里昂)
;
Dermatology department, Hôpital Lyon Sud, Hospices Civils de Lyon, Lyons, France(皮肤科部门,里昂南医院,里昂民事医院,法国里昂)
;
Cancer Research Center of Lyon, Lyons, France(里昂癌症研究中心,法国里昂)
;
eResearch Centre, Monash University(eResearch研究中心,墨尔本大学)
;
NVIDIA AI Techno(NVIDIA AI技术)
SHIELD: Suppressing Hallucinations In LVLM Encoders via Bias and Vulnerability Defense
SHIELD:通过偏差和脆弱性防御抑制LVLM编码器中的幻觉
Yiyang Huang, Liang Shi, Yitian Zhang, Yi Xu, Yun Fu
机构
*
Department of Electrical and Computer Engineering, Northeastern University(电气与计算机工程系,东北大学)
;
Khoury College of Computer Science, Northeastern University(计算机科学学院,东北大学)
机构
*
Department of Electrical and Computer Engineering, Seoul National University(首尔国立大学电子与计算机工程系)
;
Amazon(亚马逊)
;
Samsung Research(三星研究院)
;
School of Computer Science and Engineering, Soongsil University(顺天大学计算机科学与工程学院)
;
IPAI, AIIS, ASRI, INMC, and ISRC, Seoul National University(首尔国立大学IPAI、AIIS、ASRI、INMC和ISRC)
CoTZero: Annotation-Free Human-Like Vision Reasoning via Hierarchical Synthetic CoT
CoTZero:通过分层合成CoT实现无标注的人类级视觉推理
Chengyi Du, Yazhe Niu, Dazhong Shen, Luxin Xu
机构
*
University of Electronic Science and Technology of China(电子科技大学)
;
Shanghai Artificial Intelligence Laboratory(上海人工智能实验室)
;
The Chinese University of Hong Kong MMLab(香港中文大学 MMLab)
;
The College of Computer Science and Technology(计算机科学与技术学院)
;
Nanjing University of Aeronautics and Astronautics(南京航空航天大学)