MIMIR: Masked Image Modeling for Mutual Information-based Adversarial Robustness
MIMIR:基于互信息的对抗鲁棒性图像建模
Xiaoyun Xu, Shujian Yu, Zhuoran Liu, Stjepan Picek
机构
*
Radboud University Nijmegen(拉德博德大学尼姆维根分校)
;
Vrije Universiteit Amsterdam(自由大学阿姆斯特丹)
;
University of Zagreb Faculty of Electrical Engineering and Computing(Zagreb大学电子工程与计算学院)
机构
*
University of Michigan Computer Science and Engineering(密歇根大学计算机科学与工程系)
;
University of Michigan Neurosugery(密歇根大学神经外科)
;
University of Cologne Neurosugery(科隆大学神经外科)
;
University of Michigan Radiology(密歇根大学放射学)
;
University of Michigan Neurology(密歇根大学神经病学)
;
University of Michigan Computational Medicine and Bioinformatics(密歇根大学计算医学与生物信息学)
专题命中
VLM训练与架构
:vision language model(abstract);VLM(abstract);分类 cs.CV、cs.AI
SIGMMA: Hierarchical Graph-Based Multi-Scale Multi-modal Contrastive Alignment of Histopathology Image and Spatial Transcriptome
SIGMMA:基于层次图的多尺度多模态对比对齐:组织病理图像与空间转录组
Dabin Jeong, Amirhossein Vahidi, Ciro Ramírez-Suástegui, Marie Moullet, Kevin Ly, Mohammad Vali Sanian, Sebastian Birk, Yinshui Chang, Adam Boxall, Daniyal Jafree, Lloyd Steele, Vijaya Baskar MS, Muzlifah Haniffa, Mohammad Lotfollahi
机构
*
Wellcome Sanger Institute(沃森桑格研究所)
;
Cambridge Centre for AI in Medicine(剑桥人工智能医学中心)
;
Institute of AI for Health(人工智能与健康研究所)
;
Cambridge Stem Cell Institute(剑桥干细胞研究所)
VideoMem: Enhancing Ultra-Long Video Understanding via Adaptive Memory Management
VideoMem: 通过自适应内存管理增强超长视频理解
Hongbo Jin, Qingyuan Wang, Wenhao Zhang, Yang Liu, Sijie Cheng
机构
*
School of Electronic and Computer Engineering, Peking University(电子与计算机工程学院,北京大学)
;
Department of Computer Science and Technology, Tsinghua University(计算机科学与技术系,清华大学)
专题命中
其他VLM
:vision language model(abstract);分类 cs.CV