SCMM: Calibrating Cross-modal Representations for Text-Based Person Search
SCMM:基于文本的人脸搜索的跨模态表示校准
机构 * College of Future Information Technology, Fudan University(复旦大学未来信息技术学院) ; Department of Electrical and Computer Engineering, The University of British Columbia(不列颠哥伦比亚大学电气与计算机工程系) ; College of Electronic and Information Engineering, Tongji University(同济大学电子与信息工程学院) ; MEGVII Technology(MEGVII技术) ; School of Computer Science and Informatics, Cardiff University(卡迪夫大学计算机科学与信息学院) ; School of AI and CS, Nantong University(南通大学人工智能与计算机科学学院) ; Academy of Artificial Intelligence, SMBU(SMBU人工智能学院) ; College of Computer Science and Software Engineering, Shenzhen University(深圳大学计算机科学与软件工程学院)
专题命中 图文多模态 :cross-modal(title,abstract);image-text(abstract);分类 cs.CV
AI总结 SCMM通过缝校准和掩码建模方法,提升跨模态表示学习在文本驱动的人脸搜索中的性能。
Comments 11 pages, 7 figures, 7 tables