Adaptive Global and Fine-Grained Perceptual Fusion for MLLM Embeddings Compatible with Hard Negative Amplification
自适应全局与细粒度感知融合用于兼容硬负样本放大的人脸嵌入
机构 * State Key Lab of General AI, School of Intelligence Science and Technology, Peking University(人工智能国家重点实验室,智能科学与技术学院,北京大学) ; Institute for Artificial Intelligence, Peking University(人工智能研究院,北京大学)
专题命中 VLM训练与架构 :MLLM(title,abstract);VLM(abstract);分类 cs.CV、cs.LG
AI总结 本文提出AGFF-Embed方法,通过自适应融合全局和细粒度语义信息,提升多模态嵌入在一般和细粒度理解上的性能。