DVLA-RL: Dual-Level Vision-Language Alignment with Reinforcement Learning Gating for Few-Shot Learning
DVLA-RL:基于强化学习门控的双层视觉-语言对齐用于少样本学习
机构 * Software School, Shandong University(山东大学软件学院) ; Shenzhen Loop Area Institute(深圳河套学院) ; School of Computing and Artificial Intelligence, Shandong University of Finance and Economics(山东财经大学计算机与人工智能学院)
专题命中 安全训练 :alignment(title,abstract)
AI总结 DVLA-RL通过双层语义构建和强化学习门控注意力,实现少样本学习中视觉与语言的双层次对齐,提升泛化能力。
Comments Accepted by ICLR 2026