Bridging the Semantic Gap: Contrastive Rewards for Multilingual Text-to-SQL with GRPO
弥合语义鸿沟:基于GRPO的多语言文本到SQL的对比奖励
机构 * Independent Researcher(独立研究者) ; University of Michigan(密歇根大学) ; Harrisburg University of Science(哈里斯堡科学大学) ; Forward Labs AI
AI总结 本研究提出基于GRPO的对比奖励方法,提升多语言文本到SQL系统的执行与语义准确性,通过对比奖励信号实现语义对齐,实验表明小模型在少量数据下表现优于大模型。
Comments 20th International Workshop on Semantic and Social Media Adaptation & Personalization