BACH-V: Bridging Abstract and Concrete Human-Values in Large Language Models
BACH-V: 联结抽象与具体的人类价值观在大语言模型中
AI总结 BACH-V研究通过探测和引导方法揭示大语言模型中抽象与具体价值观的联结机制,发现其能稳定锚定抽象价值观以影响具体决策。
Comments 34 pagess, 16 figures, 6 tables, submitted to ACL 2026
期刊&会议
Annual Meeting of the Association for Computational Linguistics · 会议 · Natural Language Processing
BACH-V: 联结抽象与具体的人类价值观在大语言模型中
AI总结 BACH-V研究通过探测和引导方法揭示大语言模型中抽象与具体价值观的联结机制,发现其能稳定锚定抽象价值观以影响具体决策。
Comments 34 pagess, 16 figures, 6 tables, submitted to ACL 2026
IF-GEO:面向多查询生成引擎优化的冲突感知指令融合
机构 * School of Cyber Science and Technology, University of Science and Technology of China(中国科学技术大学网络科学与技术学院) ; Institute of Dataspace, Hefei Comprehensive National Science Center(合肥综合性国家科学中心数据研究所)
AI总结 IF-GEO通过冲突感知指令融合框架,提升多查询生成引擎在有限预算下的优化稳定性与性能。
Comments 9 pages, 3 figures. Submitted to ACL 2026. Corresponding author: Zhen Chen
LR-DWM: 用于扩散语言模型的高效水印技术
机构 * Bar-Ilan University(巴伊兰大学) ; NVIDIA(英伟达)
AI总结 LR-DWM是一种用于扩散语言模型的高效水印技术,通过利用左右邻居信息偏移生成令牌,实现低开销的高可检测性水印方案。
Comments Submitted to ACL Rolling Review (ARR). 7 pages, 4 figures
MedQA-CS: 用于评估大语言模型临床技能的Objective Structured Clinical Examination (OSCE)风格基准
机构 * University of Massachusetts, Amherst(马萨诸塞大学阿默斯特分校) ; Emory University(埃默里大学) ; University of Minnesota(明尼苏达大学) ; University of Massachusetts, Lowell(马萨诸塞大学洛厄尔分校) ; UMass Chan Medical School(UMass Chan医学学院)
AI总结 MedQA-CS是一种基于OSCE风格的基准,用于评估大语言模型的临床技能,通过模拟医学学生和评估者任务,提供更全面的评估方法。
Comments To appear in proceedings of the Main Conference of the European Chapter of the Association for Computational Linguistics (EACL) 2026
DeAL:大型语言模型的解码时间对齐
机构 * University of Southern California(南加州大学) ; Ω WS AI Labs(Ω WS AI实验室)
AI总结 DeAL通过允许用户自定义奖励函数和实现解码时间对齐,改进了大型语言模型对齐目标的实现。
Comments ACL 2025