It's the humans, not the data: Geopolitical bias in LLMs originates in post-training, amplified by the language of the prompt
是人类,而非数据:LLM中的地缘政治偏见源于后训练,并通过提示语言放大
机构 * Alibaba(阿里巴巴) ; seven AI labs(七家人工智能实验室)
AI总结 研究发现大语言模型的地缘政治偏见主要源于后训练阶段而非预训练,且偏见方向与模型开发者所在国一致,提示语言会放大该偏见。
Comments 12 pages, 6 figures, 2 tables, 3 appendices. Code and scenario bank: https://github.com/recozers/LLM-Bias