Quantifying LLM Biases Across Instruction Boundary in Mixed Question Forms
量化混合问题形式中跨指令边界的LLM偏见
机构 * Hong Kong University of Science and Technology (Guangzhou)(香港科技大学(广州)) ; University of Pennsylvania(宾夕法尼亚大学) ; Huazhong University of Science and Technology(华中科技大学) ; Hong Kong Polytechnic University(香港理工大学) ; Nanjing University of Posts and Telecommunications(南京邮电大学)
AI总结 本文提出BiasDetector基准,用于评估LLM在混合问题形式数据集下对稀疏标签混合的识别能力,揭示用户指令对LLM偏见的影响。