Self-Generated Text Recognition: Quality Heuristics, Cross-Task Transfer, and Downstream Bias in LLM Evaluation
自生成文本识别:LLM评估中的质量启发式、跨任务迁移与下游偏差
机构 * George Washington University(乔治·华盛顿大学) ; Geodesic Research(吉奥戴斯克研究机构) ; University of Cambridge(剑桥大学)
AI总结 该研究聚焦LLM的自生成文本识别(SGTR)能力,明确了导致过往研究结论分歧的实验设计因素,发现SGTR性能可跨任务迁移,且训练SGTR会引发下游偏差,强调需监控SGTR以保障AI安全。
Comments 31 pages, 9 figures (3 main body, 6 appendix), 18 tables (1 main body, 17 appendix)