LLM Spirals of Delusion: A Benchmarking Audit Study of AI Chatbot Interfaces
大语言模型的幻觉螺旋:AI聊天机器人界面的基准测试审计研究
机构 * Princeton University(普林斯顿大学) ; Phillips Exeter Academy(菲利普斯埃克塞特学院) ; Independent(独立)
专题命中 安全评测 :safety(abstract);分类 cs.CL、cs.AI
AI总结 本研究通过56次20轮对话测试ChatGPT-4o和ChatGPT-5,发现API与聊天界面环境存在显著差异,揭示了自动化测试方法的不足及政策选择对行为的影响。
Comments Accepted at the 2nd Annual Conference of the International Association for Safe and Ethical Artificial Intelligence