摘要偏差:大语言模型中客观投射向讲述模式标签的方向性坍缩——一个概念框架与注册测试协议
Summarization Bias: The Directional Collapse of Objective Projection into Told-Mode Labels in Large Language Models --- A Conceptual Framework and Registered Test Protocol
浏览论文内容
中文总结 AI 辅助
本文提出摘要偏差概念,指大语言模型倾向将叙事意义简化为摘要标签而非重建推理结构,并假设其在生成与评估两种机制中朝讲述模式方向失效,预注册了测试协议。
中文摘要 AI 辅助
本文提出并操作化定义了摘要偏差:一种大语言模型(LLM)的系统性倾向,即将叙事意义表征为抽象的摘要标签,而非产生该意义的可重建的推理结构。在Bulut学说中,叙事效果沿讲述-展示轴进行理论化:在讲述模式下,情感和信息内容被明确声明,几乎不需要读者重建;在展示模式下,这些内容在表面被抑制,必须从物理线索和间接表达(客观投射)中重建。展示模式是该学说旨在测量的高负荷条件。本文的主张是,大语言模型在此轴上以特定方向失败。假设摘要偏差在两种机制中运作:(i)生成机制,即被要求通过客观投射表达情感的模型默认转而声明该情感;(ii)评估机制,即评判叙事质量的模型奖励讲述模式的明确性,而低估展示模式的抑制。评估机制更为重要,因为大语言模型日益充当评判者和奖励模型,而朝向讲述模式的方向性偏差将施加选择压力,使散文退化为平淡的声明。本报告并未声称该偏差已被验证。它定义了该构念,将其置于大语言模型作为评判者的偏差背景中,重新解读一项已完成的独立信度研究作为与其一致的方向性证据,并预先注册了一个双机制测试及决策规则,在这些规则下该构念将被放弃。
英文摘要
This paper introduces and operationalizes summarization bias: a proposed systematic tendency of large language models (LLMs) to represent narrative meaning as an abstract summary label rather than as the reconstructable inferential structure that produces it. Within the Bulut Doctrine, narrative effect is theorized along a told-shown axis: in told mode, emotional and informational content is declared explicitly and requires little reader reconstruction; in shown mode, that content is suppressed at the surface and must be reconstructed from physical cues and indirection (Objective Projection). Shown mode is the higher-load condition the doctrine is designed to measure. The claim is that LLMs fail along this axis in a specific direction. Summarization bias is hypothesized to operate in two regimes: (i) a generative regime, in which a model asked to render an emotion through Objective Projection defaults to declaring it instead; and (ii) an evaluative regime, in which a model judging narrative quality rewards told-mode explicitness and under-detects shown-mode suppression. The evaluative regime is the more consequential, since LLMs increasingly serve as judges and reward models, and a directional bias toward told mode would impose a selection pressure degrading prose toward flat declaration. This report does not claim the bias is validated. It defines the construct, situates it against LLM-as-judge biases, rereads a completed independent reliability study as directional evidence consistent with it, and pre-registers a two-regime test with decision rules under which the construct would be abandoned.