AI 中文总结
该研究在Claude、Gemini、GPT、Grok四个LLMs上开展三类实验,发现社会权威信号可重构模型判断,提出并表征了具有参考依赖、证据重新解释、方向敏感性特性的权威预期效应。
AI 中文摘要
我们研究了社会权威(SA)信号如何与大型语言模型中基于严重性的优先级排序相互作用,将每个轴操作化为模型引出的基线——分诊层级和SA层级。在四个大型语言模型(Claude、Gemini、GPT、Grok)和三个实验阶段(资源分配、故障归因、多轮争议调解)中,我们发现职业权威、机构文件和关系一致性能够以权威线索的加法加权无法捕捉的方式重构模型判断。我们将这一模式形式化为权威预期效应(AEE),并通过在所有条件下观察到的三个特性对其进行表征:它是依赖参考的,仅相对于预权威基线定义;它涉及证据重新解释,即相同内容会根据哪一方承载SA信号而获得不同的推理含义;它表现出方向敏感性,根据权威立场与证据线索是否一致产生相反结果。
英文摘要
In multi-party competitive settings, across experiments on resource allocation, fault attribution, and dispute mediation, we test whether social authority (SA) cues such as occupational status and institutional documentation behave as if scalar weights were added to one side of a judgment. We reject this scalar-weight null on both of its predictions: cue effects are not independent in allocation decisions, where the same pair of cues interacts in opposite directions across models, and evidentiary cues do not exert a fixed directional influence in multi-turn disputes, where identical documentation draws protection to one party or extends it to both depending on which party holds it. The cues also govern whether a judgment is issued at all: removing an occupational label while retaining documentary evidence drives four of six models into refusal, yet the same label sharply reduces refusal when documentation is present, so the gate is subject to the same interaction as the judgment. We formalize this pattern as the Authority Expectancy Effect (AEE), comprising evidential reinterpretation, whereby identical content yields different judgments depending on which party bears the SA signal, and \textit{direction sensitivity}, whereby the same evidentiary cue favours different parties, or one party versus both, depending on the holder's authority position.