发表机构
Seoul National University; McGovern Medical School; UTHealth; Texas Institute for Restorative Neurotechnologies(首尔国立大学; 麦戈文医学院; 德克萨斯大学休斯顿健康科学中心; 德克萨斯修复性神经技术研究所)
机构由 AI 辅助整理,请以论文原文为准。AI 中文总结
提出耦合主动推断对话模型,将RSA扩展为双向推断,区分四类更新,预测适应持久性与转移,并解释澄清逆转、自我确认误解等对话现象。
AI 中文摘要
相同的言语选择可能源于不同的交际原因,相同的解释也可能在听者的学习过程中留下不同的痕迹。理性言语行为(RSA)模型将解释视为对说话者意义的推断,但标准的一次性RSA并不能内在地区分这些因果更新目标。我们开发了一个耦合的主动推断对话模型,其中听者对某一话语的似然度是归因于说话者的策略分布,将说话者的预期自由能置于听者的变分自由能之中。因此,每个话语既是对伙伴的证据,也是对伙伴的干预。该模型区分了RSA方法通常合并的四类更新:对伙伴当前语用状态的推断;伙伴特定参数的学习;对澄清或修复的前瞻性评估;以及由习惯和情境敏感的时间成本塑造的策略先验。在一步精确推断的限制下,该模型恢复RSA说话者和听者,RSA成为耦合过程的受限单轮极限。在这些限制之外,更新遵循不同的规则和时间尺度,预测哪些适应会持续、保持伙伴特定性或发生转移。实例表明,澄清后听众设计发生逆转;自我确认的误解中,双方在指称上存在分歧但自由能均较低;以及在不确定性解决前理性地结束对话。与将自然语言等同于交际的批评一致,该模型将交际视为语言结构的下游使用,并将产生与解释重新定义为耦合推断:说话既是对伙伴的干预,也是为说话者关于该伙伴的模型采样证据的认知行动。
英文摘要
Identical utterance choices can arise from different communicative causes, and identical interpretations can leave different traces in what a listener learns. Rational Speech Act (RSA) models treat interpretation as inference over speaker meaning, but standard one-shot RSA does not intrinsically distinguish these causal update targets. We develop a coupled active-inference model of dialogue in which a listener's likelihood for an utterance is the policy distribution attributed to the speaker, placing the speaker's expected free energy within the listener's variational free energy. Each utterance is therefore both evidence about and an intervention on a partner. The model separates four updates that RSA approaches typically collapse: inference about the partner's current pragmatic state; learning of partner-specific parameters; prospective evaluation of clarification or repair; and a policy prior shaped by habit and a context-sensitive price of time. Under one-step, exact-inference restrictions, the model recovers the RSA speaker and listener, with RSA as the restricted single-turn limit of the coupled process. Outside these restrictions, the updates obey distinct rules and timescales, predicting which adaptations persist, remain partner-specific, or transfer. Worked examples show audience design reversing after clarification, self-confirming misunderstanding in which both interlocutors have low free energy while disagreeing about reference, and rational closing before uncertainty is resolved. Consistent with critiques of equating natural language with communication, the model treats communication as a downstream use of linguistic structure and recasts production and interpretation as coupled inference: speaking is both an intervention on a partner and an epistemic action that samples evidence for the speaker's model of that partner.
Comments83 pages, 3 figures, 2 tables