镜像智能体模型:一种用于可解释智能体行为的贝叶斯架构
The Mirror Agent Model: a Bayesian Architecture for Interpretable Agent Behavior
浏览论文内容
中文总结 AI 辅助
该研究提出镜像智能体模型这一新型贝叶斯架构,结合显著性方法提升解释能力,前期定性结果验证其可生成可解释智能体行为。
中文摘要 AI 辅助
在本文中,我们展示了一种生成可解释行为及解释的新型架构,将其命名为镜像智能体模型,因为该架构将作为显式与隐式通信目标的观察者模型定义为智能体模型的镜像。为全面理解本研究,我们首先展示了涉及智能体意图的信息通信与清晰行为生成的相关前期成果;在论文第二部分,我们通过现成的显著性方法为该架构赋予了新的解释能力,并提供了初步的定性结果。
英文摘要
In this paper we illustrate a novel architecture generating interpretable behavior and explanations. We refer to this architecture as the Mirror Agent Model because it defines the observer model, that is the target of explicit and implicit communications, as a mirror of the agent's. With the goal of providing a general understanding of this work, we firstly show prior relevant results addressing the informative communication of agents intentions and the production of legible behavior. In the second part of the paper we furnish the architecture with novel capabilities for explanations through off-the-shelf saliency methods, followed by preliminary qualitative results.
发表机构
- Umeå University(于默奥大学)
机构由 AI 辅助整理,请以论文原文为准。