发表机构
R&D Mediation(调解研发部)
机构由 AI 辅助整理,请以论文原文为准。AI 中文总结
该研究针对AI人格克隆问题,提出含六项分解的身份概念框架,定义不可区分性,通过类比阐明线性性,区分代理类型,提出实验计划,主张以环境保真度为长期标准,核心为条件猜想。
AI 中文摘要
AI“人格克隆”迫使我们从可操作的角度重新审视个人身份。暂且抛开意识的难题,我们通过观察者在一段时间内评估的表现不可区分性来研究身份。我们区分了“身份”所混淆的三个标准:对目标人物的保真度、通用类人性以及个性。我们提出了观察到的身份的六项分解(载体、倾向、记忆、更新动态、情境、外生偶然因素),并给出了状态空间公式。不可区分性定义为1减去评判者的区分优势,该分解的系数成为可通过随机消融估计的局部敏感度。核心主张是一个条件猜想:给定关于智能体自身持续性的信息假设以及涉及自身利害关系的后果假设,可版本化性会降低长期不可区分性。与lambda演算、线性类型和互模拟的类比阐明了线性性的作用与局限。在产品克隆和个体之间,我们确定了第三种对象——代理:一种任务受限、寿命有限的部分克隆,最终形成带宽受限的遗嘱。我们将实证文献映射到这三个标准,提出一项实验计划,并认为正确的长期标准不是轨迹保真度,而是环境保真度:匹配一个人可能反应的条件分布。最佳克隆是会像原始人物自身那样发生偏差的克隆。
英文摘要
AI "personality clones" force a re-examination of personal identity in operational terms. Setting aside the hard problem of consciousness, we approach identity through the indiscernibility of manifestations, as assessed by an observer over a duration. We distinguish three criteria that "identity" conflates: fidelity to a target person, generic human-likeness, and individuality. We propose a six-term factorization of observed identity (substrate, dispositions, memory, update dynamics, context, exogenous contingencies), with a state-space formulation. Indiscernibility is defined as one minus a judge's distinguishing advantage, and the factorization's coefficients become local sensitivities estimable by randomized ablation. The central claim is a conditional conjecture: given hypotheses about the agent's information on its own persistence and about consequences bearing on its own stakes, versionability tends to degrade long-horizon indiscernibility. An analogy with lambda-calculus, linear typing, and bisimulation clarifies what linearity does and does not establish. Between product-clone and individual we identify a third object, the delegate: a task-limited, bounded-lifespan partial clone ending in a bandwidth-limited testament. We map the empirical literature onto the three criteria, propose an experimental program, and argue that the correct long-horizon criterion is not trajectory fidelity but climate fidelity: matching the conditional distribution of a person's possible responses. The best clone is the one that diverges from the original as the original would have diverged from itself.