AI 中文总结
针对持久化AI智能体状态转换验证问题,提出认知连续性测试(CCT)契约,通过限定授权、来源、确定性应用等机制区分合法与违规转换,并在IdentityLineageBench上验证有效性。
AI 中文摘要
持久化AI智能体会修订信念、整合记忆并替换执行基质。相似的后继状态可能伴随不同的授权转换声明,而合法的开发可能大幅改变状态。我们引入了认知连续性测试(CCT),这是一种基于策略的契约,用于使用限定范围授权、来源、确定性应用、语义谓词和候选持久性收据来验证提交的转换。CCT区分已验证的可接受性、明确违规和未解决的必要证据。分离结果涉及转换声明而非实时运行时身份;可靠性以指定的检查器和评估器假设为条件。IdentityLineageBench提供了24个生成的转换族。参考解析后验证器匹配所有576个规范保留标签;词汇状态相似性和仅谱系诊断基线分别接受60.0%和80.0%的无效夹具。这些比较确立了合成符合性,而非相对于策略感知部署系统的优越性。签名对抗性回归覆盖了捏造的交互计数、无支持的信念变化以及混合缺失/矛盾证据。SIT行为和实际模型迁移仍未测量。一项18,000次执行的有效路径研究测量了驻留输入上的6.21毫秒默认中位数。我们指定了部署所需的额外激活和恢复义务。
英文摘要
Persistent AI agents revise beliefs, consolidate memory, and replace execution substrates. Similar successor states can accompany differently authorized transition claims, while legitimate development can change state substantially. We introduce the Cognitive Continuity Test (CCT), a policy-relative contract for verifying submitted transitions using scoped authority, provenance, deterministic application, semantic predicates, and candidate-persistence receipts. CCT distinguishes verified admissibility, affirmative violation, and unresolved required evidence. Separation results concern transition claims rather than live runtime identity; soundness is conditional on the specified checker and evaluator assumptions. IdentityLineageBench provides 24 generated transition families. The reference post-resolution verifier matches all 576 canonical held-out labels; lexical state similarity and a lineage-only diagnostic baseline admit 60.0% and 80.0% of invalid fixtures. These comparisons establish synthetic conformance, not superiority to a policy-aware deployed system. Signed adversarial regressions cover fabricated interaction counts, unsupported belief changes, and mixed missing/contradictory evidence. SIT behavior and actual model migration remain unmeasured. An 18,000-execution valid-path study measures a 6.21 ms default median on resident inputs. We specify the additional activation and recovery obligations needed for deployment.
Comments17 pages, 2 figures, 3 tables. Includes formal proofs, transition taxonomy, and benchmark schema appendices. Reference verifier and reproducible evaluation artifacts available at https://github.com/openkedge/cctbench