arXivDaily arXiv每日学术速递 周一至周五更新
arXiv周末暂无论文更新,休息一下吧,周末愉快~~
arXiv 2608.02660cs.CYcs.AIcs.HC

AI对齐与受托义务

AI Alignment and Fiduciary Obligation

Benjamin Lange

首次发表
浏览论文内容

中文总结 AI 辅助

本文提出将受托理论应用于长期AI助手部署,以开发者对用户负有的忠诚、注意、善意和坦诚四项受托义务为基础构建AI对齐标准,为AI对齐研究提供新视角。

中文摘要 AI 辅助

先进AI助手在日益广泛的角色中与用户展开长期互动,包括建议、决策支持、协作、学习、情感支持及陪伴等。当前的对齐工作围绕应采用何种对齐标准来规范这些关系展开,借鉴了适用于人类关系的道德传统,如生物伦理学、美德伦理学、关怀伦理学及关系科学。本文探讨用户-AI-开发者三方关系中的AI对齐标准,因为每一次用户-AI互动都由开发者调节,开发者对系统的行为、记忆及互动参数拥有自由裁量权。基于商业伦理学和法律学术研究,本文主张受托理论适用于长期AI助手部署。在此基础上,受托义务的四项核心义务——忠诚、注意、善意和坦诚,可生成开发者-用户关系的对齐标准。本文将长期AI助手部署的四项用户侧风险与四项义务对应,并明确了履行每项义务所需的制度措施。该讨论补充了现有方法,其将对齐标准建立在开发者对用户负有的义务之上,而非用户-AI互动应促进的价值观之上,且表明这些义务的存在独立于对用户造成的任何实际损害。

英文摘要

Advanced AI assistants engage users in extended interactions across a widening range of roles, including advice, decision support, collaboration, learning, emotional support, and companionship among others. Current alignment efforts consider what alignment criteria should govern these relationships, drawing on moral traditions developed for human relationships such as bioethics, virtue ethics, care ethics, and relationship science. This paper considers AI alignment criteria in the user-AI-developer triad, since every user-AI interaction is mediated by a developer who exercises discretionary control over a system's behaviour, memory, and engagement parameters. Drawing on business ethics and legal scholarship, I argue that fiduciary theory applies to extended AI assistant deployment. On this basis, the four canonical fiduciary duties of loyalty, care, good faith, and candour can generate alignment criteria for the developer-user relationship. I map four user-side risks of extended AI assistant deployment to the four duties and specify institutional measures that follow from discharging each duty. The discussion complements existing approaches by grounding alignment criteria in obligations the developer owes the user, rather than in values the user-AI interaction should promote, and by showing that those obligations hold independently of any de facto harm to users.

发表机构

  • Munich Center for Machine Learning(慕尼黑机器学习中心)

机构由 AI 辅助整理,请以论文原文为准。

补充信息

↑