发表机构
Bar-Ilan University; University College London; University of Bologna(巴伊兰大学; 伦敦大学学院; 博洛尼亚大学)
机构由 AI 辅助整理,请以论文原文为准。AI 中文总结
本文针对生成式AI领域一般伦理与职业伦理的冲突问题,提出二者层级关系与情境平衡的框架,可用于系统评估AI模型职业伦理,还给出干预调整方案,为相关研究提供基础。
AI 中文摘要
近年来,AI对齐领域出现了日益明显的分歧:关于AI伦理的研究和政策建议往往假定存在一套通用的伦理价值,然而实际中法律、医疗、翻译等领域不断增多的AI系统特定用途,已有效体现出职业实践伦理。本文首先概述当代AI系统中一般伦理与职业伦理日益冲突的原因,并调研研究文献如何证实但尚未解决这一概念和实践挑战。接着,我们概念化AI模型在职业实践领域决策的主要维度,强调职业伦理与一般伦理的层级结构关系,并详细阐述二者在涉及冲突的情境中达成平衡的机制。我们认为,正是通过这种平衡,某些职业伦理会优先于其他伦理并在实践中得到实施。随后,我们展示该框架如何成为对各领域AI模型职业伦理进行系统实证评估的基础,通过考察模型在一系列相似但不同场景下的输出,确定模型偏好伦理的细微差别。最后,我们提出干预和改变AI模型在职业实践中偏好伦理的方案,同时指出在AI模型职业伦理的评估和实施中都存在固有的主观性维度。
英文摘要
Recent years have seen a growing discrepancy in the field of AI alignment: research and policy recommendations on AI ethics tend to assume a general set of ethical values, yet proliferating practice-specific uses of AI systems on the ground - in the legal, medical and translation domains, among others - have been effectively manifesting ethics of professional practice. This article begins by outlining the reasons why general and professional ethics are increasingly conflicted in contemporary AI systems, and by surveying how the research literature attests to, but has not yet resolved, this conceptual and practical challenge. We then conceptualize the main dimensions of AI models' decision-making in areas of professional practice, emphasizing professional ethics' hierarchically structured relationship with general ethics, and elaborating on the mechanisms through which they reach an equilibrium in situational contexts that involve conflict. It is through this equilibrium, we suggest, that certain professional ethics are prioritized over others and implemented in practice. We then show how our framework can be the basis for a systematic empirical assessment of AI models' professional ethics in various domains, identifying the nuances of the models' favored ethic by examining their production in a series of similar but not identical scenarios. Finally, we propose a formulation for how to intervene in and change AI models' favored ethics in professional practices - while noting the inherent dimension of subjectivity involved in both the evaluation and implementation of professional ethics in AI models.
Comments26 pages, no figures