发表机构
Stanford University; The University of Tokyo; Boston University; National University of Singapore; Massachusetts Institute of Technology(斯坦福大学; 东京大学; 波士顿大学; 新加坡国立大学; 麻省理工学院)
机构由 AI 辅助整理,请以论文原文为准。AI 中文总结
针对机器人进入人类空间时信任缺失的问题,提出整合安全性、行为可理解性和感知对齐的可信物理AI框架,强调通过物理能力、行为与期望的对齐实现可信性。
AI 中文摘要
机器人进入人类空间的速度超过了我们确定它们何时值得信任的速度。我们提出了一个可信物理AI框架,该框架在具身、控制、认知和设计层面整合了安全性、行为可理解性和感知对齐。可信性源于物理能力、可观察行为以及人们在交互过程中形成的期望之间的对齐。
英文摘要
Robots are entering human spaces faster than we can establish when they deserve trust. We propose a framework for trustworthy physical AI that integrates Safety, Behavioral Intelligibility, and Perceptual Alignment across embodiment, control, cognition, and design. Trustworthiness emerges from aligning physical capabilities, observable behavior, and expectations people form during interaction.