作为软件的智能体:面向智能体可靠性的编程语言议程
Agents as Software: A Programming Languages Agenda for Agent Reliability
- Microsoft Research(微软研究院)
- University of Wisconsin-Madison(威斯康星大学麦迪逊分校)
机构由 AI 辅助整理,请以论文原文为准。
AI总结:
本文提出以编程系统视角构建可靠AI智能体,将其视为可编程工件,通过轨迹与状态指定行为、部署前检查、执行中监控及失败改进,实现行为可控可修。
AI中文摘要:
AI智能体日益类似于软件系统:它们调用工具、记忆事实、遵循策略、委派工作,并采取具有实际后果的行动。然而,智能体的“程序”分散在提示词、工具、记忆、工作流程和执行轨迹中,使得仅通过常规测试和调试难以检查其行为。本文主张,编程系统视角为构建可靠智能体提供了自然的透镜。我们将智能体重新定义为可编程工件,其行为可以在轨迹和状态上进行指定,在部署前进行检查,在执行期间进行监控,并从观察到的失败中改进。目标并非让概率性智能体表现得像确定性程序,而是赋予它们足够的结构,使其行为能够被推理、控制和修复。
英文摘要:
AI agents increasingly resemble software systems: they call tools, remember facts, follow policies, delegate work, and take actions with real consequences. % Yet the ``program'' of an agent is scattered across prompts, tools, memories, workflows, and execution traces, making its behavior difficult to inspect through ordinary testing and debugging alone. % This essay argues that a programming-systems perspective offers a natural lens for making agents reliable. % We recast agents as programmable artifacts whose behavior can be specified over traces and state, checked before deployment, monitored during execution, and improved from observed failures. % The goal is not to make probabilistic agents behave like deterministic programs, but to give them enough structure that their behavior can be reasoned about, controlled, and repaired.