ADMITBench:用于评估工业LLM咨询建议可采纳性的安全管控参考框架
ADMITBench: A Safety-Governed Reference Framework for Evaluating the Admissibility of Industrial LLM Advisories
AI总结:
本研究提出ADMITBench框架,用于评估工业LLM咨询建议的可采纳性,该框架采用带版本的安全管控评估契约,0.1.0版本为公开参考实现,面向技术与研究评估。
AI中文摘要:
本白皮书介绍ADMITBench,一个用于在拟议行动层面评估工业LLM咨询建议的参考框架。该框架实施了一个带版本的安全管控评估契约,用于检查建议是否有可用证据支持、是否符合规定权限与流程、是否符合所选评估配置文件中编码的工厂特定后果检查。本报告中,“安全管控”指资格由源自带版本工厂配置文件的明确非补偿性检查确定,并非指评估者、模型或工厂已获安全认证。0.1.0版本是面向技术与研究评估的公开参考实现,不构成物理执行授权。
英文摘要:
This white paper presents ADMITBench, a reference framework for evaluating industrial LLM advisories at the level of the proposed action. The framework implements a versioned, safety-governed evaluation contract that checks whether a recommendation is supported by the available evidence, permitted under the stated authority and procedure, and acceptable under the plant-specific consequence checks encoded in the selected evaluation profile. In this report, \emph{safety-governed} means that eligibility is determined through explicit, non-compensatory checks derived from a versioned plant profile; it does not mean that the evaluator, model, or plant has been safety-certified. Release 0.1.0 is a public reference implementation for technical and research evaluation, not an authorisation for physical execution.