ADMITBench: A Safety-Governed Reference Framework for Evaluating the Admissibility of Industrial LLM Advisories· ADMITBench:工业 LLM 咨询的安全评估框架
This white paper presents ADMITBench, a reference framework for evaluating industrial LLM advisories at the level of the proposed action. The framework implements a versioned, safety-governed evaluation contract that checks whether a recommendation is supported by the available evidence, permitted under the stated authority and procedure, and acceptable under the plant-specific consequence checks encoded in the selected evaluation profile. In this report, \emph{safety-governed} means that eligibility is determined through explicit, non-compensatory checks derived from a versioned plant profile; it does not mean that the evaluator, model, or plant has been safety-certified. Release 0.1.0 is a public reference implementation for technical and research evaluation, not an authorisation for phy
提出 ADMITBench,用于评估工业 LLM 咨询的安全性与合规性
- 核心方法
- 通过实现一个版本化、安全治理的评估框架,检查咨询建议的证据支持、权限和程序合规性以及特定工厂的后果检查
- 适合谁读
- 研究者、工程师
- 要解决的问题
- 评估工业领域大型语言模型(LLM)生成的咨询建议是否符合安全性与合规性要求
- 关键实验
- 未提供
- 主要贡献
- 提供了一个公开的技术参考实现,用于评估工业LLM咨询的安全性和合规性
- 意义与局限
- 为工业领域使用 LLM 提供了安全评估标准,有助于提高工业应用的安全性和可靠性。然而,该框架尚未经过正式的安全认证,仅作为技术评估工具。