ai.hackcv
论文精选 65arXiv

ADMITBench: A Safety-Governed Reference Framework for Evaluating the Admissibility of Industrial LLM Advisories· ADMITBench:工业 LLM 咨询的安全评估框架

This white paper presents ADMITBench, a reference framework for evaluating industrial LLM advisories at the level of the proposed action. The framework implements a versioned, safety-governed evaluation contract that checks whether a recommendation is supported by the available evidence, permitted under the stated authority and procedure, and acceptable under the plant-specific consequence checks encoded in the selected evaluation profile. In this report, \emph{safety-governed} means that eligibility is determined through explicit, non-compensatory checks derived from a versioned plant profile; it does not mean that the evaluator, model, or plant has been safety-certified. Release 0.1.0 is a public reference implementation for technical and research evaluation, not an authorisation for phy

AI 解读论文

提出 ADMITBench,用于评估工业 LLM 咨询的安全性与合规性

核心方法
通过实现一个版本化、安全治理的评估框架,检查咨询建议的证据支持、权限和程序合规性以及特定工厂的后果检查
适合谁读
研究者、工程师
要解决的问题
评估工业领域大型语言模型(LLM)生成的咨询建议是否符合安全性与合规性要求
关键实验
未提供
主要贡献
提供了一个公开的技术参考实现,用于评估工业LLM咨询的安全性和合规性
意义与局限
为工业领域使用 LLM 提供了安全评估标准,有助于提高工业应用的安全性和可靠性。然而,该框架尚未经过正式的安全认证,仅作为技术评估工具。
领域:cs.AI作者:Yash Misra、Javal Vyas、Siddharth Gutta
相关推荐

本站内容由 LLM 精选聚合,原文版权归 arXiv 所有 · 摘录仅供参考