跳到正文
原文
Hugging Face Daily Papers·· 4 天前AI 评分34

RegLLM:面向受监管智能体 AI 有限自主性的诊断测试框架

Evaluating Bounded Autonomy in Regulated Agentic AI: A Diagnostic Harness with Constitutional Rewards, Escalation Labels, and Runtime Governance

AI 导读

研究者提出 RegLLM 诊断框架,用于评估受监管智能体工作流中的有限自主性,监测引用有效性、来源依据、schema 合规、升级正确性、宪法对齐和不安全动作率六项可信度信号。

来源:Hugging Face Daily Papers · arxiv.org