Hugging Face Daily Papers·· 4 天前AI 评分34
RegLLM:面向受监管智能体 AI 有限自主性的诊断测试框架
Evaluating Bounded Autonomy in Regulated Agentic AI: A Diagnostic Harness with Constitutional Rewards, Escalation Labels, and Runtime Governance
AI 导读
研究者提出 RegLLM 诊断框架,用于评估受监管智能体工作流中的有限自主性,监测引用有效性、来源依据、schema 合规、升级正确性、宪法对齐和不安全动作率六项可信度信号。
来源:Hugging Face Daily Papers · arxiv.org