跳到正文
原文
OpenRouter Announcements·· 11 天前AI 评分53

OpenRouter 实测 TypeSafe 决策模型 Jev 与 LLM-as-a-Judge:封闭标准打平,开放标准 LLM 判官胜出

Jev vs LLM-as-a-Judge

AI 导读

OpenRouter 在 2026-09-21 用 HaluEval 88 条标注数据和 SummEval 50 条专家评分摘要对比 typesafe/jev-1.13 与 openai/gpt-5.6-luna 两种评测判官。

来源:OpenRouter Announcements · openrouter.ai