OpenRouter Announcements·· 11 天前AI 评分53
OpenRouter 实测 TypeSafe 决策模型 Jev 与 LLM-as-a-Judge:封闭标准打平,开放标准 LLM 判官胜出
Jev vs LLM-as-a-Judge
AI 导读
OpenRouter 在 2026-09-21 用 HaluEval 88 条标注数据和 SummEval 50 条专家评分摘要对比 typesafe/jev-1.13 与 openai/gpt-5.6-luna 两种评测判官。
来源:OpenRouter Announcements · openrouter.ai