AI 导读
Claude Code 推出插件评测功能,运行 claude update 即可体验。官方提示评测会调用模型,因此消耗 token 且结果存在波动,建议先用 --runs 1 试跑再完整运行。插件中的 hooks 和 MCP 服务器会以用户身份运行,所以只应评测自己信任的插件。
正文
Evals call the model, so they use tokens and results vary. Pilot with `--runs 1` before a full run. Your plugin's hooks and MCP servers run as you, so only evaluate plugins you trust.
Run claude update to try it. Docs: https://t.co/K3fKEzO7bI
来源:@ClaudeDevs · x.com