跳到正文
@omarsar0· @omarsar0 · X·· 2026-08-25AI 评分38
AI 导读

@agentsky_dev 团队发布的 Agent Playground 非常酷。它让你在一个浏览器里给 Claude Code、Codex 和 DeepSeek 分配完全相同的任务,并排比较时间、成本和 token。 我经常跑同任务测试,我认为这是首批能在真正相同条件下比较智能体的平台之一。这很棒,因为你可以更好地判断哪个智能体框架最适合目标任务。

正文

Very cool launch from the @agentsky_dev team. Agent Playground lets you give Claude Code, Codex, and DeepSeek the identical task in one browser and compare time, cost, and tokens side by side.

I run same-task harness tests constantly, and I think this is one of the first places agents can be compared under truly identical conditions. It's great because you can make a better decision about which agent harness is best for the desired task.

来源:@omarsar0 · x.com