阿里发布旗舰模型 Qwen3.7-Max,面向智能体场景,官方称其在一次内核优化任务中自主运行 35 小时、发起 1,158 次工具调用,并在单个注意力内核上取得 10 倍加速,模型已上线阿里云 Model Studio,也可在 Qwen Studio 试用。
作者把 35 小时自主优化的传播印象与实际范围区分开,并单独讨论智能体能力泛化这一论断。
Alibaba released Qwen 3.7 max. Benchmarks incredible.
Their new model ran autonomously for 35 hours, made 1,158 tool calls, and achieved a 10x speedup - on a single attention kernel.
This isn't "AI improving itself across the board." It's a model grinding through compile-profile-rewrite loops on one well-defined optimization target.
Impressive? Absolutely. The kind of self-improvement people will imagine when they see the headline? Not yet.
The actually interesting claim is buried deeper: Qwen says agentic capabilities generalize from diverse training environments the same way language capabilities generalize from diverse text. If that holds, it's a bigger deal than any benchmark number.
📣Meet Qwen3.7-Max — our latest flagship, made for the Agent Era. A versatile foundation for agents that actually get things done: 🧑💻 Coding agent, end to end. Frontend prototypes, multi-file refactors, real debugging — nails it. 🗂️ A reliable office and productivity assistant. Get your work done through MCP integrations and multi-agent orchestration. ⏱️ Long-horizon autonomy. 35 hours straight on a kernel optimization task — 1,000+ tool calls, zero hand-holding. 🔌 Scaffold-agnostic. Claude Code, OpenClaw, Qwen Code, or your own stack. Consistent reliability everywhere. API's up on Alibaba Model Studio. You can also take it for a spin on Qwen Studio. Go build something wild!🏃🏃♂️ 📖 Blog: qwen.ai/blog?id=qwen3.7 ✅ Qwen Studio: chat.qwen.ai/?models=qwen3.7… ⚡️ API:modelstudio.console.alibabac…在 X 查看被引用的帖子
来源:@kimmonismus · x.com