AI 导读
Greg Isenberg 表示不会在播客 startupideaspod 专门讲 Claude Opus 4.8,因为它相比 GPT-5.5 没有值得花一小时的身位优势。
正文
Greg Isenberg 说了句挺多人不爱听的话, Claude Opus 4.8 发布,他不打算在自己的播客 startupideaspod 里专门讲一期, 理由很简单,它没比 GPT-5.5 强出一个值得你花一小时的身位。
他拿 iPhone 打了个比方,早期每代都是大跃进, 现在变成相机好了一点点、边框圆了一点点, benchmark 说进步明显,真上手的人 vibes 却说不太清。
4.6 到 4.7 再到 4.8,模型这条线大概率已经卷到边际收益递减, 真正能把活儿撬动的,基本都是模型外面那层东西, Claude Code 同周上线的 Dynamic Workflows,能让 Claude 自己写编排脚本、并行拉一堆子代理互相验证, Codex 那个带内置浏览器的桌面 App,把写代码和查资料缝进了同一个界面。
说白了,模型现在越来越像发动机, 你上一次打车,问过司机这车装的什么发动机吗, 没有吧,你只关心它能不能准时把你送到公司。
Greg 赌六个月内没人会在乎你用哪个模型, 就跟没人在乎 Uber 用什么引擎一个道理。
也就是说,模型正在变成电,谁家发出来的电都一样亮, 真正决定你能干成什么的,是你家里装了哪些电器。
说白了,聪明是模型的事,能不能帮你交活,是它外面那层壳的事。
I didn't cover Claude Opus 4.8 on my pod because I don't think it's MEANINGFULLY better than GPT 5.5 as of May 29th. We're entering the era where model releases start to feel like iPhone releases. Remember when every new iPhone was a genuine leap? Now it's a slightly better camera and you can't really tell the difference. That's where models are heading. 4.6 to 4.7 to 4.8. Each one is a little different. Nobody can agree if it's better or worse. The benchmarks say one thing, the vibes say another. The thing that actually matters right now is what's happening around the models. Claude Code shipped dynamic workflows this same week and that genuinely changes what one person can build. Codex shipped a desktop app with an in app browser that combines coding and knowledge work in one surface. Those are the releases that move the needle for people. The model underneath is becoming interchangeable. I think we're maybe 6 months from nobody caring which model they're using the way nobody cares which engine is in their Uber. You just want to get where you're going. When something genuinely changes the game for builders, I'll cover it on @startupideaspod. Opus 4.8 wasn't that. Dynamic workflows was. I'd rather save you the hour.在 X 查看被引用的帖子
来源:@AYi_AInotes · x.com