推荐理由:这份评测用两个智能体知识工作基准给出 Claude Opus 5.5 的领先幅度和子项表现,便于横向比较。
X
关注 AI 研究者、开发者与机构的动态
按账号或来源筛选(541)
@ArtificialAnlys@ArtificialAnlys精选AI 评分7070 
@ArtificialAnlys@ArtificialAnlysAI 评分5454 
@ArtificialAnlys@ArtificialAnlys精选AI 评分8282 
推荐理由:实测数据呈现 Claude Opus 5.5 在智能指数登顶并降价 20%,可据此比较头部模型的智能水平与单任务成本。
@rohanpaul_ai@rohanpaul_ai精选AI 评分6767 
推荐理由:特朗普在联大提出用 super intelligence 替换 artificial intelligence 的说法,并同步给出放松监管、加快开发的立场。
@trq212@trq212精选AI 评分7575 引用@claudeai@claudeaiIntroducing Claude Opus 5.5, the first model in our new Claude 5.5 family. It performs at the level of Claude Fable 5.1 for most tasks, and costs 40% less to run than Opus 5. https://t.co/Q9C2VKQ79f
推荐理由:官方给出 Opus 5.5 与 Opus 5、Fable 5.1 的性能与成本对比,可用于判断升级的取舍。
@testingcatalog@testingcatalog精选AI 评分8181 
引用@claudeai@claudeaiIntroducing Claude Opus 5.5, the first model in our new Claude 5.5 family. It performs at the level of Claude Fable 5.1 for most tasks, and costs 40% less to run than Opus 5. https://t.co/Q9C2VKQ79f
推荐理由:Opus 5.5 多数任务达到 Fable 5.1 水平且运行成本比 Opus 5 低 40%,这组官方对比可用于判断升级性价比。
@kimmonismus@kimmonismusAI 评分33
Anthropic@AnthropicAI精选AI 评分7171引用Claude@claudeaiIntroducing Claude Opus 5.5, the first model in our new Claude 5.5 family. It performs at the level of Claude Fable 5.1 for most tasks, and costs 40% less to run than Opus 5.
推荐理由:官方宣布 Claude Opus 5.5 上线,可对照引用内容了解其性能对标与成本变化。
@kimmonismus@kimmonismusAI 评分4646 基准测试已上线:https://t.co/GqUC5Ula4U
引用@kimmonismus@kimmonismusAnd here are the benchmarks! Opus 5.5 excels Fable 5.1 in literally every single eval and even GPT-6 Astra in almost every single one. Holy moly, this is going to be interesting! h/t @synthwavedd
@claudeai@claudeaiAI 评分6060 Claude 官方宣布上调 Pro、Max 和 Team 套餐的五小时用量上限。订阅用户还将获得一次速率限制重置,可以自行保存并在需要时启用。
@claudeai@claudeai精选AI 评分6565 Anthropic 的 Claude 官方账号表示,这些改进让 Opus 5.5 成为明显更好的协作伙伴,并附上详情链接。推文未列出具体能力细节或版本参数。
@claudeai@claudeaiAI 评分5555 
@claudeai@claudeai精选AI 评分7171 
推荐理由:官方给出 Opus 5.5 与 Opus 5 的逐项 token 定价对比,可据此估算典型工作负载的成本变化。
@claudeai@claudeaiAI 评分6262
Claude@claudeai精选AI 评分7171
推荐理由:官方宣布 Claude 5.5 家族首模型,给出与上代性能对标和 40% 成本下降,可据此评估迁移价值。
@kimmonismus@kimmonismusAI 评分3535 基准测试来了!Opus 5.5 在几乎每一项评测中都超越了 Fable 5.1,甚至几乎全面胜过 GPT-6 Astra。 天呐,这下有意思了!h/t @synthwavedd
引用@kimmonismus@kimmonismusOpus 5.5 already in Claude Code! Release imminent! https://t.co/tXYHbp9AEQ
Andrew Milich@milichabAI 评分4747在 Tesla 里对 @bot 说话!点咖啡、管理日程,甚至预订行程
引用Tesla@Tesla.@Grok in your Tesla can now do meaningful work for you With Connectors, you can manage your inbox, clean up your calendar, or talk through existing files/chat/tasks – all hands-free
@kimmonismus@kimmonismusAI 评分2626 Opus 5.5 已经出现在 Claude Code 里了!发布在即!https://t.co/tXYHbp9AEQ

@rohanpaul_ai@rohanpaul_aiAI 评分4848 
@alexandr_wang@alexandr_wangAI 评分2222 很多人都在说 muse 连接器平台开放营业了,各位! https://t.co/HG9110wV4I https://t.co/3y4gKNR9DJ
@alexandr_wang@alexandr_wangAI 评分77 @OpenRouter@OpenRouterAI 评分55 @OpenRouter@OpenRouterAI 评分3939 @PixVerse_@PixVerse_AI 评分1010 @rohanpaul_ai@rohanpaul_aiAI 评分5959 阿里 CEO 吴泳铭公布 AI 路线图,称智能本身正成为可扩展的公共资源,机器生成思考目前不足人类认知总量的 3%,但阿里看到通往 1000 倍人类容量的路径。

@elonmusk@elonmuskAI 评分55 @omarsar0@omarsar0AI 评分5959 
@natolambert@natolambertAI 评分3939 引用@natolambert@natolambertNew podcast with @datagenproc of @EpochAIResearch digging into the open questions determining the future of frontier AI! We cover: 00:00 Predictions for RSI 18:15 The role of robotics in an AI acceleration 24:20 How far behind are Chinese models? 27:39 Does distillation explain the gap? 40:58 What Chinese job postings reveal about their labs 48:13 Are open or closed models safer? 58:10 How Epoch AI ticks 1:00:55 What a frontier post-training recipe looks like He's one of the people who gives the best feedback on my writing, so I was stoked to have him on.
@cb_doge@cb_dogeAI 评分2828 Grok Theft Auto,由 Grok 4.7 制作 一行提示词。几分钟内构建完成。🔥 https://t.co/8gz9Greu6d

@elonmusk@elonmuskAI 评分22 @OpenRouter@OpenRouterAI 评分3434 @OpenRouter@OpenRouterAI 评分2727 @OpenRouter@OpenRouterAI 评分2828 @OpenRouter@OpenRouterAI 评分2222 @OpenRouter@OpenRouterAI 评分3030 @OpenRouter@OpenRouterAI 评分5252 OpenRouter 的 token 定价五折优惠同时适用于输入和输出 token,目前覆盖 70+ 模型。网页搜索和其他工具调用仍按标准费率计费,部分供应商折扣更少,具体折扣以各模型页面为准。
@OpenRouter@OpenRouterAI 评分2020 
@OpenRouter@OpenRouterAI 评分3838 
@OpenRouter@OpenRouterAI 评分6262 
karminski-牙医@karminski3AI 评分5252karminski 发布小米 MiMo-v2.6-pro 测试速报,向量数据库测试得分从 MiMo-v2.5-Pro 的 2505 升至 7810,接近 Claude Fable-5。



