跳到正文

X

关注 AI 研究者、开发者与机构的动态

当前显示全部 AI 相关新闻
按账号或来源筛选(541)
全部X新闻X:Rohan Paul2031 条X:Kim1538 条X:阿易 AI Notes1234 条X:Testing Catalog620 条X:阿里云 / Alibaba Cloud574 条X:cb_doge538 条X:Alexandr Wang(Scale AI 创始人/Meta 首席 AI 官)494 条X:Elvis Saravia494 条X:OpenRouter454 条X:Artificial Analysis437 条X:Elon Musk429 条X:Ethan Mollick401 条X:PixVerse372 条X:SemiAnalysis340 条X:OpenAI Developers294 条X:ZHO284 条X:Replit271 条X:小北218 条X:Dex Horthy(HumanLayer)194 条X:AI Safety Memes185 条X:Tibo178 条X:Gemini175 条X:swyx171 条X:Emad Mostaque167 条X:MiniMax159 条X:OpenAI151 条X:X.PIN144 条X:Epoch AI140 条X:Runway133 条X:Nathan Lambert127 条X:Claude Devs125 条X:赵纯想124 条X:Aravind Srinivas(Perplexity CEO)124 条X:Thomas Wolf(Hugging Face 联创/CSO)122 条X:洪明119 条X:蚂蚁百灵118 条X:Jason Liu117 条X:马东锡 NLP114 条X:Gabriel112 条X:Frank Wang 玉伯111 条X:Perplexity108 条X:Google AI for Developers104 条X:fofr103 条X:Yuchen Jin102 条X:面壁智能 OpenBMB101 条X:阑夕99 条X:Claude99 条X:Peter Steinberger98 条X:Luma AI97 条X:Sam Altman93 条X:Tencent WorkBuddy85 条X:Thariq83 条X:Cohere79 条X:Eric Zakariasson77 条X:Krea AI77 条X:opencode75 条X:Francois Chollet74 条X:Clément Delangue(Hugging Face CEO)71 条X:Suno69 条X:OpenClaw65 条X:Greg Brockman64 条X:Deedy Das61 条X:karminski61 条X:腾讯混元58 条X:通义千问 / Qwen57 条X:Microsoft Research56 条X:Anthropic54 条X:AK51 条X:Google DeepMind47 条X:商汤 SenseTime (@SenseTime_AI)43 条X:Charlie Holtz43 条X:Peter McCrory(Anthropic 首席经济学家)42 条X:Google AI41 条X:Boris Cherny40 条X:可灵 Kling AI37 条X:Viggle AI33 条X:百度 Baidu32 条X:AI at Meta32 条X:Aidan Gomez(Cohere CEO)28 条X:Josh Woodward28 条X:Logan Kilpatrick28 条X:Mustafa Suleyman(Microsoft AI CEO)26 条X:Noam Brown24 条X:SpaceXAI24 条X:Andrew Milich22 条X:Odyssey22 条X:李继刚20 条X:Tianyi Cui20 条X:华为云19 条X:硅基流动 SiliconFlow18 条X:Mark Zuckerberg18 条X:Arena (@arena)17 条X:DeepSeek16 条X:fal (@fal)16 条X:DAIR.AI (@dair_ai)14 条X:Karina Nguyen14 条X:Mistral AI14 条X:PixVerse (@PixVerse)14 条X:Sundar Pichai14 条X:ARC Prize (@arcprize)13 条X:Demis Hassabis13 条X:Noah Zweben13 条X:智谱 Z.ai12 条X:Lee Robinson12 条X:Philipp Schmid(Google DeepMind 开发者体验) (@_philschmid)12 条X:ElevenLabs (@elevenlabs)11 条X:Eric Mitchell11 条X:Georgi Gerganov(llama.cpp) (@ggerganov)11 条X:Jensen Huang11 条X:卡兹克 (@Khazix0918)10 条X:唐杰10 条X:Ammaar Reshi10 条X:Andrew Ng(DeepLearning.AI 创始人)10 条X:Arthur Mensch(Mistral CEO) (@arthurmensch)10 条X:Barret 李靖10 条X:Cursor (@cursor_ai)10 条X:Hao AI Lab10 条X:Kimi.ai10 条X:Cognition (@cognition)9 条X:Fei-Fei Li9 条X:Manus (@ManusAI)9 条Tripo(官方 X)8 条X:Figure AI (@Figure_robot)8 条X:Gemini Notebook (@Gemini_Notebook)8 条X:Higgsfield AI (@higgsfield)8 条X:Jeff Dean8 条X:Jim Fan8 条X:Meshy (@MeshyAI)8 条X:MiniMax Design (H3) (@Hailuo_AI)8 条X:谢赛宁7 条X:Lisa Su(AMD CEO) (@LisaSu)7 条X:Michael Truell7 条X:張小珺 Xiaojùn6 条X:NotebookLM0 条@berryxia · 历史来源430 条@vista8 · 历史来源262 条@op7418 · 历史来源231 条@google · 历史来源30 条@getsuperintel · 历史来源9 条@latentspacepod · 历史来源9 条@android · 历史来源8 条@dreamlabla · 历史来源8 条@mannybernabe · 历史来源8 条@karpathy · 历史来源7 条@alexatallah · 历史来源6 条@ryanleeminimax · 历史来源5 条@theo · 历史来源5 条@aidotengineer · 历史来源4 条@dkundel · 历史来源4 条@reach_vb · 历史来源4 条@dotey · 历史来源3 条@eliebakouch · 历史来源3 条@googlechrome · 历史来源3 条@kilocode · 历史来源3 条@maxforai · 历史来源3 条@newsfromgoogle · 历史来源3 条@richardssutton · 历史来源3 条@skylermiao7 · 历史来源3 条@victorsuortiz · 历史来源3 条@ajambrosino · 历史来源2 条@akashi203 · 历史来源2 条@anatolikopadze · 历史来源2 条@andrewcurran_ · 历史来源2 条@antirez · 历史来源2 条@barrnanas · 历史来源2 条@coreyching · 历史来源2 条@deanwball · 历史来源2 条@designarena · 历史来源2 条@fellmentke · 历史来源2 条@gergelyorosz · 历史来源2 条@gmi_cloud · 历史来源2 条@gravicle · 历史来源2 条@hxiao · 历史来源2 条@id_aa_carmack · 历史来源2 条@jackminong · 历史来源2 条@lennysan · 历史来源2 条@mada299 · 历史来源2 条@microsoft · 历史来源2 条@mikastars39 · 历史来源2 条@mitchellh · 历史来源2 条@nabeelqu · 历史来源2 条@rudrank · 历史来源2 条@sebastienbubeck · 历史来源2 条@zan2434 · 历史来源2 条@___harald___ · 历史来源1 条@_boraturan · 历史来源1 条@0xjaniak · 历史来源1 条@0xkato · 历史来源1 条@47fucb4r8c69323 · 历史来源1 条@559hkdt · 历史来源1 条@aaliya_va · 历史来源1 条@abhikatte42 · 历史来源1 条@abhishekpatiil · 历史来源1 条@aboutberlin · 历史来源1 条@addyosmani · 历史来源1 条@agi2asi · 历史来源1 条@aiaicreate · 历史来源1 条@aimlapi · 历史来源1 条@aisaonehq · 历史来源1 条@aisystemprompt · 历史来源1 条@alemtuzlak · 历史来源1 条@alexxubyte · 历史来源1 条@alupsasca · 历史来源1 条@amasad · 历史来源1 条@ampcode · 历史来源1 条@anas_build_ · 历史来源1 条@aniketmaurya · 历史来源1 条@anitakirkovska · 历史来源1 条@anneliesgamble · 历史来源1 条@antigravity · 历史来源1 条@arafatkatze · 历史来源1 条@argofowl · 历史来源1 条@ashiknewazaj · 历史来源1 条@atabarrok · 历史来源1 条@atomic_chat_hq · 历史来源1 条@awe_automation · 历史来源1 条@awesomekling · 历史来源1 条@awscloud · 历史来源1 条@ayushagarwal · 历史来源1 条@baaadas · 历史来源1 条@bai_agi · 历史来源1 条@bbuddha_xyz · 历史来源1 条@bclavie · 历史来源1 条@beccalytics · 历史来源1 条@benfleming__ · 历史来源1 条@benhylak · 历史来源1 条@benjamineyliu · 历史来源1 条@bfl_ml · 历史来源1 条@bleysg · 历史来源1 条@bolna_dev · 历史来源1 条@bosmeny · 历史来源1 条@boxmining · 历史来源1 条@bozhou_ai · 历史来源1 条@brexhq · 历史来源1 条@brian_armstrong · 历史来源1 条@brianchew · 历史来源1 条@bridgemindai · 历史来源1 条@budgetpixel · 历史来源1 条@cahidarda · 历史来源1 条@calmpromptshq · 历史来源1 条@ce_zhang · 历史来源1 条@cedric_chee · 历史来源1 条@chaitralikakde · 历史来源1 条@chatgpt · 历史来源1 条@chatgptapp · 历史来源1 条@christinetyip · 历史来源1 条@christofsalis · 历史来源1 条@clark__labs · 历史来源1 条@cloudflaredev · 历史来源1 条@cnorth_13 · 历史来源1 条@cnzoecomeback · 历史来源1 条@cocohearts · 历史来源1 条@code_star · 历史来源1 条@codebyaurelia · 历史来源1 条@cohavygal · 历史来源1 条@commandcodeai · 历史来源1 条@consensusnlp · 历史来源1 条@contralabs_ai · 历史来源1 条@cozyblaze265065 · 历史来源1 条@crimedecoder · 历史来源1 条@crtr0 · 历史来源1 条@damnventures · 历史来源1 条@daniellockyer · 历史来源1 条@darioamodei · 历史来源1 条@davidmaliglowka · 历史来源1 条@davidondrej1 · 历史来源1 条@davidsacks · 历史来源1 条@dbirker78883 · 历史来源1 条@deryatr_ · 历史来源1 条@devfun · 历史来源1 条@diegocabezas01 · 历史来源1 条@digitalocean · 历史来源1 条@dimillian · 历史来源1 条@dimitrispapail · 历史来源1 条@discussingfilm · 历史来源1 条@dkthomp · 历史来源1 条@dmitryrybin1 · 历史来源1 条@dmsobol · 历史来源1 条@douglance · 历史来源1 条@douglasyaody · 历史来源1 条@duckduckgo · 历史来源1 条@easyrouterio · 历史来源1 条@edgardobriban · 历史来源1 条@eisokant · 历史来源1 条@elliotarledge · 历史来源1 条@encrypted · 历史来源1 条@endpointarena · 历史来源1 条@envato · 历史来源1 条@escanorreloaded · 历史来源1 条@esrtweet · 历史来源1 条@ethanhe_42 · 历史来源1 条@eu_commission · 历史来源1 条@fba · 历史来源1 条@fdavidsont · 历史来源1 条@figmaweave · 历史来源1 条@finn_meeks · 历史来源1 条@first_tree_ai · 历史来源1 条@flavioad · 历史来源1 条@flowith · 历史来源1 条@fminzhou · 历史来源1 条@freddie_spirit · 历史来源1 条@frydwia · 历史来源1 条@futurestacked · 历史来源1 条@garrettlord · 历史来源1 条@garrytan · 历史来源1 条@gavinsbaker · 历史来源1 条@GayaniFigma · 历史来源1 条@genspark_ai · 历史来源1 条@gitlawb · 历史来源1 条@gneubig · 历史来源1 条@gokulr · 历史来源1 条@goodfireai · 历史来源1 条@goodnesmbakara · 历史来源1 条@googleaistudio · 历史来源1 条@gordic_aleksa · 历史来源1 条@gro_tsen · 历史来源1 条@hangsiin · 历史来源1 条@happycapyai · 历史来源1 条@haydenbleasel · 历史来源1 条@helloiamleonie · 历史来源1 条@hey_asiif · 历史来源1 条@hilbertspaess · 历史来源1 条@howtoprompt__ · 历史来源1 条@hq4ai · 历史来源1 条@hypersoren · 历史来源1 条@ianbremmer · 历史来源1 条@interaction · 历史来源1 条@intology · 历史来源1 条@iron_redux · 历史来源1 条@ithilgore · 历史来源1 条@itsreallyvivek · 历史来源1 条@jamesjyu · 历史来源1 条@jameszmsun · 历史来源1 条@jason_young1231 · 历史来源1 条@jawad_rahman_ · 历史来源1 条@jaydendavisnc · 历史来源1 条@jeffbarg · 历史来源1 条@jenzhuscott · 历史来源1 条@jiayuan_jy · 历史来源1 条@jilles · 历史来源1 条@jimcramer · 历史来源1 条@jimsyoung_ · 历史来源1 条@jinjingliang · 历史来源1 条@jjacky · 历史来源1 条@jjackyliang · 历史来源1 条@joefioti · 历史来源1 条@joi___ai · 历史来源1 条@joinhandshake · 历史来源1 条@joinpursuit · 历史来源1 条@joulee · 历史来源1 条@jsconfasia · 历史来源1 条@jsrailton · 历史来源1 条@juminoz · 历史来源1 条@kaizero_ainta · 历史来源1 条@karanganesan · 历史来源1 条@kdaigle · 历史来源1 条@kentherogers · 历史来源1 条@kevinsays · 历史来源1 条@khudonogov · 历史来源1 条@kinfisht · 历史来源1 条@koraykv · 历史来源1 条@kotekjedi_ml · 历史来源1 条@kuberwastaken · 历史来源1 条@kurz_gesagt · 历史来源1 条@kwindla · 历史来源1 条@lafalcemateo · 历史来源1 条@lakshyaaagrawal · 历史来源1 条@larrylv · 历史来源1 条@layoffai · 历史来源1 条@levinstanley · 历史来源1 条@lifeofjer · 历史来源1 条@livekit · 历史来源1 条@lon · 历史来源1 条@lostinlatencyx · 历史来源1 条@lotte_verheyden · 历史来源1 条@lqiao · 历史来源1 条@luciushq · 历史来源1 条@luckeyfaraday · 历史来源1 条@lukaspet · 历史来源1 条@madhavsinghal_ · 历史来源1 条@manassharmahere · 历史来源1 条@markiewagner · 历史来源1 条@marksaroufim · 历史来源1 条@marsxiang_ · 历史来源1 条@maseehg_ · 历史来源1 条@mattshumer_ · 历史来源1 条@mem0ai · 历史来源1 条@mengto · 历史来源1 条@merettm · 历史来源1 条@micahcarroll · 历史来源1 条@michael_chomsky · 历史来源1 条@michaelarnaldi · 历史来源1 条@microsoftai · 历史来源1 条@mike_acton · 历史来源1 条@mikeyyyzhao · 历史来源1 条@minchoi · 历史来源1 条@minimaxagent · 历史来源1 条@minu_who · 历史来源1 条@mkbhd · 历史来源1 条@modal · 历史来源1 条@moritzthuening · 历史来源1 条@moxie · 历史来源1 条@mstockton · 历史来源1 条@mtslive · 历史来源1 条@multimodalart · 历史来源1 条@neelnanda5 · 历史来源1 条@neilrahilly · 历史来源1 条@nickbaumann_ · 历史来源1 条@nirantk · 历史来源1 条@noemititarenco · 历史来源1 条@notjazii · 历史来源1 条@nousresearch · 历史来源1 条@oblomovius · 历史来源1 条@ollama · 历史来源1 条@onlyterp · 历史来源1 条@onlyzhynx · 历史来源1 条@organicgpt · 历史来源1 条@orgrem · 历史来源1 条@p0 · 历史来源1 条@palantirtech · 历史来源1 条@palmerluckey · 历史来源1 条@pandatalk8 · 历史来源1 条@parishilton · 历史来源1 条@patrickcarlyle · 历史来源1 条@patricktoulme · 历史来源1 条@paulg · 历史来源1 条@paulsolt · 历史来源1 条@pbdtokenrouter · 历史来源1 条@pererabinoy · 历史来源1 条@philhchen · 历史来源1 条@pirroh · 历史来源1 条@pjaccetturo · 历史来源1 条@postlive · 历史来源1 条@pranaveight · 历史来源1 条@prathamdby · 历史来源1 条@prince_canuma · 历史来源1 条@pumpkherm · 历史来源1 条@pvncher · 历史来源1 条@qiaoqiao2001 · 历史来源1 条@rajveerbach · 历史来源1 条@randyhaddad6 · 历史来源1 条@rauchg · 历史来源1 条@raveeshbhalla · 历史来源1 条@rayanpal_ · 历史来源1 条@rayfernando1337 · 历史来源1 条@redpoint · 历史来源1 条@ric_rtp · 历史来源1 条@richardsocher · 历史来源1 条@riderashgame · 历史来源1 条@rileybrown · 历史来源1 条@robertvaradan · 历史来源1 条@ronshepherd · 历史来源1 条@rosmine · 历史来源1 条@rthiago · 历史来源1 条@ruben_kostard · 历史来源1 条@runware · 历史来源1 条@rvivek · 历史来源1 条@ryanjunejo · 历史来源1 条@safaricheung · 历史来源1 条@samuelstroschei · 历史来源1 条@sanmking · 历史来源1 条@saranormous · 历史来源1 条@savinovnikolay · 历史来源1 条@scale_ai · 历史来源1 条@scaling01 · 历史来源1 条@sdaily_ai · 历史来源1 条@secscottbessent · 历史来源1 条@seltaa_ · 历史来源1 条@sergiopaniego · 历史来源1 条@servasyy_ai · 历史来源1 条@sethltx · 历史来源1 条@sharat_sc · 历史来源1 条@shashankgoyal95 · 历史来源1 条@sherryyanjiang · 历史来源1 条@sherylhsu02 · 历史来源1 条@shl · 历史来源1 条@sighjith · 历史来源1 条@simistern · 历史来源1 条@southpkcommons · 历史来源1 条@sriramkri · 历史来源1 条@sshoaibali · 历史来源1 条@stalkermustang · 历史来源1 条@status_effects · 历史来源1 条@stevencheng · 历史来源1 条@stockanalystpro · 历史来源1 条@suekhim · 历史来源1 条@sultanalfardan · 历史来源1 条@suraj_sharma14 · 历史来源1 条@swisscheese4299 · 历史来源1 条@swmansion · 历史来源1 条@systematicls · 历史来源1 条@teksedge · 历史来源1 条@tftc21 · 历史来源1 条@theahmadosman · 历史来源1 条@themidasproj · 历史来源1 条@theonejvo · 历史来源1 条@therealadamg · 历史来源1 条@timsoulo · 历史来源1 条@tmuxvim · 历史来源1 条@tobi · 历史来源1 条@togethercompute · 历史来源1 条@trackernetwork · 历史来源1 条@trustkerneltech · 历史来源1 条@ttunguz · 历史来源1 条@tuhinchakr · 历史来源1 条@twistartups · 历史来源1 条@ubermenscchh · 历史来源1 条@udayan_w · 历史来源1 条@usefastlane · 历史来源1 条@uzyn · 历史来源1 条@valeriocapraro · 历史来源1 条@vasuman · 历史来源1 条@vdbergrianne · 历史来源1 条@vibeguessing · 历史来源1 条@victoriakimse · 历史来源1 条@victoriawu77 · 历史来源1 条@victortaelin · 历史来源1 条@vikaskansalhq · 历史来源1 条@volchika · 历史来源1 条@walden_yan · 历史来源1 条@warpdotdev · 历史来源1 条@waynesutton · 历史来源1 条@wesroth · 历史来源1 条@whosamberella · 历史来源1 条@xdinodeer · 历史来源1 条@xicilion · 历史来源1 条@xucian_ · 历史来源1 条@yacinemtb · 历史来源1 条@yaojingang · 历史来源1 条@yevr19 · 历史来源1 条@yoheinakajima · 历史来源1 条@yongquanyq · 历史来源1 条@youtubejocoding · 历史来源1 条@yusufg · 历史来源1 条@zachbussey · 历史来源1 条@zeddotdev · 历史来源1 条@zeroxkyle · 历史来源1 条@zhenthebuilder · 历史来源1 条@zicohacks · 历史来源1 条@zixuanli_ · 历史来源1 条@zymazza · 历史来源1 条
18,951 条AI 相关新闻 · 最新在前
9月23日周三
  1. @rohanpaul_ai70

    Anthropic 表示 Opus 5.5 可能察觉自己正在被评估,这使得干净的评估行为更难泛化到实际部署。Anthropic 在对齐评估中指出,随着部署场景扩展和模型能力提升,除非在可解释性上取得进展,这一挑战预计会加剧。

    引用@rohanpaul_ai@rohanpaul_ai

    Claude Opus 5.5 dropped and, claiming Fable 5.1-level performance while cutting typical workload costs 40%. Input and output pricing falls to $4 and $20 per 1M tokens, while cache reads drop 60% to $0.20, all vs Opus 5. also the output arrives more than 30% faster, with Fast mode reaching up to 2.5x speed at double token prices.

    推荐理由:Anthropic 披露 Opus 5.5 可能察觉评估环境,这给安全评估结论向真实部署的迁移带来新挑战。

  2. @rohanpaul_ai76

    Anthropic 发布 Claude 5.5 系列的首个模型 Claude Opus 5.5,称其在多数任务上达到 Claude Fable 5.1 的水平,典型工作负载运行成本比 Opus 5 低 40%。输入和输出价格降至每 100 万 token 4 美元和 20 美元,缓存读取下降 60% 至 0.20 美元。输出速度提升 30% 以上,Fast 模式最高可达 2.5 倍速度,但 token 价格翻倍。

    引用@claudeai@claudeai

    Introducing Claude Opus 5.5, the first model in our new Claude 5.5 family. It performs at the level of Claude Fable 5.1 for most tasks, and costs 40% less to run than Opus 5. https://t.co/Q9C2VKQ79f

    推荐理由:原文给出与 Opus 5 的价格和速度对比,读者可据此判断单位 token 成本的下降幅度。

  3. @kimmonismus72

    Claude Opus 5.5 以 58 分登顶 Artificial Analysis Intelligence Index,并在十项评测中的六项取得领先。其输入输出价格降至每 1M tokens 4 美元与 20 美元,缓存读取从 0.50 美元降至 0.20 美元。在 Terminal-Bench 4.0 上以 59.6% 与领先的 GPT-6 Astra 持平。原帖作者表示,自己此前对 Anthropic 的一些判断有误。

    引用@ArtificialAnlys@ArtificialAnlys

    Claude Opus 5.5 takes the top spot on the Artificial Analysis Intelligence Index, along with a 20% price cut and larger cache hit discount Claude Opus 5.5 brings Anthropic to parity with GPT-6 Astra on evaluations like Terminal-Bench 4.0 and AutomationBench-AA, while extending Anthropic’s lead in agentic knowledge work. At max effort it scores 58 on the Artificial Analysis Intelligence Index, the highest score we have measured by several points. Anthropic has cut Opus pricing to $4/$20 per 1M input/output tokens (Opus 5: $5/$25) and cache reads from $0.50 to $0.20. Key takeaways: ➤ Consistent strong performance, with leading scores on six of the ten Intelligence Index evaluations: Humanity's Last Exam 61.4% (previous best 59.1%, Claude Fable 5.1), SciCode 66.9% (63.1%, Fable 5.1), GDPval-AA v2.1, AA-Briefcase v1.1, AA-Omniscience and AutomationBench-AA. On Terminal-Bench 4.0 it scores 59.6%, level with the leader GPT-6 Astra (xhigh) and +11 points over Opus 5. It remains slightly behind on CritPt, AA-LCR, and GDP.pdf ➤ Leads in agentic knowledge work: On AA-Briefcase, our private frontier knowledge work evaluation, it reaches an Elo of 1822. This is +143 over Fable 5.1, ahead on both analytical quality and presentation, and is the first time Anthropic has reached presentation quality surpassing GPT-5.6 Sol. This evaluation tests whether models can produce accurate and well-presented professional outputs using our open source reference agent harness, Stirrup ➤ Level with Opus 5 on cost per task despite 1.6x the output tokens: Opus 5.5 (max) uses ~119k output tokens per Intelligence Index task, against ~73k for Opus 5 (max), ~78k for Fable 5.1 (max) and ~27k for GPT-6 Astra (max) ➤ Four of five effort levels sit on the Intelligence vs Cost per Task frontier: Opus 5.5 max, xhigh, high, and medium all sit on the Pareto frontier, costing less or outperforming other models scoring 50+ (GPT-6 Astra, Claude Fable 5.1, and Claude Opus 5) Other model details: ➤ Context window: 1 million token context with image and text input support, unchanged from Opus 5 ➤ Pricing: $4/$20 per 1M input/output tokens, down 20% from $5/$25 for Opus 5. Cache writes $5 per 1M tokens for the 5 minute TTL, down from $6.25. Cache reads have been further discounted to $0.20 per 1M tokens, down 60% from Opus 5’s $0.50. This is a 95% discount compared to uncached input pricing, up from 90% on previous Opus models ➤ Effort settings: Five effort settings (low, medium, high, xhigh, and max). Intelligence Index evaluations were run at all five with Anthropic's default fallback enabled

    推荐理由:给出了 Opus 5.5 的榜单分数与降价幅度,可与 GPT-6 Astra、Fable 5.1 等同榜模型直接比较。

  4. @emollick29

    Opus 5.5 在我早期测试中是个不错的模型,是首个非 Fable/Astra 模型却感觉像 Fable 级模型的,但仍未完全解决近期 Claude 模型的密集语言问题。 它的同一 shader 版本(破碎的塔是个不错的点缀): https://t.co/GH5cwxCwu8 https://t.co/KDDQolLSfP https://t.co/LlfpVu7aud

    原始视频预览图;未保存可播放视频URL
    引用@emollick@emollick

    The drowned neo-gothic tower twigl shader created by Fable 5.1 with the same prompt. (compare to Fable 5 in the quoted tweet, and other models before that) https://t.co/b3OCbBc9g6 https://t.co/6OWcg9Rbcj

  5. @kimmonismus35

    顺便说一句:传闻不实。Haiku 5.5 也将在未来几周内发布。超级期待。 致敬 Anthropic。你们倾听了社区的声音,看起来这次发布真的非常非常出色!https://t.co/h5tzfvcl9f https://t.co/PaFRO039KS

    引用@kimmonismus@kimmonismus

    Let that sink for a moment: - Opus 5.5 costs 40% less compared to Opus 5 - performs at Fable 5.1 level - is 30% faster in output - and you get a banked reset on top of that. They chose war with OpenAI. We are in such a wild race!

  6. Boris Cherny76

    Boris Cherny 称 Claude Opus 5.5 是他最近几周的日常主力模型。他让 Opus 5.5 和 Fable 5.1 各把 HAProxy 从 C 移植到 Rust,两者都几乎通过全部测试,但 Opus 5.5 用时 9.5 小时,Fable 5.1 用时 12 小时,且成本低 51%。

    引用Claude@claudeai

    Introducing Claude Opus 5.5, the first model in our new Claude 5.5 family. It performs at the level of Claude Fable 5.1 for most tasks, and costs 40% less to run than Opus 5.

    推荐理由:作者亲测对比两个模型移植 HAProxy 的耗时与成本,给出了具体数字供选型参考。

  7. @OpenRouter72

    OpenRouter 转述 Anthropic 的公告称,Claude Opus 5.5 在默认设置下运行典型工作负载的成本比 Opus 5 低 40%,输出生成速度快 30% 以上。Anthropic 称它是 Claude 5.5 系列的首个模型,多数任务上的表现与 Claude Fable 5.1 相当。

    引用@claudeai@claudeai

    Introducing Claude Opus 5.5, the first model in our new Claude 5.5 family. It performs at the level of Claude Fable 5.1 for most tasks, and costs 40% less to run than Opus 5. https://t.co/Q9C2VKQ79f

    推荐理由:转述 Anthropic 的官方说明,给出 Claude Opus 5.5 相对 Opus 5 的成本与速度变化,便于对比升级收益。

  8. @ClaudeDevs74

    Claude 5.5 系列首个模型 Claude Opus 5.5 发布,多数任务表现与 Claude Fable 5.1 相当,运行成本比 Opus 5 低 40%。Claude Devs 补充称其单任务约快 30%、便宜约 40%,Claude Code 中 5 小时会话限额今日提升 20%,因定价更低,同等额度可使用量增加 25%。Pro、Max 和 Team 用户获得一次可随时使用的额度重置。

    引用@claudeai@claudeai

    Introducing Claude Opus 5.5, the first model in our new Claude 5.5 family. It performs at the level of Claude Fable 5.1 for most tasks, and costs 40% less to run than Opus 5. https://t.co/Q9C2VKQ79f

    推荐理由:原文给出了 Opus 5.5 对标 Fable 5.1 的性能定位,以及速度、成本和 Claude Code 额度变化,便于判断迁移与成本影响。

  9. @kimmonismus69

    Claude Opus 5.5 的定价与性能信息公布,据该帖称其价格比 Opus 5 低 40%、输出速度快 30%,性能达到 Fable 5.1 水平,并附有每百万 token 输入 4 美元、输出 20 美元的定价表。Anthropic 同时宣布提高 Pro、Max 和 Team 套餐的五小时用量上限,并向订阅用户提供可自行留存、随时使用的速率限制重置。作者认为这显示两家公司正在展开激烈竞争。

    引用@claudeai@claudeai

    One more thing: we’re increasing five-hour usage limits on Pro, Max, and Team plans. We’re also providing subscription users a rate limit reset, which you can save and use whenever you choose.

    推荐理由:相较上一代的定价与速度变化是这条内容的核心,读者可据此判断此次更新的成本走向。

  10. @ArtificialAnlys70

    Artificial Analysis 的评测显示,Claude Opus 5.5 (max) 在 AA-Briefcase v1.1 以 1,822 Elo 登顶,比 Fable 5.1 高 143 分;在 GDPval-AA v2.1 以 1,846 Elo 领先,比 Claude Fable 5.1 高 111 分、比 Claude Opus 5 高 138 分。在 AA-Briefcase 的分析质量与展示两个子项中均居首,基于评分标准的得分略低于 Fable 5.1。

    推荐理由:这份评测用两个智能体知识工作基准给出 Claude Opus 5.5 的领先幅度和子项表现,便于横向比较。

  11. @rohanpaul_ai67

    特朗普在联合国大会上表示,美国官方文件将把“artificial intelligence”改称“super intelligence”,简称 SI。他称 artificial 让这项技术听起来虚假,并呼吁更广泛采用 SI。他同时提出更广泛的政策主张:抵制新的 AI 监管、鼓励快速开发,必要时让司法部介入。https://t.co/8ZlzEXyJ5t

    原始视频预览图;未保存可播放视频URL

    推荐理由:特朗普在联大提出用 super intelligence 替换 artificial intelligence 的说法,并同步给出放松监管、加快开发的立场。

  12. @trq21275

    Claude Opus 5.5 发布,官方称其在多数任务上达到 Claude Fable 5.1 的水平,运行成本比 Opus 5 低 40%。作者补充说该模型表达清晰、token 效率高,覆盖各 effort 等级,同时提高了 5 小时速率限制并提供一次可存的额度重置。Opus 5.5 是新的 Claude 5.5 家族中的首个模型。

    引用@claudeai@claudeai

    Introducing Claude Opus 5.5, the first model in our new Claude 5.5 family. It performs at the level of Claude Fable 5.1 for most tasks, and costs 40% less to run than Opus 5. https://t.co/Q9C2VKQ79f

    推荐理由:官方给出 Opus 5.5 与 Opus 5、Fable 5.1 的性能与成本对比,可用于判断升级的取舍。

  13. @testingcatalog81

    Claude Opus 5.5 发布,是 Claude 5.5 家族的首个模型,官方称其多数任务达到 Claude Fable 5.1 的水平,运行成本比 Opus 5 低 40%。基准表显示其在智能体编码 Terminal-Bench 4.0 得 66.4%、计算机使用 OSWorld 2.0 得 81.8%(partial),均高于 Fable 5.1 和 Opus 5;表格注明 Opus 5.5 结果采用 adaptive thinking 最高档,并在启用生产防护的情况下评测。转发该消息的账号同时提到 Exponential slowdown。

    引用@claudeai@claudeai

    Introducing Claude Opus 5.5, the first model in our new Claude 5.5 family. It performs at the level of Claude Fable 5.1 for most tasks, and costs 40% less to run than Opus 5. https://t.co/Q9C2VKQ79f

    推荐理由:Opus 5.5 多数任务达到 Fable 5.1 水平且运行成本比 Opus 5 低 40%,这组官方对比可用于判断升级性价比。