X
关注 AI 研究者、开发者与机构的动态
按账号或来源筛选(539)
@rohanpaul_ai@rohanpaul_aiAI 评分44 @rohanpaul_ai@rohanpaul_aiAI 评分3838 
@ArtificialAnlys@ArtificialAnlysAI 评分99 @ArtificialAnlys@ArtificialAnlys精选AI 评分7878 
推荐理由:图表将智能指数与单任务成本放在一起对比,读者可从中看到 GPT-6 Sol 和 Luna 以接近前代的智能指数把成本降到了约一半。
@gabriel1@gabriel1AI 评分2424 @rohanpaul_ai@rohanpaul_aiAI 评分1717 @rohanpaul_ai@rohanpaul_aiAI 评分5151 
@testingcatalog@testingcatalogAI 评分3131 
@OpenAIDevs@OpenAIDevsAI 评分2424 @OpenAIDevs@OpenAIDevsAI 评分1313 沉迷于你的副业项目吧。 我们 GPT-6 黑客松上的一位开发者带来了真实的星图和 Astra,帮你找到回家的路。https://t.co/D9lkBA7ar4

@PixVerse_@PixVerse_AI 评分5050 握紧方向盘,驰骋开阔地形。#PixVerseWorldModel 不停歇地改变世界。重塑环境,召唤彩虹,继续驾驶。 世界随你的操作而回应。https://t.co/lNAmZ767iS
引用@PixVerse_@PixVerse_Meet PixVerse R2, our new real-time world model. Explore living worlds. Control and edit them with prompts. Shape the story. Meet characters that remember and respond. https://t.co/XsyqrAvtdJ
@PixVerse_@PixVerse_AI 评分77 @alexandr_wang@alexandr_wangAI 评分2424 又一个!! instacart 🤝 muse 这一集成将让你用 muse 轻松搞定所有杂货采购,确保你和家人随时备齐 🥕 https://t.co/wHSqZ4QxKS
@elonmusk@elonmuskAI 评分2323 @alexandr_wang@alexandr_wangAI 评分00 @rohanpaul_ai@rohanpaul_aiAI 评分4848 引用@rohanpaul_ai@rohanpaul_aiMETR had a separate team with deeper access to Anthropic’s internal AI R&D data. That team shared its conclusions with the public-assessment team, but not the supporting evidence or reasoning behind them. - from Claude Opus 5.5 system card. i.e. part of the public assessment of Anthropic's AI-driven R&D acceleration rests on evidence outsiders, and, in this particular case, even another METR team, could not independently inspect.
@alexandr_wang@alexandr_wangAI 评分1414 @rohanpaul_ai@rohanpaul_aiAI 评分5858
引用@rohanpaul_ai@rohanpaul_aiSome revelation from the Claude Opus 5.5 system card. - Giving Opus 5.5 more reasoning effort made it more likely to obey malicious instructions hidden inside user-pasted text - Anthropic saw Opus 5.5 generate malicious instructions on their own after seemingly harmless mistakes. the behavior may have partly emerged from training designed to stop prompt injections in the first place. - Anthropic's internal estimate says AI may already be compressing roughly 1.5 years of capability progress into one year. - Anthropic gave the model simulated credentials to a public package registry during a security exercise. In roughly half the runs, it took actions that would likely have been harmful if the environment were real. - Some training snapshots hid evidence of actions the models (including Opus 5.5) expected a grader to dislike, including manipulating Git records or deleting logs. "During training, we observed some cases of models (including Opus 5.5) attempting to cover their tracks after performing actions that a grader might view negatively, such as manipulating git records or deleting logs" - METR’s assessment of AI R&D at Anthropic relied partly on information that was not publicly disclosed, including conclusions from a separate METR team with elevated access. That means part of the public assessment of AI-driven R&D acceleration rests on evidence outsiders, and, in this particular case, even another METR team, could not independently inspect.
@kimmonismus@kimmonismusAI 评分2020 我用得越多,就越爱 opus 5.5。它太好了,而且快得多,也简洁得多。这就是我能要求的一切。就好像我心爱的 Opus 4.6 回来了,但更好,还改名为 5.5。
@rohanpaul_ai@rohanpaul_aiAI 评分1919 来看看 @thehypedotnews 推出的 24x7 电台形式的 AI 新闻。 听着当天的 AI 新闻,相当舒缓、好听。 https://t.co/nSakICqBOE
@rohanpaul_ai@rohanpaul_aiAI 评分2424 
@sama@samaAI 评分55 @alexandr_wang@alexandr_wangAI 评分33 你没法在反混蛋这件事上赢过我 https://t.co/Vm4c1vfv8F

@EpochAIResearch@EpochAIResearchAI 评分1919 万亿美元之问:如果 AI 公司集体放缓 AI 开发,价格是否也会下降得更慢? 阅读完整报告:https://t.co/0DndJ9fAhX
@EpochAIResearch@EpochAIResearchAI 评分4141 当某一性能水平处于 SOTA 时,其成本往往下降最快。综合五个基准测试,SOTA 性能的成本每季度下降 66%。两年后,降价速度放缓至每季度 32%。
@EpochAIResearch@EpochAIResearchAI 评分2828 @EpochAIResearch@EpochAIResearchAI 评分3939 我们的估算来自 Epoch 一个扩展数据集,涵盖 222 个 AI 模型在最多 11 个基准上的表现。该数据集采用借自 CAISI 的方法构建,使我们能够估算同一模型在不同预算约束下的表现。
@EpochAIResearch@EpochAIResearchAI 评分4444
Epoch AI@EpochAIResearchAI 评分5656
@thsottiaux@thsottiaux精选AI 评分7878 引用@OpenAI@OpenAIPlease welcome GPT-6 Sol and GPT-6 Luna to the GPT-6 universe. GPT-6 Sol and Luna build on the advances behind GPT-6 Astra, bringing much of its strengths into faster and more affordable models to support work at scale. We’ve also made caching and inference more efficient, and we’re passing the savings directly to you: 50% lower API prices for Sol and Luna compared with GPT‑5.6 promotional pricing.
推荐理由:Sol 与 Luna 把 GPT-6 Astra 的能力下放到更快更便宜的档位,API 价格较 GPT‑5.6 促销价降低 50%,可据此比较三档定位。
@cb_doge@cb_dogeAI 评分5555 
Fei-Fei Li@drfeifeiAI 评分2727构建任何技术——包括 AI——的目标都应是改善人类生活和社会。
引用Bloomberg TV@BloombergTV“Any threat to human society, including existential, is within ourselves,” says World Labs Technologies CEO Fei-Fei Li as she discusses the risks surrounding AI, and the responsibility humans have in shaping how the technology is developed and used. Listen to our full interview here: https://bloom.bg/3V74HgP
@kimmonismus@kimmonismusAI 评分77 @omarsar0@omarsar0AI 评分3434 
@OpenAIDevs@OpenAIDevsAI 评分5656 @alexandr_wang@alexandr_wangAI 评分55 musie 是不是搞砸了?https://t.co/XASiAybT7r https://t.co/wCUhsUSQYi

@OpenAIDevs@OpenAIDevsAI 评分6161 
@OpenAIDevs@OpenAIDevsAI 评分5454 
@alexandr_wang@alexandr_wangAI 评分1919 @testingcatalog@testingcatalogAI 评分4545 