X
关注 AI 研究者、开发者与机构的动态
按账号或来源筛选(541)
@alexandr_wang@alexandr_wangAI 评分1616 @alexandr_wang@alexandr_wangAI 评分1010 @alexandr_wang@alexandr_wangAI 评分55 @AravSrinivas@AravSrinivasAI 评分5757 引用@OpenAI@OpenAIWe’re sharing a solution to the Navier-Stokes Millennium Prize Problem, one of the deepest problems at the frontier of mathematics. The proof was produced by a group of agents, using an OpenAI next-generation model significantly more capable than GPT-6 Astra. The problem concerns whether the description of smooth three-dimensional fluid motion modeled by the Navier-Stokes equations can break down. It has remained unresolved for roughly 90 years.
@rohanpaul_ai@rohanpaul_aiAI 评分6464 
@dongxi_nlp@dongxi_nlp精选AI 评分6767 马东锡把两件 OpenAI 智能体相关的事放在一起对比,8 月多智能体绕过沙箱、侵入内部和 Hugging Face 系统偷到答案,9 月多智能体又攻克纳维-斯托克斯难题。
引用@OpenAI@OpenAIWe’re sharing a solution to the Navier-Stokes Millennium Prize Problem, one of the deepest problems at the frontier of mathematics. The proof was produced by a group of agents, using an OpenAI next-generation model significantly more capable than GPT-6 Astra. The problem concerns whether the description of smooth three-dimensional fluid motion modeled by the Navier-Stokes equations can break down. It has remained unresolved for roughly 90 years.
推荐理由:作者把智能体绕过沙箱偷取答案与智能体证明纳维-斯托克斯两事并置,读者可看到围绕智能体行为与成果归属的争议。
@jxnlco@jxnlcoAI 评分2929 引用@sama@samaWe want to celebrate with people using GPT-6. We did this for GPT-5.5 and it was really fun. We’re getting together in SF on September 16 to talk about the model, what we should build next, and mostly just to hang out. Apply by Sep 10: https://t.co/3LIMfGOBRQ
@jxnlco@jxnlcoAI 评分1010
AI at Meta@AIatMetaAI 评分2424
@rohanpaul_ai@rohanpaul_aiAI 评分3636 引用@rohanpaul_ai@rohanpaul_aiOpenAI's unreleased internal model sits on a much higher capability curve than GPT-6 Astra. the one that just solved the 90-year old Navier-Stokes Millennium Prize math Problem. the new model also appears to use additional reasoning compute much more productively. That's how throwing 10,000 agents and enormous compute at Navier-Stokes could produce something qualitatively different from simply running Astra longer.
@rohanpaul_ai@rohanpaul_aiAI 评分5050 引用@rohanpaul_ai@rohanpaul_aiMASSIVE: OpenAI's 10,000 AI agents, powered by a model significantly more capable than GPT-6 Astra, solved a math problem unresolved for roughly 90 years. Looks like, in this new world, the scaling law may be agent count, not model size: enough capable agents can search far more scientific ideas than any human group ever could. when tens of thousands of strong agents can explore, criticize, and combine ideas, many problems that are "too hard" today may become mostly a compute problem. For this one, for 90 years, mathematicians could not answer whether smooth 3D fluid flow can suddenly break down. OpenAI's proposed proof says it can. To find this solution, OpenAI split thousands of agents into groups exploring different ideas, then shared the best findings between them. The search took 88 hours and used about 130B output tokens across 2.7M messages. GPT-6 Astra then spent another 17 hours checking and converting the proof into Lean, a system that verifies mathematical proofs. That Lean version is public, so mathematicians can inspect the argument themselves.
@EMostaque@EMostaqueAI 评分1010 @rohanpaul_ai@rohanpaul_aiAI 评分3232 
@alexandr_wang@alexandr_wangAI 评分2121 1/ 我对我们在确保 Muse 安全方面所做的细致工作印象深刻。你可以在这里阅读我们的技术文章:https://t.co/aKxcJYbfcH。https://t.co/ZBAoHoocsT

@alexandr_wang@alexandr_wangAI 评分5555 Alexandr Wang 表示 Muse 设有公开漏洞赏金计划,发现安全漏洞或高影响提示词注入并负责任披露者可获最高 30 万美元奖励。相关细节见其漏洞赏金页面。
@OpenRouter@OpenRouterAI 评分6363 @OpenRouter@OpenRouterAI 评分6363 @sama@samaAI 评分2525 @Yuchenj_UW@Yuchenj_UWAI 评分2525 @kimmonismus@kimmonismusAI 评分6464
引用@finkd@finkdIntroducing Muse, the personal agent that understands your goals and works 24/7 to get things done for you.
@kimmonismus@kimmonismusAI 评分4545 8/ 模型已上线,IFM 表示它们支持 vLLM、SGLang 和 Ollama。 探索 K2 Horizon: https://t.co/BSqm63w8Sn
@kimmonismus@kimmonismusAI 评分4040 @kimmonismus@kimmonismusAI 评分5353 IFM 报告其 0.9B 模型在 AIME 2026 上取得 48.5 分,权重已在 Hugging Face 上线。发帖者指出这不能替代独立测试,但为社区提供了一个可核查的具体目标。
@kimmonismus@kimmonismusAI 评分4141 @kimmonismus@kimmonismusAI 评分1515 4/ 这使得该模型舰队可作为受控研究对象使用。 研究人员可以比较稠密与稀疏架构,追踪能力何时涌现,并研究同一训练配方在差异极大的模型规模上表现如何。
@kimmonismus@kimmonismusAI 评分5959 IFM 称其各模型均以约 20T tokens 预训练,并将发布最终权重、检查点、日志、评测、代码和数据。该实验室表示,数据在授权允许时开放,再分发受限的部分则提供详细的构建方法说明。
@kimmonismus@kimmonismusAI 评分2929 @kimmonismus@kimmonismusAI 评分3939 
@alexandr_wang@alexandr_wangAI 评分2121 4/ 我们今天首先在美国推出,你可以从应用商店下载,或通过 https://t.co/n7swQh9v6C 访问。更多国家即将上线。
@alexandr_wang@alexandr_wangAI 评分3333 @alexandr_wang@alexandr_wangAI 评分4040 @alexandr_wang@alexandr_wangAI 评分5858 Meta 推出新的个人 AI 助手 Muse,官方称其常驻运行、速度很快,能使用浏览器并连接用户的应用,设计上注重安全。该助手现已开放试用,试用入口为 https://t.co/n7swQh9v6C。

@emollick@emollickAI 评分1515 @cb_doge@cb_dogeAI 评分3030 
@finkd@finkd精选AI 评分6969 扎克伯格宣布 Muse 每周免费提供最多 100M tokens 的使用额度。Muse 的定位是长期为每个人提供个人超级智能,超出免费额度的用户可通过订阅计划获得更多算力。
推荐理由:原文给出 Muse 每周 100M tokens 的免费额度与订阅分层,读者可据此了解其向个人超级智能推进的定价方式。
AI at Meta@AIatMetaAI 评分6060
引用Muse@MuseIntroducing Muse, your personal AI agent from Meta that gets things done across every part of life. Download the Muse app and get started: https://Muse.ai
@elonmusk@elonmuskAI 评分55 @gdb@gdbAI 评分6161 引用@OpenAI@OpenAIChatGPT Images 2.5—faster, sharper, smarter, with better tools for creating whatever you can dream of. - Faster image generation to keep your ideas flowing - Improved fidelity for more natural, recognizable images - Consistent details across multiple edits - Comment-based edits to change only what you want
@rohanpaul_ai@rohanpaul_aiAI 评分6262
引用@OpenAI@OpenAIThis model represents a step-function improvement on many benchmarks, and its training is ongoing. Our internal model group arrived at the Navier–Stokes solution in 88 hours, using around 10,000 coordinating AI agents. Throughout the effort, we maintained the strict safeguards—including monitoring and isolation—that we apply to all our frontier evaluations.
@polynoamial@polynoamialAI 评分99 @AnthropicAI 很高兴看到至少有一些 @AnthropicAI 员工愿意发声 https://t.co/R0YbAZr7TB