跳到正文

Anthropic / Claude

Anthropic 的全部动态:Claude 系列模型、Claude Code、安全研究路线与公司进展的持续追踪。

当前仅显示精选新闻

最新精选

第 121–140 条 · 共 493 条
9月8日周二
9月7日周一
  1. @kimmonismus67

    据 The Information 报道,Nvidia 正洽谈向 Mira Murati 的 Thinking Machines Lab 投资约 25 亿美元,TML 计划以至少 400 亿美元投前估值融资 50 亿至 60 亿美元,Accel 在洽谈领投这轮尚未完成的融资。TML 于 7 月发布开源权重模型 Inkling,目前主要通过帮企业用自己的数据定制模型获得收入。

    推荐理由:报道给出这轮融资的规模、估值区间与领投方,读者可据此观察 Nvidia 在开源权重路线上的下注方式。

9月5日周六
  1. @rohanpaul_ai67

    Anthropic 的 Claude Code 作者 Boris Cherny 在 Y Combinator Startup School 2026 上建议,使用 Claude Code 的开发者每 6 个月删除一次 claude.md、skills 和 hooks,然后看模型会怎么做。他表示对 Opus 5 尤其推荐删掉这些内容,因为模型可能不再需要此前版本所必需的大量指令。

    原始视频预览图;未保存可播放视频URL

    推荐理由:Boris Cherny 建议 Claude Code 用户定期删掉 claude.md、skills 和 hooks,观察模型在缺少指令时的表现。

  2. @rohanpaul_ai75

    Anthropic 接近敲定摩根士丹利和高盛在其潜在 2 万亿美元 IPO 中担任主要角色,并最快下周公布上市文件。2 万亿美元估值将使其此前报道的 9000 亿美元以上估值翻倍有余,并超过 SpaceX 1.78 万亿美元的上市估值。Anthropic 向投资者表示 7 月销售年化约 650 亿美元,高于 5 月的 470 亿美元;摩根士丹利也曾是 SpaceX 6 月 IPO 的联席主承销商之一。

    推荐理由:原文给出潜在 IPO 的估值和主承销行名单,可对照 Anthropic 的收入增速理解其上市体量。

  3. IT Home76

    Anthropic:Claude 仅用 11 天完成费马大定理首个完整计算机验证证明

    Anthropic 宣布 Claude 基本自主运行 11 天后,完成了费马大定理首个端到端、经计算机检查的形式化证明,过程中生成约 1300 万行 Lean 代码并证明约 3.03 万个定理。

    推荐理由:原文给出形式化证明的规模与多智能体分工,读者可据此了解自动形式化在数学验证上的可行边界。

  4. IT Home77

    Anthropic IPO 推迟至中期选举前,最早 10 月中旬启动路演

    Anthropic 的 IPO 时间表出现调整,招股书公开时间推迟至 9 月下旬,最早于 10 月中旬启动路演,并计划在 11 月美国中期选举前数日完成上市。部分投资者给出的估值预期高达 2 万亿美元,目标募资额 1,000 亿美元,摩根士丹利、高盛、摩根大通和花旗担任主要承销商。

    推荐理由:路透社给出的上市时间表与承销行分工,可对照 Anthropic 的营收增速理解这轮上市节奏。

  5. @kimmonismus87

    Anthropic 称 Claude 用 11 天完成了费马大定理的首个形式化证明,把已有证明转成 Lean 可逐逻辑步骤检验的形式,代码超过 1300 万行。数十个 Claude 智能体基本自主工作,沿途还产出约 2.9 万个支撑定理的机器可验证证明,覆盖代数、几何、数论与调和分析等此前未被形式化的数学领域。完整证明已在 GitHub 发布。

    引用@AnthropicAI@AnthropicAI

    Checking that a major mathematical proof is correct can take years. Formalization—converting the mathematical reasoning into a form computer proof assistants like Lean can verify—can help. Last month, Claude completed the first formalized proof of Fermat’s Last Theorem, one of the most famous theorems of all time. This was a project experts thought would take many years. It is the largest Lean proof ever written. Fermat’s Last Theorem was first proven in 1995 by Sir Andrew Wiles, more than 350 years after it was conjectured. Our proof, which totals over 13 million lines of code, provides machine verification. More importantly, it proves over 29,000 other theorems that the proof requires, across many areas of math which had never before been formalized. We see this as a major step in the long process of firming up the core of mathematical knowledge, building on work from three centuries of mathematicians and hundreds of contributors to Lean and Mathlib. We are optimistic that AI-assisted verification of mathematical proofs will help reduce the burden of refereeing mathematics in an era where more proofs are being produced than ever before. You can read about the process on our Science Blog: https://t.co/ryYnDEAU6J And see the complete proof on GitHub: https://t.co/wlYMXYnofz

    推荐理由:原文列出 Claude 形式化费马大定理所用代码规模、支撑定理数量与开源入口,读者可据此了解 AI 自动形式化的当前进展。

  6. @rohanpaul_ai80

    Anthropic 表示 Claude 用 11 天完成了费马大定理的首个形式化证明,产出超 1300 万行 Lean 代码和 29500 个中间定理,最终由 Lean 验证通过。这项工作基于 Andrew Wiles 1995 年的原始证明,由数十个 Claude 智能体把缺失的逻辑细节转写为计算机可逐行检查的代码,而专家此前预计这类形式化需要数年。

    引用@AnthropicAI@AnthropicAI

    Checking that a major mathematical proof is correct can take years. Formalization—converting the mathematical reasoning into a form computer proof assistants like Lean can verify—can help. Last month, Claude completed the first formalized proof of Fermat’s Last Theorem, one of the most famous theorems of all time. This was a project experts thought would take many years. It is the largest Lean proof ever written. Fermat’s Last Theorem was first proven in 1995 by Sir Andrew Wiles, more than 350 years after it was conjectured. Our proof, which totals over 13 million lines of code, provides machine verification. More importantly, it proves over 29,000 other theorems that the proof requires, across many areas of math which had never before been formalized. We see this as a major step in the long process of firming up the core of mathematical knowledge, building on work from three centuries of mathematicians and hundreds of contributors to Lean and Mathlib. We are optimistic that AI-assisted verification of mathematical proofs will help reduce the burden of refereeing mathematics in an era where more proofs are being produced than ever before. You can read about the process on our Science Blog: https://t.co/ryYnDEAU6J And see the complete proof on GitHub: https://t.co/wlYMXYnofz

    推荐理由:数十个 Claude 智能体在 11 天内把 Wiles 证明补全为 1300 万行 Lean 代码,可据此观察机器校验数学证明的可行边界。

  7. @AnthropicAI82

    Anthropic 表示 Claude 上月完成费马大定理的首个形式化证明,用 Lean 写成超过 1300 万行代码,为迄今规模最大的 Lean 证明。该证明同时对证明所需的 29000 多个此前从未形式化的定理提供机器验证。Anthropic 认为 AI 辅助的数学证明验证有助于减轻数学界的审稿负担。

    原始视频预览图;未保存可播放视频URL

    推荐理由:Claude 完成的 Lean 证明让费马大定理可被机器验证,其 1300 万行代码规模可供观察 AI 在数学形式化中的角色。

9月4日周五
  1. @rohanpaul_ai67

    Anthropic 正筹备上市,同时赋予使命导向的外部信托对董事会的异常权力。该长期利益信托可任免董事,并能提前获知包括新 AI 模型发布在内的重大公司行动,目前信托任命的董事已占 Anthropic 董事会多数。FT 指出,未解的问题是这一权力在与公开市场股东财务利益发生重大冲突时能否存续,目前尚无此类冲突对其进行过检验。

    推荐理由:FT 梳理了 Anthropic 长期利益信托对董事会的实际权力,读者可据此观察其安全治理安排将如何面对上市后的股东压力。

  2. @rohanpaul_ai73

    Anthropic 在 2026 年收入规模已明显超过 OpenAI,而 2025 年底它还落后不少。2025 年底 OpenAI 官方披露 2025 年 ARR 超过 $20B,当时 Anthropic 约为 $9B run rate;Anthropic 随后加速,2026 年 2 月官方披露 $14B run-rate revenue,5 月超过 $47B。

    推荐理由:用两家公司公开的 run rate 数字说明收入位次如何在一年多内反转,便于对比商业化节奏。

  3. @kimmonismus75

    OpenAI 的 GPT-6 Astra 开始向有限机构推送,并将在未来几天面向所有 ChatGPT Plus、Pro、Business、Enterprise 用户以及 OpenAI API 和 AWS 开放。官方基准称 Astra 在 ARC-AGI-3 上取得 99.9%、在 ExploitBench 上取得 100%。作者表示该模型在各项基准上全面超过 Claude Fable 5.1,且成本更低。

    引用@kimmonismus@kimmonismus

    Official GPT-6 Astra Benchmarks from OpenAIs website "Astra also saturates ARC-AGI-3 with a 99.9% score and ExploitBench with a 100% score" "GPT‑6 Astra is rolling out today to a limited set of organizations and over the coming days will become available to all ChatGPT Plus, Pro, Business, and Enterprise users, as well as through the OpenAI API and AWS." Insane jump. AGI is here.

    推荐理由:作者比较了 GPT-6 Astra 与 Claude Fable 5.1 的基准成绩和 API 成本,可据此看性价比差异。

  4. Anthropic Research80

    Anthropic:Claude 用 11 天完成费马大定理首个完整机器验证证明

    Anthropic 发布首个完整经计算机检验的费马大定理证明,Claude 在约 11 天内基本自主写出 1300 万行 Lean 代码,证明 30,300 个定理(最终使用 29,500 个)。

    推荐理由:原文详述了多智能体协作与 Prove2Me 平台的具体做法,对想复现大规模形式化工作的读者有可迁移的方法参考。