跳到正文

#Anthropic

今日 33 条
9月29日周二
  1. Thariq72

    Anthropic 发布 Claude Sonnet 5.5,为 Claude 5.5 家族第二款模型,较 Sonnet 5 速度提升超 30%,多数工作成本降低最高 30%。作者 Thariq 表示 Sonnet 与 Opus 5.5 让高阶抽象如 projects、claude tag 和动态工作流的 token 成本顾虑更小,建议在构建工作流时优先试用 Sonnet 5.5。

    引用Claude@claudeai

    Introducing Claude Sonnet 5.5, the second model in the Claude 5.5 family. It’s a clear upgrade over Sonnet 5, runs more than 30% faster, and costs up to 30% less for most work.

    推荐理由:作者结合 token 成本这一常见顾虑,指出 Sonnet 5.5 与 Opus 5.5 让更高层智能更易负担,适合在构建工作流时选用。

  2. ClaudeDevs74

    Anthropic 推出 Claude 5.5 家族第二个模型 Claude Sonnet 5.5,称其相比 Sonnet 5 更聪明、更高效,速度快 30% 以上,多数工作成本最多降低 30%。作者建议用于修复 bug、快速迭代功能等边界清晰度的日常任务,Claude Code 用量也能更省。

    引用Claude@claudeai

    Introducing Claude Sonnet 5.5, the second model in the Claude 5.5 family. It’s a clear upgrade over Sonnet 5, runs more than 30% faster, and costs up to 30% less for most work.

    推荐理由:原文给出 Sonnet 5.5 相对 Sonnet 5 的速度、成本与适用任务,开发者可据此判断是否切换日常 Claude Code 用法。

  3. Boris Cherny61

    Anthropic 发布 Claude Sonnet 5.5,是 Claude 5.5 家族的第二款模型,官方称相比 Sonnet 5 是明显升级,运行速度提升超过 30%,多数任务成本最多降低 30%。作者 Boris Cherny 演示用 Sonnet 5.5 修复 Claude Code 的一个 bug,并强调其快 30%、用量费用省 30%。

    引用Claude@claudeai

    Introducing Claude Sonnet 5.5, the second model in the Claude 5.5 family. It’s a clear upgrade over Sonnet 5, runs more than 30% faster, and costs up to 30% less for most work.

  4. Dongxi 东锡 NLP75

    Anthropic 发布 Claude Sonnet 5.5,为 Claude 5.5 家族的第二款模型。官方称其相比 Sonnet 5 是明显升级,速度提升超过 30%,多数工作场景成本最多降低 30%。

    引用Claude@claudeai

    Introducing Claude Sonnet 5.5, the second model in the Claude 5.5 family. It’s a clear upgrade over Sonnet 5, runs more than 30% faster, and costs up to 30% less for most work.

    推荐理由:官方发布说明给出了相对 Sonnet 5 的速度提升与降价幅度,读者可据此权衡换用成本。

  5. Anthropic76

    Anthropic 宣布 Claude Sonnet 5.5 现已可用,这是 Claude 5.5 家族的第二个模型。相比 Sonnet 5 是明显升级,运行速度提升超过 30%,多数工作的成本降低最多 30%。

    引用Claude@claudeai

    Introducing Claude Sonnet 5.5, the second model in the Claude 5.5 family. It’s a clear upgrade over Sonnet 5, runs more than 30% faster, and costs up to 30% less for most work.

    推荐理由:Anthropic 官宣 Claude Sonnet 5.5 上线,直接给出比 Sonnet 5 快 30%、多数工作成本低 30% 的关键变化。

  6. Claude Blog44

    Asana 如何用 Claude 打造可训练的人机协作团队

    Asana 让 AI 智能体直接运行在其 Work Graph 模型内,与人类同事一样拥有角色、任务、消息读写和活动流记录,并由 Claude 驱动复杂任务。每个智能体按内容撰写、洞察分析、项目管理等角色预置技能与 Hubspot 等集成,其实际访问权限受触发者权限约束。智能体的共享记忆仅允许管理员和编辑者写入永久记忆,普通成员反馈只作用于当前任务。

  7. Anthropic Research46

    Anthropic 启动新研究:用 Anthropic Interviewer 征集你对 AI 的真实想法

    Anthropic 发起新研究,通过 Anthropic Interviewer 收集人们与 AI 相处的真实经历,参与者可自行决定是否将完整访谈公开,供任何人阅读研究。研究关注最有意义的 AI 体验、希望 AI 改变的现实领域,以及对 AI 公司的期待。此前去年 12 月的同类研究有 81,000 人参与,成果曾用于 Anthropic Institute 议程并在世界经济论坛上展示。

  8. Anthropic Research82

    Anthropic 评测 GLM-5.3:可自主构建端到端漏洞利用且防护易被绕过

    Anthropic 发布对智谱 GLM-5.3 的网络安全能力分析,认为它是首个在无实质防护下开放权重的强网络攻击能力模型,与 NIST CAISI 评估结论大致一致。

    推荐理由:Anthropic 以一手评测数据说明 GLM-5.3 的漏洞利用能力与防护绕过率,并解释攻击者可及性与 Claude 的访问限制差异。

9月28日周一
  1. Simon Willison54

    Simon Willison 发布 Bluesky 回复机器人检测工具

    Simon Willison 用 Opus 5.5 编写了一个 Bluesky 回复机器人检测工具,通过分析打字速度、发帖时间、互动模式等行为信号判断账号是否为自动回复机器人。工具展示每项测量和规则及触发信号最多的示例回复;作者称 Bluesky 开放 API 使此类调查比 Twitter 更可行,检测信号包括秒级连发回复、从不发布原创内容、专门回复高粉丝用户及使用问号等。

9月27日周日
9月26日周六
  1. ByteByteGo29

    Jev 最值得替代 LLM 的 9 个场景

    TypeSafe AI 的首个 System One Model Jev 比前沿 LLM 快 100 倍、成本低 100 倍,适合承担 LLM 周边的决策类任务。文章列出 9 个替代场景:模型路由、护栏检测、工具调用权限分类、收件箱分类、重排序、LLM 评测打分、批量打标、实时决策和置信度门控。核心思路是让 LLM 负责生成,Jev 负责围绕生成的决策。

  2. Boris Cherny55

    Anthropic 官方账号 ClaudeDevs 宣布推出新门户,开发者可提交 Claude 插件、跟踪审核状态并查看用量。插件打包 MCP 与技能,正成为面向 Claude 构建的方式,MCP 在 Claude 产品中的用量今年增长 110 倍。Boris Cherny 转发并表示期待开发者构建的作品。详情见 https://claude.com/blog/build-plugins-for-claude

    引用ClaudeDevs@ClaudeDevs

    It’s now easier to build plugins for Claude. We built a new portal to submit your plugin, track review, and see usage. Plugins package MCP and skills, and are becoming the way to build for Claude. MCP usage across Claude products is up 110x this year! https://claude.com/blog/build-plugins-for-claude

  3. Ars Technica · AI84

    美国上诉法院裁定国防部可将拒绝开放 Claude 功能的 Anthropic 列入黑名单

    美国哥伦比亚特区联邦巡回上诉法院以 2-1 裁定,国防部有权因 Anthropic 拒绝向军方开放部分 Claude AI 功能而将其列入黑名单,即使 Anthropic 并无恶意。

    推荐理由:判决同时呈现政府与 Anthropic 各自的风险论证,有助于理解 AI 供应商与军方合作边界的法律走向。

  4. ClaudeDevs70

    Opus 5.5 的输入和输出 token 比 Opus 5 便宜 20%,cache reads 便宜 60%。作者据此测算了在 Claude Code 中完成一个任务的实际成本变化,并发布了博客和计算器,读者可从 /usage 运行自己的数据:https://claude.dev/blog/what-a-task-costs-on-opus-5-5/

    推荐理由:作者用具体数字拆解了 Opus 5.5 降价对 Claude Code 单任务成本的实际影响,并给出可复用的成本计算器入口。

  5. Anthropic65

    Anthropic 发文称,Claude 在收到单个九圈问题提示词后,在 Claude Science 中基本无人监督地运行数天,用 Dixon 等人的方法完成求解,总成本几千美元,突破了此前八圈的纪录(平面 N=4 超杨-米尔斯简化模型)。物理学家 Lance Dixon 独立验证了结果,von Hippel 为该博客撰写了经历回顾。

    推荐理由:九圈散射振幅计算由 Dixon 独立验证,计算成本仅几千美元,为学界评估 AI 科研能力提供了一个可核验的案例。

  6. Boris Cherny50

    Boris Cherny 称 Claude Tag 每天写他超 50% 的 PR,完成约 100% 的数据分析,并修复大部分产品反馈和 bug。他介绍 Claude Tag 不同于普通 Slack bot,具备主动、可编程、有记忆和连接器访问能力,配合 Opus 5.5 和 Fable 5.1 有较强判断力,并给出自动复现 bug 并提 PR、深挖数据假设、生成讲解游戏等示例提示词。

    引用Noah Zweben@noahzweben

    Claude Tag in Slack can now use your personal connectors! You can now securely access that Drive doc, Salesforce account, or Warehouse table that you have personal access to right where the work happens. Avail. on Teams today and Enterprise next week https://claude.com/blog/claude-tag-now-supports-personal-connectors-in-channels

  7. Noah Zweben49

    今年 2 月我们首次推出 /remote-control 时,我用 Opus 4.6 做这些视频玩得很开心。那么,这是 Opus 5.5 的粘土动画版本。 Remote Control with 5.5,当你不得不去的时候!

    引用Noah Zweben@noahzweben

    Rolling out Claude Code Remote Control to Pro users - because they deserve to use the bathroom too . (Team and Enterprise coming soon). 🧻 Rolling out to 10% and ramping 1. Update to claude v2.1.58+ 2. Try log-out and log-in to get fresh flag values. 3. /remote-control

  8. Claude66

    Claude 官方表示 Claude Opus 5.5 发布数日,汇总了用户用其探索和发现的喜爱案例。引用案例中,@RyanSael 让 Opus 5.5 通过构建交互式镜头实验室讲解相机对焦,一次生成耗时 1 小时 26 分钟,API 成本 $25.66,成品见 https://lens.lab.sael.net。

    引用Ryan Sael@RyanSael

    I asked Opus 5.5 to explain camera focus by building an interactive lens lab Here's what it came up with after 1 hour 26 minutes in one shot, $25.66 API cost https://lens.lab.sael.net Move the focus ring and you can see the glass elements shift the sharp plane through the scene

    推荐理由:官方汇总用户用 Opus 5.5 探索的成果,引用案例给出了单次生成时长与成本,可作实际使用参考。

9月25日周五
  1. Claude Code GitHub Releases35

    Claude Code v2.1.282 发布

    Claude Code 发布 v2.1.282,新增 maxProseWidth 设置,可限制宽终端中 Claude 正文宽度,表格与代码块仍保持全宽。该版本还新增启动提示及 /status、claude doctor 条目,列出项目设置文件中被忽略或关闭遥测的变量,并修复了会话续接重发旧消息、扩展思考丢失及多处 API 报错等问题。

  2. GitHub Blog · AI & ML63

    GitHub Security Lab 发布 Fuzzing Taskflow:用 LLM 智能体自动化 C/C++ 模糊测试

    GitHub Security Lab 的 Antonio Morales 基于自家的 Taskflow Agent 框架构建了 Fuzzing Taskflow,一个面向 C/C++ 项目的自主模糊测试流水线。

    推荐理由:原文给出完整的架构设计、覆盖反馈循环和分层判断方法,读者可以据此把 LLM 智能体接到自己的模糊测试流程里。