跳到正文

模型发布

新模型的发布、开源与迭代:大模型厂商的旗舰更新、开源权重放出、性能与价格变化的第一时间记录。

当前仅显示精选新闻
561条精选相关主题产品更新论文研究开源生态

最新精选

第 381–400 条 · 共 561 条
6月10日周三
  1. @berryxia67

    Google 将 Gemini 3.5 Live Translate 推向公开预览,通过 Gemini API 提供低延迟语音到语音翻译,覆盖 70 多种语言和 2000 种语言对。作者指出它把小众语言对也纳入覆盖,可用于实时对话、客服、直播和跨国会议等场景,并提到目前尚不清楚它与阿里部分小语种模型的对比效果。

    引用Google for Developers (@googledevs)@googledevs

    Gemini 3.5 Live Translate is now in Public Preview via the Gemini API, delivering low-latency speech-to-speech translation across 70+ languages and 2,000 language pairs! 🌍 Challenge time: What is the most niche, unique, or complex language pair your application needs to translate? Tell us in the comments, and we’ll let you know if the model supports it! Full blog link: goo.gle/3QzaHwN Video

    推荐理由:Google 将 Gemini 3.5 实时语音翻译推向公开预览,可了解其对小众语言对的覆盖范围与 API 接入方式。

  2. IT Home78

    Anthropic 推出 Claude Fable 5 与 Claude Mythos 5 两款模型

    Anthropic 发布 Claude Fable 5 与 Claude Mythos 5 两款模型,二者基本为同一模型,只是安全护栏不同。Fable 5 面向普通用户,官方称其为公开可用能力最强的 Claude 模型;Mythos 5 继续通过 Project Glasswing 先向少量网络安全防御方和基础设施提供商开放。

    推荐理由:材料给出两款模型的能力定位、统一价格与安全分流机制,可供判断 Claude 在高敏感场景的开放边界。

  3. 量子位 · 微信公众号88

    Anthropic 发布 Claude Fable 5 与 Claude Mythos 5,5000 万行代码库迁移一天完成

    Anthropic 发布 Claude Fable 5 与 Claude Mythos 5 两个版本,Fable 5 面向所有用户开放,Mythos 5 只给少数受信任用户使用,两者统一定价为每百万输入 Token 10 美元、每百万输出 Token 50 美元,比之前的预览版便宜一半以上。

    推荐理由:两款同内核模型按权限分级,Fable 5 用分类器降级路由兼顾能力与安全,可观察前沿模型的形态变化。

  4. @berryxia69

    Anthropic 发布 Claude Fable 5,这是一款 Mythos 级模型,已被做成安全版本面向通用场景开放,能力超过其此前任何公开模型。基准评测近乎全线 SOTA,在软件工程、知识工作、科研和视觉等硬任务上领先,长任务越复杂领先越多。

    引用Claude (@claudeai)@claudeai

    Introducing Claude Fable 5: a Mythos-class model that we’ve made safe for general use. Its capabilities exceed those of any model we’ve ever made generally available. Video

    推荐理由:原文列出该模型基准几乎全线 SOTA,并说明敏感领域会自动回退到 Opus 4.8 的机制。

  5. @swyx70

    karpathy 称 Claude Fable 5 与 Mythos 是同一底层模型,只是增加了防护措施,在各项基准上都以一定优势达到 SOTA,并认为这是值得大版本号跃迁的阶跃式进步,尤其擅长在极难问题上的长时间解题会话。swyx 转发表示自己重跑了历史图表上的 FC Diamond,认为官方表格和图表都没有体现出这种起飞幅度,因为 Fable 属于不同级别的模型。

    引用Andrej Karpathy (@karpathy)@karpathy

    This is a super exciting release - Claude Fable 5 is the same underlying model as Mythos but with added safeguards. The benchmarks are great and it's SOTA on everything by a margin but I'll add that *qualitatively* also, this is a major-version-bump-deserving step change forward (imo of the same order as Claude 4.5 was in November), peaking especially for long problem-solving sessions on very difficult problems. You can give it a lot more ambitious tasks than what you're used to, the model "gets it" and it will just go, and it's never felt this tempting to stop looking at the code at all (but don't do this in prod!). The model still has quirks that people will run into and the safeguards are configured to be a little too trigger happy for launch, which can hopefully be tuned over time. I feel a lot of things changing as working software increasingly comes out on a tap. The Jevon's paradox kicks in and I feel my own demand for software growing substantially. You can ask for anything - explainers, visualizers, dashboards, bespoke single-use apps (e.g. a full wandb that is hyper-specific just for your project), you can 10X your test suite, auto-optimize code, run giant research projects with custom HTML for the results, anything! "Free your mind" (Matrix ref). Really looking forward to all the things people build!

    推荐理由:转发的评测者认为官方榜单未体现 Fable 5 的进步幅度,可作为判断这次模型跃迁的定性参考。

  6. @kimmonismus79

    Claude 5 Fable 在几乎所有已测试的 AI 能力基准上达到 SOTA,在软件工程、知识工作、视觉与科研方面表现突出。作者称其比过往 Claude 模型更省 token,能在数百万 token 的长任务中保持专注并用自记笔记改进输出。Stripe 早期测试称该模型把数月工程压缩到数天,在 5000 万行 Ruby 代码库中用一天完成原本需团队两个多月的全库迁移。

    引用Chubby♨️ (@kimmonismus)@kimmonismus

    Claude 5 Fable Benchmarks! Holy moly, significant jump even to Mythos

    推荐理由:原文汇总了 Fable 5 的基准成绩和 token 效率变化,并借 Stripe 的迁移案例展示其长任务表现。

  7. @swyx81

    Anthropic 发布 Claude Fable 5,称这是一款面向通用使用做了安全处理的 Mythos 级模型,并称其能力超过以往所有通用发布的模型。swyx 转发该发布并附上 Anthropic 的博客链接 anthropic.com/news/claude-fa…。

    引用Claude (@claudeai)@claudeai

    Introducing Claude Fable 5: a Mythos-class model that we’ve made safe for general use. Its capabilities exceed those of any model we’ve ever made generally available. Video

  8. @op741872

    Anthropic 发布 Mythos 低配版 Claude Fable 5,向 API、Pro、Max、Team 及企业用户开放。它采用与 Mythos 5 相同的底层模型,API 输入每百万 Token 10 美元、输出 50 美元,比 Mythos Preview 便宜一半。Fable 加强了安全防护,涉及网络攻击、生化攻击或大规模能力蒸馏的请求会被拒绝,并回退到 4.8 版本。

    引用Claude (@claudeai)@claudeai

    Introducing Claude Fable 5: a Mythos-class model that we’ve made safe for general use. Its capabilities exceed those of any model we’ve ever made generally available. Video

    推荐理由:梳理了 Fable 5 与 Mythos 5 的开放范围、API 定价减半及安全回退机制,可据此评估其可用性。

  9. Claude Code GitHub Releases70

    Claude Code v2.1.170 发布:引入 Claude Fable 5 并修复会话转录保存问题

    Claude Code 发布 v2.1.170,宣布引入 Claude Fable 5,称其为面向通用场景开放、能力超过此前所有公开模型的 Mythos 级模型,升级该版本即可使用,并附 Anthropic 公告链接。该版本还修复了从 VS Code 集成终端或继承 Claude Code 环境变量的 shell 启动时,会话不保存转录且不出现在 --resume 中的问题。

    推荐理由:原文给出新模型的获取方式和一个会话转录保存的修复,Claude Code 用户可据此决定是否升级到该版本。

  10. @dotey79

    Anthropic 发布 Claude Fable 5 和 Claude Mythos 5 两个模型,二者共用同一底座,Fable 5 加装安全分类器面向所有用户开放,Mythos 5 去掉部分安全限制只给 Project Glasswing 网络安全合作伙伴使用。

    引用Claude (@claudeai)@claudeai

    Introducing Claude Fable 5: a Mythos-class model that we’ve made safe for general use. Its capabilities exceed those of any model we’ve ever made generally available. Video

    推荐理由:同底座的两个模型分走通用与受限两条路线,安全降级机制与定价变化构成理解这次发布的关键。

  11. @OpenRouter67

    Anthropic 的 Claude Fable 5 已在 OpenRouter 上线,被称为 Anthropic 最强的编码模型。该模型面向长时程、模糊任务,包括 legacy migrations、生产环境疑难 bug 以及运行数小时到数天的异步会话,在几乎所有测试基准上达到 SOTA。

    推荐理由:Anthropic 这款主打长时程异步编码任务的模型已接入 OpenRouter,读者可据此了解它的能力定位与使用入口。

  12. @kimmonismus80

    Claude 5 Fable 上线,作者称即便在德国也已可用。所引 @claudeai 内容称 Fable 5 在几乎所有测试基准上达到 SOTA,在软件工程、知识工作、科学研究和视觉方面表现突出,任务越长越复杂,其相对其他模型的领先幅度越大。

    引用Claude (@claudeai)@claudeai

    Fable 5 is state-of-the-art on nearly all tested benchmarks, with exceptional performance in software engineering, knowledge work, scientific research, and vision. The longer and more complex the task, the larger Fable 5’s lead over our other models.

    推荐理由:原文转述 Claude 官方发布的基准声明,读者可据此了解 Fable 5 在长任务上相对其他模型的领先幅度。

6月9日周二
  1. Google DeepMind67

    Google DeepMind 发布 Gemma 4 12B:无编码器统一架构的多模态模型

    Google DeepMind 发布 Gemma 4 12B,采用无编码器统一架构,视觉与音频输入直接进入 LLM backbone,是该系列首个支持原生音频输入的中等规模模型。

    推荐理由:官方说明新模型以无编码器统一架构把性能接近 26B 的多模态能力压到 16GB 内存笔记本可跑,开发者可据此评估端侧部署选择。

6月5日周五
  1. Hugging Face Blog61

    NVIDIA 发布 Nemotron 3.5 Content Safety 多模态安全模型

    NVIDIA 发布 Nemotron 3.5 Content Safety,基于 Google Gemma 3 4B IT 微调,把多模态输入、自定义企业策略与可审计推理链统一到单次推理调用中,并保持 12 种语言显式训练和约 140 种语言的零样本泛化。

    推荐理由:相比 Nemotron 3,3.5 版把多模态审核、自定义策略与可审计推理链合并到一次调用中。

  2. @googleaidevs69

    Google 发布开放权重的实时音乐模型 Magenta RealTime 2(MRT2),支持 MIDI 与提示词控制,可在 MacBook 上本地运行且延迟低于 200ms。官方同时提供开放权重、开源推理引擎及一系列应用和插件,演示中还展示了用 MIDI 键盘、实时文本提示乃至手势来演奏。

    引用Google Magenta Project (@GoogleMagenta)@GoogleMagenta

    Introducing Magenta RealTime 2 (MRT2): the live music model you can play as an instrument. MRT2 offers MIDI and prompt controls, and runs natively on a MacBook with <200ms latency. Open weights. Open source inference engine. Suite of apps and plugins. Hear what it can do and try it out for yourself below 🧵 Video

    推荐理由:开放权重与开源推理引擎一并给出,读者可了解实时音乐模型在笔记本上的本地延迟表现。