跳到正文

模型发布

新模型的发布、开源与迭代:大模型厂商的旗舰更新、开源权重放出、性能与价格变化的第一时间记录。

当前仅显示精选新闻
530条精选相关主题产品更新论文研究开源生态

最新精选

第 461–480 条 · 共 530 条
5月20日周三
  1. @OpenRouter79

    Google DeepMind 发布 Gemini 3.5 模型家族,称其将前沿智能与现实世界行动结合,3.5 Flash 是该家族发布的第一款模型,官方称其为面向智能体与编码的最强模型。OpenRouter 转发该消息,并附上阅读模型详情的链接。

    引用Google DeepMind (@GoogleDeepMind)@GoogleDeepMind

    Introducing Gemini 3.5: our newest family of models combining frontier intelligence with real-world action. The first release is 3.5 Flash, our strongest model yet for agents and coding 🧵

    推荐理由:Google DeepMind 公布 Gemini 3.5 家族,3.5 Flash 主打智能体与编码,可看到新模型的能力侧重。

  2. @GeminiApp69

    Gemini 3.5 Flash 发布,官方称这是其目前最快最高效的模型,可用于日常任务和多步骤创意项目,帮助应对真实世界的复杂性。

    引用Google Gemini (@GeminiApp)@GeminiApp

    Gemini 3.5 Flash is here and it's our best model yet for getting things done quickly and efficiently. Whether you need help with everyday tasks or multi-step creative projects, Gemini 3.5 Flash navigates real-world complexity to help you take action. #GoogleIO

  3. @kimmonismus82

    Gemini Omni 发布,官方介绍这是一款可从任意输入生成内容的新模型,首批聚焦视频,已在 Gemini App、Flow 和 YouTube 上线,API 支持即将推出。转发的作者 @kimmonismus 称这是真正的惊艳时刻,认为它是迈向 AGI 的世界模型,能从任何输入生成任何内容。

    引用Logan Kilpatrick (@OfficialLoganK)@OfficialLoganK

    Introducing Gemini Omni 🔮........ Omni is our new model that can create anything from any input — starting with video (think Nano Banana but for video). Available in the Gemini App, Flow, and YouTube, with API support coming soon! Video

    推荐理由:官方说明了 Gemini Omni 从任意输入生成内容的能力和首批视频场景,读者可据此了解当前可用渠道与范围。

5月19日周二
  1. @berryxia66

    Odyssey 发布 Agora-1,一个多智能体世界模型,人类与 AI 可同时进入同一模拟世界并实时互动、互相影响。官方推出可游玩的研究预览,用 Agora-1 模拟多人 GoldenEye 死亡竞赛,模型实时生成画面和声音,整个世界持续更新。

    引用Odyssey (@odysseyml)@odysseyml

    Introducing Agora-1, a multi-agent world model. Multiple participants—human or AI—can now interact inside the same world simulation, all in real-time. Try our playable research preview today, with Agora-1 simulating a multiplayer GoldenEye deathmatch! Video

    推荐理由:世界模型从单人视频生成扩展到多人实时共享模拟,读者可据此了解人机共处同一模拟世界的当前形态。

5月18日周一
5月16日周六
5月15日周五
  1. @vista866

    面壁智能发布 1.3B 参数的视觉模型 MiniCPM-V 4.6,面向消费级和移动硬件,已在 Hugging Face、GitHub 和 ModelScope 上线。该模型采用 LLaVA-UHD v4 技术,官方称将视觉编码成本降低 55%。官方还称其在多模态和 Artificial Analysis 基准上超过 Gemma4-E2B-it 和 Qwen3.5-0.8B,TTFT 为 75.7ms、比 Qwen3.5-0.8B 快 2.2 倍。原文作者表示在 Hugging Face 看到该模型论文,准备抽空测试。

    引用OpenBMB (@OpenBMB)@OpenBMB

    1/5 MiniCPM-V 4.6 (1.3B) is now live 🚀🚀 High-res visual processing, optimized for consumer-grade and mobile hardware. We’ve leveraged the latest LLaVA-UHD v4 technique to cut vision encoding costs by 55%, enabling native edge deployment with extreme efficiency. 🔥 Beats Gemma4-E2B-it and Qwen3.5-0.8B across key multimodal and Artificial Analysis benchmarks — scoring higher than Qwen3.5-0.8B using just 2.5% of its token budget. ⚡ TTFT (75.7ms) 2.2x Faster than Qwen3.5-0.8B even with 3136² high-res images. 🏗️ ~1.5x Token Throughput compared with Qwen3.5-0.8B on a single RTX 4090. Try the model here: 🤗 Hugging Face: huggingface.co/openbmb/MiniC… 💻 GitHub: github.com/OpenBMB/MiniCPM-V 🔭 Modelscope: modelscope.cn/models/OpenBMB… 🌐 Web Demo: huggingface.co/spaces/openbm… 📱 App Demo: github.com/OpenBMB/MiniCPM-V… Video

    推荐理由:官方给出 1.3B 小模型处理高分辨率图像的编码成本与吞吐数据,可供端侧多模态选型参考。