跳到正文

NVIDIA 英伟达

英伟达的 AI 芯片与生态动态:GPU 新品、CUDA 生态、机器人平台与 AI 算力市场的风向。

165条精选相关主题部署工程行业动态具身智能

最新精选

第 141–160 条 · 共 165 条
6月8日周一
  1. Hugging Face Blog61

    OpenEnv 转为委员会共同治理,成为开源智能体 RL 的互操作层

    OpenEnv 宣布由 Meta-PyTorch、Reflection、Unsloth、Modal、Prime Intellect、Nvidia、Mercor、Fleet AI、Microsoft、Hugging Face 和 RadixArk 组成的委员会共同协调,项目地址迁移至 huggingface/OpenEnv。

    推荐理由:OpenEnv 由多家机构组成的委员会共同协调,并明确为 RL 环境的互操作层,读者可了解开源智能体训练的协作与协议设计。

6月6日周六
  1. IT Home77

    谷歌每月向 SpaceX 支付 9.2 亿美元租用 AI 算力,含约 11 万张英伟达 GPU

    谷歌与 SpaceX 达成云计算合作,计划自 2026 年 10 月起至 2029 年 6 月每月支付 9.2 亿美元租用数据中心算力。租赁内容涵盖至少 11 万张英伟达 GPU、CPU 等芯片对应的计算能力,主要面向训练和推理等 AI 高密度场景。华尔街日报认为,这既能缓解谷歌的算力供应紧张与扩容周期压力,也为 SpaceX 的 AI 业务新增一条收入来源,为其 IPO 提供叙事筹码。

    推荐理由:协议披露了按月支付的算力租赁规模与起止时间,读者可据此观察谷歌的算力缺口与 SpaceX 的 AI 收入布局。

6月5日周五
  1. Hugging Face Blog61

    NVIDIA 发布 Nemotron 3.5 Content Safety 多模态安全模型

    NVIDIA 发布 Nemotron 3.5 Content Safety,基于 Google Gemma 3 4B IT 微调,把多模态输入、自定义企业策略与可审计推理链统一到单次推理调用中,并保持 12 种语言显式训练和约 140 种语言的零样本泛化。

    推荐理由:相比 Nemotron 3,3.5 版把多模态审核、自定义策略与可审计推理链合并到一次调用中。

6月4日周四
  1. @kimmonismus69

    Artificial Analysis 发布了对 NVIDIA Nemotron 3 Ultra 的评测,其智能指数为 47.7,领先 Gemma 4 31B 的 39.2、Nemotron 3 Super 的 36.0 和 gpt-oss-120b 的 33.3,但低于 Kimi K2.6 的 53.9。该模型约 5500 亿总参数、550 亿激活参数,评测使用 NVIDIA 推荐的 NVFP4 权重,比 BF16 测试小幅下降(48.2 对 47.7)。它在 BlackBox AI 上以每秒超 400 output tokens 的速度提供,略快于 gpt-oss-120b,但体量超过后者 4 倍以上。

    引用Artificial Analysis (@ArtificialAnlys)@ArtificialAnlys

    NVIDIA has just released Nemotron 3 Ultra, the new most intelligent US open weights model, with leading speed for its intelligence Nemotron 3 Ultra scores 47.7 on the Artificial Analysis Intelligence Index, well ahead of the next strongest US open weights models, Gemma 4 31B (39.2), Nemotron 3 Super (36.0) and gpt-oss-120b (33.3), but behind the Chinese-led open weights frontier (Kimi K2.6 at 53.9). We partnered with @NVIDIA to evaluate this model for intelligence and speed ahead of its public release. These figures use the final NVFP4 weights that NVIDIA recommends for inference, but our tests show minimal intelligence impact compared to BF16 testing, with higher precision resulting in an Artificial Analysis Intelligence Index score of 48.2 vs. the NVFP4 score of 47.7. Key Takeaways: ➤ Nemotron 3 Ultra leads in speed for its intelligence: through BlackBox AI ahead of release, Nemotron 3 Ultra is served at over 400 output tokens per second - this is slightly faster than the typical serving speed of gpt-oss-120b despite being >4X larger, and comes with significantly greater intelligence ➤ Largest Nemotron 3 model so far: with approximately 550 billion total parameters and 55 billion active, Nemotron 3 Ultra is significantly larger than its siblings and is the largest and most intelligent US open weights model release ever ➤ Nemotron 3 Ultra is the leading US open weights model on the Artificial Analysis Intelligence and Agentic Indexes by far, but Gemma 4 31B scores ~1 point higher on the Coding Index (comprised of Terminal-Bench Hard and SciCode)

    推荐理由:Artificial Analysis 的评测让读者能横向比较 Nemotron 3 Ultra 与美国及中国开源权重模型的智能与速度表现。

  2. @kimmonismus81

    NVIDIA 发布 Nemotron 3 Ultra,一款完全开源的 550B MoE 模型,激活参数 55B,权重、训练数据与完整配方全部公开。该模型采用混合 Mamba-Attention MoE 架构,NVIDIA 称其在长输出智能体任务上的吞吐量约为同类开源模型的 6 倍,同时保持相同准确率。

    引用NVIDIA AI (@NVIDIAAI)@NVIDIAAI

    Today we're shipping Nemotron 3 Ultra. A 550B MoE frontier-intelligence open model built for long-running agents. It delivers 5x faster inference and lowers the cost of complex agentic tasks by up to 30% versus other open frontier models. Video

    推荐理由:原文给出 550B 开源模型的权重、训练数据与完整配方,并说明其在长任务智能体上的吞吐表现,读者可据此判断开源前沿模型的可复现程度。

  3. Hugging Face Blog66

    NVIDIA Nemotron 3.5 ASR 如何微调适配特定语言、领域或口音

    NVIDIA 在 Hugging Face 博客给出 Nemotron 3.5 ASR 的微调流程,用约 2000 小时希腊语和保加利亚语数据微调后,在 80ms 流式设置下 WER 分别从 35 降到 24、从 22 降到 15。

    推荐理由:文章概述了从基础检查点微调流式多语种 ASR 的流程,并用希腊语和保加利亚语的 WER 变化说明微调对低资源语言的效果。

6月3日周三
  1. AI前线 · 微信公众号83

    Alphabet 启动 800 亿美元股权融资,加码 AI 基础设施

    Alphabet 当地时间 6 月 1 日公布总额 800 亿美元的股权融资计划,资金用于 AI 底层基础设施与全球算力集群扩建。融资由 300 亿美元公开发行、400 亿美元 ATM 持续增发计划和向伯克希尔·哈撒韦定向配售 100 亿美元三部分组成,后者的 A 类普通股较 Alphabet 收盘价 376 美元约有 6% 折扣。

    推荐理由:账面现金超千亿仍启动800亿美元增发,可作为观察AI基建投入如何改变科技巨头财务结构的样本。

6月2日周二
  1. @kimmonismus67

    NVIDIA 发布 DGX Station for Windows 桌面级 AI 超算,搭载 GB300 Grace Blackwell Ultra 桌面超级芯片,最高 748GB 一致内存、20 petaflops FP4 算力,可本地运行最高 1 万亿参数模型,Q4 出货。

    引用NVIDIA Newsroom (@nvidianewsroom)@nvidianewsroom

    Introducing NVIDIA DGX Station for Windows, the world's most powerful deskside AI supercomputer with Windows powered by NVIDIA GB300. ✅ Run frontier AI models with up to 1 trillion parameters locally ✅ Build and run secure AI agents on Windows with NVIDIA OpenShell ✅ Built by @ASUS, @Dell, @GIGABYTE, @HP, @msigaming, and @Supermicro #NVIDIAGTC nvda.ws/3RIkjpc

    推荐理由:原文列出 GB300 桌面超级芯片的完整规格与出货时间,可据此判断本地运行前沿模型与智能体的硬件门槛。

6月1日周一
  1. @kimmonismus78

    NVIDIA 在 GTC Taipei 开源物理 AI 模型 Cosmos 3,作者称其为首个完全开放的 omnimodel,可理解现实世界、预测后续变化并生成机器人动作。该模型开放权重、代码与数据集,NVIDIA 同步发布 Super(32B)与 Nano(8B)两个版本。

    引用NVIDIA AI (@NVIDIAAI)@NVIDIAAI

    Introducing Cosmos 3: Our latest frontier model for Physical AI Cosmos 3 is the world’s first fully open omnimodel with native vision reasoning, world and action generation. Today we’re releasing Super (32B) and Nano (8B) variants. Video

    推荐理由:英伟达开源 Cosmos 3 的权重、代码与数据集,读者可了解物理 AI 全开放模型的两种规模与开放范围。

  2. Hugging Face Blog80

    NVIDIA 发布 Cosmos 3:首个面向物理 AI 推理与动作的开源全模态模型

    NVIDIA 在 Hugging Face 上发布 Cosmos 3,这是一个面向物理 AI 的开源全模态模型。它基于 Mixture-of-Transformers 架构,将世界生成、物理推理与动作生成整合到单一模型中。

    推荐理由:NVIDIA 将世界生成、物理推理与动作生成整合到一个模型中,并给出开源模型与 Diffusers 接入方式。

5月23日周六
5月21日周四
  1. IT Home78

    英伟达 2027 财年第一财季归母净利润 583.21 亿美元,同比增长 211%

    英伟达发布 2027 财年第一财季财报,营收 816.15 亿美元、归母净利润 583.21 亿美元,同比分别增长 85% 和 211%。数据中心业务收入 752 亿美元同比增长 92%,网络业务收入 148 亿美元同比增长 199%,主要受 Blackwell 300 放量及 InfiniBand、Spectrum-X、NVLink 需求带动。

    推荐理由:财报给出数据中心收入 752 亿美元、网络业务同比增长 199% 等数据,可据此观察 AI 算力需求的节奏。

5月20日周三
  1. @berryxia65

    NVIDIA 研究员 Yukang Chen 开源了 LongLive 2.0,一套端到端长视频生成基础设施,训练与推理都以 FP4 量化和并行加速为核心。该方案在 5B 模型上达到 45.7 FPS,并支持真实视频训练、few-step 蒸馏、多 shot 训练与推理、序列并行、NVFP4 KV cache 和异步 VAE 解码部署。

    引用Yukang Chen (@yukangchen_)@yukangchen_

    🚀 Excited to release LongLive 2.0! 🎬 An end-to-end infrastructure for long video generation, with FP4 and parallelism at the core of both training and inference. ⚡45.7 FPS generation speed on 5B model⚡ ✨ LongLive 2.0 supports real-video training, few-step distillation, multi-shot training/inference, sequence-parallel acceleration, NVFP4 KV cache, and async VAE decoding deployment. 🧩 To our knowledge, this is the first open-source 4-bit long video generation infra that covers both training and inference. 🙌 Welcome to check it out, try it, and share feedback! 🔗 Code: github.com/NVlabs/LongLive 📰 Paper: huggingface.co/papers/2605.1… 🎥 Demo: nvlabs.github.io/LongLive/Lo… #LongVideoGeneration #VideoGeneration #Realtime #AIInfra #EfficientAI #FP4 #Parallel #NVIDIA Video

    推荐理由:开源方案把 FP4 量化与并行加速同时用在训练和推理,读者可据此了解长视频实时生成的技术路线。

5月19日周二
  1. @op741865

    英伟达开始交付自研通用 CPU NVIDIA Vera,主要面向长期高并发高吞吐场景,用于 Agent 编排和工具调用的中枢。作者指出模型在 GPU 上推理,而调度编排和调用工具放在该 CPU 上,密集 Agent 常驻带来的强 IO、内存和调度压力由 CPU 承担。此次交付由英伟达上门送至 Anthropic、OpenAI、xAI、OCI,其中 xAI 由马斯克接待。

    引用NVIDIA (@nvidia)@nvidia

    NVIDIA’s Ian Buck hand-delivered the first-ever NVIDIA Vera CPUs to our partners @AnthropicAI, @OpenAI, @SpaceX, and @OracleCloud. 🎉 Vera is NVIDIA's first custom CPU, purpose-built for the age of agentic AI. This is just the beginning. The road to Vera-powered systems starts here. Thank you to our partners for being on this journey with us. The best is yet to come. 💚 Video

    推荐理由:英伟达把自研 CPU Vera 定位为 Agent 编排与工具调用的中枢,交付对象透露出这类常驻负载的硬件思路。

5月15日周五
  1. @AYi_AInotes68

    Anthropic 发布一份阐述美中 AI 竞争观点的论文,主张继续收紧对华算力出口。转发该消息的博主指出,报告没提 NVIDIA 高度依赖中国市场,而 Anthropic 几乎没有中国业务,出口管制反而间接保护其闭源优势和 9000 亿美元估值。博主认为这场博弈的关键在于谁能把自己的商业模式包装成国家利益叙事。

    引用Anthropic (@AnthropicAI)@AnthropicAI

    We've published a paper that explains our views on AI competition between the US and China. The US and democratic allies hold the lead in frontier AI today. Read more on what it’ll take to keep that lead: anthropic.com/research/2028-…

    推荐理由:作者从商业利益角度拆解对华算力出口报告,点出NVIDIA与Anthropic在管制立场上的利益差异。

5月6日周三
4月28日周二