Nvidia 承诺向 AI 交易投入 900 亿美元
Semafor 报道称,Nvidia 承诺向 AI 交易投入 900 亿美元($90B)。该链接在 Hacker News 上被分享,目前讨论量较少,原文未披露交易的具体对象与结构。
英伟达的 AI 芯片与生态动态:GPU 新品、CUDA 生态、机器人平台与 AI 算力市场的风向。
Semafor 报道称,Nvidia 承诺向 AI 交易投入 900 亿美元($90B)。该链接在 Hacker News 上被分享,目前讨论量较少,原文未披露交易的具体对象与结构。
英伟达发布 2027 财年第一财季财报,营收 816.15 亿美元、归母净利润 583.21 亿美元,同比分别增长 85% 和 211%。数据中心业务收入 752 亿美元同比增长 92%,网络业务收入 148 亿美元同比增长 199%,主要受 Blackwell 300 放量及 InfiniBand、Spectrum-X、NVLink 需求带动。
推荐理由:财报给出数据中心收入 752 亿美元、网络业务同比增长 199% 等数据,可据此观察 AI 算力需求的节奏。
🚀 Excited to release LongLive 2.0! 🎬 An end-to-end infrastructure for long video generation, with FP4 and parallelism at the core of both training and inference. ⚡45.7 FPS generation speed on 5B model⚡ ✨ LongLive 2.0 supports real-video training, few-step distillation, multi-shot training/inference, sequence-parallel acceleration, NVFP4 KV cache, and async VAE decoding deployment. 🧩 To our knowledge, this is the first open-source 4-bit long video generation infra that covers both training and inference. 🙌 Welcome to check it out, try it, and share feedback! 🔗 Code: github.com/NVlabs/LongLive 📰 Paper: huggingface.co/papers/2605.1… 🎥 Demo: nvlabs.github.io/LongLive/Lo… #LongVideoGeneration #VideoGeneration #Realtime #AIInfra #EfficientAI #FP4 #Parallel #NVIDIA Video
推荐理由:开源方案把 FP4 量化与并行加速同时用在训练和推理,读者可据此了解长视频实时生成的技术路线。




NVIDIA’s Ian Buck hand-delivered the first-ever NVIDIA Vera CPUs to our partners @AnthropicAI, @OpenAI, @SpaceX, and @OracleCloud. 🎉 Vera is NVIDIA's first custom CPU, purpose-built for the age of agentic AI. This is just the beginning. The road to Vera-powered systems starts here. Thank you to our partners for being on this journey with us. The best is yet to come. 💚 Video
推荐理由:英伟达把自研 CPU Vera 定位为 Agent 编排与工具调用的中枢,交付对象透露出这类常驻负载的硬件思路。
We've published a paper that explains our views on AI competition between the US and China. The US and democratic allies hold the lead in frontier AI today. Read more on what it’ll take to keep that lead: anthropic.com/research/2028-…
推荐理由:作者从商业利益角度拆解对华算力出口报告,点出NVIDIA与Anthropic在管制立场上的利益差异。
Anthropic 宣布与 SpaceX 达成合作,将使用其 Colossus 1 数据中心的全部算力,本月内获得超过 300 兆瓦、逾 220,000 块 NVIDIA GPU 的新增容量。
推荐理由:Anthropic 列出了 SpaceX 及多家云厂商的算力来源,读者可据此理解其近期上调 Claude 用量上限的算力基础。
NVIDIA 发布开源全能多模态理解模型 Nemotron 3 Nano Omni 30B-A3B,在 OCRBenchV2-En 得分 65.8、MMLongBench-Doc 57.5,并新增音频与视频理解能力。
推荐理由:原文列出文档、视频与音频榜单成绩及吞吐对比,读者可据此判断该开源全能模型的能力边界。
SGLang 宣布 Day 0 支持 NVIDIA Nemotron 3 Super。该模型为 120B 参数混合 MoE,每次前向仅激活 12B 参数,采用混合 Transformer-Mamba 架构,支持 1M token 上下文、多 token 预测和思考预算,相比上一代 Nemotron Super 1.5 吞吐提升至 5 倍。
推荐理由:原文给出模型架构、吞吐与准确率对比和 SGLang 部署命令,读者可以据此评估在多智能体工作流中的实际可用性。