跳到正文

NVIDIA 英伟达

英伟达的 AI 芯片与生态动态:GPU 新品、CUDA 生态、机器人平台与 AI 算力市场的风向。

最新精选

第 81–90 条 · 共 90 条
6月1日周一
  1. Hugging Face Blog80

    NVIDIA 发布 Cosmos 3:首个面向物理 AI 推理与动作的开源全模态模型

    NVIDIA 在 Hugging Face 上发布 Cosmos 3,这是一个面向物理 AI 的开源全模态模型。它基于 Mixture-of-Transformers 架构,将世界生成、物理推理与动作生成整合到单一模型中。

    推荐理由:NVIDIA 将世界生成、物理推理与动作生成整合到一个模型中,并给出开源模型与 Diffusers 接入方式。

5月23日周六
5月21日周四
  1. IT Home78

    英伟达 2027 财年第一财季归母净利润 583.21 亿美元,同比增长 211%

    英伟达发布 2027 财年第一财季财报,营收 816.15 亿美元、归母净利润 583.21 亿美元,同比分别增长 85% 和 211%。数据中心业务收入 752 亿美元同比增长 92%,网络业务收入 148 亿美元同比增长 199%,主要受 Blackwell 300 放量及 InfiniBand、Spectrum-X、NVLink 需求带动。

    推荐理由:财报给出数据中心收入 752 亿美元、网络业务同比增长 199% 等数据,可据此观察 AI 算力需求的节奏。

5月20日周三
  1. @berryxia65

    NVIDIA 研究员 Yukang Chen 开源了 LongLive 2.0,一套端到端长视频生成基础设施,训练与推理都以 FP4 量化和并行加速为核心。该方案在 5B 模型上达到 45.7 FPS,并支持真实视频训练、few-step 蒸馏、多 shot 训练与推理、序列并行、NVFP4 KV cache 和异步 VAE 解码部署。

    引用Yukang Chen (@yukangchen_)@yukangchen_

    🚀 Excited to release LongLive 2.0! 🎬 An end-to-end infrastructure for long video generation, with FP4 and parallelism at the core of both training and inference. ⚡45.7 FPS generation speed on 5B model⚡ ✨ LongLive 2.0 supports real-video training, few-step distillation, multi-shot training/inference, sequence-parallel acceleration, NVFP4 KV cache, and async VAE decoding deployment. 🧩 To our knowledge, this is the first open-source 4-bit long video generation infra that covers both training and inference. 🙌 Welcome to check it out, try it, and share feedback! 🔗 Code: github.com/NVlabs/LongLive 📰 Paper: huggingface.co/papers/2605.1… 🎥 Demo: nvlabs.github.io/LongLive/Lo… #LongVideoGeneration #VideoGeneration #Realtime #AIInfra #EfficientAI #FP4 #Parallel #NVIDIA Video

    推荐理由:开源方案把 FP4 量化与并行加速同时用在训练和推理,读者可据此了解长视频实时生成的技术路线。

5月19日周二
  1. @op741865

    英伟达开始交付自研通用 CPU NVIDIA Vera,主要面向长期高并发高吞吐场景,用于 Agent 编排和工具调用的中枢。作者指出模型在 GPU 上推理,而调度编排和调用工具放在该 CPU 上,密集 Agent 常驻带来的强 IO、内存和调度压力由 CPU 承担。此次交付由英伟达上门送至 Anthropic、OpenAI、xAI、OCI,其中 xAI 由马斯克接待。

    引用NVIDIA (@nvidia)@nvidia

    NVIDIA’s Ian Buck hand-delivered the first-ever NVIDIA Vera CPUs to our partners @AnthropicAI, @OpenAI, @SpaceX, and @OracleCloud. 🎉 Vera is NVIDIA's first custom CPU, purpose-built for the age of agentic AI. This is just the beginning. The road to Vera-powered systems starts here. Thank you to our partners for being on this journey with us. The best is yet to come. 💚 Video

    推荐理由:英伟达把自研 CPU Vera 定位为 Agent 编排与工具调用的中枢,交付对象透露出这类常驻负载的硬件思路。

5月15日周五
  1. @AYi_AInotes68

    Anthropic 发布一份阐述美中 AI 竞争观点的论文,主张继续收紧对华算力出口。转发该消息的博主指出,报告没提 NVIDIA 高度依赖中国市场,而 Anthropic 几乎没有中国业务,出口管制反而间接保护其闭源优势和 9000 亿美元估值。博主认为这场博弈的关键在于谁能把自己的商业模式包装成国家利益叙事。

    引用Anthropic (@AnthropicAI)@AnthropicAI

    We've published a paper that explains our views on AI competition between the US and China. The US and democratic allies hold the lead in frontier AI today. Read more on what it’ll take to keep that lead: anthropic.com/research/2028-…

    推荐理由:作者从商业利益角度拆解对华算力出口报告,点出NVIDIA与Anthropic在管制立场上的利益差异。

5月6日周三
4月28日周二
3月11日周三
  1. LMSYS Blog61

    SGLang 支持 NVIDIA Nemotron 3 Super,面向高效多智能体系统

    SGLang 宣布 Day 0 支持 NVIDIA Nemotron 3 Super。该模型为 120B 参数混合 MoE,每次前向仅激活 12B 参数,采用混合 Transformer-Mamba 架构,支持 1M token 上下文、多 token 预测和思考预算,相比上一代 Nemotron Super 1.5 吞吐提升至 5 倍。

    推荐理由:原文给出模型架构、吞吐与准确率对比和 SGLang 部署命令,读者可以据此评估在多智能体工作流中的实际可用性。