跳到正文
@MiniMax_AI· @MiniMax_AI · X·· 2026-08-25AI 评分39
AI 导读

NVIDIA SANA 团队在 MiniMax H3 上的 Sol Engine 工作将单张 GB200 上 10s 768p 视频生成延迟从 414s 降至 14.93s,实现 27.7 倍加速。

正文

The real breakthrough in the NVIDIA SANA team’s Sol Engine work on MiniMax H3!

By splitting generation into a 4-step low-res H3 draft and a 3-step LTX refinement pass at target resolution with Sol-Attn, they’ve crushed 10s 768p latency on a single GB200 from 414s down to 14.93s (27.7x speedup). Replacing heavy VAE decodes with TAEH3/TAEHV while holding the latents stable for refinement is a masterclass in co-designing sampling topology with hardware kernel acceleration.

When inference latency collapses this dramatically, unit economics fundamentally shift: a single node can suddenly serve 378K videos a month at 97%+ GPU margins. This is how high-fidelity AI video moves from asynchronous batch rendering to near-instant, interactive infrastructure. Huge respect to the team for setting a new engineering bar for our open-weights ecosystem! 🫡🩵

来源:@MiniMax_AI · x.com