SemiAnalysis 发布 ClusterMAX 3.0,用三个月实测 77 家 GPU 云并给出新排名,覆盖健康检查、缓存命中率、NCCL、集群上的智能体、Linux 内核 bug、安全与融资等维度。播客中三人称行业正面临 7 万亿美元 GPU 云支出、每家供应商被要求一夜之间扩大 100 倍,并讨论 H100 与托管式 RL 训练。
$7 trillion is about to get spent on GPU clouds, and every provider is being asked to get 100x bigger overnight. ClusterMAX 3.0 spent three months finding out who actually can.
Ep. 033 of SemiAnalysis Weekly is live. Jordan Nanos (@JordanNanos) sits down with Sam Harshe (@sharshe02) and Pratt Bhatt (@PrathmeshBhat19) fresh off testing 77 GPU clouds: the new rankings, health checks, cache hit rates, NCCL, agents on clusters, a Linux kernel bug, security, financing and backstops, H100s, and hosted RL training.
"It's my first time seeing sunlight today, actually, since testing began."
"This time we rank 77 providers."
"Just about everyone in the industry looks like a genius right now."
"No way. I thought you could just earn 80% margins on the inference business, right?"
"Who are you going to trust for your Vera Rubin going into next year?"
来源:@SemiAnalysis_ · x.com