据 The Information 报道,字节跳动正在研发一款模仿 Groq LPU 架构的推理芯片,Nvidia 曾于 12 月为这一架构支付约 20B 美元的授权费。LPU 将模型保存在片上 SRAM 中并跳过 HBM,而 HBM 正是美国对华出口限制最严的部件。字节的存储合作方 InnoStar 在台积电成熟制程代工,同样不在管制范围内,这些选择都绕开了美国的出口限制。
ByteDance is reportedly building its own inference chip modeled on Groq's LPU, the same architecture Nvidia paid roughly $20B to license in December.
The LPU keeps the model in on-chip SRAM and skips high-bandwidth memory. HBM is the component the US restricts most tightly for export to China. ByteDance's memory partner InnoStar fabs at TSMC's mature nodes, which also sit outside the controls.
Each of those choices routes around a US restriction. What's left is the architecture Nvidia just spent $20B to own.
China is increasingly moving toward developing its own chips and is succeeding in becoming ever more independent of the USA.
That is truly impressive.
Source: The Information.
来源:@kimmonismus · x.com