AI 导读
IFM 称其各模型均以约 20T tokens 预训练,并将发布最终权重、检查点、日志、评测、代码和数据。该实验室表示,数据在授权允许时开放,再分发受限的部分则提供详细的构建方法说明。
正文
3/ Each model was pretrained on roughly 20T tokens, according to IFM.
The lab says it is publishing final weights, checkpoints, logs, evaluations, code, data where licensing permits, and detailed construction recipes where redistribution is restricted.
来源:@kimmonismus · x.com