跳到正文
@Thom_Wolf· @Thom_Wolf · X·· 26 天前AI 评分52
AI 导读

TRL v1.13 发布,这是一个用于后训练基础模型的开源 RL 训练库,本次更新聚焦长上下文训练。新版本附带 1M+ token 上下文后训练指南,并延续了速度与内存占用方面的改进。

正文

TRL v1.13 is out! our open-source RL training library to to post-train foundation models

this new release is focusing on "long context training" with a new guide on how to post-train model with 1M+ token context https://t.co/a9GQcYNS8w + various improvements on speed and memory usage as usual

check it out at https://t.co/wYHO7vqp6m

来源:@Thom_Wolf · x.com