AI 导读
TRL v1.13 发布,这是一个用于后训练基础模型的开源 RL 训练库,本次更新聚焦长上下文训练。新版本附带 1M+ token 上下文后训练指南,并延续了速度与内存占用方面的改进。
正文
TRL v1.13 is out! our open-source RL training library to to post-train foundation models
this new release is focusing on "long context training" with a new guide on how to post-train model with 1M+ token context https://t.co/a9GQcYNS8w + various improvements on speed and memory usage as usual
check it out at https://t.co/wYHO7vqp6m
来源:@Thom_Wolf · x.com