模型发布 / 更新普通
TRL v1.13 发布:长上下文训练
原始标题:TRL v1.13 is out! our open-source RL training library to to post-train foundation models this new r…
内容摘要
TRL v1.13 发布!我们的开源 RL 训练库,用于对基础模型进行后训练 本次新版本聚焦"长上下文训练",新增了一份关于如何用 1M+ token 上下文对模型进行后训练的指南 https://huggingface.co/docs/trl/long_context_training ,以及一如既往的速度和内存使用方面的多项改进 详见 https://github.com/huggingface/trl
内容分类AI 模型发布与更新
内容层级普通情报
发布时间(北京时间)
本站收录时间(北京时间)
信息来源X:Thomas Wolf(Hugging Face 联创/CSO) (@Thom_Wolf)
站内情报编号intel-d868cd933adc8036c06cc99a