模型精选

Hugging Face 发布 TRL v1.13,支持百万级 token 长上下文训练

TRL v1.13 is out! our open-source RL training library to to post-train foundation models this new r...

精选理由

Hugging Face 新出的 TRL 库版本,能直接训练百万级 token 的模型,比之前版本更强大,适合做长文本任务。

Hugging Face 推出开源强化学习训练库 TRL v1.13,新增长上下文训练指南,可处理超过 1M token 的数据(约等于《哈利波特》全集),同时优化了训练速度和内存使用。

原文 · Thomas Wolf

TRL v1.13 is out! our open-source RL training library to to post-train foundation models this new r...

TRL v1.13 is out! our open-source RL training library to to post-train foundation models this new release is focusing on "long context training" with a new guide on how to post-train model with 1M+ token context huggingface.co/docs/trl/long_… + various improvements on speed and memory usage as usual check it out at github.com/huggingface/trl Quentin Gallouédec @QGallouedec 1M token is basically the entire Harry Potter series ⚡️🪄 Your browser does not support the video tag. 🔗 View on Twitter 🔗 View Quoted Tweet 💬 2 🔄 2 ❤️ 15 👀 1689 📊 3 ⚡