TRL v1.13 is out! our open-source RL training library to to post-train foundation models
this new release is focusing on "long context training" with a new guide on how to post-train model with 1M+ token context https://huggingface.co/docs/trl/long_context_training + various improvements on speed and memory usage as usual
check it out at https://github.com/huggingface/trl