TRL
Open Source
Library
Transformer reinforcement learning
About
TRL is Hugging Face's library for post-training transformer models with supervised fine-tuning, reward modeling, and preference optimization (DPO, GRPO, PPO), integrated with the Transformers and PEFT ecosystem.
Compatibility
Supported Languages
python
Details
- Category
- Library
- License
- Apache-2.0