TRL

Post-training library: SFT, DPO, GRPO and RL for LLMs.

Hugging Face’s Transformer Reinforcement Learning library with trainers for supervised fine-tuning, preference optimisation and RL with verifiable rewards — the toolkit behind many open reasoning models.

Vendor
Hugging Face
Category
Data & ML platforms
Pricing
Open source
Open source
Yes
License
Apache 2.0
Platforms
Python
Website
huggingface.co/docs/trl

Features

Best for

More data & ml platforms

Links