Preparing for system design interviews?  Try bugzed.com →

TRL

JSON →
library 0.29.1 ·python
verified Jun 30, 2026 install stale

Hugging Face library for post-training LLMs: SFT, DPO, GRPO, PPO, reward modeling. Current version is 0.29.1 (Mar 2026). Requires Python >=3.10. Extremely high API churn — major parameter renames across versions. tokenizer= renamed to processing_class= in 0.12. Still pre-1.0 (Development Status: Pre-Alpha).

total hits 38
actors 8 distinct systems
last hit 8d ago human
Amazonbot
4
ByteDance
4
ClaudeBot
4
GPTBot
3
OAI-SearchBot
3
ChatGPT-User
2
Search engines
7
Humans
8

top countries 🇺🇸 United States · 🇸🇬 Singapore · 🇨🇦 Canada · VE · VN