LOADING
56 words
1 minute
HF TRL

HF TRL

Category: Fine-Tuning · GitHub · Docs ⭐ 19,000 · 1.12.0 · Apache-2.0 · snapshot 2026-08-31 Full profile (deep dive, contenders, matrix): [[AI Tools & Platforms Landscape]]

One-liner: The standard post-training library: SFT/GRPO/DPO trainers in the HF ecosystem

What problem does it solve?

How does it work?

When would I reach for it?

  • [[Axolotl]]
  • [[Unsloth]]
  • [[LLaMA-Factory]]

My exploration notes

Blog post checklist

  • Outline
  • Draft → Blog/posts/tech-explore/hf-trl/

Some information may be outdated