trl

by huggingface · Developer Tools

Train transformer language models with reinforcement learning.

New0 ratings19,235 starsActive
Developer ToolsDeveloper ToolPythonJinja
Open project ↗↓ Download v1.12.0GitHub⚑ Report
trl preview

About this project

TRL - Transformers Reinforcement Learning A comprehensive library to post-train foundation models 🎉 What's New 📜 Training beyond 1M tokens: A new long context guide walks through the four things that break as sequences grow — the loss, the positions, the activations and the memory of a single GPU — and ends on an example that trains Qwen3-8B on million-token sequences on one 8-GPU node. Overview TRL is a cutting-edge library designed for post-training foundation models using advanced techniques like Supervised Fine-Tuning (SFT), Group Relative Policy Optimization (GRPO), and Direct Preference Optimization (DPO). Built on top of the 🤗 Transformers…

Technologies

PythonDockerfileMakefileJinja

Project health

Actively maintained
Last update2 days ago
Contributors465
Latest releasev1.12.0
Open issues & PRs330
LicenseApache-2.0
On GitHubsince 2020

GitHub

19,235
stars
2,965
forks
465
contributors
330
open issues & PRs
Python
language
2 days ago
last commit
View on GitHub ↗All releases ↗

Reviews

out of 5 · 0 ratings
★★★★★
0%
★★★★
0%
★★★
0%
★★
0%
0%
Sign in to write a review
No reviews yet
Be the first to review trl.

Built by

huggingface
Imported from GitHub · not yet claimed on GitPalace
View developer pageSign in with GitHub to claim

Maintain huggingface/trl? Claiming verifies admin access through your GitHub account and gives you control of this listing.

You might also like