alignment-handbook

by huggingface · AI

Robust recipes to align language models with human and AI preferences

New0 ratings5,674 starsActive
AIAI ApplicationPythonShell
Open project ↗GitHub⚑ Report
alignment-handbook preview

About this project

🤗 Models & Datasets 📃 Technical Report The Alignment Handbook Robust recipes to continue pretraining and to align language models with human and AI preferences. What is this? Just one year ago, chatbots were out of fashion and most people hadn't heard about techniques like Reinforcement Learning from Human Feedback (RLHF) to align language models with human preferences. Then, OpenAI broke the internet with ChatGPT and Meta followed suit by releasing the Llama series of language models which enabled the ML community to build their very own capable chatbots. This has led to a rich ecosystem of datasets and models that have mostly focused on teaching language models to follow…

Technologies

ShellPythonMakefilellmtransformersrlhf

Project health

Low recent activity
Last update3 months ago
Contributors31
Latest release
Open issues & PRs98
LicenseApache-2.0
On GitHubsince 2023

GitHub

5,674
stars
490
forks
31
contributors
98
open issues & PRs
Python
language
3 months ago
last commit
View on GitHub ↗

Reviews

out of 5 · 0 ratings
★★★★★
0%
★★★★
0%
★★★
0%
★★
0%
0%
Sign in to write a review
No reviews yet
Be the first to review alignment-handbook.

Built by

huggingface
Imported from GitHub · not yet claimed on GitPalace
View developer pageSign in with GitHub to claim

Maintain huggingface/alignment-handbook? Claiming verifies admin access through your GitHub account and gives you control of this listing.

You might also like