alignment-handbook

Robust recipes to align language models with human and AI preferences

What is alignment-handbook?

Robust recipes to align language models with human and AI preferences