2024

Tulu 3: Pushing Frontiers in Open Language Model Post-Training

Lambert, Nathan, Morrison, Jacob, Pyatkin, Valentina et al.

Understand

Language model post-training is applied to refine behaviors and unlock new skills across a wide range of recent language models, but open recipes for applying these techniques lag behind proprietary ones.

  • The underlying training data and recipes for post-training are simultaneously the most important pieces of the puzzle and the portion with the least transparency.
  • To bridge this gap, we introduce Tulu 3, a family of fully-open state-of-the-art post-trained models, alongside its data, code, and training recipes, serving as a comprehensive guide for modern post-training techniques.
  • Tulu 3, which builds on Llama 3.1 base models, achieves results surpassing the instruct versions of Llama 3.1, Qwen 2.5, Mistral, and even closed models such as GPT-4o-mini and Claude 3.5-Haiku.

Reading the bibliography…