Fetching the paper…

Aligning Language Models with Preferences through f-divergence Minimization · Around