2023

Orca 2: Teaching Small Language Models How to Reason

Mitra, Arindam, Del Corro, Luciano, Mahajan, Shweti et al.

Understand

Orca 1 learns from rich signals, such as explanation traces, allowing it to outperform conventional instruction-tuned models on benchmarks like BigBench Hard and AGIEval.

  • In Orca 2, we continue exploring how improved training signals can enhance smaller LMs' reasoning abilities.
  • Research on training small LMs has often relied on imitation learning to replicate the output of more capable models.
  • We contend that excessive emphasis on imitation may restrict the potential of smaller models.

Reading the bibliography…