2025

O1-Pruner: Length-Harmonizing Fine-Tuning for O1-Like Reasoning Pruning

Luo, Haotian, Shen, Li, He, Haiying et al.

Understand

Recently, long-thought reasoning LLMs, such as OpenAI's O1, adopt extended reasoning processes similar to how humans ponder over complex problems.

  • This reasoning paradigm significantly enhances the model's problem-solving abilities and has achieved promising results.
  • However, long-thought reasoning process leads to a substantial increase in inference time.
  • A pressing challenge is reducing the inference overhead of long-thought LLMs while ensuring accuracy.

Reading the bibliography…