Fetching the paper…

Shortened LLaMA: Depth Pruning for Large Language Models with Comparison of Retraining Methods · Around