Fetching the paper…

Minitron-SSM: Efficient Hybrid Language Model Compression through Group-Aware SSM Pruning · Around