Fetching the paper…

Llamba: Scaling Distilled Recurrent Models for Efficient Language Processing · Around