Fetching the paper…

Greedy-layer Pruning: Speeding up Transformer Models for Natural Language Processing · Around