Fetching the paper…

Train Flat, Then Compress: Sharpness-Aware Minimization Learns More Compressible Models · Around