One Shot vs. Iterative: Rethinking Pruning Strategies for Model Compression
arXiv:2508.13836v2 Announce Type: replace-cross Abstract: Pruning is a core technique for compressing neural networks to improve computational efficiency. This process is typically approached in two w