vix.ing · top · new · best · stats · spec

Effective Model Pruning

2025/09/30 by Wang, Yixuan, Guralnik, Dan, Akbari, Saiedeh +1
#FOS: Computer and information sciences #Machine Learning (cs.LG)

paper · doi:10.48550/arxiv.2509.25606

Abstract

We introduce Effective Model Pruning (EMP), a context-agnostic, parameter-free rule addressing a fundamental question about pruning: how many entries to keep. EMP does not prescribe how to score the parameters or prune the models; instead, it supplies a universal adaptive threshold that can be applied to any pruning criterion: weight magnitude, attention score, KAN importance score, or even feature-level signals such as image pixel, and used on structural parts or weights of the models. Given any score vector s, EMP maps s to a built-in effective number Neff which is inspired by the Inverse Simpson index of contributors. Retaining the Neff highest scoring entries and zeroing the remainder yields sparse models with performance comparable to the original dense networks across MLPs, CNNs, Transformers/LLMs, and KAN, in our experiments. By leveraging the geometry of the simplex, we derive a tight lower bound on the preserved mass seff (the sum of retained scores) over the corresponding ordered probability simplex associated with the score vector s. We further verify the effectiveness of Neff by pruning the model with a scaled threshold \beta*Neff across a variety of criteria and models. Experiments suggest that the default \beta = 1 yields a robust threshold for model pruning while \beta not equal to 1 still serves as an optional adjustment to meet specific sparsity requirements.

Citations

Related