arXiv · 2503.01542
Revisiting Large Language Model Pruning using Neuron Semantic Attribution
Abstract
Model pruning technique is vital for accelerating large language models by reducing their size and computational requirements. However, the generalizability of existing pruning methods across diverse datasets and tasks remains unclear. Thus, we conduct extensive evaluations on 24 datasets and 4 tasks using popular pruning methods. Based on these evaluations, we find and then investigate that calibration set greatly affect the performance of pruning methods. In addition, we surprisingly find a significant performance drop of existing pruning methods in sentiment classification tasks. To understand the link between performance drop and pruned neurons, we propose Neuron Semantic Attribution, which learns to associate each neuron with specific semantics. This method first makes the unpruned neurons of LLMs explainable.
Explore related subjects
Keep this discovery
Explore connections, maps & timelines
Yizhuo Ding, Xinwei Sun, Yanwei Fu, Guosheng Hu. 2025-03-03. Revisiting Large Language Model Pruning using Neuron Semantic Attribution. https://arxiv.org/abs/2503.01542
Cite the original work for its findings. Save a collection to share your selection of sources.