Machine learningLLMs & Text
A Simple and Effective Pruning Approach for Large Language Models
Wanda, a new method, efficiently prunes weights in Large Language Models without retraining, offering a more efficient approach to inducing sparsity in pretrained models.
Featured in No. 48 on 8 May 2024 · · 981 citations today · published in International Conference on Learning Representations
- Released
- 20 Jun 2023
- First featured
- No. 48 · 8 May 2024
- Citations (Semantic Scholar)
- 981
- Influential citations
- 198
- Published in
- International Conference on Learning Representations
- Shares when featured
- 721
- Identifier
- arXiv:2306.11695
Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).