ML-QuantSubscribe

Machine learningLLMs & Text

ShiftAddLLM: Accelerating Pretrained LLMs via Post-Training Multiplication-Less Reparameterization

ShiftAddLLM is a new method developed to speed up large language models on devices with limited resources by replacing complex multiplications with simpler operations, thus reducing memory usage and latency and enhancing model performance.

Featured in No. 59 on 31 Jul 2024 · 51 days after release · 47 citations today · published in Neural Information Processing Systems

Released
10 Jun 2024
First featured
No. 59 · 31 Jul 2024
Citations (Semantic Scholar)
47
Influential citations
2
Published in
Neural Information Processing Systems
Shares when featured
170
Identifier
arXiv:2406.05981

Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).

    Type to search. Try rough volatility, LLM agents or FinGPT.

    ↑↓ move↵ openesc closeFull search page