ML-QuantSubscribe

Machine learningLLMs & Text

Extreme Compression of Large Language Models via Additive Quantization

The article discusses a new algorithm that enhances the compression of large language models, providing better accuracy and is now available for future research.

Featured in No. 36 on 7 Feb 2024 · 27 days after release · 268 citations today · published in International Conference on Machine Learning

Released
11 Jan 2024
First featured
No. 36 · 7 Feb 2024
Citations (Semantic Scholar)
268
Influential citations
44
Published in
International Conference on Machine Learning
Shares when featured
45
Identifier
arXiv:2401.06118

Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).

    Type to search. Try rough volatility, LLM agents or FinGPT.

    ↑↓ move↵ openesc closeFull search page