Machine learningLLMs & Text
Extreme Compression of Large Language Models via Additive Quantization
The article discusses a new algorithm that enhances the compression of large language models, providing better accuracy and is now available for future research.
Featured in No. 36 on 7 Feb 2024 · 27 days after release · 268 citations today · published in International Conference on Machine Learning
- Released
- 11 Jan 2024
- First featured
- No. 36 · 7 Feb 2024
- Citations (Semantic Scholar)
- 268
- Influential citations
- 44
- Published in
- International Conference on Machine Learning
- Shares when featured
- 45
- Identifier
- arXiv:2401.06118
Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).