Machine learningLLMs & Text
Low-Rank Quantization-Aware Training for LLMs
The article introduces LR-QAT, a memory-efficient training algorithm for Large Language Models (LLMs), demonstrating its effectiveness and memory efficiency over common post-training quantization methods.
Featured in No. 64 on 5 Sep 2024 · · 60 citations today
- Released
- 10 Jun 2024
- First featured
- No. 64 · 5 Sep 2024
- Citations (Semantic Scholar)
- 60
- Influential citations
- 2
- Published in
- Not yet, as far as Semantic Scholar knows
- Shares when featured
- 30
- Identifier
- arXiv:2406.06385
Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).