Machine learningLLMs & Text
ShiftAddLLM: Accelerating Pretrained LLMs via Post-Training Multiplication-Less Reparameterization
ShiftAddLLM is a new method developed to speed up large language models on devices with limited resources by replacing complex multiplications with simpler operations, thus reducing memory usage and latency and enhancing model performance.
Featured in No. 59 on 31 Jul 2024 · 51 days after release · 47 citations today · published in Neural Information Processing Systems
- Released
- 10 Jun 2024
- First featured
- No. 59 · 31 Jul 2024
- Citations (Semantic Scholar)
- 47
- Influential citations
- 2
- Published in
- Neural Information Processing Systems
- Shares when featured
- 170
- Identifier
- arXiv:2406.05981
Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).