Machine learningLLMs & Text
Simple and Scalable Strategies to Continually Pre-train Large Language Models
The research shows that large language models can be updated efficiently with new data, saving significant computational resources and matching the performance of re-training from scratch.
Featured in No. 42 on 27 Mar 2024 · 14 days after release · 132 citations today · published in Trans. Mach. Learn. Res.
- Released
- 13 Mar 2024
- First featured
- No. 42 · 27 Mar 2024
- Citations (Semantic Scholar)
- 132
- Influential citations
- 12
- Published in
- Trans. Mach. Learn. Res.
- Shares when featured
- 1,205
- Identifier
- arXiv:2403.08763
Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).