Machine learningLLMs & Text
Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters
The research investigates enhancing Large Language Models' (LLMs) performance using more test-time computation, suggesting a compute-optimal scaling strategy based on prompt difficulty.
Featured in No. 60 on 7 Aug 2024 · 1 day after release · 2,189 citations today
- Released
- 6 Aug 2024
- First featured
- No. 60 · 7 Aug 2024
- Citations (Semantic Scholar)
- 2,189
- Influential citations
- 153
- Published in
- Not yet, as far as Semantic Scholar knows
- Shares when featured
- 214
- Identifier
- arXiv:2408.03314
Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).