ML-QuantSubscribe

Machine learningLLMs & Text

Scaling LLM Test-Time Compute Optimally can be More Effective than Scaling Model Parameters

The research investigates enhancing Large Language Models' (LLMs) performance using more test-time computation, suggesting a compute-optimal scaling strategy based on prompt difficulty.

Featured in No. 60 on 7 Aug 2024 · 1 day after release · 2,189 citations today

Released
6 Aug 2024
First featured
No. 60 · 7 Aug 2024
Citations (Semantic Scholar)
2,189
Influential citations
153
Published in
Not yet, as far as Semantic Scholar knows
Shares when featured
214
Identifier
arXiv:2408.03314

Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).

    Type to search. Try rough volatility, LLM agents or FinGPT.

    ↑↓ move↵ openesc closeFull search page