ML-QuantSubscribe

Machine learningLLMs & Text

s1: Simple test-time scaling

The research presents a method called budget forcing, which uses a small dataset to achieve test-time scaling and improved reasoning performance in language modeling, particularly in competition math questions.

Featured in No. 84 on 5 Feb 2025 · 5 days after release · 1,462 citations today · published in Conference on Empirical Methods in Natural Language Processing

Released
31 Jan 2025
First featured
No. 84 · 5 Feb 2025
Citations (Semantic Scholar)
1,462
Influential citations
214
Published in
Conference on Empirical Methods in Natural Language Processing
Shares when featured
220
Identifier
arXiv:2501.19393

Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).

    Type to search. Try rough volatility, LLM agents or FinGPT.

    ↑↓ move↵ openesc closeFull search page