Machine learningLLMs & Text
Critical Tokens Matter: Token-Level Contrastive Estimation Enhances LLM's Reasoning Capability
Enhancing Reasoning: The cDPO method identifies and rewards 'critical tokens' that cause incorrect reasoning in Large Language Models, showing effectiveness in two popular models.
Featured in No. 77 on 4 Dec 2024 · 5 days after release · 89 citations today · published in International Conference on Machine Learning
- Released
- 29 Nov 2024
- First featured
- No. 77 · 4 Dec 2024
- Citations (Semantic Scholar)
- 89
- Influential citations
- 11
- Published in
- International Conference on Machine Learning
- Shares when featured
- 9
- Identifier
- arXiv:2411.19943
Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).