ML-QuantSubscribe

Machine learningLLMs & Text

Adaptive Decoding via Latent Preference Optimization

Preference Optimization: Adaptive Decoding is a technique that dynamically selects the sampling temperature during language model decoding, optimizing performance across various tasks that require different temperatures.

Featured in No. 75 on 20 Nov 2024 · 6 days after release · 13 citations today

Released
14 Nov 2024
First featured
No. 75 · 20 Nov 2024
Citations (Semantic Scholar)
13
Influential citations
2
Published in
Not yet, as far as Semantic Scholar knows
Shares when featured
19
Identifier
arXiv:2411.09661

Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).

    Type to search. Try rough volatility, LLM agents or FinGPT.

    ↑↓ move↵ openesc closeFull search page