ML-QuantSubscribe

Machine learningLLMs & Text

Stronger Models are NOT Stronger Teachers for Instruction Tuning

A study introduces Compatibility-Adjusted Reward (CAR), a new metric to evaluate the effectiveness of language models, challenging the belief that larger models are better for instruction tuning.

Featured in No. 74 on 13 Nov 2024 · 2 days after release · 17 citations today

Released
11 Nov 2024
First featured
No. 74 · 13 Nov 2024
Citations (Semantic Scholar)
17
Influential citations
0
Published in
Not yet, as far as Semantic Scholar knows
Shares when featured
19
Identifier
arXiv:2411.07133

Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).

    Type to search. Try rough volatility, LLM agents or FinGPT.

    ↑↓ move↵ openesc closeFull search page