Machine learningLLMs & Text
Stronger Models are NOT Stronger Teachers for Instruction Tuning
A study introduces Compatibility-Adjusted Reward (CAR), a new metric to evaluate the effectiveness of language models, challenging the belief that larger models are better for instruction tuning.
Featured in No. 74 on 13 Nov 2024 · 2 days after release · 17 citations today
- Released
- 11 Nov 2024
- First featured
- No. 74 · 13 Nov 2024
- Citations (Semantic Scholar)
- 17
- Influential citations
- 0
- Published in
- Not yet, as far as Semantic Scholar knows
- Shares when featured
- 19
- Identifier
- arXiv:2411.07133
Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).