ML-QuantSubscribe

Machine learningLLMs & Text

Attention heads of large language models

The article delves into the reasoning processes of Large Language Models, focusing on the interpretability of attention heads, and suggests future research areas.

Featured in No. 65 on 10 Sep 2024 · 5 days after release · 99 citations today · published in Patterns

Released
5 Sep 2024
First featured
No. 65 · 10 Sep 2024
Citations (Semantic Scholar)
99
Influential citations
2
Published in
Patterns
Shares when featured
36
Identifier
arXiv:2409.03752

Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).

    Type to search. Try rough volatility, LLM agents or FinGPT.

    ↑↓ move↵ openesc closeFull search page