Machine learningLLMs & Text
FlashBack: Efficient Retrieval-Augmented Language Modeling for Fast Inference
Efficient LM: The paper introduces FlashBack, a Retrieval-Augmented Language Modeling system that enhances inference efficiency by adding retrieved documents to the context, leading to quicker inference speed and lower costs.
Featured in No. 50 on 22 May 2024 · 15 days after release · 2 citations today · published in Annual Meeting of the Association for Computational Linguistics
- Released
- 7 May 2024
- First featured
- No. 50 · 22 May 2024
- Citations (Semantic Scholar)
- 2
- Influential citations
- 0
- Published in
- Annual Meeting of the Association for Computational Linguistics
- Shares when featured
- 82
- Identifier
- arXiv:2405.04065
Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).