ML-QuantSubscribe

Machine learningLLMs & Text

FlashBack: Efficient Retrieval-Augmented Language Modeling for Fast Inference

Efficient LM: The paper introduces FlashBack, a Retrieval-Augmented Language Modeling system that enhances inference efficiency by adding retrieved documents to the context, leading to quicker inference speed and lower costs.

Featured in No. 50 on 22 May 2024 · 15 days after release · 2 citations today · published in Annual Meeting of the Association for Computational Linguistics

Released
7 May 2024
First featured
No. 50 · 22 May 2024
Citations (Semantic Scholar)
2
Influential citations
0
Published in
Annual Meeting of the Association for Computational Linguistics
Shares when featured
82
Identifier
arXiv:2405.04065

Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).

    Type to search. Try rough volatility, LLM agents or FinGPT.

    ↑↓ move↵ openesc closeFull search page