Machine learningLLMs & Text
FABLES: Evaluating faithfulness and content selection in book-length summarization
Research shows large language models (LLMs) often misrepresent events and character states when summarizing long documents, highlighting the need for improved evaluation methods.
Featured in No. 68 on 3 Oct 2024 · · 79 citations today
- Released
- 1 Apr 2024
- First featured
- No. 68 · 3 Oct 2024
- Citations (Semantic Scholar)
- 79
- Influential citations
- 7
- Published in
- Not yet, as far as Semantic Scholar knows
- Shares when featured
- 495
- Identifier
- arXiv:2404.01261
Citations and venue from Semantic Scholar (ODC-BY), refreshed weekly. Summary: Quant Letter (CC BY 4.0).