---
title: RAGGED: Towards Informed Design of Scalable and Stable RAG Systems
url: https://www.ml-quant.com/papers/arxiv/2403.09040/
site: ML-Quant (https://www.ml-quant.com)
updated: 2026-09-26
license: Summaries CC BY 4.0; links go to the original sources
index: https://www.ml-quant.com/llms.txt
identifier: arXiv:2403.09040
source_url: https://arxiv.org/abs/2403.09040
featured: 2024-08-15
citations: 24
topic: LLMs & Text
---


# RAGGED: Towards Informed Design of Scalable and Stable RAG Systems

RAGGED, a new framework, optimizes language models for document-based question answering by analyzing Retrieval-augmented generation configurations.

- Source: https://arxiv.org/abs/2403.09040
- Identifier: arXiv:2403.09040
- Released: 2024-03-14
- First featured: Quant Letter No. 61 (2024-08-15): https://www.ml-quant.com/issues/2024-08-15/
- Citations (Semantic Scholar): 24
- Published in: International Conference on Machine Learning
- Topic: LLMs & Text

## Related

- [RAFT: Adapting Language Model to Domain Specific RAG](https://www.ml-quant.com/papers/arxiv/2403.10131/): Retrieval Augmented FineTuning (RAFT) is a new training method that enhances large language models' ability to answer domain-specific questions by training them to ignore irrelevant documents and cite relevant ones.
- [HybridRAG: Integrating Knowledge Graphs and Vector Retrieval Augmented Generation for Efficient Information Extraction](https://www.ml-quant.com/papers/arxiv/2408.04948/): Q&A Systems for Financial Data: HybridRAG, a new method combining Knowledge Graphs and VectorRAG techniques, improves question-answer systems for extracting information from financial documents, offering better accuracy and answer generation.
- [ChatQA 2: Bridging the Gap to Proprietary LLMs in Long Context and RAG Capabilities](https://www.ml-quant.com/papers/arxiv/2407.14482/): Bridging the Gap: ChatQA 2 is a model that improves long-context understanding and retrieval-augmented generation, matching the accuracy of top proprietary models.
- [KnowPO: Knowledge-aware Preference Optimization for Controllable Knowledge Selection in Retrieval-Augmented Language Models](https://www.ml-quant.com/papers/arxiv/2408.03297/): The study introduces a Knowledge-aware Preference Optimization method to improve large language models' knowledge selection, showing enhanced performance in managing knowledge conflicts and robust generalization across different datasets.
- [FACTS About Building Retrieval Augmented Generation-based Chatbots](https://www.ml-quant.com/papers/arxiv/2407.07858/): The article introduces the FACTS framework for developing Retrieval Augmented Generation (RAG)-based chatbots, and presents empirical results on the balance between accuracy and latency in large and small LLMs.
- [From RAGs to rich parameters: Probing how language models utilize external knowledge over parametric information for factual queries](https://www.ml-quant.com/papers/arxiv/2406.12824/): Retrieval Augmented Generation (RAG) enhances language models' reasoning abilities using external context, but models tend to rely heavily on this context and less on their parametric memory.
