Skip to content
Preprint

How retriever redundancy and diversity impact RAG effectiveness

Aug 2026 · 2 citations · 18 references
Computer Science

TL;DR

It is shown that duplicate redundancy and LLM paraphrasing does not significantly improve answer correctness, however, providing diverse documents is highly beneficial, improving answer correctness by 17%-47%.

Abstract

In RAG, while the retriever typically ranks documents by their individual relevance to the query, the generator instead produces an answer based on the retrieved documents as a whole. This paper investigates how redundancy and diversity from the retrieved document set impact the generator in terms of answer correctness. Previous work has provided a mix of findings: some showing that redundancy improves generation by reinforcing relevant information, others that LLM-based paraphrasing of the same content may be beneficial. Many of these studies did not control for confounding factors like whether the documents contained the exact answer or not, and if parametric knowledge plays a role. We conduct a carefully controlled experiment investigating three key scenarios of retrieved document sets: 1) Duplicate (exact copies of the same document), 2) Paraphrased (LLM rephrased versions of one document) and 3) Diverse (documents from different genres each containing relevant information in different forms). We control for which documents contain the answer in exact match or rephrased form. Evaluation is done with FictionalQA, a synthetic, fictional question-answer dataset that ensures the LLM generator prior knowledge cannot answer the question; the answer must come from retrieved documents. We show that duplicate redundancy and LLM paraphrasing does not significantly improve answer correctness. However, providing diverse documents is highly beneficial, improving answer correctness by 17%-47%. We further show this improvement is driven by diverse forms of document genre (news, blogs, etc.) alone and not a consequence of more relevant answer being available to generator. Our findings help to direct more attention to how new retrieval methods might improve RAG by catering to the generator preference for diversity in retrieval results.

View source

Similar papers

Preprint Sep 2026

A Systematic Multi-Domain Evaluation of Document Retrievers

Document retrieval is a crucial component of many modern AI systems, directly influencing their effectiveness, robustness, and fairness in downstream tasks. While recent years have seen a growing number of retrievers, comparative studies in the literature are typically limited in scope or focused on singular benchmarks...

Valentin Velev, Andreas Spitz · 0 citations
#artificial intelligence Preprint Sep 2026

Generated Query Expansion Still Helps Strong Sparse Retrieval: A Controlled Study with SPLADE-v3

Scientific queries are often brief, while relevant papers use specialized vocabulary. Generated query expansion can bridge this mismatch, but earlier work suggests that its value shrinks as the underlying retriever becomes stronger. We test the four generated formats of term lists, a pseudo-document, multiple pseudo-re...

Ryan C. Barron, C. Trotter, M. Eren et al. · 0 citations
Preprint Aug 2026

Robustness of IR Models to Collection Growth

This study empirically evaluates the robustness of an IR model to the addition of non-relevant documents by merging two collections with negligible topic overlap and finds that MDA is more effective than MDD for retrieval, whereas MDD and MDA rerankers are equally effective.

Emmanouil Georgios Lionis, Sean MacAvaney, Debasis Ganguly · 0 citations
#artificial intelligence Preprint Aug 2026

MUDDLE: Measuring Understanding of Documents under Distractor and Length Effects

In the complete markdown sweep, hard negatives lower accuracy more than length-matched random documents at both context sizes for gpt-5-mini, while random documents stay near the no-distractor baseline, and for gpt-5-mini hard negatives significantly underperform length-matched random distractors when pooled across con...

Jason Luo, Saibilila Abudukelimu, Judy Song et al. · 1 citation
#small language model Review Open access Aug 2026

Query Optimization Techniques in Retrieval-Augmented Generation (RAG) Systems

Retrieval-Augmented Generation (RAG) helps language models give more accurate answers by using information from external documents. Instead of depending only on what the model learned during training, it can search for relevant information when needed. The quality of the retrieved information is important for getting g...

Ayush Aryan, Sudhakar Ranjan · 0 citations

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.