EncouRAGe: Evaluating RAG Local, Fast, and Reliable

Title:EncouRAGe: Evaluating RAG Local, Fast, and Reliable

Abstract:We introduce EncouRAGe, a comprehensive Python framework designed to streamline the development and evaluation of Retrieval-Augmented Generation (RAG) systems using Large Language Models (LLMs) and Embedding Models. EncouRAGe comprises five modular and extensible components: Type Manifest, RAG Factory, Inference, Vector Store, and Metrics, facilitating flexible experimentation and extensible development. The framework emphasizes scientific reproducibility, diverse evaluation metrics, and local deployment, enabling researchers to efficiently assess datasets within RAG workflows. This paper presents implemen…

Title:EncouRAGe: Evaluating RAG Local, Fast, and Reliable

View PDF HTML (experimental)

Abstract:We introduce EncouRAGe, a comprehensive Python framework designed to streamline the development and evaluation of Retrieval-Augmented Generation (RAG) systems using Large Language Models (LLMs) and Embedding Models. EncouRAGe comprises five modular and extensible components: Type Manifest, RAG Factory, Inference, Vector Store, and Metrics, facilitating flexible experimentation and extensible development. The framework emphasizes scientific reproducibility, diverse evaluation metrics, and local deployment, enabling researchers to efficiently assess datasets within RAG workflows. This paper presents implementation details and an extensive evaluation across multiple benchmark datasets, including 25k QA pairs and over 51k documents. Our results show that RAG still underperforms compared to the Oracle Context, while Hybrid BM25 consistently achieves the best results across all four datasets. We further examine the effects of reranking, observing only marginal performance improvements accompanied by higher response latency.


Comments:	Currently under review
Subjects:	Computation and Language (cs.CL); Artificial Intelligence (cs.AI); Information Retrieval (cs.IR)
Cite as:	arXiv:2511.04696 [cs.CL]
	(or arXiv:2511.04696v1 [cs.CL] for this version)
	https://doi.org/10.48550/arXiv.2511.04696 arXiv-issued DOI via DataCite

Submission history

From: Jan Strich [view email] [v1] Fri, 31 Oct 2025 15:19:29 UTC (821 KB)

Title:EncouRAGe: Evaluating RAG Local, Fast, and Reliable

Title:EncouRAGe: Evaluating RAG Local, Fast, and Reliable

Submission history

Similar Posts