English

DRAGOn: Designing RAG On Periodically Updated Corpus

Computation and Language 2026-02-10 v3 Artificial Intelligence

Abstract

This paper introduces DRAGOn, method to design a RAG benchmark on a regularly updated corpus. It features recent reference datasets, a question generation framework, an automatic evaluation pipeline, and a public leaderboard. Specified reference datasets allow for uniform comparison of RAG systems, while newly generated dataset versions mitigate data leakage and ensure that all models are evaluated on unseen, comparable data. The pipeline for automatic question generation extracts the Knowledge Graph from the text corpus and produces multiple question-answer pairs utilizing modern LLM capabilities. A set of diverse LLM-as-Judge metrics is provided for a comprehensive model evaluation. We used Russian news outlets to form the datasets and demonstrate our methodology. We launch a public leaderboard to track the development of RAG systems and encourage community participation.

Keywords

Cite

@article{arxiv.2507.05713,
  title  = {DRAGOn: Designing RAG On Periodically Updated Corpus},
  author = {Fedor Chernogorskii and Sergei Averkiev and Liliya Kudraleeva and Zaven Martirosian and Maria Tikhonova and Valentin Malykh and Alena Fenogenova},
  journal= {arXiv preprint arXiv:2507.05713},
  year   = {2026}
}

Comments

EACL 2026

R2 v1 2026-07-01T03:50:53.432Z