English

A Comparative Study of Retrieval Methods in Azure AI Search

Information Retrieval 2025-12-10 v1

Abstract

Increasingly, attorneys are interested in moving beyond keyword and semantic search to improve the efficiency of how they find key information during a document review task. Large language models (LLMs) are now seen as tools that attorneys can use to ask natural language questions of their data during document review to receive accurate and concise answers. This study evaluates retrieval strategies within Microsoft Azure's Retrieval-Augmented Generation (RAG) framework to identify effective approaches for Early Case Assessment (ECA) in eDiscovery. During ECA, legal teams analyze data at the outset of a matter to gain a general understanding of the data and attempt to determine key facts and risks before beginning full-scale review. In this paper, we compare the performance of Azure AI Search's keyword, semantic, vector, hybrid, and hybrid-semantic retrieval methods. We then present the accuracy, relevance, and consistency of each method's AI-generated responses. Legal practitioners can use the results of this study to enhance how they select RAG configurations in the future.

Keywords

Cite

@article{arxiv.2512.08078,
  title  = {A Comparative Study of Retrieval Methods in Azure AI Search},
  author = {Qiang Mao and Han Qin and Robert Neary and Charles Wang and Fusheng Wei and Jianping Zhang and Nathaniel Huber-Fliflet},
  journal= {arXiv preprint arXiv:2512.08078},
  year   = {2025}
}
R2 v1 2026-07-01T08:15:48.838Z