How Good are LLMs at Retrieving Documents in a Specific Domain?

doi:https://doi.org/10.1007/978-3-032-05607-8_26

How Good are LLMs at Retrieving Documents in a Specific Domain?

Authors	Nafis Tanveer Islam Zhiming Zhao
Publication date	2026
Host editors	G. De Tré S. Sotirov J. Kacprzyk G. Psaila G. Smits T. Andreasen G. Bordogna H. Legind Larsen
Book title	Flexible Query Answering Systems
Book subtitle	16th International Conference, FQAS 2025, Burgas, Bulgaria, September 11–13, 2025 : proceedings
ISBN	9783032056061
ISBN (electronic)	9783032056078
Series	Lecture Notes in Computer Science
Event	16th International Conference on Flexible Query Answering Systems, FQAS 2025
Pages (from-to)	277-288
Publisher	Cham: Springer
Organisations	Faculty of Science (FNWI) - Informatics Institute (IVI)
Abstract	Classical search engines using indexing methods in data infrastructures primarily allow keyword-based queries to retrieve content. While these indexing-based methods are highly scalable and efficient, due to a lack of an appropriate evaluation dataset and a limited understanding of semantics, they often fail to capture the user’s intent and generate incomplete responses during evaluation. This problem also extends to domain-specific search systems that utilize a Knowledge Base (KB) to access data from various research infrastructures. Research infrastructures (RIs) from the environmental and earth science domain, which encompass the study of ecosystems, biodiversity, oceanography, and climate change, generate, share, and reuse large volumes of data. While there are attempts to provide a centralized search service (https://search.envri.eu/) using Elasticsearch as a knowledge base, they also face similar challenges in understanding queries with multiple intents. To address these challenges, we proposed an automated method to curate a domain-specific evaluation dataset to analyze the capability of a search system. Furthermore, we incorporate the Retrieval of Augmented Generation (RAG), powered by Large Language Models (LLMs), for high-quality retrieval of environmental domain data using natural language queries. Our quantitative and qualitative analysis of the evaluation dataset shows that LLM-based systems for information retrieval return results with higher precision when understanding queries with multiple intents, compared to Elasticsearch-based systems.
Document type	Conference contribution
Language	English
Published at	https://doi.org/10.1007/978-3-032-05607-8_26
Other links	https://www.scopus.com/pages/publications/105016222976
Permalink to this page

Back

UvA-DARE

Digital Academic Repository

How Good are LLMs at Retrieving Documents in a Specific Domain?