Hybrid Pooling with LLMs via Relevance Context Learning

UDC.coleccionInvestigación
UDC.conferenceTitleSIGIR 2026
UDC.departamentoCiencias da Computación e Tecnoloxías da Información
UDC.endPage1439
UDC.grupoInvInformation Retrieval Lab (IRlab)
UDC.institutoCentroCITIC - Centro de Investigación de Tecnoloxías da Información e da Comunicación
UDC.startPage1429
dc.contributor.authorOtero, David
dc.contributor.authorParapar, Javier
dc.date.accessioned2026-07-21T08:10:23Z
dc.date.available2026-07-21T08:10:23Z
dc.date.issued2026
dc.descriptionPresented at: SIGIR '26, 49th International ACM SIGIR Conference on Research and Development in Information Retrieval, July 20–24, 2026, Melbourne, Australia
dc.description.abstract[Abstract]: High-quality relevance judgements over large query sets are essential for evaluating Information Retrieval (IR) systems, yet manual annotation remains costly and time-consuming. Large Language Models (LLMs) have recently shown promise as automatic relevance assessors, but their reliability is still limited. Most existing approaches rely on zero-shot prompting or in-context learning (ICL) with a small number of labelled examples. However, standard ICL treats examples as independent instances and fails to explicitly capture the underlying relevance criteria of a topic, restricting its ability to generalise to unseen query-document pairs. To address this limitation, we introduce Relevance Context Learning (RCL), a novel framework that leverages human relevance judgements to explicitly model topic-specific relevance criteria. Rather than directly using labelled examples for in-context prediction, RCL first prompts an LLM (Instructor LLM) to analyse sets of judged query-document pairs and generate explicit narratives that describe what constitutes relevance for a given topic. These relevance narratives are then used as structured prompts to guide a second LLM (Assessor LLM) in producing relevance judgements. To evaluate RCL in a realistic data collection setting, we propose a hybrid pooling strategy in which a shallow depth-k pool from participating systems is judged by human assessors, while the remaining documents are labelled by LLMs. Experimental results demonstrate that RCL substantially outperforms zero-shot prompting and consistently improves over standard ICL. Overall, our findings indicate that transforming relevance examples into explicit, context-aware relevance narratives is a more effective way of exploiting human judgements for LLM-based IR dataset construction.
dc.description.sponsorshipAll authors acknowledge funding from the Ministry of Science, Innovation and Universities of the Government of Spain (project PID2022-137061OB-C21, MCIN/AEI/10.13039/501100011033), as well as from the Department of Education, Science, Universities, andVocational Training of the Xunta de Galicia(grantGRCED431C 2025/49). CITIC, as a center accredited for excellence within the Galician University System and a member of the CIGUS Network, receives subsidies from the Department of Education, Science, Universities, and Vocational Training of the Xunta de Galicia. Additionally, it is co-financed by the EU through the FEDER Galicia 2021-27 operational program (Ref. ED431G 2023/01).
dc.description.sponsorshipXunta de Galicia; GRC ED431C 2025/49
dc.description.sponsorshipXunta de Galicia; ED431G 2023/01
dc.identifier.citationDavid Otero and Javier Parapar. 2026. Hybrid Pooling with LLMs via Relevance Context Learning. In Proceedings of the 49th International ACM SIGIR Conference on Research and Development in Information Retrieval (SIGIR ’26), July 20–24, 2026, Melbourne, VIC, Australia. ACM, New York, NY, USA, pp. 1429-1439. https://doi.org/10.1145/3805712.3809669
dc.identifier.doi10.1145/3805712.3809669
dc.identifier.isbn979-8-4007-2599-9
dc.identifier.urihttps://hdl.handle.net/2183/48901
dc.language.isoeng
dc.publisherACM
dc.relation.projectIDinfo:eu-repo/grantAgreement/AEI/Plan Estatal de Investigación Científica, Técnica y de Innovación 2021-2023/PID2022-137061OB-C21/ES/BUSQUEDA, SELECCION Y ORGANIZACION DE CONTENIDOS PARA NECESIDADES DE INFORMACION RELACIONADAS CON LA SALUD - CONSTRUCCION DE RECURSOS Y PERSONALIZACION
dc.relation.urihttps://doi.org/10.1145/3805712.3809669
dc.rightsAttribution 4.0 Internationalen
dc.rights.accessRightsopen access
dc.rights.urihttp://creativecommons.org/licenses/by/4.0/
dc.subjectRelevance Judgements
dc.subjectInformation Retrieval Evaluation
dc.subjectLarge Language Models
dc.subjectIn-Context Learning
dc.titleHybrid Pooling with LLMs via Relevance Context Learning
dc.typeconference output
dspace.entity.typePublication
relation.isAuthorOfPublication00d04042-9b75-419e-9aab-33fd14b201af
relation.isAuthorOfPublicationfef1a9cb-e346-4e53-9811-192e144f09d0
relation.isAuthorOfPublication.latestForDiscovery00d04042-9b75-419e-9aab-33fd14b201af

Files

Original bundle

Now showing 1 - 1 of 1
Loading...
Thumbnail Image
Name:
Otero_David_2026_Hybrid_Pooling_with_LLMs.pdf
Size:
7.32 MB
Format:
Adobe Portable Document Format