Skip to content
Book Open access

Improving Ad-hoc Search Effectiveness for Conversational Information Retrieval via Model Merging

Jul 2026 · Annual International ACM SIGIR Conference on Research and Development in Information Retrieval · pp. 3860-3865 · 0 citations · 30 references
Computer Science

TL;DR

This paper introduces model merging as a training-free strategy enabling the design of a single retrieval model that operates across both ad-hoc and conversational settings with no additional fine-tuning, demonstrating that model merging significantly enhances the ad-hoc search capabilities of conversational retrievers while improving generalizability across task-specific datasets.

Abstract

Conversational information retrieval is challenging since it requires the consideration of the conversation history which potentially gives rise to topic shifts and coreference resolution across previous turns. To address these challenges, previous work mainly rely on traditional fine-tuning of ad-hoc retrievers on conversational datasets or extrapolates their generalizability through multi-tasking. However, this mainstream approach is costly—since it requires model re-training—and exhibits catastrophic forgetting, where the model loses its foundational ad-hoc retrieval performance. In this paper, we fill this gap by introducing model merging as a training-free strategy enabling the design of a single retrieval model that operates across both ad-hoc and conversational settings with no additional fine-tuning. We conduct experiments using linear and non-linear parameter-wise merging strategies—namely Model Soup and Slerp—on standard ad-hoc search and conversational retrieval datasets. Our results demonstrate that model merging significantly enhances the ad-hoc search capabilities of conversational retrievers while improving generalizability across task-specific datasets, achieving up to 15% higher NDCG@3 under zero-shot conditions.

Read PDF

Similar papers

#artificial intelligence Preprint Sep 2026

CompCQR: Compositional Query Generation for Training-Free Conversational Search

Multi-turn interactions with LLMs are becoming increasingly common in information-seeking scenarios. However, user queries are often ambiguous and context-dependent, making them ill-suited for direct use as retriever queries. Conversational query reformulation (CQR) addresses this issue by rewriting the current utteran...

Yunah Jang, Kang-il Lee, Joongbo Shin et al. · 0 citations
Conference Aug 2026

Questioning Matters: A Controlled Study of Question Generation in Conversational Image Retrieval

Conversational Image Retrieval (CIR) refines image search through multi-turn interaction, where the Questioner plays a central role in eliciting information about the user’s target. However, existing CIR systems are commonly evaluated in end-to-end settings, making it difficult to determine whether performance gains or...

Bui Tay, Son Nguyen Thanh, Phuc Nguyen Vu et al. · 0 citations
Conference Jul 2026

Conversational Query Reformulation Using Fine-Grained Retrieval and Keyword Augmentation

Conversational Query Reformulation (CQR) is an important component in Conversational Question Answering (ConvQA), where user queries are often incomplete, ambiguous, and dependent on previous dialogue turns. Recent CQR approaches have shown the effectiveness of large language models (LLMs) in generating standalone quer...

Andhika Putra Bagaskara, Arie Ardiyanti Suryani · 0 citations
#large language models Book Open access Sep 2026

Overview and Analysis of the RecSys Challenge 2026: Conversational Music Recommendation

The RecSys Challenge 2026 studies conversational music recommendation as a joint item recommendation and response generation problem: given a multi-turn dialogue, systems must retrieve relevant tracks from a large catalog and produce a grounded natural-language response. This paper presents the challenge task, dataset,...

Seungheon Doh, Sergio Oramas, B. Sguerra et al. · 0 citations
#natural language process... Preprint Sep 2026

Where to Look and What to Use: Retrieve-Localize-Generate for Long-Term Conversational Memory Question Answering

Retrieval-augmented generation (RAG) enables large language models (LLMs) to answer questions by accessing external knowledge and has been widely adopted for long-term conversational memory question answering. However, existing methods suffer from two key challenges: (1) fragmented evidence scattered across temporally...

Yi-Fan Wang, Xin-Kui Lin, Yong-Xiu Xu et al. · 0 citations
#artificial intelligence Preprint Sep 2026

When Users Don't Ask: Benchmarking Context-Driven Memory Retrieval in Conversational Agents

Large language models (LLMs) are increas- ingly deployed as long-horizon conversational agents, motivating growing interest in mem- ory systems. However, existing benchmarks primarily evaluate memory through QA-style probing rather than in-situ conversational usage. We introduce LOCOMO-CONV, a conversa- tional memory b...

Wen-Yu Chang, Yun-Nung Chen · 3 citations · ⚡1

We use cookies to run the site and, with your consent, for analytics and to show ads. See our Cookie Policy.