Accommodate Knowledge Conflicts in Retrieval-augmented LLMs: Towards Reliable Response Generation in the Wild

📅 2025-04-17
📈 Citations: 0
Influential: 0
📄 PDF
🤖 AI Summary
In retrieval-augmented generation (RAG), conflicts between large language models’ (LLMs’) internal knowledge and externally retrieved information induce unreliable responses, yet the underlying uncertainty dynamics remain poorly understood. Method: This paper models such conflicts through an information-theoretic lens, revealing—for the first time—the anomalous drop in LLM confidence under knowledge ambiguity. Building on this insight, we propose Swin-VIB: a cascaded framework grounded in the variational information bottleneck (VIB) that adaptively filters and injects retrieved content while jointly optimizing LLM preference modeling and response generation. Contribution/Results: Evaluated across single-choice QA, open-ended QA, and standard RAG benchmarks, Swin-VIB significantly improves response reliability—achieving ≥7.54% absolute accuracy gain over the strongest baseline in single-choice QA. Our work establishes a novel, conflict-aware paradigm for trustworthy RAG.

Technology Category

Application Category

📝 Abstract
The proliferation of large language models (LLMs) has significantly advanced information retrieval systems, particularly in response generation (RG). Unfortunately, LLMs often face knowledge conflicts between internal memory and retrievaled external information, arising from misinformation, biases, or outdated knowledge. These conflicts undermine response reliability and introduce uncertainty in decision-making. In this work, we analyze how LLMs navigate knowledge conflicts from an information-theoretic perspective and reveal that when conflicting and supplementary information exhibit significant differences, LLMs confidently resolve their preferences. However, when the distinction is ambiguous, LLMs experience heightened uncertainty. Based on this insight, we propose Swin-VIB, a novel framework that integrates a pipeline of variational information bottleneck models into adaptive augmentation of retrieved information and guiding LLM preference in response generation. Extensive experiments on single-choice, open-ended question-answering (QA), and retrieval augmented generation (RAG) validate our theoretical findings and demonstrate the efficacy of Swin-VIB. Notably, our method improves single-choice task accuracy by at least 7.54% over competitive baselines.
Problem

Research questions and friction points this paper is trying to address.

Resolving knowledge conflicts in retrieval-augmented LLMs
Improving response reliability in dynamic information environments
Reducing uncertainty in LLM decision-making processes
Innovation

Methods, ideas, or system contributions that make the work stand out.

Variational information bottleneck models integration
Adaptive augmentation of retrieved information
Guiding LLM preference in response generation
💼 Related Jobs
No related jobs found.
J
Jiatai Wang
College of Computer Science, Nankai University, Tianjin, China
Z
Zhiwei Xu
Haihe Lab of ITAI, Tianjin, China
D
Di Jin
Meta AI, USA
X
Xuewen Yang
InnoPeak Technology, Inc, Palo Alto, USA
T
Tao Li
College of Computer Science, Nankai University, Tianjin, China