QUMem: Personalized Memory for Query-Conditioned User-State Inference in LLM Agents

📅 2026-08-17
📈 Citations: 0
Influential: 0
📄 PDF
🤖 AI Summary
This study addresses the challenges of rigid memory boundaries, information coupling, and retrieval difficulties in capturing user state evolution within LLM agent personalization. We propose QUMem, a framework that employs semantic continuous segmentation and typological decoupling to enable independent storage of factual, preference-based, and insight memories. Furthermore, it introduces a three-stage multi-agent collaborative planning mechanism that supports dynamic user state reasoning under query conditions to optimize retrieval. Experimental results demonstrate that QUMem achieves state-of-the-art performance on both PersonaMem and KnowU-Bench benchmarks, effectively validating its superiority in long-term personalization tasks.
📝 Abstract
Large language model (LLM) agents increasingly use external memory systems to support personalization by drawing on long and evolving interaction histories, in which user preferences may be distributed across time, change with context, and conflict with earlier evidence. However, existing systems face three limitations: fixed-turn, fixed-token, or session-based boundaries can mix unrelated dialogue or split an event from its causes, decisions, and outcomes; storing multiple pieces of user information from the same interaction as a single memory binds together items that serve different functions and should be independently retrievable; and treating the current task as a single top-$k$ retrieval query can return fragments that are individually relevant but fail to jointly capture preference evolution, temporal validity, and contextual applicability. We introduce \textsc{QUMem}, a structured memory framework for query-conditioned user-state inference. \textsc{QUMem} first segments interaction histories into variable-length episodes according to semantic continuity, then decomposes each episode into independently retrievable factual, preference, and transferable insight memories while preserving temporal positions and source evidence. At inference time, three sequential agents identify task-specific information needs, plan multi-query retrieval over the typed memory stores, and jointly infer a temporally and contextually valid user state for downstream response generation. \textsc{QUMem} achieves state-of-the-art performance on both PersonaMem and KnowU-Bench, demonstrating the effectiveness of query-conditioned user-state inference for long-term personalization.
Problem

Research questions and friction points this paper is trying to address.

LLM Agents
Personalized Memory
User-State Inference
Memory Retrieval
Long-term Personalization
Innovation

Methods, ideas, or system contributions that make the work stand out.

Query-Conditioned User-State Inference
Structured Memory Framework
Semantic Episode Segmentation
Multi-Query Retrieval Planning
Decomposed Memory Units
🔎 Similar Papers
No similar papers found.
Heng Wang
Heng Wang
University of Science and Technology Beijing
Robust ControlFault detectionSLAM
Yifei Li
Yifei Li
Xi'an Jiaotong university
NLP、Relation extraction
Lingling Zhang
Lingling Zhang
Assistant Professor, Xi'an Jiaotong University
Computer visionFew-shot learningZero-shot learning
P
Pengyu Li
School of Computer Science and Technology, Xi’an Jiaotong University; MOE KLNN Lab, Xi’an Jiaotong University
X
Xinyu Che
School of Computer Science and Technology, Xi’an Jiaotong University; MOE KLNN Lab, Xi’an Jiaotong University
X
Xinyu Zhang
School of Computer Science and Technology, Xi’an Jiaotong University; MOE KLNN Lab, Xi’an Jiaotong University
Z
Zesheng Yang
School of Computer Science and Technology, Xi’an Jiaotong University; MOE KLNN Lab, Xi’an Jiaotong University