Emerging from Ground: Addressing Intent Deviation in Tool-Using Agents via Deriving Real Calls into Virtual Trajectories

๐Ÿ“… 2026-01-21
๐Ÿ“ˆ Citations: 0
โœจ Influential: 0
๐Ÿ“„ PDF
๐Ÿค– AI Summary
Tool-calling agents often exhibit unreliable behaviors and are difficult to evaluate due to misaligned intentions. To address this, this work proposes RISE, a novel approach featuring a โ€œreal-to-syntheticโ€ data synthesis mechanism. It first extracts verified tool primitives from real-world tool invocation logs and generates diverse negative samples via parameter perturbation to construct synthetic trajectories. A two-stage preference fine-tuning strategy is then employed to align agent intentions with user requirements. This method effectively mitigates distributional shift, achieving significant improvements of 35.28% and 23.27% on the AccTask and AccIntent metrics, respectively, and outperforming existing state-of-the-art methods across all eight evaluation benchmarks.

Technology Category

Application Category

๐Ÿ“ Abstract
LLMs have advanced tool-using agents for real-world applications, yet they often lead to unexpected behaviors or results. Beyond obvious failures, the subtle issue of"intent deviation"severely hinders reliable evaluation and performance improvement. Existing post-training methods generally leverage either real system samples or virtual data simulated by LLMs. However, the former is costly due to reliance on hand-crafted user requests, while the latter suffers from distribution shift from the real tools in the wild. Additionally, both methods lack negative samples tailored to intent deviation scenarios, hindering effective guidance on preference learning. We introduce RISE, a"Real-to-Virtual"method designed to mitigate intent deviation. Anchoring on verified tool primitives, RISE synthesizes virtual trajectories and generates diverse negative samples through mutation on critical parameters. With synthetic data, RISE fine-tunes backbone LLMs via the two-stage training for intent alignment. Evaluation results demonstrate that data synthesized by RISE achieve promising results in eight metrics covering user requires, execution trajectories and agent responses. Integrating with training, RISE achieves an average 35.28% improvement in Acctask (task completion) and 23.27% in Accintent (intent alignment), outperforming SOTA baselines by 1.20--42.09% and 1.17--54.93% respectively.
Problem

Research questions and friction points this paper is trying to address.

intent deviation
tool-using agents
negative samples
distribution shift
preference learning
Innovation

Methods, ideas, or system contributions that make the work stand out.

intent deviation
tool-using agents
real-to-virtual synthesis
negative sampling
preference learning
๐Ÿ”Ž Similar Papers
No similar papers found.
Q
Qian Xiong
Beijing Forestry University, State Key Laboratory of Complex System Modeling and Simulation Technology, Beijing, China, Institute of Software, Chinese Academy of Sciences, Beijing, China, University of Chinese Academy of Sciences, Beijing, China
Y
Yuekai Huang
State Key Laboratory of Complex System Modeling and Simulation Technology, Beijing, China, Institute of Software, Chinese Academy of Sciences, Beijing, China, University of Chinese Academy of Sciences, Beijing, China
B
Bo Yang
Beijing Forestry University
Yujia Zheng
Yujia Zheng
Carnegie Mellon University
Machine LearningCausal Discovery and InferenceLatent Variable ModelsGenerative Models
T
Tianhao Li
Duke University
Ziyou Jiang
Ziyou Jiang
Institute of Software Chinese Academy of Sciences
software engineering
Zhiyuan Chang
Zhiyuan Chang
Institute of Software Chinese Academy of Science
LLM SecurityMultimodal TestingRequirements Engineering
Zhaoyang Li
Zhaoyang Li
Ph.D student, University of Science and Technology of China
Computer Vision
H
Huanxiang Feng
State Key Laboratory of Complex System Modeling and Simulation Technology, Beijing, China, Institute of Software, Chinese Academy of Sciences, Beijing, China, University of Chinese Academy of Sciences, Beijing, China
Mingyang Li
Mingyang Li
Associate Professor, Industrial and Management Systems Engineering, The University of South Florida
data sciencereliability and qualitysystem informaticscomplex systems modeling and optimizationcomputational intelligence