Enhancing LLM Instruction Following: An Evaluation-Driven Multi-Agentic Workflow for Prompt Instructions Optimization

πŸ“… 2026-01-06
πŸ›οΈ arXiv.org
πŸ“ˆ Citations: 0
✨ Influential: 0
πŸ“„ PDF
πŸ€– AI Summary
This work addresses the challenge that large language models often struggle to simultaneously satisfy content relevance and formal constraints, leading to procedural errors. To overcome this, the authors propose a multi-agent workflow that, for the first time, decouples the primary task description from fine-grained constraints and iteratively refines prompts through an evaluation-driven collaborative mechanism. By integrating automated scoring feedback, prompt rewriting, and multi-agent coordination, the approach significantly enhances adherence to formal constraints in model outputs. Experiments on Llama 3.1 8B and Mixtral-8x 7B demonstrate substantial improvements, validating the effectiveness of constraint decoupling and evaluation-guided refinement in boosting instruction-following performance.

Technology Category

Application Category

πŸ“ Abstract
Large Language Models (LLMs) often generate substantively relevant content but fail to adhere to formal constraints, leading to outputs that are conceptually correct but procedurally flawed. Traditional prompt refinement approaches focus on rephrasing the description of the primary task an LLM has to perform, neglecting the granular constraints that function as acceptance criteria for its response. We propose a novel multi-agentic workflow that decouples optimization of the primary task description from its constraints, using quantitative scores as feedback to iteratively rewrite and improve them. Our evaluation demonstrates this method produces revised prompts that yield significantly higher compliance scores from models like Llama 3.1 8B and Mixtral-8x 7B.
Problem

Research questions and friction points this paper is trying to address.

instruction following
prompt optimization
formal constraints
large language models
compliance
Innovation

Methods, ideas, or system contributions that make the work stand out.

multi-agentic workflow
prompt optimization
instruction following
constraint decoupling
evaluation-driven feedback
πŸ”Ž Similar Papers
No similar papers found.
Alberto Purpura
Alberto Purpura
Capital One
Generative AIInformation RetrievalNatural Language ProcessingSentiment Analysis
Li Wang
Li Wang
Ant Group
machine learning、MPC
S
Sahil Badyal
Card Intelligence, Capital One
E
Eugenio Beaufrand
Card Intelligence, Capital One
A
Adam Faulkner
Card Intelligence, Capital One