BayesPrompt: human readable prompts that make sense

📅 2026-08-18
📈 Citations: 0
Influential: 0
📄 PDF
🤖 AI Summary
为解决生成难以理解的伪提示问题,本文通过贝叶斯后验推理方法提出一种既能降低困惑度又可读性强的高效算法。
📝 Abstract
Reconstructing prompts that can elicit a desired answer or behaviour in an LLM is an open and important research topic. Optimisation methods which aim at minimising the perplexity of a given answer, however, consistently yield so-called pseudoprompts, unintelligible strings of tokens which can lack human interpretability. We argue that this is a consequence of the ill-posedness of the prompt optimisation task. By reframing the task as a Bayesian posterior inference over prompts, we propose an efficient algorithm to sample prompts which are both efficient (in terms of perplexity) and human readable. We compare our approach with state of the art alternatives showing on a real data set a marked improvement over a range of metrics.
Problem

Research questions and friction points this paper is trying to address.

prompts
perplexity
pseudoprompts
interpretability
Innovation

Methods, ideas, or system contributions that make the work stand out.

Bayesian posterior inference
human readable prompts
perplexity
🔎 Similar Papers
F
Franky Kevin Nando Tezoh
Department of Physics, Scuola Internazionale Superiore di Studi Avanzati (SISSA), Trieste, Italy
A
Ali Hussaini Umar
Department of Physics, Scuola Internazionale Superiore di Studi Avanzati (SISSA), Trieste, Italy
Alessandro Laio
Alessandro Laio
SISSA
molecular dynamicsatomistic simulationsmachine learning
Guido Sanguinetti
Guido Sanguinetti
Reader in Informatics, University of Edinburgh
Machine LearningSystems BiologyStatistical modelling
R
Riccardo Rende
Center for Computational Quantum Physics, Flatiron Institute, 162 5th Avenue, New York, NY 10010