When Noise Fabricates Bias: The Fragility of LLM-as-a-Judge Bias Measurement under Noisy Text

📅 2026-09-10
📈 Citations: 0
Influential: 0
📄 PDF
🤖 AI Summary
研究探讨了文本噪声对大型语言模型作为偏见测量工具的影响,发现噪声使中立判断更易被误判为有偏见,并导致偏见被系统性高估。
📝 Abstract
Large language models are increasingly used as judges to measure social bias in text, yet the passages they judge are often noisy, containing typos, informal spelling, and broken punctuation. The consequences of such surface noise for social bias measurement remain unclear. To investigate this question, we apply five realistic noise conditions at multiple intensity levels to 3,822 stereotype-related responses and compare the resulting bias judgments with those on the original text. We find that such surface noise does not degrade bias measurement symmetrically: it is far more likely to turn neutral judgments into biased ones than biased judgments into neutral ones, by up to a 120x margin. We further observe two non-obvious effects across four LLM judges: in the most fragile judge the distortion is at its purest at mild, realistic noise levels, where erasure is scarcest, and as judges grow robust it attenuates toward parity rather than reversing. Bias measured on noisy text is therefore systematically overestimated, most in the categories that matter most for fairness.
Problem

Research questions and friction points this paper is trying to address.

surface noise
social bias measurement
large language models
Innovation

Methods, ideas, or system contributions that make the work stand out.

noisy text
bias measurement
large language models
social bias
🔎 Similar Papers
No similar papers found.
💼 Related Jobs
No related jobs found.
D
DongHyun Ryu
Sungkyunkwan University, Suwon, South Korea
J
Jaehyeok Lee
Sungkyunkwan University, Suwon, South Korea
Y
YeongJun Hwang
Sungkyunkwan University, Suwon, South Korea
JinYeong Bak
JinYeong Bak
College of Computing, Sungkyunkwan University
Artificial IntelligenceConversation Modeling