Skip to content
AI Atlas
PaperActive

When Noise Fabricates Bias: The Fragility of LLM-as-a-Judge Bias Measurement under Noisy Text

arxiv.org/abs/2609.11067

Updated 23 min ago · first seen 11 Sept 2026

paper_01M294FQJ80DDFHEJJV2PYF00G

Published
11 Sept 2026
T1 · 23 min ago
arXiv
2609.11067
T1 · 23 min ago
Category
cs.CL
T1 · 23 min ago

Abstract

Large language models are increasingly used as judges to measure social bias in text, yet the passages they judge are often noisy, containing typos, informal spelling, and broken punctuation. The consequences of such surface noise for social bias measurement remain unclear. To investigate this question, we apply five realistic noise conditions at multiple intensity levels to 3,822 stereotype-related responses and compare the resulting bias judgments with those on the original text. We find that such surface noise does not degrade bias measurement symmetrically: it is far more likely to turn neutral judgments into biased ones than biased judgments into neutral ones, by up to a 120x margin. We further observe two non-obvious effects across four LLM judges: in the most fragile judge the distortion is at its purest at mild, realistic noise levels, where erasure is scarcest, and as judges grow robust it attenuates toward parity rather than reversing. Bias measured on noisy text is therefore systematically overestimated, most in the categories that matter most for fairness.

Authors 4

DongHyun Ryu, Jaehyeok Lee, YeongJun Hwang, JinYeong Bak

Specification

Official page

Source:arXiv (Atom API + RSS)T1observed 23 min agohigh

Arxiv announce type
new

Source:arXiv (Atom API + RSS)T1observed 23 min agohigh

arXiv id
2609.11067

Source:arXiv (Atom API + RSS)T1observed 23 min agohigh

Categories
cs.CL, cs.LG

Source:arXiv (Atom API + RSS)T1observed 23 min agohigh

PDF

Source:arXiv (Atom API + RSS)T1observed 23 min agohigh

Primary category
cs.CL

Source:arXiv (Atom API + RSS)T1observed 23 min agohigh

Published
11 Sept 2026

Source:arXiv (Atom API + RSS)T1observed 23 min agohigh

Each value shows its source, tier and observation time. Conflicting claims are kept side by side and flagged — never averaged. How AI Atlas records facts →

Provenance

Attributed facts

9

Source tiers

T19

Freshest observation

23 min ago

Conflicts

None