Demographic Leakage Persists in De-Identified Résumés Despite Language Redaction
September 16, 2026
Testing across nine open-weight models shows that non-language prose sustains demographic inference in de-identified résumés, with target-group recovery averaging 0.757. The study also reveals that LLM-as-a-judge evaluations are highly sensitive to design flaws like forbidden ties and position effects.
HOW THIS AFFECTS YOU
●
researcherYou should consider salience and evaluation design as critical axes when auditing LLM bias.
●
policyYou must account for unstructured prose leakage when auditing AI for bias in hiring.