Algorithmic Gender Prediction Is Illegitimate, But Gender Imputation Can Yield Valid Measurements

📅 2026-08-13
📈 Citations: 0
Influential: 0
📄 PDF
🤖 AI Summary
This study addresses the tension between the ethical illegitimacy of gender prediction algorithms and their instrumental utility in fairness research. By decoupling “legitimacy” from “effectiveness” and drawing on trans feminist theory, it distinguishes between two forms of gender-based discrimination: one targeting cisgender women and feminine-coded individuals, and another affecting transgender and non-binary populations, thereby clarifying conceptual confusions in current debates. Integrating computational gender inference, case studies—including audits of generative image models, assessments of cinematic gender gaps, and analyses of name-based gender annotation—and critical theory, the paper demonstrates that gender inference can serve fairness objectives only under strictly constrained conditions and exhibits fundamental limitations regarding transgender and non-binary individuals. The work concludes by advocating for more inclusive methodological alternatives.
📝 Abstract
Machine learning ethics researchers and critical HCI scholars have argued that algorithmically predicting gender is wrong. At the same time, other researchers rely on predicted gender labels to study gender disparities and develop algorithmic fairness techniques. How do we reconcile these two seemingly contradictory intuitions? We differentiate two ways gender prediction may be wrong: being illegitimate, thereby contributing to harm; and being invalid, thereby producing unusable measurements. Our analysis translates arguments against gender prediction into these terms of legitimacy and validity and shows how gender imputation applied for fairness purposes can be illegitimate yet still yield valid disparity measurements. We clarify this bind by drawing upon transfeminist literature to distinguish sexism that targets women and femininity from sexism that targets transgender and nonbinary people. While gender imputation can produce valid measurements for the former, it is illegitimate and harmful for the latter. We argue that practitioners should deploy gender imputation only when it would achieve anti-discrimination benefits that cannot be achieved through other reasonable means, while harms are minimized to the extent possible. We examine this tension in three case studies: auditing gender bias in generative image models, measuring gender disparities in film, and imputing gender from personal names. By disentangling legitimacy from validity, and differentiating these two forms of sexism, we show how debates over gender prediction have conflated distinct concerns, obscuring both the settings in which gender imputation can support fairness efforts and the harms towards transgender and nonbinary people that it fundamentally cannot capture. We conclude by recommending the development of more inclusive methods that address all kinds of sexism.
Problem

Research questions and friction points this paper is trying to address.

gender prediction
algorithmic fairness
legitimacy
validity
transfeminism
Innovation

Methods, ideas, or system contributions that make the work stand out.

gender imputation
algorithmic fairness
transfeminist theory
legitimacy vs validity
gender bias measurement
🔎 Similar Papers
No similar papers found.