Institution profile

SIFT

Research institutionnorthamerica · us
Research library1linked papers
Opportunities0open roles
Selected work

Representative Papers

Relating Word Embedding Gender Biases to Gender Gaps: A Cross-Cultural Analysis

Jan 23, 2026Proceedings of the First Workshop on Gender Bias in Natural Language Processing

This study quantifies gender bias in word embeddings and investigates its association with real-world gender disparities across societies. Leveraging Twitter data from 51 U.S. regions and 99 countries in 2018, the authors construct a metric of gender bias in word embeddings and systematically correlate it with 18 international and 5 U.S.-specific gender gap indicators spanning education, politics, economics, and health. The analysis reveals significant cross-cultural correlations and strong predictive power, demonstrating for the first time that gender bias embedded in language models serves as a valid proxy for societal gender inequality. These findings establish a novel paradigm for monitoring social biases through computational linguistic methods at a global scale.

23 citations2 influentialRead paper
Recent publications

Latest Papers

Relating Word Embedding Gender Biases to Gender Gaps: A Cross-Cultural Analysis

Jan 23, 2026Proceedings of the First Workshop on Gender Bias in Natural Language Processing

This study quantifies gender bias in word embeddings and investigates its association with real-world gender disparities across societies. Leveraging Twitter data from 51 U.S. regions and 99 countries in 2018, the authors construct a metric of gender bias in word embeddings and systematically correlate it with 18 international and 5 U.S.-specific gender gap indicators spanning education, politics, economics, and health. The analysis reveals significant cross-cultural correlations and strong predictive power, demonstrating for the first time that gender bias embedded in language models serves as a valid proxy for societal gender inequality. These findings establish a novel paradigm for monitoring social biases through computational linguistic methods at a global scale.

23 citations2 influentialRead paper