Papers
Topics
Authors
Recent
Gemini 2.5 Flash
Gemini 2.5 Flash
41 tokens/sec
GPT-4o
59 tokens/sec
Gemini 2.5 Pro Pro
41 tokens/sec
o3 Pro
7 tokens/sec
GPT-4.1 Pro
50 tokens/sec
DeepSeek R1 via Azure Pro
28 tokens/sec
2000 character limit reached

Welcome to the Modern World of Pronouns: Identity-Inclusive Natural Language Processing beyond Gender (2202.11923v1)

Published 24 Feb 2022 in cs.CL

Abstract: The world of pronouns is changing. From a closed class of words with few members to a much more open set of terms to reflect identities. However, NLP is barely reflecting this linguistic shift, even though recent work outlined the harms of gender-exclusive language technology. Particularly problematic is the current modeling 3rd person pronouns, as it largely ignores various phenomena like neopronouns, i.e., pronoun sets that are novel and not (yet) widely established. This omission contributes to the discrimination of marginalized and underrepresented groups, e.g., non-binary individuals. However, other identity-expression phenomena beyond gender are also ignored by current NLP technology. In this paper, we provide an overview of 3rd person pronoun issues for NLP. Based on our observations and ethical considerations, we define a series of desiderata for modeling pronouns in language technology. We evaluate existing and novel modeling approaches w.r.t. these desiderata qualitatively, and quantify the impact of a more discrimination-free approach on established benchmark data.

User Edit Pencil Streamline Icon: https://streamlinehq.com
Authors (3)
  1. Anne Lauscher (58 papers)
  2. Archie Crowley (2 papers)
  3. Dirk Hovy (57 papers)
Citations (47)