Counterfactual Data Augmentation for Mitigating Gender Stereotypes in Languages with Rich Morphology

Identifikátory výsledku

Kód výsledku v IS VaVaI
<a href="https://www.isvavai.cz/riv?ss=detail&h=RIV%2F00216208%3A11320%2F19%3A10427128" target="_blank" >RIV/00216208:11320/19:10427128 - isvavai.cz</a>
Výsledek na webu
<a href="https://www.aclweb.org/anthology/P19-1161" target="_blank" >https://www.aclweb.org/anthology/P19-1161</a>
DOI - Digital Object Identifier
—

Alternativní jazyky

Jazyk výsledku
angličtina
Název v původním jazyce
Counterfactual Data Augmentation for Mitigating Gender Stereotypes in Languages with Rich Morphology
Popis výsledku v původním jazyce
Gender stereotypes are manifest in most of the world's languages and are consequently propagated or amplified by NLP systems. Although research has focused on mitigating gender stereotypes in English, the approaches that are commonly employed produce ungrammatical sentences in morphologically rich languages. We present a novel approach for converting between masculine-inflected and feminine-inflected sentences in such languages. For Spanish and Hebrew, our approach achieves F1 scores of 82% and 73% at the level of tags and accuracies of 90% and 87% at the level of forms. By evaluating our approach using four different languages, we show that, on average, it reduces gender stereotyping by a factor of 2.5 without any sacrifice to grammaticality.
Název v anglickém jazyce
Counterfactual Data Augmentation for Mitigating Gender Stereotypes in Languages with Rich Morphology
Popis výsledku anglicky
Gender stereotypes are manifest in most of the world's languages and are consequently propagated or amplified by NLP systems. Although research has focused on mitigating gender stereotypes in English, the approaches that are commonly employed produce ungrammatical sentences in morphologically rich languages. We present a novel approach for converting between masculine-inflected and feminine-inflected sentences in such languages. For Spanish and Hebrew, our approach achieves F1 scores of 82% and 73% at the level of tags and accuracies of 90% and 87% at the level of forms. By evaluating our approach using four different languages, we show that, on average, it reduces gender stereotyping by a factor of 2.5 without any sacrifice to grammaticality.

Klasifikace

Druh
O - Ostatní výsledky
CEP obor
—
OECD FORD obor
10201 - Computer sciences, information science, bioinformathics (hardware development to be 2.2, social aspect to be 5.8)

Návaznosti výsledku

Projekt
—
Návaznosti
—

Ostatní

Rok uplatnění
2019
Kód důvěrnosti údajů
S - Úplné a pravdivé údaje o projektu nepodléhají ochraně podle zvláštních právních předpisů

Podobné výsledky(10)

Areálové a typologické rysy ve struktuře češtiny a polštiny Don't Forget About Pronouns: Removing Gender Bias in Language Models Without Losing Factual Gender Information Odraz genderových stereotypů ve výuce angličtiny na druhém stupni ZŠ: Výsledky kvantitativních analýz

Co hledáte?

Rychlé hledání

Chytré vyhledávání

Counterfactual Data Augmentation for Mitigating Gender Stereotypes in Languages with Rich Morphology

Identifikátory výsledku

Alternativní jazyky

Klasifikace

Návaznosti výsledku

Ostatní

Podobné výsledky(10)

Co hledáte?

Rychlé hledání

Chytré vyhledávání

Popis výsledku

Identifikátory výsledku

Identifikátory výsledku

Alternativní jazyky

Alternativní jazyky

Klasifikace

Klasifikace

Návaznosti výsledku

Návaznosti výsledku

Ostatní

Ostatní

Podobné výsledky(10)