DSpace at KOASAS: Neutralizing Gender Bias in Word Embeddings with Latent Disentanglement and Counterfactual Generation

DSpace at KOASAS

College of Engineering(공과대학)Dept. of Industrial and Systems Engineering(산업및시스템공학과)IE-Conference Papers(학술회의논문)

Neutralizing Gender Bias in Word Embeddings with Latent Disentanglement and Counterfactual Generation

Cited 0 time in webofscience

Cited 0 time in

Hit : 241
Download : 0

Export

Shin, Seungjae / Song, Kyungwoo / Jang, JoonHo / Kim, Hyemi / Joo, Weonyoung / Moon, Il-Chul researcher

Recent research demonstrates that word embeddings, trained on the human-generated corpus, have strong gender biases in embedding spaces, and these biases can result in the discriminative results from the various downstream tasks. Whereas the previous methods project word embeddings into a linear subspace for debiasing, we introduce a Latent Disentanglement method with a siamese auto-encoder structure with an adapted gradient reversal layer. Our structure enables the separation of the semantic latent information and gender latent information of given word into the disjoint latent dimensions. Afterwards, we introduce a Counterfactual Generation to convert the gender information of words, so the original and the modified embeddings can produce a gender-neutralized word embedding after geometric alignment regularization, without loss of semantic information. From the various quantitative and qualitative debiasing experiments, our method shows to be better than existing debiasing methods in debiasing word embeddings. In addition, Our method shows the ability to preserve semantic information during debiasing by minimizing the semantic information losses for extrinsic NLP downstream tasks.

Publisher: Association for Computational Linguistics

Issue Date: 2020-11-20

Language: English

Citation: Empirical Methods in Natural Language Processing conference (EMNLP) 2020, pp.3126 - 3140

DOI: 10.18653/v1/2020.findings-emnlp.280

URI: http://hdl.handle.net/10203/282437

Appears in Collection: IE-Conference Papers(학술회의논문)

Files in This Item: There are no files associated with this item.

Display Full Item Record

qr_code

트윗하기

KOASAS

Knowledge Service Development Team, KAIST 291 Daehak-ro, Yuseong-gu, Daejeon 34141, Republic of Korea. T. 82-42-350-4493 Email. koasas@kaist.ac.kr
Copyright © 2016. Korea Advanced Institute of Science and Technology. All Rights Reserved.

KOASAS

KOASAS

Browse

Neutralizing Gender Bias in Word Embeddings with Latent Disentanglement and Counterfactual Generation

KOASAS

Communities & Collections