SOTAVerified

Casteism in India, but Not Racism - a Study of Bias in Word Embeddings of Indian Languages

2022-06-01LATERAISSE (LREC) 2022Unverified0· sign in to hype

Senthil Kumar B, Pranav Tiwari, Aman Chandra Kumar, Aravindan Chandrabose

Unverified — Be the first to reproduce this paper.

Reproduce

Abstract

In this paper, we studied the gender bias in monolingual word embeddings of two Indian languages Hindi and Tamil. Tamil is one of the classical languages of India from the Dravidian language family. In Indian society and culture, instead of racism, a similar type of discrimination called casteism is against the subgroup of peoples representing lower class or Dalits. The word embeddings measurement to evaluate bias using the WEAT score reveals that the embeddings are biased with gender and casteism which is in line with the common stereotypical human biases.

Tasks

Reproductions