SOTAVerified

Understanding Linguistic Accommodation in Code-Switched Human-Machine Dialogues

2020-11-01CONLLCode Available0· sign in to hype

Tanmay Parekh, Emily Ahn, Yulia Tsvetkov, Alan W Black

Code Available — Be the first to reproduce this paper.

Reproduce

Code

Abstract

Code-switching is a ubiquitous phenomenon in multilingual communities. Natural language technologies that wish to communicate like humans must therefore adaptively incorporate code-switching techniques when they are deployed in multilingual settings. To this end, we propose a Hindi-English human-machine dialogue system that elicits code-switching conversations in a controlled setting. It uses different code-switching agent strategies to understand how users respond and accommodate to the agent's language choice. Through this system, we collect and release a new dataset CommonDost, comprising of 439 human-machine multilingual conversations. We adapt pre-defined metrics to discover linguistic accommodation from users to agents. Finally, we compare these dialogues with Spanish-English dialogues collected in a similar setting, and analyze the impact of linguistic and socio-cultural factors on code-switching patterns across the two language pairs.

Reproductions