Learning Trilingual Dictionaries for Urdu -- Roman Urdu -- English
2019-08-01WS 2019Code Available0· sign in to hype
Moiz Rauf, Sebastian Pad{\'o}
Code Available — Be the first to reproduce this paper.
ReproduceCode
- github.com/MoizRauf/Urdu--Roman-Urdu--English--DictionaryOfficialIn papernone★ 0
Abstract
In this paper, we present an effort to generate a joint Urdu, Roman Urdu and English trilingual lexicon using automated methods. We make a case for using statistical machine translation approaches and parallel corpora for dictionary creation. To this purpose, we use word alignment tools on the corpus and evaluate translations using human evaluators. Despite different writing script and considerable noise in the corpus our results show promise with over 85\% accuracy of Roman Urdu--Urdu and 45\% English--Urdu pairs.