Milimili. Collecting Parallel Data via Crowdsourcing
2023-07-23Code Available0· sign in to hype
Alexander Antonov
Code Available — Be the first to reproduce this paper.
ReproduceCode
- github.com/alantonov/milimiliOfficialIn papernone★ 0
Abstract
We present a methodology for gathering a parallel corpus through crowdsourcing, which is more cost-effective than hiring professional translators, albeit at the expense of quality. Additionally, we have made available experimental parallel data collected for Chechen-Russian and Fula-English language pairs.