Neural Network Architectures for Arabic Dialect Identification

2018-08-01COLING 2018Unverified0· sign in to hype

Elise Michon, Minh Quang Pham, Josep Crego, Jean Senellart

Unverified — Be the first to reproduce this paper.

Abstract

SYSTRAN competes this year for the first time to the DSL shared task, in the Arabic Dialect Identification subtask. We participate by training several Neural Network models showing that we can obtain competitive results despite the limited amount of training data available for learning. We report our experiments and detail the network architecture and parameters of our 3 runs: our best performing system consists in a Multi-Input CNN that learns separate embeddings for lexical, phonetic and acoustic input features (F1: 0.5289); we also built a CNN-biLSTM network aimed at capturing both spatial and sequential features directly from speech spectrograms (F1: 0.3894 at submission time, F1: 0.4235 with later found parameters); and finally a system relying on binary CNN-biLSTMs (F1: 0.4339).

Tasks

Automatic Speech Recognition (ASR)Dialect Identification Feature Engineering Language Identification Machine Translation Sentence Classification Speech Recognition

Neural Network Architectures for Arabic Dialect Identification

Abstract

Tasks

Reproductions