SOTAVerified

Language Identification

Language identification is the task of determining the language of a text.

Papers

Showing 101–150 of 794 papers

TitleStatusHype
Automatic Language Identification for Romance Languages using Stop Words and Diacritics—0
Anlirika: An LSTM–CNN Flow Twister for Spoken Language Identification—0
A Text-to-Text Model for Multilingual Offensive Language Identification—0
AUTOMATIC LANGUAGE IDENTIFICATION USING DEEP NEURAL NETWORKS—0
Amrita_CEN_NLP@DravidianLangTech-EACL2021: Deep Learning-based Offensive Language Identification in Malayalam, Tamil and Kannada—0
Automatic Spoken Language Identification Utilizing Acoustic and Phonetic Speech Information—0
Automatic Spoken Language Identification using a Time-Delay Neural Network—0
Automatic Token and Turn Level Language Identification for Code-Switched Text Dialog: An Analysis Across Language Pairs and Corpora—0
Babler - Data Collection from the Web to Support Speech Recognition and Keyword Search—0
Beefmoves: Dissemination, Diversity, and Dynamics of English Borrowings in a German Hip Hop Forum—0
BERT-based Multi-Task Model for Country and Province Level Modern Standard Arabic and Dialectal Arabic Identification—0
BERT-based Multi-Task Model for Country and Province Level MSA and Dialectal Arabic Identification—0
Active learning and negative evidence for language identification—0
Beware Haters at ComMA@ICON: Sequence and Ensemble Classifiers for Aggression, Gender Bias and Communal Bias Identification in Indian Languages—0
BFCAI at ComMA@ICON 2021: Support Vector Machines for Multilingual Gender Biased and Communal Language Identification—0
BhamNLP at SemEval-2020 Task 12: An Ensemble of Different Word Embeddings and Emotion Transfer Learning for Arabic Offensive Language Identification in Social Media—0
Accurate Language Identification of Twitter Messages—0
BigSSL: Exploring the Frontier of Large-Scale Semi-Supervised Learning for Automatic Speech Recognition—0
Bilingual Streaming ASR with Grapheme units and Auxiliary Monolingual Loss—0
BNU-HKBU UIC NLP Team 2 at SemEval-2019 Task 6: Detecting Offensive Language Using BERT model—0
Bootstrapping a historical commodities lexicon with SKOS and DBpedia—0
BRUMS at SemEval-2020 Task 12 : Transformer based Multilingual Offensive Language Identification in Social Media—0
BRUMS at SemEval-2020 Task 12: Transformer Based Multilingual Offensive Language Identification in Social Media—0
bs,hr,srWaC - Web Corpora of Bosnian, Croatian and Serbian—0
Building a learner corpus for Russian—0
Building a TOCFL Learner Corpus for Chinese Grammatical Error Diagnosis—0
Code Switched and Code Mixed Speech Recognition for Indic languages—0
A survey on phrase structure learning methods for text classification—0
A Study on Spoken Language Identification using Deep Neural Networks—0
American Sign Language Identification Using Hand Trackpoint Analysis—0
A Compact End-to-End Model with Local and Global Context for Spoken Language Identification—0
ASIREM Participation at the Discriminating Similar Languages Shared Task 2016—0
Advancing Linguistic Features and Insights by Label-informed Feature Grouping: An Exploration in the Context of Native Language Identification—0
Neighbors and relatives: How do speech embeddings reflect linguistic connections across the world?—0
A Simple and Efficient Probabilistic Language model for Code-Mixed Text—0
A Shallow Neural Network for Native Language Identification with Character N-grams—0
A Mandarin-English Code-Switching Corpus—0
ALT at SemEval-2020 Task 12: Arabic and English Offensive Language Identification in Social Media—0
Character Level Convolutional Neural Network for Indo-Aryan Language Identification—0
Characterizing Stylistic Elements in Syntactic Structure—0
Advanced accent/dialect identification and accentedness assessment with multi-embedding models and automatic speech recognition—0
Acoustic characterization of speech rhythm: going beyond metrics with recurrent neural networks—0
CN-HIT-MI.T at SemEval-2019 Task 6: Offensive Language Identification Based on BiLSTM with Double Attention—0
Challenges of Computational Processing of Code-Switching—0
Chinese Native Language Identification—0
CIC-FBK Approach to Native Language Identification—0
Classification of Closely Related Sub-dialects of Arabic Using Support-Vector Machines—0
Classifier Stacking for Native Language Identification—0
CLUZH at VarDial GDI 2017: Testing a Variety of Machine Learning Tools for the Classification of Swiss German Dialects—0
Challenges in Neural Language Identification: NRC at VarDial 2020—0
Show:102550
← PrevPage 3 of 16Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1wav2vec 2.0 LV-60KError rate7.2—Unverified
2XLS-RError rate5.7—Unverified
#ModelMetricClaimedVerifiedStatus
1GlotLIDMacro F10.98—Unverified
#ModelMetricClaimedVerifiedStatus
1FastTextAccuracy0.97—Unverified
#ModelMetricClaimedVerifiedStatus
1Apple bi-LSTMAccuracy91.37—Unverified
#ModelMetricClaimedVerifiedStatus
1Apple bi-LSTMAccuracy86.93—Unverified
#ModelMetricClaimedVerifiedStatus
1ConformerG-PAccuracy99.8—Unverified