SOTAVerified

Language Identification

Language identification is the task of determining the language of a text.

Papers

Showing 201–250 of 794 papers

TitleStatusHype
DELab@IIITSM at ICON-2021 Shared Task: Identification of Aggression and Biasness Using Decision Tree—0
MUM at ComMA@ICON: Multilingual Gender Biased and Communal Language Identification Using Supervised Learning Approaches—0
ComMA@ICON: Multilingual Gender Biased and Communal Language Identification Task at ICON-2021—0
MUCIC at ComMA@ICON: Multilingual Gender Biased and Communal Language Identification Using N-grams and Multilingual Sentence Encoders—0
Unsupervised Preference-Aware Language Identification—0
Developing Successful Shared Tasks on Offensive Language Identification for Dravidian Languages—0
Language Clustering for Multilingual Named Entity Recognition—0
An Investigation into the Contribution of Locally Aggregated Descriptors to Figurative Language IdentificationCode0
Native Language Identification and Reconstruction of Native Language Relationship Using Japanese Learner Corpus—0
Tackling the Score Shift in Cross-Lingual Speaker Verification by Exploiting Language Information—0
Ceasing hate withMoH: Hate Speech Detection in Hindi-English Code-Switched Language—0
Mandarin-English Code-switching Speech Recognition with Self-supervised Speech Representation Models—0
Pretrained Transformers for Offensive Language Identification in TanglishCode0
Is Attention always needed? A Case Study on Language Identification from Speech—0
BigSSL: Exploring the Frontier of Large-Scale Semi-Supervised Learning for Automatic Speech Recognition—0
Language Identification with a Reciprocal Rank ClassifierCode0
UPV at CheckThat! 2021: Mitigating Cultural Differences for Identifying Multilingual Check-worthy ClaimsCode0
Unsupervised Personality-Aware Language Identification—0
The futility of STILTs for the classification of lexical borrowings in Spanish—0
On the Language-specificity of Multilingual BERT and the Impact of Fine-tuningCode0
FBERT: A Neural Transformer for Identifying Offensive Content—0
Cross-lingual Offensive Language Identification for Low Resource Languages: The Case of MarathiCode0
A Pre-trained Transformer and CNN Model with Joint Language ID and Part-of-Speech Tagging for Code-Mixed Social-Media Text—0
Fiction in Russian Translation: A Translationese Study—0
Corpus Creation and Language Identification in Low-Resource Code-Mixed Telugu-English Text—0
Offensive Language Identification in Low-resourced Code-mixed Dravidian languages using Pseudo-labelingCode0
Towards Offensive Language Identification for Tamil Code-Mixed YouTube Comments and PostsCode0
A Dual-Decoder Conformer for Multilingual Speech Recognition—0
Dyn-ASR: Compact, Multilingual Speech Recognition via Spoken Language and Accent Identification—0
OLR 2021 Challenge: Datasets, Rules and Baselines—0
Improved Language Identification Through Cross-Lingual Self-Supervised Learning—0
Oriental Language Recognition (OLR) 2020: Summary and Analysis—0
Language Identification of Hindi-English tweets using code-mixed BERT—0
Language Lexicons for Hindi-English Multilingual Text Processing—0
A Simple and Efficient Probabilistic Language model for Code-Mixed Text—0
BERT-based Multi-Task Model for Country and Province Level Modern Standard Arabic and Dialectal Arabic Identification—0
SIGTYP 2021 Shared Task: Robust Spoken Language Identification—0
Self-Contextualized Attention for Abusive Language Identification—0
Anlirika: An LSTM–CNN Flow Twister for Spoken Language Identification—0
Data Filtering using Cross-Lingual Word Embeddings—0
Transliteration for Low-Resource Code-Switching Texts: Building an Automatic Cyrillic-to-Latin Converter for Tatar—0
Active learning and negative evidence for language identification—0
Language ID Prediction from Speech Using Self-Attentive Pooling—0
Much Gracias: Semi-supervised Code-switch Detection for Spanish-English: How far can we get?—0
Singing Language Identification using a Deep Phonotactic ApproachCode0
Low-Resource Spoken Language Identification Using Self-Attentive Pooling and Deep 1D Time-Channel Separable Convolutions—0
An Exploratory Analysis of the Relation Between Offensive Language and Mental Health—0
Multilingual Offensive Language Identification for Low-resource Languages—0
Cross-Corpora Language Recognition: A Preliminary Investigation with Indian Languages—0
Language ID Prediction from Speech Using Self-Attentive Pooling and 1D-Convolutions—0
Show:102550
← PrevPage 5 of 16Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1wav2vec 2.0 LV-60KError rate7.2—Unverified
2XLS-RError rate5.7—Unverified
#ModelMetricClaimedVerifiedStatus
1GlotLIDMacro F10.98—Unverified
#ModelMetricClaimedVerifiedStatus
1FastTextAccuracy0.97—Unverified
#ModelMetricClaimedVerifiedStatus
1Apple bi-LSTMAccuracy91.37—Unverified
#ModelMetricClaimedVerifiedStatus
1Apple bi-LSTMAccuracy86.93—Unverified
#ModelMetricClaimedVerifiedStatus
1ConformerG-PAccuracy99.8—Unverified