SOTAVerified

Language Identification

Language identification is the task of determining the language of a text.

Papers

Showing 101–150 of 794 papers

TitleStatusHype
AfriHuBERT: A self-supervised speech representation model for African languagesCode0
Automatic Language Identification in Texts: A SurveyCode0
Hierarchical Character-Word Models for Language IdentificationCode0
IIITK@DravidianLangTech-EACL2021: Offensive Language Identification and Meme Classification in Tamil, Malayalam and KannadaCode0
Hate-Alert@DravidianLangTech-EACL2021: Ensembling strategies for Transformer-based Offensive language DetectionCode0
Ghmerti at SemEval-2019 Task 6: A Deep Word- and Character-based Approach to Offensive Language IdentificationCode0
Aggressive Language Identification Using Word Embeddings and Sentiment FeaturesCode0
HeLI, a Word-Based Backoff Method for Language IdentificationCode0
Spoken Language Identification System for English-Mandarin Code-Switching Child-Directed SpeechCode0
STIL -- Simultaneous Slot Filling, Translation, Intent Classification, and Language Identification: Initial Results using mBART on MultiATIS++Code0
The WiLI benchmark dataset for written language identificationCode0
Topics to Avoid: Demoting Latent Confounds in Text ClassificationCode0
indicnlp@kgp at DravidianLangTech-EACL2021: Offensive Language Identification in Dravidian LanguagesCode0
JU\_ETCE\_17\_21 at SemEval-2019 Task 6: Efficient Machine Learning and Neural Network Approaches for Identifying and Categorizing Offensive Language in TweetsCode0
TweetCaT: a tool for building Twitter corpora of smaller languagesCode0
GeezSwitch: Language Identification in Typologically Related Low-resourced East African LanguagesCode0
Finding Structure in Text, Genome and Other Symbolic SequencesCode0
Approaches to Corpus Creation for Low-Resource Language Technology: the Case of Southern Kurdish and LakiCode0
Fleurs-SLU: A Massively Multilingual Benchmark for Spoken Language UnderstandingCode0
English Please: Evaluating Machine Translation with Large Language Models for Multilingual Bug ReportsCode0
End-to-end Language Identification using NetFV and NetVLADCode0
FBK-DH at SemEval-2020 Task 12: Using Multi-channel BERT for Multilingual Offensive Language DetectionCode0
Geographic Adaptation of Pretrained Language ModelsCode0
Distilled Non-Semantic Speech Embeddings with Binary Neural Networks for Low-Resource DevicesCode0
CyberTronics at SemEval-2020 Task 12: Multilingual Offensive Language Identification over Social MediaCode0
DocLangID: Improving Few-Shot Training to Identify the Language of Historical DocumentsCode0
Comparing the Performance of CNNs and Shallow Models for Language IdentificationCode0
Code-Switched Language Identification is Harder Than You ThinkCode0
AdelaideCyC at SemEval-2020 Task 12: Ensemble of Classifiers for Offensive Language Detection in Social MediaCode0
Combination of multiple Deep Learning architectures for Offensive Language Detection in TweetsCode0
Cross-Domain Adaptation of Spoken Language Identification for Related Languages: The Curious Case of Slavic LanguagesCode0
Cross-lingual Offensive Language Identification for Low Resource Languages: The Case of MarathiCode0
Discriminating between Similar Languages using Weighted Subword FeaturesCode0
Discriminating Between Similar Nordic LanguagesCode0
Crawling microblogging services to gather language-classified URLs. Workflow and case studyCode0
Embeddia at SemEval-2019 Task 6: Detecting Hate with Neural Network and Transfer Learning ApproachesCode0
DOSA: Dravidian Code-Mixed Offensive Span Identification DatasetCode0
Enhance Language Identification using Dual-mode Model with Knowledge DistillationCode0
Geographically-Informed Language IdentificationCode0
Building a TOCFL Learner Corpus for Chinese Grammatical Error Diagnosis—0
Building a learner corpus for Russian—0
Arabic Native Language Identification—0
bs,hr,srWaC - Web Corpora of Bosnian, Croatian and Serbian—0
BRUMS at SemEval-2020 Task 12: Transformer Based Multilingual Offensive Language Identification in Social Media—0
Arabic Language WEKA-Based Dialect Classifier for Arabic Automatic Speech Recognition Transcripts—0
AlexU-BackTranslation-TL at SemEval-2020 Task 12: Improving Offensive Language Detection Using Data Augmentation and Transfer Learning—0
BRUMS at SemEval-2020 Task 12 : Transformer based Multilingual Offensive Language Identification in Social Media—0
Bootstrapping a historical commodities lexicon with SKOS and DBpedia—0
Arabic Dialect Identification in the Context of Bivalency and Code-Switching—0
BNU-HKBU UIC NLP Team 2 at SemEval-2019 Task 6: Detecting Offensive Language Using BERT model—0
Show:102550
← PrevPage 3 of 16Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1wav2vec 2.0 LV-60KError rate7.2—Unverified
2XLS-RError rate5.7—Unverified
#ModelMetricClaimedVerifiedStatus
1GlotLIDMacro F10.98—Unverified
#ModelMetricClaimedVerifiedStatus
1FastTextAccuracy0.97—Unverified
#ModelMetricClaimedVerifiedStatus
1Apple bi-LSTMAccuracy91.37—Unverified
#ModelMetricClaimedVerifiedStatus
1Apple bi-LSTMAccuracy86.93—Unverified
#ModelMetricClaimedVerifiedStatus
1ConformerG-PAccuracy99.8—Unverified