SOTAVerified

Language Identification

Language identification is the task of determining the language of a text.

Papers

Showing 151–175 of 794 papers

TitleStatusHype
LAMASSU: Streaming Language-Agnostic Multilingual Speech Recognition and Translation Using Neural Transducers—0
A Compact End-to-End Model with Local and Global Context for Spoken Language Identification—0
Italian Language and Dialect Identification and Regional French Variety Detection using Adaptive Naive BayesCode0
Neural Networks for Cross-domain Language Identification. Phlyers @Vardial 2022—0
OcWikiDisc: a Corpus of Wikipedia Talk Pages in Occitan—0
The Curious Case of Logistic Regression for Italian Languages and Dialects IdentificationCode0
Streaming End-to-End Multilingual Speech Recognition with Joint Language Identification—0
Evaluation of Off-the-Shelf Language Identification Tools on Bulgarian Social Media Posts—0
Unravelling Interlanguage Facts via Explainable Machine Learning—0
Extending RNN-T-based speech recognition systems with emotion and language classification—0
Distilled Non-Semantic Speech Embeddings with Binary Neural Networks for Low-Resource DevicesCode0
Huqariq: A Multilingual Speech Corpus of Native Languages of Peru for Speech Recognition—0
TechSSN at SemEval-2022 Task 6: Intended Sarcasm Detection using Transformer Models—0
Language Identification for Austronesian LanguagesCode0
HeLI-OTS, Off-the-shelf Language Identifier for Text—0
Dialects Identification of Armenian Language—0
MHE: Code-Mixed Corpora for Similar Language Identification—0
Deep learning-based end-to-end spoken language identification system for domain-mismatched scenario—0
GeezSwitch: Language Identification in Typologically Related Low-resourced East African LanguagesCode0
Universal Dependencies Treebank for Tatar: Incorporating Intra-Word Code-Switching Information—0
CoSwID, a Code Switching Identification Method Suitable for Under-Resourced Languages—0
Huqariq: A Multilingual Speech Corpus of Native Languages of Peru forSpeech Recognition—0
Adversarial synthesis based data-augmentation for code-switched spoken language identification—0
FLEURS: Few-shot Learning Evaluation of Universal Representations of SpeechCode0
Modernizing Open-Set Speech Language Identification—0
Show:102550
← PrevPage 7 of 32Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1wav2vec 2.0 LV-60KError rate7.2—Unverified
2XLS-RError rate5.7—Unverified
#ModelMetricClaimedVerifiedStatus
1GlotLIDMacro F10.98—Unverified
#ModelMetricClaimedVerifiedStatus
1FastTextAccuracy0.97—Unverified
#ModelMetricClaimedVerifiedStatus
1Apple bi-LSTMAccuracy91.37—Unverified
#ModelMetricClaimedVerifiedStatus
1Apple bi-LSTMAccuracy86.93—Unverified
#ModelMetricClaimedVerifiedStatus
1ConformerG-PAccuracy99.8—Unverified