SOTAVerified

Automatic Speech Recognition

Papers

Showing 126150 of 3174 papers

TitleStatusHype
A Reference-less Quality Metric for Automatic Speech Recognition via Contrastive-Learning of a Multi-Language Model with Self-SupervisionCode1
NoRefER: a Referenceless Quality Metric for Automatic Speech Recognition via Semi-Supervised Language Model Fine-Tuning with Contrastive LearningCode1
Quilt-1M: One Million Image-Text Pairs for HistopathologyCode1
Pushing the Limits of Unsupervised Unit Discovery for SSL Speech RepresentationCode1
SGEM: Test-Time Adaptation for Automatic Speech Recognition via Sequential-Level Generalized Entropy MinimizationCode1
Improved DeepFake Detection Using Whisper FeaturesCode1
Can Contextual Biasing Remain Effective with Whisper and GPT-2?Code1
Scaling Speech Technology to 1,000+ LanguagesCode1
CopyNE: Better Contextual ASR by Copying Named EntitiesCode1
Making More of Little Data: Improving Low-Resource Automatic Speech Recognition Using Data AugmentationCode1
Cross-Modal Global Interaction and Local Alignment for Audio-Visual Speech RecognitionCode1
Back Translation for Speech-to-text Translation Without TranscriptsCode1
CB-Conformer: Contextual biasing Conformer for biased word recognitionCode1
When Good and Reproducible Results are a Giant with Feet of Clay: The Importance of Software Quality in NLPCode1
Gradient Remedy for Multi-Task Learning in End-to-End Noise-Robust Speech RecognitionCode1
A Sidecar Separator Can Convert a Single-Talker Speech Recognition System to a Multi-Talker OneCode1
Complex Dynamic Neurons Improved Spiking Transformer Network for Efficient Automatic Speech RecognitionCode1
Cross-modal information fusion for voice spoofing detectionCode1
Knowledge Transfer from Pre-trained Language Models to Cif-based Speech Recognizers via Hierarchical DistillationCode1
Audio-Visual Efficient Conformer for Robust Speech RecognitionCode1
Towards Voice Reconstruction from EEG during Imagined SpeechCode1
Skit-S2I: An Indian Accented Speech to Intent datasetCode1
BASPRO: a balanced script producer for speech corpus collection based on the genetic algorithmCode1
SoftCTC -- Semi-Supervised Learning for Text Recognition using Soft Pseudo-LabelsCode1
A Persian ASR-based SER: Modification of Sharif Emotional Speech Database and Investigation of Persian Text CorporaCode1
Show:102550
← PrevPage 6 of 127Next →

No leaderboard results yet.