SOTAVerified

Speech-to-Text

Papers

Showing 201225 of 403 papers

TitleStatusHype
Improving Metrics for Speech Translation0
Application-Agnostic Language Modeling for On-Device ASR0
Hybrid Transducer and Attention based Encoder-Decoder Modeling for Speech-to-Text Tasks0
Improving Autoregressive NLP Tasks via Modular Linearized Attention0
Enhancing Speech-to-Speech Translation with Multiple TTS Targets0
ESPnet-ST-v2: Multipurpose Spoken Language Translation Toolkit0
Natural Language Robot Programming: NLP integrated with autonomous robotic grasping0
Improving the previous state-of-the-art Frisian ASR by fine-tuning XLS-R0
wav2vec and its current potential to Automatic Speech Recognition in German for the usage in Digital History: A comparative assessment of available ASR-technologies for the use in cultural heritage contexts0
Google USM: Scaling Automatic Speech Recognition Beyond 100 Languages0
Improving Medical Speech-to-Text Accuracy with Vision-Language Pre-training Model0
PATCorrect: Non-autoregressive Phoneme-augmented Transformer for ASR Error Correction0
Characterizing Financial Market Coverage using Artificial Intelligence0
Using External Off-Policy Speech-To-Text Mappings in Contextual End-To-End Automated Speech Recognition0
Pushing the performances of ASR models on English and Spanish accents0
WACO: Word-Aligned Contrastive Learning for Speech TranslationCode0
M3ST: Mix at Three Levels for Speech Translation0
MMSpeech: Multi-modal Multi-task Encoder-Decoder Pre-training for Speech Recognition0
Handling and extracting key entities from customer conversations using Speech recognition and Named Entity recognition0
Multilingual Speech Emotion Recognition With Multi-Gating Mechanism and Neural Architecture Search0
Phonemic Representation and Transcription for Speech to Text Applications for Under-resourced Indigenous African Languages: The Case of Kiswahili0
Efficient Speech Translation with Dynamic Latent PerceiversCode0
Don't Discard Fixed-Window Audio Segmentation in Speech-to-Text TranslationCode0
Named Entity Detection and Injection for Direct Speech Translation0
Improving Semi-supervised End-to-end Automatic Speech Recognition using CycleGAN and Inter-domain Losses0
Show:102550
← PrevPage 9 of 17Next →

No leaderboard results yet.