SOTAVerified

Visual Speech Recognition

Papers

Showing 111120 of 182 papers

TitleStatusHype
Leveraging Modality-specific Representations for Audio-visual Speech Recognition via Reinforcement Learning0
VATLM: Visual-Audio-Text Pre-Training with Unified Masked Prediction for Speech Representation Learning0
Streaming Audio-Visual Speech Recognition with Alignment Regularization0
Visual Speech Recognition in a Driver Assistance System0
Kaggle Competition: Cantonese Audio-Visual Speech Recognition for In-car Commands0
Lip-Listening: Mixing Senses to Understand Lips using Cross Modality Knowledge Distillation for Word-Based Models0
RUSAVIC Corpus: Russian Audio-Visual Speech in Cars0
Is Lip Region-of-Interest Sufficient for Lipreading?0
Deep Learning for Visual Speech Analysis: A Survey0
Learning Contextually Fused Audio-visual Representations for Audio-visual Speech Recognition0
Show:102550
← PrevPage 12 of 19Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1VTP with more dataWord Error Rate (WER)30.7Unverified
2CTC/AttentionWord Error Rate (WER)19.1Unverified
#ModelMetricClaimedVerifiedStatus
1VTP with more dataWord Error Rate (WER)22.6Unverified