SOTAVerified

Visual Speech Recognition

Papers

Showing 91100 of 182 papers

TitleStatusHype
A Multi-Purpose Audio-Visual Corpus for Multi-Modal Persian Speech Recognition: the Arman-AV Dataset0
OLKAVS: An Open Large-Scale Korean Audio-Visual Speech DatasetCode1
ReVISE: Self-Supervised Speech Resynthesis With Visual Input for Universal and Generalized Speech Regeneration0
ReVISE: Self-Supervised Speech Resynthesis with Visual Input for Universal and Generalized Speech Enhancement0
Jointly Learning Visual and Auditory Speech Representations from Raw DataCode1
Leveraging Modality-specific Representations for Audio-visual Speech Recognition via Reinforcement Learning0
VATLM: Visual-Audio-Text Pre-Training with Unified Masked Prediction for Speech Representation Learning0
Streaming Audio-Visual Speech Recognition with Alignment Regularization0
Visual Speech Recognition in a Driver Assistance System0
Visual Context-driven Audio Feature Enhancement for Robust End-to-End Audio-Visual Speech RecognitionCode1
Show:102550
← PrevPage 10 of 19Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1VTP with more dataWord Error Rate (WER)30.7Unverified
2CTC/AttentionWord Error Rate (WER)19.1Unverified
#ModelMetricClaimedVerifiedStatus
1VTP with more dataWord Error Rate (WER)22.6Unverified