SOTAVerified

Visual Speech Recognition

Papers

Showing 101110 of 182 papers

TitleStatusHype
SynthVSR: Scaling Up Visual Speech Recognition With Synthetic Supervision0
The NPU-ASLP System for Audio-Visual Speech Recognition in MISP 2022 Challenge0
Deep Visual Forced Alignment: Learning to Align Transcription with Talking Face Video0
Conformers are All You Need for Visual Speech Recognition0
Audio-Visual Speech and Gesture Recognition by Sensors of Mobile Devices0
Prompt Tuning of Deep Neural Networks for Speaker-adaptive Visual Speech Recognition0
AV-data2vec: Self-supervised Learning of Audio-Visual Speech Representations with Contextualized Target Representations0
A Multi-Purpose Audio-Visual Corpus for Multi-Modal Persian Speech Recognition: the Arman-AV Dataset0
ReVISE: Self-Supervised Speech Resynthesis With Visual Input for Universal and Generalized Speech Regeneration0
ReVISE: Self-Supervised Speech Resynthesis with Visual Input for Universal and Generalized Speech Enhancement0
Show:102550
← PrevPage 11 of 19Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1VTP with more dataWord Error Rate (WER)30.7Unverified
2CTC/AttentionWord Error Rate (WER)19.1Unverified
#ModelMetricClaimedVerifiedStatus
1VTP with more dataWord Error Rate (WER)22.6Unverified