SOTAVerified

Visual Speech Recognition

Papers

Showing 5160 of 182 papers

TitleStatusHype
SlideAVSR: A Dataset of Paper Explanation Videos for Audio-Visual Speech Recognition0
The NPU-ASLP-LiAuto System Description for Visual Speech Recognition in CNVSRC 2023Code1
Multichannel AV-wav2vec2: A Framework for Learning Multichannel Multi-Modal Speech RepresentationCode0
MLCA-AVSR: Multi-Layer Cross Attention Fusion based Audio-Visual Speech Recognition0
LiteVSR: Efficient Visual Speech Recognition by Learning from Speech Representations of Unlabeled Data0
The GUA-Speech System Description for CNVSRC Challenge 20230
Do VSR Models Generalize Beyond LRS3?Code1
LIP-RTVE: An Audiovisual Database for Continuous Spanish in the WildCode0
Speaker-Adapted End-to-End Visual Speech Recognition for Continuous Spanish0
Analysis of Visual Features for Continuous Lipreading in Spanish0
Show:102550
← PrevPage 6 of 19Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1VTP with more dataWord Error Rate (WER)30.7Unverified
2CTC/AttentionWord Error Rate (WER)19.1Unverified
#ModelMetricClaimedVerifiedStatus
1VTP with more dataWord Error Rate (WER)22.6Unverified