SOTAVerified

Speaker Recognition

Speaker Recognition is the process of identifying or confirming the identity of a person given his speech segments.

Source: Margin Matters: Towards More Discriminative Deep Neural Network Embeddings for Speaker Recognition

Papers

Showing 301–325 of 435 papers

TitleStatusHype
Towards Relevance and Sequence Modeling in Language Recognition—0
End-to-end Recurrent Denoising Autoencoder Embeddings for Speaker Identification—0
Real-time, Universal, and Robust Adversarial Attacks Against Speaker Recognition Systems—0
An Open-set Recognition and Few-Shot Learning Dataset for Audio Event Classification in Domestic EnvironmentsCode0
Disentangled Speech Embeddings using Cross-modal Self-supervision—0
Deep Speaker Embeddings for Far-Field Speaker Recognition on Short Utterances—0
x-vectors meet emotions: A study on dependencies between emotion and speaker recognition—0
LEAP System for SRE19 CTS Challenge -- Improvements and Error Analysis—0
Multi-task Learning for Speaker Verification and Voice Trigger Detection—0
Robust Speaker Recognition Using Speech Enhancement And Attention Model—0
Deep Representation Learning in Speech Processing: Challenges, Recent Advances, and Future Trends—0
THUEE system description for NIST 2019 SRE CTS Challenge—0
End-to-end training of time domain audio separation and recognition—0
Short-duration Speaker Verification (SdSV) Challenge 2021: the Challenge Evaluation Plan—0
VoxSRC 2019: The first VoxCeleb Speaker Recognition Challenge—0
Deep learning methods in speaker recognition: a review—0
Who is Real Bob? Adversarial Attacks on Speaker Recognition SystemsCode0
Robust speaker recognition using unsupervised adversarial invarianceCode0
CN-CELEB: a challenging Chinese speaker recognition datasetCode0
Mockingjay: Unsupervised Speech Representation Learning with Deep Bidirectional Transformer EncodersCode0
Structural sparsification for Far-field Speaker Recognition with GNA—0
Delving into VoxCeleb: environment invariant speaker recognition—0
Filterbank design for end-to-end speech separationCode0
Contextual Joint Factor Acoustic Embeddings—0
Frequency and temporal convolutional attention for text-independent speaker recognition—0
Show:102550
← PrevPage 13 of 18Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1w2v2-aamEER1.88—Unverified
2WavLM+ECAPA-TDNNEER0.39—Unverified