SOTAVerified

Speaker Recognition

Speaker Recognition is the process of identifying or confirming the identity of a person given his speech segments.

Source: Margin Matters: Towards More Discriminative Deep Neural Network Embeddings for Speaker Recognition

Papers

Showing 401–425 of 435 papers

TitleStatusHype
Contextual Joint Factor Acoustic Embeddings—0
Convexity-based Pruning of Speech Representation Models—0
CopyPaste: An Augmentation Method for Speech Emotion Recognition—0
Cosine Scoring with Uncertainty for Neural Speaker Embedding—0
Cross-modal Speaker Verification and Recognition: A Multilingual Perspective—0
Data augmentation versus noise compensation for x- vector speaker recognition systems in noisy environments—0
Attention and DCT based Global Context Modeling for Text-independent Speaker Recognition—0
Deep CNN based feature extractor for text-prompted speaker recognition—0
Deep factorization for speech signal—0
Deep Learning for Single and Multi-Session i-Vector Speaker Recognition—0
Deep learning methods in speaker recognition: a review—0
DeepMSRF: A novel Deep Multimodal Speaker Recognition framework with Feature selection—0
Deep neural network based i-vector mapping for speaker verification using short utterances—0
Deep Neural Networks for Automatic Speaker Recognition Do Not Learn Supra-Segmental Temporal Features—0
Deep Representation Learning in Speech Processing: Challenges, Recent Advances, and Future Trends—0
DeepSonar: Towards Effective and Robust Detection of AI-Synthesized Fake Voices—0
Deep Speaker Embeddings for Far-Field Speaker Recognition on Short Utterances—0
Deep Speaker Vectors for Semi Text-independent Speaker Verification—0
DeepVOX: Discovering Features from Raw Audio for Speaker Recognition in Non-ideal Audio Signals—0
Delving into VoxCeleb: environment invariant speaker recognition—0
Detecting Agreement in Multi-party Conversational AI—0
Detection and Analysis of Content Creator Collaborations in YouTube Videos using Face- and Speaker-Recognition—0
Differences in Speaker Individualising Information between Case Particles and Fillers in Spoken Japanese—0
Discriminatively Re-trained i-vector Extractor for Speaker Recognition—0
Disentangled representation learning for multilingual speaker recognition—0
Show:102550
← PrevPage 17 of 18Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1w2v2-aamEER1.88—Unverified
2WavLM+ECAPA-TDNNEER0.39—Unverified