SOTAVerified

Speaker Identification

Papers

Showing 201–248 of 248 papers

TitleStatusHype
Characteristic-Specific Partial Fine-Tuning for Efficient Emotion and Speaker Adaptation in Codec Language Text-to-Speech Models—0
Comparison of Gender- and Speaker-adaptive Emotion Recognition—0
Comparison of Multiple Features and Modeling Methods for Text-dependent Speaker Verification—0
Computer-assisted Speaker Diarization: How to Evaluate Human Corrections—0
Computing with Hypervectors for Efficient Speaker Identification—0
Cosine similarity-based adversarial process—0
Cross-Lingual Speaker Identification from Weak Local Evidence—0
Curie: A method for protecting SVM Classifier from Poisoning Attack—0
DASB -- Discrete Audio and Speech Benchmark—0
Deep Neural Networks for Automatic Speech Processing: A Survey from Large Corpora to Limited Data—0
Deep versus Wide: An Analysis of Student Architectures for Task-Agnostic Knowledge Distillation of Self-Supervised Speech Models—0
Delving into VoxCeleb: environment invariant speaker recognition—0
Discrimination between Similar Languages, Varieties and Dialects using CNN- and LSTM-based Deep Neural Networks—0
Effect of utterance duration and phonetic content on speaker identification using second-order statistical methods—0
Efficiency-oriented approaches for self-supervised speech representation learning—0
Emirati-Accented Speaker Identification in Stressful Talking Conditions—0
End-to-End Diarization for Variable Number of Speakers with Local-Global Networks and Discriminative Speaker Embeddings—0
End-to-end Multichannel Speaker-Attributed ASR: Speaker Guided Decoder and Input Feature Analysis—0
End-to-end Recurrent Denoising Autoencoder Embeddings for Speaker Identification—0
End-to-End Speaker-Attributed ASR with Transformer—0
Enhancing Open-Set Speaker Identification through Rapid Tuning with Speaker Reciprocal Points and Negative Sample—0
Ensemble knowledge distillation of self-supervised speech models—0
Evaluating Speaker Identity Coding in Self-supervised Models and Humans—0
Evaluation of Automatic Formant Trackers—0
ExARN: self-attending RNN for target speaker extraction—0
Experiments on Open-Set Speaker Identification with Discriminatively Trained Neural Networks—0
Exploring VQ-VAE with Prosody Parameters for Speaker Anonymization—0
Face Recognition with Machine Learning in OpenCV_ Fusion of the results with the Localization Data of an Acoustic Camera for Speaker Identification—0
Few-Shot Speaker Identification Using Depthwise Separable Convolutional Network with Channel Attention—0
Few-Shot Speaker Identification Using Lightweight Prototypical Network with Feature Grouping and Interaction—0
French Listening Tests for the Assessment of Intelligibility, Quality, and Identity of Body-Conducted Speech Enhancement—0
From Benedict Cumberbatch to Sherlock Holmes: Character Identification in TV series without a Script—0
From Dialect Gaps to Identity Maps: Tackling Variability in Speaker Verification—0
From Speaker Identification to Affective Analysis: A Multi-Step System for Analyzing Children's Stories—0
Fusion of Embeddings Networks for Robust Combination of Text Dependent and Independent Speaker Recognition—0
Graph-based Label Propagation for Semi-Supervised Speaker Identification—0
Graph-based Multi-View Fusion and Local Adaptation: Mitigating Within-Household Confusability for Speaker Identification—0
HiSSNet: Sound Event Detection and Speaker Identification via Hierarchical Prototypical Networks for Low-Resource Headphones—0
Histogram Transform-based Speaker Identification—0
How Far Are We from Robust Voice Conversion: A Survey—0
How Redundant Is the Transformer Stack in Speech Representation Models?—0
HPP-Voice: A Large-Scale Evaluation of Speech Embeddings for Multi-Phenotypic Classification—0
H-VECTORS: Utterance-level Speaker Embedding Using A Hierarchical Attention Model—0
Hypothesis Stitcher for End-to-End Speaker-attributed ASR on Long-form Multi-talker Recordings—0
Identification of Speakers in Novels—0
Identifying Source Speakers for Voice Conversion based Spoofing Attacks on Speaker Verification Systems—0
Identifying Speakers and Addressees in Dialogues Extracted from Literary Fiction—0
Identifying Speakers and Listeners of Quoted Speech in Literary Works—0
Show:102550
← PrevPage 5 of 5Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1MSM-MAETop-1 (%)96.6—Unverified
2M2D/0.6Top-1 (%)96.5—Unverified
3M2D/0.7Top-1 (%)96.3—Unverified
4M2D ratio=0.6Top-1 (%)94.8—Unverified
5AudioMAE (local)Top-1 (%)94.8—Unverified
6ATST Base (ours)Top-1 (%)94.3—Unverified
7AudioMAE (global)Top-1 (%)94.1—Unverified
8AutoSpeech (N=8,C=128)Top-1 (%)87.66—Unverified
9SSAST-FRAMETop-1 (%)80.8—Unverified
10SSAMBATop-1 (%)70.1—Unverified
#ModelMetricClaimedVerifiedStatus
1Fuzzy RetrievalTop-1 (%)67.77—Unverified
#ModelMetricClaimedVerifiedStatus
1Fuzzy RetrievalTop-1 (%)80.83—Unverified
#ModelMetricClaimedVerifiedStatus
1Fuzzy RetrievalTop-1 (%)95.13—Unverified