SOTAVerified

Speaker Identification

Papers

Showing 201–248 of 248 papers

TitleStatusHype
Cosine similarity-based adversarial process—0
Large-Scale Speaker Diarization of Radio Broadcast Archives—0
A user study to compare two conversational assistants designed for people with hearing impairments—0
Many-to-Many Voice Conversion with Out-of-Dataset Speaker Support—0
Experiments on Open-Set Speaker Identification with Discriminatively Trained Neural Networks—0
Advanced Rich Transcription System for Estonian Speech—0
Histogram Transform-based Speaker Identification—0
Weakly Supervised Training of Speaker Identification Models—0
On Learning Associations of Faces and VoicesCode0
Computer-assisted Speaker Diarization: How to Evaluate Human Corrections—0
Matics Software Suite: New Tools for Evaluation and Data Exploration—0
Identifying Speakers and Addressees in Dialogues Extracted from Literary Fiction—0
Evaluation of Automatic Formant Trackers—0
VAST: A Corpus of Video Annotation for Speech Technologies—0
Seeing Voices and Hearing Faces: Cross-modal biometric matching—0
Neural Predictive Coding using Convolutional Neural Networks towards Unsupervised Learning of Speaker Characteristics—0
From Benedict Cumberbatch to Sherlock Holmes: Character Identification in TV series without a Script—0
Speaker identification from the sound of the human breath—0
Identifying Speakers and Listeners of Quoted Speech in Literary Works—0
基於聽覺感知模型之類神經網路及其在語者識別上之應用 (Two-stage Attentional Auditory Model Inspired Neural Network and Its Application to Speaker Identification) [In Chinese]—0
Story Comprehension for Predicting What Happens Next—0
Comparison of Multiple Features and Modeling Methods for Text-dependent Speaker Verification—0
Face Recognition with Machine Learning in OpenCV_ Fusion of the results with the Localization Data of an Acoustic Camera for Speaker Identification—0
Text-based Speaker Identification on Multiparty Dialogues Using Multi-document Convolutional Neural Networks—0
Speaker Identification in each of the Neutral and Shouted Talking Environments based on Gender-Dependent Approach Using SPHMMs—0
Deep Speaker: an End-to-End Neural Speaker Embedding SystemCode0
Can Musical Emotion Be Quantified With Neural Jitter Or Shimmer? A Novel EEG Based Study With Hindustani Classical Music—0
An Unsupervised Speaker Clustering Technique based on SOM and I-vectors for Speech Recognition Systems—0
Discrimination between Similar Languages, Varieties and Dialects using CNN- and LSTM-based Deep Neural Networks—0
A domain-agnostic approach for opinion prediction on speechCode0
Monaural Multi-Talker Speech Recognition using Factorial Speech Processing Models—0
Curie: A method for protecting SVM Classifier from Poisoning Attack—0
Look, Listen and Learn - A Multimodal LSTM for Speaker Identification—0
A Novel Minimum Divergence Approach to Robust Speaker Identification—0
Speaker Identification From Youtube Obtained Data—0
Invited Talk: IBM Cognitive Computing - An NLP Renaissance!—0
基於稀疏表示之語者識別 (Sparse Representation Based Speaker Identification) [In Chinese]—0
On the Use of Different Feature Extraction Methods for Linear and Non Linear kernels—0
A Multi Level Data Fusion Approach for Speaker Identification on Telephone Speech—0
The DIRHA simulated corpus—0
Comparison of Gender- and Speaker-adaptive Emotion Recognition—0
The RATS Collection: Supporting HLT Research with Degraded Audio Data—0
A Joint Model for Quotation Attribution and Coreference Resolution—0
From Speaker Identification to Affective Analysis: A Multi-Step System for Analyzing Children's Stories—0
A Generative Product-of-Filters Model of AudioCode0
Identification of Speakers in Novels—0
MKPLS: Manifold Kernel Partial Least Squares for Lipreading and Speaker Identification—0
L'identification du locuteur : 20 ans de t\'emoignage dans les cours de Justice. Le cas du LIPSADON laboratoire ind\'ependant de police scientifique (Forensic speaker identification: 20 years of scientific testimonies in courts of Justice. The case of LIPSADON ``forensics independent laboratory'') [in French]—0
Show:102550
← PrevPage 5 of 5Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1MSM-MAETop-1 (%)96.6—Unverified
2M2D/0.6Top-1 (%)96.5—Unverified
3M2D/0.7Top-1 (%)96.3—Unverified
4M2D ratio=0.6Top-1 (%)94.8—Unverified
5AudioMAE (local)Top-1 (%)94.8—Unverified
6ATST Base (ours)Top-1 (%)94.3—Unverified
7AudioMAE (global)Top-1 (%)94.1—Unverified
8AutoSpeech (N=8,C=128)Top-1 (%)87.66—Unverified
9SSAST-FRAMETop-1 (%)80.8—Unverified
10SSAMBATop-1 (%)70.1—Unverified
#ModelMetricClaimedVerifiedStatus
1Fuzzy RetrievalTop-1 (%)67.77—Unverified
#ModelMetricClaimedVerifiedStatus
1Fuzzy RetrievalTop-1 (%)80.83—Unverified
#ModelMetricClaimedVerifiedStatus
1Fuzzy RetrievalTop-1 (%)95.13—Unverified