SOTAVerified

Voice Cloning

Voice cloning is a highly desired feature for personalized speech interfaces. Neural voice cloning system learns to synthesize a person’s voice from only a few audio samples.

Papers

Showing 101–112 of 112 papers

TitleStatusHype
Enhancing Synthetic Training Data for Speech Commands: From ASR-Based Filtering to Domain Adaptation in SSL Latent Space—0
Evaluating Voice Conversion-based Privacy Protection against Informed Attackers—0
Exploring Timbre Disentanglement in Non-Autoregressive Cross-Lingual Text-to-Speech—0
Expressive Neural Voice Cloning—0
High-Fidelity Speech Synthesis with Minimal Supervision: All Using Diffusion Models—0
Hindi audio-video-Deepfake (HAV-DF): A Hindi language-based Audio-video Deepfake Dataset—0
Improve Cross-lingual Voice Cloning Using Low-quality Code-switched Data—0
Improve few-shot voice cloning using multi-modal learning—0
Investigating on Incorporating Pretrained and Learnable Speaker Representations for Multi-Speaker Multi-Style Text-to-Speech—0
"It's not a representation of me": Examining Accent Bias and Digital Exclusion in Synthetic AI Voice Services—0
Just Because We Camp, Doesn't Mean We Should: The Ethics of Modelling Queer Voices—0
Latent linguistic embedding for cross-lingual text-to-speech and voice conversion—0
Show:102550
← PrevPage 5 of 5Next →

No leaderboard results yet.