SOTAVerified

Voice Cloning

Voice cloning is a highly desired feature for personalized speech interfaces. Neural voice cloning system learns to synthesize a person’s voice from only a few audio samples.

Papers

Showing 2650 of 112 papers

TitleStatusHype
DubWise: Video-Guided Speech Duration Control in Multimodal LLM-based Text-to-Speech for Dubbing0
DMOSpeech: Direct Metric Optimization via Distilled Diffusion Model in Zero-Shot Speech Synthesis0
Beyond Face Swapping: A Diffusion-Based Digital Human Benchmark for Multimodal Deepfake Detection0
Empowering Global Voices: A Data-Efficient, Phoneme-Tone Adaptive Approach to High-Fidelity Speech Synthesis0
Deepfake Technology Unveiled: The Commoditization of AI and Its Impact on Digital Trust0
Latent linguistic embedding for cross-lingual text-to-speech and voice conversion0
Augmentation through Laundering Attacks for Audio Spoof Detection0
Advancing NAM-to-Speech Conversion with Novel Methods and the MultiNAM Dataset0
Meta-Voice: Fast few-shot style transfer for expressive voice cloning using meta learning0
De-AntiFake: Rethinking the Protective Perturbations Against Voice Cloning Attacks0
Data Efficient Voice Cloning for Neural Singing Synthesis0
Empowering Communication: Speech Technology for Indian and Western Accents through AI-powered Speech Synthesis0
Improve few-shot voice cloning using multi-modal learning0
CUHK-EE Voice Cloning System for ICASSP 2021 M2VoC Challenge0
Improve Cross-lingual Voice Cloning Using Low-quality Code-switched Data0
Hindi audio-video-Deepfake (HAV-DF): A Hindi language-based Audio-video Deepfake Dataset0
Cross-lingual Multi-speaker Text-to-speech Synthesis for Voice Cloning without Using Parallel Corpus for Unseen Speakers0
A multi-speaker multi-lingual voice cloning system based on vits2 for limmits 2024 challenge0
MARS6: A Small and Robust Hierarchical-Codec Text-to-Speech Model0
"It's not a representation of me": Examining Accent Bias and Digital Exclusion in Synthetic AI Voice Services0
Just Because We Camp, Doesn't Mean We Should: The Ethics of Modelling Queer Voices0
High-Fidelity Speech Synthesis with Minimal Supervision: All Using Diffusion Models0
Collaborative Watermarking for Adversarial Speech Synthesis0
Expressive Neural Voice Cloning0
Algorithms For Automatic Accentuation And Transcription Of Russian Texts In Speech Recognition Systems0
Show:102550
← PrevPage 2 of 5Next →

No leaderboard results yet.