Voice Conversion

I remember all the summer days Drinking wine in the sunshine I hope it never leaves And I remember all the summer nights Staring at you in the moonlight I hope you never leave 'cause baby You're so good to me You have all that all that I ever need It's easy to love you So easy to love you Ooh you know it's true The best part of being with you To know you're with me It's not so hard to say It's easy to love you I remember all those winter days frozen In the cold tryin' to get you home Should I be moving in, we can be together then Remember spending all those winter nights Stayin' inside by the warm fire Yeah you gotta know that I can never let you go You and I have the rest of our lives to say It's easy to love you So easy to love you Ooh you know it's true The best part of being with you To know you're with me It's not so hard to say It's easy to love you Can anybody else see it? Mm, can anybody else see what I do? Can anybody else feel it? Oh, can anybody else feel the way I do? But now I'm with you Hard to forget all the moments when We'd be sitting there hoping it would never end 'Cause this is meant to be So baby, will you marry me? It's easy to love you So easy to love you Ooh, you know it's true The best part of being with you To know you are with me It's not so hard to say It's easy to love you You and me will be together I know our love will last forever You and me will be together I know our love will last forever You know it's true The best part of being with you You're easy to love

Source: Joint training framework for text-to-speech and voice conversion using multi-source Tacotron and WaveNet

Papers

Recently Added Most Hyped Most Active Needs Verification Most Verified

Showing 51–75 of 520 papers

Title	Date	Tasks	Status	Hype	Score
Disentanglement in a GAN for Unconditional Speech Synthesis	Jul 4, 2023	DisentanglementGenerative Adversarial Network	CodeCode Available	1	5
Assem-VC: Realistic Voice Conversion by Assembling Modern Speech Synthesis Techniques	Apr 2, 2021	DecoderRhythm	CodeCode Available	1	5
FMFCC-A: A Challenging Mandarin Dataset for Synthetic Speech Detection	Oct 18, 2021	Speech SynthesisSynthetic Speech Detection	CodeCode Available	1	5
MaskCycleGAN-VC: Learning Non-parallel Voice Conversion with Filling in Frames	Feb 25, 2021	Voice Conversion	CodeCode Available	1	5
GAN You Hear Me? Reclaiming Unconditional Speech Synthesis from Diffusion Models	Oct 11, 2022	DisentanglementGenerative Adversarial Network	CodeCode Available	1	5
FSD: An Initial Chinese Dataset for Fake Song Detection	Sep 5, 2023	Audio Deepfake DetectionDeepFake Detection	CodeCode Available	1	5
Improving fairness for spoken language understanding in atypical speech with Text-to-Speech	Nov 16, 2023	Data AugmentationFairness	CodeCode Available	1	5
Neural Analysis and Synthesis: Reconstructing Speech from Self-Supervised Representations	Oct 27, 2021	Voice Conversion	CodeCode Available	1	5
Any-to-Many Voice Conversion with Location-Relative Sequence-to-Sequence Modeling	Sep 6, 2020	feature selectionspeech-recognition	CodeCode Available	1	5
LDNet: Unified Listener Dependent Modeling in MOS Prediction for Synthetic Speech	Oct 18, 2021	Voice Conversion	CodeCode Available	1	5
CycleTransGAN-EVC: A CycleGAN-based Emotional Voice Conversion Model with Transformer	Nov 30, 2021	Voice Conversion	CodeCode Available	1	5
crank: An Open-Source Software for Nonparallel Voice Conversion Based on Vector-Quantized Variational Autoencoder	Mar 4, 2021	Voice Conversion	CodeCode Available	1	5
kNN-SVC: Robust Zero-Shot Singing Voice Conversion with Additive Synthesis and Concatenation Smoothness Optimization	Apr 8, 2025	Voice Conversion	CodeCode Available	1	5
Limited Data Emotional Voice Conversion Leveraging Text-to-Speech: Two-stage Sequence-to-Sequence Training	Mar 31, 2021	text-to-speechText to Speech	CodeCode Available	1	5
Baseline System of Voice Conversion Challenge 2020 with Cyclic Variational Autoencoder and Parallel WaveGAN	Oct 9, 2020	Generative Adversarial NetworkTask 2	CodeCode Available	1	5
Deep Learning Based Assessment of Synthetic Speech Naturalness	Apr 23, 2021	Deep LearningPrediction	CodeCode Available	1	5
Controllable and Interpretable Singing Voice Decomposition via Assem-VC	Oct 25, 2021	Voice Conversion	CodeCode Available	1	5
Anonymizing Speech: Evaluating and Designing Speaker Anonymization Techniques	Aug 5, 2023	QuantizationSpeaker anonymization	CodeCode Available	1	5
Converting Anyone's Emotion: Towards Speaker-Independent Emotional Voice Conversion	May 13, 2020	DecoderVoice Conversion	CodeCode Available	1	5
ControlVC: Zero-Shot Voice Conversion with Time-Varying Controls on Pitch and Speed	Sep 23, 2022	Pitch controlSpeech Synthesis	CodeCode Available	1	5
A Comparative Study of Self-supervised Speech Representation Based Voice Conversion	Jul 10, 2022	Voice Conversion	CodeCode Available	1	5
BiSinger: Bilingual Singing Voice Synthesis	Sep 25, 2023	Singing Voice Synthesistext-to-speech	CodeCode Available	1	5
CSLP-AE: A Contrastive Split-Latent Permutation Autoencoder Framework for Zero-Shot Electroencephalography Signal Conversion	Nov 13, 2023	Contrastive LearningEEG	CodeCode Available	1	5
CycleGAN-VC3: Examining and Improving CycleGAN-VCs for Mel-spectrogram Conversion	Oct 22, 2020	Voice Conversion	CodeCode Available	1	5
Emotionless: Privacy-Preserving Speech Analysis for Voice Assistants	Aug 9, 2019	Emotion RecognitionPrivacy Preserving	CodeCode Available	1	5

Show:10 25 50

← PrevPage 3 of 21Next →

All datasets ZeroSpeech 2019 English LibriSpeech test-clean VCTK

Benchmark Results

#	Model	Metric	Claimed	Verified	Status
1	VQ-CPC	Speaker Similarity	3.8	—	Unverified
2	VQ-VAE	Speaker Similarity	3.49	—	Unverified

#	Model	Metric	Claimed	Verified	Status
1	kNN-VC (prematched HiFiGAN)	Character Error Rate (CER)	2.96	—	Unverified

#	Model	Metric	Claimed	Verified	Status
1	DISSC	Total Length Error (TLE)	0.83	—	Unverified