Voice Conversion

I remember all the summer days Drinking wine in the sunshine I hope it never leaves And I remember all the summer nights Staring at you in the moonlight I hope you never leave 'cause baby You're so good to me You have all that all that I ever need It's easy to love you So easy to love you Ooh you know it's true The best part of being with you To know you're with me It's not so hard to say It's easy to love you I remember all those winter days frozen In the cold tryin' to get you home Should I be moving in, we can be together then Remember spending all those winter nights Stayin' inside by the warm fire Yeah you gotta know that I can never let you go You and I have the rest of our lives to say It's easy to love you So easy to love you Ooh you know it's true The best part of being with you To know you're with me It's not so hard to say It's easy to love you Can anybody else see it? Mm, can anybody else see what I do? Can anybody else feel it? Oh, can anybody else feel the way I do? But now I'm with you Hard to forget all the moments when We'd be sitting there hoping it would never end 'Cause this is meant to be So baby, will you marry me? It's easy to love you So easy to love you Ooh, you know it's true The best part of being with you To know you are with me It's not so hard to say It's easy to love you You and me will be together I know our love will last forever You and me will be together I know our love will last forever You know it's true The best part of being with you You're easy to love

Source: Joint training framework for text-to-speech and voice conversion using multi-source Tacotron and WaveNet

Papers

Recently Added Most Hyped Most Active Needs Verification Most Verified

Showing 51–100 of 520 papers

Title	Date	Tasks	Status	Hype	Score
HM-Conformer: A Conformer-based audio deepfake detection system with hierarchical pooling and multi-level classification token aggregation methods	Sep 15, 2023	Audio Deepfake DetectionDeepFake Detection	CodeCode Available	1	5
Assem-VC: Realistic Voice Conversion by Assembling Modern Speech Synthesis Techniques	Apr 2, 2021	DecoderRhythm	CodeCode Available	1	5
FragmentVC: Any-to-Any Voice Conversion by End-to-End Extracting and Fusing Fine-Grained Voice Fragments With Attention	Oct 27, 2020	DisentanglementSpeaker Verification	CodeCode Available	1	5
MediumVC: Any-to-any voice conversion using synthetic specific-speaker speeches as intermedium features	Oct 6, 2021	Voice Conversion	CodeCode Available	1	5
FMFCC-A: A Challenging Mandarin Dataset for Synthetic Speech Detection	Oct 18, 2021	Speech SynthesisSynthetic Speech Detection	CodeCode Available	1	5
MOSNet: Deep Learning based Objective Assessment for Voice Conversion	Apr 17, 2019	Deep LearningVoice Conversion	CodeCode Available	1	5
SpeechLMScore: Evaluating speech generation using speech language model	Dec 8, 2022	Language ModelingLanguage Modelling	CodeCode Available	1	5
Neural Analysis and Synthesis: Reconstructing Speech from Self-Supervised Representations	Oct 27, 2021	Voice Conversion	CodeCode Available	1	5
Robust Training of Vector Quantized Bottleneck Models	May 18, 2020	ClusteringDisentanglement	CodeCode Available	1	5
One-class learning towards generalized voice spoofing detection	Oct 27, 2020	Speaker Verificationtext-to-speech	CodeCode Available	1	5
FSD: An Initial Chinese Dataset for Fake Song Detection	Sep 5, 2023	Audio Deepfake DetectionDeepFake Detection	CodeCode Available	1	5
One-shot Voice Conversion by Separating Speaker and Content Representations with Instance Normalization	Apr 10, 2019	Voice Conversion	CodeCode Available	1	5
Emotionless: Privacy-Preserving Speech Analysis for Voice Assistants	Aug 9, 2019	Emotion RecognitionPrivacy Preserving	CodeCode Available	1	5
F0-consistent many-to-many non-parallel voice conversion via conditional autoencoder	Apr 15, 2020	Style TransferVoice Conversion	CodeCode Available	1	5
S2VC: A Framework for Any-to-Any Voice Conversion with Self-Supervised Pretrained Representations	Apr 7, 2021	Self-Supervised LearningVoice Conversion	CodeCode Available	1	5
Seen and Unseen emotional style transfer for voice conversion with a new emotional speech dataset	Oct 28, 2020	DecoderEmotion Recognition	CodeCode Available	1	5
Efficient Non-Autoregressive GAN Voice Conversion using VQWav2vec Features and Dynamic Convolution	Mar 31, 2022	Voice Conversion	CodeCode Available	1	5
Disentanglement in a GAN for Unconditional Speech Synthesis	Jul 4, 2023	DisentanglementGenerative Adversarial Network	CodeCode Available	1	5
Retriever: Learning Content-Style Representation as a Token-Level Bipartite Graph	Feb 24, 2022	DecoderQuantization	CodeCode Available	1	5
DeID-VC: Speaker De-identification via Zero-shot Pseudo Voice Conversion	Sep 9, 2022	De-identificationSpeaker Verification	CodeCode Available	1	5
Diffusion-Based Voice Conversion with Fast Maximum Likelihood Sampling Scheme	Sep 28, 2021	Speech SynthesisVoice Conversion	CodeCode Available	1	5
BiSinger: Bilingual Singing Voice Synthesis	Sep 25, 2023	Singing Voice Synthesistext-to-speech	CodeCode Available	1	5
A Comparative Study of Self-supervised Speech Representation Based Voice Conversion	Jul 10, 2022	Voice Conversion	CodeCode Available	1	5
Defending Your Voice: Adversarial Attack on Voice Conversion	May 18, 2020	Adversarial AttackVoice Conversion	CodeCode Available	1	5
Building Bilingual and Code-Switched Voice Conversion with Limited Training Data Using Embedding Consistency Loss	Apr 22, 2021	Voice CloningVoice Conversion	CodeCode Available	1	5
Any-to-Many Voice Conversion with Location-Relative Sequence-to-Sequence Modeling	Sep 6, 2020	feature selectionspeech-recognition	CodeCode Available	1	5
DuTa-VC: A Duration-aware Typical-to-atypical Voice Conversion Approach with Diffusion Probabilistic Model	Jun 18, 2023	Data AugmentationDecoder	CodeCode Available	1	5
Emo-StarGAN: A Semi-Supervised Any-to-Many Non-Parallel Emotion-Preserving Voice Conversion	Sep 14, 2023	Voice Conversion	CodeCode Available	1	5
Rhythm Modeling for Voice Conversion	Jul 12, 2023	RhythmVoice Conversion	CodeCode Available	1	5
CinC-GAN for Effective F0 prediction for Whisper-to-Normal Speech Conversion	Aug 18, 2020	PredictionVoice Conversion	CodeCode Available	1	5
End-to-End Zero-Shot Voice Conversion with Location-Variable Convolutions	May 19, 2022	Speech SynthesisStyle Transfer	CodeCode Available	1	5
Evaluating Methods for Ground-Truth-Free Foreign Accent Conversion	Sep 5, 2023	Voice Conversion	CodeCode Available	1	5
FastSVC: Fast Cross-Domain Singing Voice Conversion with Feature-wise Linear Modulation	Nov 11, 2020	Voice Conversion	CodeCode Available	1	5
Deep Learning Based Assessment of Synthetic Speech Naturalness	Apr 23, 2021	Deep LearningPrediction	CodeCode Available	1	5
Baseline System of Voice Conversion Challenge 2020 with Cyclic Variational Autoencoder and Parallel WaveGAN	Oct 9, 2020	Generative Adversarial NetworkTask 2	CodeCode Available	1	5
Phonetic Posteriorgrams based Many-to-Many Singing Voice Conversion via Adversarial Training	Dec 3, 2020	Audio GenerationDisentanglement	CodeCode Available	1	5
CycleTransGAN-EVC: A CycleGAN-based Emotional Voice Conversion Model with Transformer	Nov 30, 2021	Voice Conversion	CodeCode Available	1	5
Anonymizing Speech: Evaluating and Designing Speaker Anonymization Techniques	Aug 5, 2023	QuantizationSpeaker anonymization	CodeCode Available	1	5
Controllable and Interpretable Singing Voice Decomposition via Assem-VC	Oct 25, 2021	Voice Conversion	CodeCode Available	1	5
ControlVC: Zero-Shot Voice Conversion with Time-Varying Controls on Pitch and Speed	Sep 23, 2022	Pitch controlSpeech Synthesis	CodeCode Available	1	5
Pretraining Techniques for Sequence-to-Sequence Voice Conversion	Aug 7, 2020	Automatic Speech RecognitionAutomatic Speech Recognition (ASR)	CodeCode Available	1	5
AraBERT: Transformer-based Model for Arabic Language Understanding	Feb 28, 2020	modelnamed-entity-recognition	CodeCode Available	1	5
AutoVisual Fusion Suite: A Comprehensive Evaluation of Image Segmentation and Voice Conversion Tools on HuggingFace Platform	Dec 17, 2023	Image SegmentationSegmentation	CodeCode Available	1	5
CycleGAN-VC3: Examining and Improving CycleGAN-VCs for Mel-spectrogram Conversion	Oct 22, 2020	Voice Conversion	CodeCode Available	1	5
Emotional Voice Conversion: Theory, Databases and ESD	May 31, 2021	Voice Conversion	CodeCode Available	1	5
Blow: a single-scale hyperconditioned flow for non-parallel raw-audio voice conversion	Jun 3, 2019	Audio GenerationVoice Conversion	CodeCode Available	1	5
A unified one-shot prosody and speaker conversion system with self-supervised discrete speech units	Nov 12, 2022	RhythmVoice Conversion	CodeCode Available	1	5
crank: An Open-Source Software for Nonparallel Voice Conversion Based on Vector-Quantized Variational Autoencoder	Mar 4, 2021	Voice Conversion	CodeCode Available	1	5
CSLP-AE: A Contrastive Split-Latent Permutation Autoencoder Framework for Zero-Shot Electroencephalography Signal Conversion	Nov 13, 2023	Contrastive LearningEEG	CodeCode Available	1	5
Improving fairness for spoken language understanding in atypical speech with Text-to-Speech	Nov 16, 2023	Data AugmentationFairness	CodeCode Available	1	5

Show:10 25 50

← PrevPage 2 of 11Next →

All datasets ZeroSpeech 2019 English LibriSpeech test-clean VCTK

Benchmark Results

#	Model	Metric	Claimed	Verified	Status
1	VQ-CPC	Speaker Similarity	3.8	—	Unverified
2	VQ-VAE	Speaker Similarity	3.49	—	Unverified

#	Model	Metric	Claimed	Verified	Status
1	kNN-VC (prematched HiFiGAN)	Character Error Rate (CER)	2.96	—	Unverified

#	Model	Metric	Claimed	Verified	Status
1	DISSC	Total Length Error (TLE)	0.83	—	Unverified