SOTAVerified

Text to Speech

import gTTS import os def text_to_speech_kurdish(text, output_file="output.mp3"): # گۆڕینی نووسین بۆ دەنگ بە زمانی کوردی (هەڵبژاردنی زمانی "ku" بۆ کوردی) tts = gTTS(text=text, lang='ku', slow=False) tts.save(output_file) os.system(f"start {output_file}") # کردنەوەی فایلە دەنگییەکە (لە Windows) # نموونە: text_to_speech_kurdish("سڵاو، ئەمە دەنگی منە بە زمانی کوردی.")

Papers

Showing 551575 of 1419 papers

TitleStatusHype
SeamlessM4T: Massively Multilingual & Multimodal Machine TranslationCode2
Multi-GradSpeech: Towards Diffusion-based Multi-Speaker Text-to-speech Using Consistent Diffusion Models0
AffectEcho: Speaker Independent and Language-Agnostic Emotion and Affect Transfer for Speech Synthesis0
SpeechX: Neural Codec Language Model as a Versatile Speech Transformer0
Text-to-Video: a Two-stage Framework for Zero-shot Identity-agnostic Talking-head GenerationCode0
AudioLDM 2: Learning Holistic Audio Generation with Self-supervised PretrainingCode4
Towards an AI to Win Ghana's National Science and Maths QuizCode1
Let's Give a Voice to Conversational Agents in Virtual RealityCode0
Textless Unit-to-Unit training for Many-to-Many Multilingual Speech-to-Speech TranslationCode1
SALTTS: Leveraging Self-Supervised Speech Representations for improved Text-to-Speech Synthesis0
Multilingual context-based pronunciation learning for Text-to-Speech0
DiffProsody: Diffusion-based Latent Prosody Generation for Expressive Speech Synthesis with Prosody Conditional Adversarial TrainingCode1
VITS2: Improving Quality and Efficiency of Single-Stage Text-to-Speech with Adversarial Learning and Architecture DesignCode2
Improving grapheme-to-phoneme conversion by learning pronunciations from speech recordings0
Comparing normalizing flows and diffusion models for prosody and acoustic modelling in text-to-speech0
Improving TTS for Shanghainese: Addressing Tone Sandhi via Word SegmentationCode1
METTS: Multilingual Emotional Text-to-Speech by Cross-speaker and Cross-lingual Emotion Transfer0
ÌròyìnSpeech: A multi-purpose Yorùbá Speech CorpusCode1
Minimally-Supervised Speech Synthesis with Conditional Diffusion Model and Language Model: A Comparative Study of Semantic Coding0
SC VALL-E: Style-Controllable Zero-Shot Text to Speech SynthesizerCode1
SLMGAN: Exploiting Speech Language Model Representations for Unsupervised Zero-Shot Voice Conversion in GANs0
Mega-TTS 2: Boosting Prompting Mechanisms for Zero-Shot Speech Synthesis0
Controllable Emphasis with zero data for text-to-speech0
On the Use of Self-Supervised Speech Representations in Spontaneous Speech Synthesis0
Artificial Eye for the Blind0
Show:102550
← PrevPage 23 of 57Next →

No leaderboard results yet.