SOTAVerified

Text to Speech

import gTTS import os def text_to_speech_kurdish(text, output_file="output.mp3"): # گۆڕینی نووسین بۆ دەنگ بە زمانی کوردی (هەڵبژاردنی زمانی "ku" بۆ کوردی) tts = gTTS(text=text, lang='ku', slow=False) tts.save(output_file) os.system(f"start {output_file}") # کردنەوەی فایلە دەنگییەکە (لە Windows) # نموونە: text_to_speech_kurdish("سڵاو، ئەمە دەنگی منە بە زمانی کوردی.")

Papers

Showing 501–525 of 1419 papers

TitleStatusHype
An overview of text-to-speech systems and media applications—0
Evaluating Text-to-Speech Synthesis from a Large Discrete Token-based Speech Language Model—0
Efficient Generative Modeling with Residual Vector Quantization-Based Tokens—0
Explicit Intensity Control for Accented Text-to-speech—0
Efficient data selection employing Semantic Similarity-based Graph Structures for model training—0
Exploiting Transliterated Words for Finding Similarity in Inter-Language News Articles using Machine Learning—0
Exploring an Inter-Pausal Unit (IPU) based Approach for Indic End-to-End TTS Systems—0
Exploring Machine Speech Chain for Domain Adaptation and Few-Shot Speaker Adaptation—0
Exploring Speech Enhancement for Low-resource Speech Synthesis—0
Exploring speech style spaces with language models: Emotional TTS without emotion labels—0
Boosting Diffusion Model for Spectrogram Up-sampling in Text-to-speech: An Empirical Study—0
Effect of choice of probability distribution, randomness, and search methods for alignment modeling in sequence-to-sequence text-to-speech synthesis using hard alignment—0
BOFFIN TTS: Few-Shot Speaker Adaptation by Bayesian Optimization—0
An Overview of Affective Speech Synthesis and Conversion in the Deep Learning Era—0
Adversarial Speaker-Consistency Learning Using Untranscribed Speech Data for Zero-Shot Multi-Speaker Text-to-Speech—0
Effectiveness of text to speech pseudo labels for forced alignment and cross lingual pretrained models for low resource speech recognition—0
BiVocoder: A Bidirectional Neural Vocoder Integrating Feature Extraction and Waveform Generation—0
Effective Decoder Masking for Transformer Based End-to-End Speech Recognition—0
BitTTS: Highly Compact Text-to-Speech Using 1.58-bit Quantization and Weight Indexing—0
A Novel Data Augmentation Approach for Automatic Speaking Assessment on Opinion Expressions—0
Easy, Interpretable, Effective: openSMILE for voice deepfake detection—0
E3 TTS: Easy End-to-End Diffusion-based Text to Speech—0
A Novel Chinese Dialect TTS Frontend with Non-Autoregressive Neural Machine Translation—0
Adversarial Attacks and Robust Defenses in Speaker Embedding based Zero-Shot Text-to-Speech System—0
Scheduled Interleaved Speech-Text Training for Speech-to-Speech Translation with LLMs—0
Show:102550
← PrevPage 21 of 57Next →

No leaderboard results yet.