SOTAVerified

Text to Speech

import gTTS import os def text_to_speech_kurdish(text, output_file="output.mp3"): # گۆڕینی نووسین بۆ دەنگ بە زمانی کوردی (هەڵبژاردنی زمانی "ku" بۆ کوردی) tts = gTTS(text=text, lang='ku', slow=False) tts.save(output_file) os.system(f"start {output_file}") # کردنەوەی فایلە دەنگییەکە (لە Windows) # نموونە: text_to_speech_kurdish("سڵاو، ئەمە دەنگی منە بە زمانی کوردی.")

Papers

Showing 601625 of 1419 papers

TitleStatusHype
Building a Luganda Text-to-Speech Model From Crowdsourced Data0
Faces that Speak: Jointly Synthesising Talking Face and Speech from Text0
Towards Evaluating the Robustness of Automatic Speech Recognition Systems via Audio Style Transfer0
PolyGlotFake: A Novel Multilingual and Multimodal DeepFake DatasetCode0
Real-Time Pill Identification for the Visually Impaired Using Deep Learning0
Attention-Constrained Inference for Robust Decoder-Only Text-to-Speech0
TI-ASU: Toward Robust Automatic Speech Understanding through Text-to-speech Imputation Against Missing Speech Modality0
StoryTTS: A Highly Expressive Text-to-Speech Dataset with Rich Textual Expressiveness Annotations0
Retrieval-Augmented Audio Deepfake Detection0
Prior-agnostic Multi-scale Contrastive Text-Audio Pre-training for Parallelized TTS Frontend Modeling0
Voice-Assisted Real-Time Traffic Sign Recognition System Using Convolutional Neural Network0
The X-LANCE Technical Report for Interspeech 2024 Speech Processing Using Discrete Speech Unit Challenge0
Cross-Domain Audio Deepfake Detection: Dataset and Analysis0
RALL-E: Robust Codec Language Modeling with Chain-of-Thought Prompting for Text-to-Speech Synthesis0
CLaM-TTS: Improving Neural Codec Language Model for Zero-Shot Text-to-Speech0
PSCodec: A Series of High-Fidelity Low-bitrate Neural Speech Codecs Leveraging Prompt Encoders0
Humane Speech Synthesis through Zero-Shot Emotion and Disfluency GenerationCode0
A Review of Multi-Modal Large Language and Vision Models0
Isometric Neural Machine Translation using Phoneme Count Ratio Reward-based Reinforcement Learning0
Creating an African American-Sounding TTS: Guidelines, Technical Challenges,and Surprising Evaluations0
EM-TTS: Efficiently Trained Low-Resource Mongolian Lightweight Text-to-Speech0
Attempt Towards Stress Transfer in Speech-to-Speech Machine Translation0
AttentionStitch: How Attention Solves the Speech Editing Problem0
Towards Accurate Lip-to-Speech Synthesis in-the-Wild0
Extending Multilingual Speech Synthesis to 100+ Languages without Transcribed Data0
Show:102550
← PrevPage 25 of 57Next →

No leaderboard results yet.