SOTAVerified

Speech-to-Text

Papers

Showing 101–125 of 403 papers

TitleStatusHype
Audio Adversarial Examples: Targeted Attacks on Speech-to-TextCode0
LibriS2S: A German-English Speech-to-Speech Translation CorpusCode0
Pre-training on high-resource speech recognition improves low-resource speech-to-text translationCode0
Attentively Embracing Noise for Robust Latent Representation in BERTCode0
Joint CTC-Attention based End-to-End Speech Recognition using Multi-task LearningCode0
Investigating Zero-Shot Generalizability on Mandarin-English Code-Switched ASR and Speech-to-text Translation of Recent Foundation Models with Self-Supervision and Weak SupervisionCode0
InstaIndoor and Multi-modal Deep Learning for Indoor Scene RecognitionCode0
mask-Net: Learning Context Aware Invariant Features using Adversarial Forgetting (Student Abstract)Code0
Code-Switched Urdu ASR for Noisy Telephonic Environment using Data Centric Approach with Hybrid HMM and CNN-TDNNCode0
Kurdish (Sorani) Speech to Text: Presenting an Experimental DatasetCode0
Fleurs-SLU: A Massively Multilingual Benchmark for Spoken Language UnderstandingCode0
FunnyNet-W: Multimodal Learning of Funny Moments in Videos in the WildCode0
Finstreder: Simple and fast Spoken Language Understanding with Finite State Transducers using modern Speech-to-Text modelsCode0
Greek2MathTex: A Greek Speech-to-Text Framework for LaTeX Equations GenerationCode0
End-to-End Automatic Speech Translation of AudiobooksCode0
Direct speech-to-speech translation with a sequence-to-sequence modelCode0
Don't Discard Fixed-Window Audio Segmentation in Speech-to-Text TranslationCode0
End-to-End Learning of Speech 2D Feature-Trajectory for Prosthetic HandsCode0
fairseq S2T: Fast Speech-to-Text Modeling with fairseqCode0
Challenges and Opportunities of Speech Recognition for Bengali Language—0
Efficient Monotonic Multihead Attention—0
Effectively pretraining a speech translation decoder with Machine Translation data—0
Can We Achieve High-quality Direct Speech-to-Speech Translation without Parallel Speech Data?—0
Application of Audio Fingerprinting Techniques for Real-Time Scalable Speech Retrieval and Speech Clusterization—0
A General Multi-Task Learning Framework to Leverage Text Data for Speech to Text Tasks—0
Show:102550
← PrevPage 5 of 17Next →

No leaderboard results yet.