SOTAVerified

Speech-to-Text

Papers

Showing 301–350 of 403 papers

TitleStatusHype
Label-Synchronous Speech-to-Text Alignment for ASR Using Forward and Backward Transformers—0
SPGISpeech: 5,000 hours of transcribed financial audio for fully formatted end-to-end speech recognitionCode0
Multi-Discriminator Sobolev Defense-GAN Against Adversarial Attacks for End-to-End Speech Systems—0
Towards the evaluation of automatic simultaneous speech translation from a communicative perspective—0
Towards Robust Speech-to-Text Adversarial Attack—0
Inductive biases, pretraining and fine-tuning jointly account for brain responses to speech—0
NUVA: A Naming Utterance Verifier for Aphasia Treatment—0
Audio Adversarial Examples: Attacks Using Vocal Masks—0
Graph Neural Networks to Predict Customer Satisfaction Following Interactions with a Corporate Call Center—0
BCN2BRNO: ASR System Fusion for Albayzin 2020 Speech to Text Challenge—0
WER-BERT: Automatic WER Estimation with BERT in a Balanced Ordinal Classification Paradigm—0
Exploring Transfer Learning For End-to-End Spoken Language Understanding—0
Incorporating Domain Knowledge To Improve Topic Segmentation Of Long MOOC Lecture Videos—0
End to End ASR System with Automatic Punctuation InsertionCode0
Attentively Embracing Noise for Robust Latent Representation in BERTCode0
mask-Net: Learning Context Aware Invariant Features using Adversarial Forgetting (Student Abstract)Code0
A low latency ASR-free end to end spoken language understanding system—0
Effectively pretraining a speech translation decoder with Machine Translation data—0
Bridging the Modality Gap for Speech-to-Text Translation—0
Multilingual Speech Translation with Efficient Finetuning of Pretrained Models—0
MAM: Masked Acoustic Modeling for End-to-End Speech-to-Text Translation—0
Class-Conditional Defense GAN Against End-to-End Speech Attacks—0
A General Multi-Task Learning Framework to Leverage Text Data for Speech to Text Tasks—0
Towards End-to-End Training of Automatic Speech Recognition for Nigerian PidginCode0
Subtitles to Segmentation: Improving Low-Resource Speech-to-Text Translation Pipelines—0
Ensemble Chinese End-to-End Spoken Language Understanding for Abnormal Event Detection from audio stream—0
fairseq S2T: Fast Speech-to-Text Modeling with fairseqCode0
End-to-End Learning of Speech 2D Feature-Trajectory for Prosthetic HandsCode0
Contextualized Translation of Automatically Segmented Speech—0
Adversarial Attacks against Neural Networks in Audio Domain: Exploiting Principal Components—0
Contextualized Spoken Word Representations from Convolutional Autoencoders—0
SimulSpeech: End-to-End Simultaneous Speech to Text Translation—0
End-to-End Simultaneous Translation System for IWSLT2020 Using Modality Agnostic Meta-Learning—0
End-to-End Offline Speech Translation System for IWSLT 2020 using Modality Agnostic Meta-Learning—0
Self-Supervised Representations Improve End-to-End Speech Translation—0
Exploration of End-to-End ASR for OpenSTT -- Russian Open Speech-to-Text Dataset—0
Improving Cross-Lingual Transfer Learning for End-to-End Speech Recognition with Speech Translation—0
ON-TRAC Consortium for End-to-End and Simultaneous Speech Translation Challenge Tasks at IWSLT 2020—0
Speech to Text Adaptation: Towards an Efficient Cross-Modal Distillation—0
SpiCE: A New Open-Access Corpus of Conversational Bilingual Speech in Cantonese and English—0
Subtitles to Segmentation: Improving Low-Resource Speech-to-TextTranslation Pipelines—0
Crossing the SSH Bridge with Interview Data—0
Jointly Trained Transformers models for Spoken Language Translation—0
Cloud-Based Face and Speech Recognition for Access Control Applications—0
Learnings from Technological Interventions in a Low Resource Language: A Case-Study on Gondi—0
Speak2Label: Using Domain Knowledge for Creating a Large Scale Driver Gaze Zone Estimation Dataset—0
The Spotify Podcast Dataset—0
A.I. based Embedded Speech to Text Using Deepspeech—0
Synchronous Speech Recognition and Speech-to-Text Translation with Interactive DecodingCode0
Re-Translation Strategies For Long Form, Simultaneous, Spoken Language TranslationCode0
Show:102550
← PrevPage 7 of 9Next →

No leaderboard results yet.