SOTAVerified

Speech-to-Text Translation

Translate audio signals of speech in one language into text in a foreign language, either in an end-to-end or cascade manner.

Papers

Showing 76100 of 146 papers

TitleStatusHype
Cross-Modal Multi-Tasking for Speech-to-Text Translation via Hard Parameter Sharing0
CTC Alignments Improve Autoregressive Translation0
Data Efficient Direct Speech-to-Text Translation with Modality Agnostic Meta-Learning0
Decision Attentive Regularization to Improve Simultaneous Speech Translation Systems0
Direct Simultaneous Speech-to-Text Translation Assisted by Synchronized Streaming ASR0
Efficient Monotonic Multihead Attention0
End-to-End Offline Speech Translation System for IWSLT 2020 using Modality Agnostic Meta-Learning0
End-to-End Speech-to-Text Translation: A Survey0
End-to-End Speech Translation for Low-Resource Languages Using Weakly Labeled Data0
Enhanced Direct Speech-to-Speech Translation Using Self-supervised Pre-training and Data Augmentation0
Enhancing Speech-to-Speech Translation with Multiple TTS Targets0
Enhancing Transformer for End-to-end Speech-to-Text Translation0
ESPnet-ST-v2: Multipurpose Spoken Language Translation Toolkit0
Europarl-ST: A Multilingual Corpus For Speech Translation Of Parliamentary Debates0
Finetuning End-to-End Models for Estonian Conversational Spoken Language Translation0
Google USM: Scaling Automatic Speech Recognition Beyond 100 Languages0
How "Real" is Your Real-Time Simultaneous Speech-to-Text Translation System?0
Hybrid Transducer and Attention based Encoder-Decoder Modeling for Speech-to-Text Tasks0
Improved Cross-Lingual Transfer Learning For Automatic Speech Translation0
SimulSeamless: FBK at IWSLT 2024 Simultaneous Speech Translation0
SimulSpeech: End-to-End Simultaneous Speech to Text Translation0
SpeechAlign: a Framework for Speech Translation Alignment Evaluation0
Speech is More Than Words: Do Speech-to-Text Translation Systems Leverage Prosody?0
Speech to Speech Translation with Translatotron: A State of the Art Review0
Speech-to-Text Translation with Phoneme-Augmented CoT: Enhancing Cross-Lingual Transfer in Low-Resource Scenarios0
Show:102550
← PrevPage 4 of 6Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1Task Modulation + Multitask Learning(ASR/MT) + Data AugmentationCase-sensitive sacreBLEU28.88Unverified
2Wav2Vec2.0+mBART+AdaptorsCase-sensitive sacreBLEU28.22Unverified
3Transformer + Meta Learning(ASR/MT) + Data AugmentationCase-sensitive sacreBLEU27.51Unverified
4Transformer with AdaptersCase-sensitive sacreBLEU24.63Unverified
5Dual-decoder TransformerCase-sensitive sacreBLEU23.63Unverified
6SpeechformerCase-sensitive sacreBLEU23.6Unverified
7Transformer + ASR PretrainCase-sensitive sacreBLEU22.8Unverified
8Transformer + ASR PretrainCase-sensitive sacreBLEU22.7Unverified
#ModelMetricClaimedVerifiedStatus
1Transformer with AdaptersCase-sensitive sacreBLEU28.73Unverified
2SpeechformerCase-sensitive sacreBLEU28.5Unverified
3Dual-decoder TransformerCase-sensitive sacreBLEU28.12Unverified
4Transformer + ASR Pretrain + SpecAugCase-sensitive sacreBLEU27.4Unverified
5Transformer + ASR PretrainCase-sensitive sacreBLEU26.8Unverified
#ModelMetricClaimedVerifiedStatus
1Dual-decoder TransformerCase-sensitive sacreBLEU33.45Unverified
2Transformer + ASR Pretrain + SpecAugCase-sensitive sacreBLEU33.3Unverified
3Transformer + ASR PretrainCase-sensitive sacreBLEU32.3Unverified
#ModelMetricClaimedVerifiedStatus
1SeamlessM4T LargeBLEU30.6Unverified
2SeamlessM4T MediumBLEU26.6Unverified
#ModelMetricClaimedVerifiedStatus
1SeamlessM4T LargeBLEU34.1Unverified
2SeamlessM4T MediumBLEU29.8Unverified
#ModelMetricClaimedVerifiedStatus
1SeamlessM4T LargeBLEU21.5Unverified
2SeamlessM4T MediumBLEU19.2Unverified
#ModelMetricClaimedVerifiedStatus
1SeamlessM4T LargeBLEU24Unverified
2SeamlessM4T MediumBLEU20.9Unverified
#ModelMetricClaimedVerifiedStatus
1Transformer + ASR Pretrain + SpecAugCase-insensitive sacreBLEU17.2Unverified
2Transformer + ASR PretrainCase-insensitive sacreBLEU16.5Unverified
#ModelMetricClaimedVerifiedStatus
1MediBeng Whisper TinyBleu0.98Unverified
2Whisper TinyBleu0.3Unverified
#ModelMetricClaimedVerifiedStatus
1Transformer with AdaptersSacreBLEU26.61Unverified
2Dual-decoder TransformerSacreBLEU25.62Unverified
#ModelMetricClaimedVerifiedStatus
1SpeechformerCase-sensitive sacreBLEU27.7Unverified