| LightHuBERT: Lightweight and Configurable Speech Representation Learning with Once-for-All Hidden-Unit BERT | Mar 29, 2022 | AllAutomatic Speech Recognition | CodeCode Available | 1 |
| Robust Speaker Recognition with Transformers Using wav2vec 2.0 | Mar 28, 2022 | Data AugmentationRepresentation Learning | —Unverified | 0 |
| Leveraging unsupervised and weakly-supervised data to improve direct speech-to-speech translation | Mar 24, 2022 | Representation LearningSpeech Representation Learning | —Unverified | 0 |
| XTREME-S: Evaluating Cross-lingual Speech Representations | Mar 21, 2022 | Representation LearningRetrieval | —Unverified | 0 |
| A^3T: Alignment-Aware Acoustic and Text Pretraining for Speech Synthesis and Editing | Mar 18, 2022 | Representation LearningSpeaker Verification | CodeCode Available | 1 |
| Privacy-Preserving Speech Representation Learning using Vector Quantization | Mar 15, 2022 | Privacy PreservingQuantization | —Unverified | 0 |
| Language Adaptive Cross-lingual Speech Representation Learning with Sparse Sharing Sub-networks | Mar 9, 2022 | Representation Learningspeech-recognition | —Unverified | 0 |
| A Brief Overview of Unsupervised Neural Speech Representation Learning | Mar 1, 2022 | Representation LearningSpeech Representation Learning | —Unverified | 0 |
| A Noise-Robust Self-supervised Pre-training Model Based Speech Representation Learning for Automatic Speech Recognition | Jan 22, 2022 | Automatic Speech RecognitionAutomatic Speech Recognition (ASR) | —Unverified | 0 |
| A Deep Paradigm for Articulatory Speech Representation Learning via Neural Convolutive Sparse Matrix Factorization | Jan 16, 2022 | Phoneme RecognitionRepresentation Learning | —Unverified | 0 |
| Robust Self-Supervised Audio-Visual Speech Recognition | Jan 5, 2022 | Audio-Visual Speech RecognitionAutomatic Speech Recognition | CodeCode Available | 2 |
| Learning Audio-Visual Speech Representation by Masked Multimodal Cluster Prediction | Jan 5, 2022 | Automatic Speech RecognitionAutomatic Speech Recognition (ASR) | CodeCode Available | 2 |
| Robust Speech Representation Learning via Flow-based Embedding Regularization | Dec 7, 2021 | Deep LearningLanguage Identification | —Unverified | 0 |
| XLS-R: Self-supervised Cross-lingual Speech Representation Learning at Scale | Nov 17, 2021 | Language IdentificationRepresentation Learning | CodeCode Available | 1 |
| Characterizing the adversarial vulnerability of speech self-supervised learning | Nov 8, 2021 | Adversarial RobustnessBenchmarking | —Unverified | 0 |
| Improving Noise Robustness of Contrastive Speech Representation Learning with Speech Reconstruction | Oct 28, 2021 | Automatic Speech RecognitionAutomatic Speech Recognition (ASR) | —Unverified | 0 |
| Speech Representation Learning Through Self-supervised Pretraining And Multi-task Finetuning | Oct 18, 2021 | Multi-Task LearningRepresentation Learning | —Unverified | 0 |
| Conformer-Based Self-Supervised Learning for Non-Speech Audio Tasks | Oct 14, 2021 | Audio ClassificationRepresentation Learning | —Unverified | 0 |
| UniSpeech-SAT: Universal Speech Representation Learning with Speaker Aware Pre-Training | Oct 12, 2021 | Data AugmentationMulti-Task Learning | CodeCode Available | 1 |
| DistilHuBERT: Speech Representation Learning by Layer-wise Distillation of Hidden-unit BERT | Oct 5, 2021 | Multi-Task LearningRepresentation Learning | CodeCode Available | 0 |
| W2v-BERT: Combining Contrastive Learning and Masked Language Modeling for Self-Supervised Speech Pre-Training | Aug 7, 2021 | Contrastive LearningLanguage Modeling | CodeCode Available | 3 |
| An Adapter Based Pre-Training for Efficient and Scalable Self-Supervised Speech Representation Learning | Jul 26, 2021 | Automatic Speech RecognitionAutomatic Speech Recognition (ASR) | —Unverified | 0 |
| Pretext Tasks selection for multitask self-supervised speech representation learning | Jul 1, 2021 | Automatic Speech RecognitionAutomatic Speech Recognition (ASR) | CodeCode Available | 0 |
| HuBERT: Self-Supervised Speech Representation Learning by Masked Prediction of Hidden Units | Jun 14, 2021 | ClusteringLanguage Modelling | CodeCode Available | 1 |
| PARP: Prune, Adjust and Re-Prune for Self-Supervised Speech Recognition | Jun 10, 2021 | Automatic Speech RecognitionAutomatic Speech Recognition (ASR) | —Unverified | 0 |