| Improved Self-Supervised Multilingual Speech Representation Learning Combined with Auxiliary Language Information | Dec 7, 2022 | Automatic Speech RecognitionAutomatic Speech Recognition (ASR) | —Unverified | 0 | 0 |
| Improving Noise Robustness of Contrastive Speech Representation Learning with Speech Reconstruction | Oct 28, 2021 | Automatic Speech RecognitionAutomatic Speech Recognition (ASR) | —Unverified | 0 | 0 |
| Improving Speech Representation Learning via Speech-level and Phoneme-level Masking Approach | Oct 25, 2022 | Representation LearningSpeaker Recognition | —Unverified | 0 | 0 |
| Improving the Robustness of DistilHuBERT to Unseen Noisy Conditions via Data Augmentation, Curriculum Learning, and Multi-Task Enhancement | Nov 12, 2022 | Data AugmentationEmotion Recognition | —Unverified | 0 | 0 |
| Improving Unsupervised Subword Modeling via Disentangled Speech Representation Learning and Transformation | Jun 17, 2019 | ClusteringRepresentation Learning | —Unverified | 0 | 0 |
| INTapt: Information-Theoretic Adversarial Prompt Tuning for Enhanced Non-Native Speech Recognition | May 25, 2023 | Automatic Speech RecognitionAutomatic Speech Recognition (ASR) | —Unverified | 0 | 0 |
| JOOCI: a Framework for Learning Comprehensive Speech Representations | Oct 14, 2024 | Representation LearningSpeech Representation Learning | —Unverified | 0 | 0 |
| Label Aware Speech Representation Learning For Language Identification | Jun 7, 2023 | Language IdentificationMissing Labels | —Unverified | 0 | 0 |
| Language Adaptive Cross-lingual Speech Representation Learning with Sparse Sharing Sub-networks | Mar 9, 2022 | Representation Learningspeech-recognition | —Unverified | 0 | 0 |
| A Convolutional Deep Markov Model for Unsupervised Speech Representation Learning | Jun 3, 2020 | Representation LearningSelf-Supervised Learning | —Unverified | 0 | 0 |
| Learning Cross-lingual Visual Speech Representations | Mar 14, 2023 | Representation LearningSelf-Supervised Learning | —Unverified | 0 | 0 |
| Learning Disentangled Speech Representations | Nov 4, 2023 | BenchmarkingDisentanglement | —Unverified | 0 | 0 |
| Learning Robust and Multilingual Speech Representations | Jan 29, 2020 | Representation Learningspeech-recognition | —Unverified | 0 | 0 |
| Leveraging unsupervised and weakly-supervised data to improve direct speech-to-speech translation | Mar 24, 2022 | Representation LearningSpeech Representation Learning | —Unverified | 0 | 0 |
| Towards the Next Frontier in Speech Representation Learning Using Disentanglement | Jul 2, 2024 | DisentanglementRepresentation Learning | —Unverified | 0 | 0 |
| MASR: Multi-label Aware Speech Representation | Jul 20, 2023 | Emotion RecognitionLanguage Identification | —Unverified | 0 | 0 |
| Towards Unsupervised Speech Recognition and Synthesis with Quantized Speech Representation Learning | Oct 28, 2019 | ClusteringPhoneme Recognition | —Unverified | 0 | 0 |
| VATLM: Visual-Audio-Text Pre-Training with Unified Masked Prediction for Speech Representation Learning | Nov 21, 2022 | Audio-Visual Speech RecognitionLanguage Modelling | —Unverified | 0 | 0 |
| On-Device Constrained Self-Supervised Speech Representation Learning for Keyword Spotting via Knowledge Distillation | Jul 6, 2023 | Keyword SpottingKnowledge Distillation | —Unverified | 0 | 0 |
| On the Use of Semantically-Aligned Speech Representations for Spoken Language Understanding | Oct 11, 2022 | Representation LearningSentence | —Unverified | 0 | 0 |
| PARP: Prune, Adjust and Re-Prune for Self-Supervised Speech Recognition | Jun 10, 2021 | Automatic Speech RecognitionAutomatic Speech Recognition (ASR) | —Unverified | 0 | 0 |
| Privacy-preserving Representation Learning for Speech Understanding | Oct 26, 2023 | ClassificationEmotion Recognition | —Unverified | 0 | 0 |
| Privacy-Preserving Speech Representation Learning using Vector Quantization | Mar 15, 2022 | Privacy PreservingQuantization | —Unverified | 0 | 0 |
| Progressive Residual Extraction based Pre-training for Speech Representation Learning | Aug 31, 2024 | Emotion RecognitionRepresentation Learning | —Unverified | 0 | 0 |
| TranUSR: Phoneme-to-word Transcoder Based Unified Speech Representation Learning for Cross-lingual Speech Recognition | May 23, 2023 | Automatic Speech RecognitionAutomatic Speech Recognition (ASR) | —Unverified | 0 | 0 |