| Are Music Foundation Models Better at Singing Voice Deepfake Detection? Far-Better Fuse them with Speech Foundation Models | Sep 21, 2024 | DeepFake DetectionFace Swapping | —Unverified | 0 | 0 |
| ATCSpeechNet: A multilingual end-to-end speech recognition framework for air traffic control systems | Feb 17, 2021 | Automatic Speech RecognitionAutomatic Speech Recognition (ASR) | —Unverified | 0 | 0 |
| Automatic Pronunciation Assessment using Self-Supervised Speech Representation Learning | Apr 8, 2022 | Representation LearningSelf-Supervised Learning | —Unverified | 0 | 0 |
| Characterizing the adversarial vulnerability of speech self-supervised learning | Nov 8, 2021 | Adversarial RobustnessBenchmarking | —Unverified | 0 | 0 |
| UniWav: Towards Unified Pre-training for Speech Representation Learning and Generation | Mar 2, 2025 | DecoderRepresentation Learning | —Unverified | 0 | 0 |
| Combining Adversarial Training and Disentangled Speech Representation for Robust Zero-Resource Subword Modeling | Jun 17, 2019 | Representation LearningSpeech Representation Learning | —Unverified | 0 | 0 |
| Conformer-Based Self-Supervised Learning for Non-Speech Audio Tasks | Oct 14, 2021 | Audio ClassificationRepresentation Learning | —Unverified | 0 | 0 |
| An Adapter Based Pre-Training for Efficient and Scalable Self-Supervised Speech Representation Learning | Jul 26, 2021 | Automatic Speech RecognitionAutomatic Speech Recognition (ASR) | —Unverified | 0 | 0 |
| CSTNet: Contrastive Speech Translation Network for Self-Supervised Speech Representation Learning | Jun 4, 2020 | BIG-bench Machine LearningContrastive Learning | —Unverified | 0 | 0 |
| Speech-XLNet: Unsupervised Acoustic Model Pretraining For Self-Attention Networks | Oct 23, 2019 | Representation LearningSpeech Representation Learning | —Unverified | 0 | 0 |