SOTAVerified

Speech Representation Learning

Papers

Showing 2650 of 131 papers

TitleStatusHype
CLARA: Multilingual Contrastive Learning for Audio Representation AcquisitionCode1
MUST&P-SRL: Multi-lingual and Unified Syllabification in Text and Phonetic Domains for Speech Representation LearningCode0
Spatial HuBERT: Self-supervised Spatial Speech Representation Learning for a Single Talker from Multi-channel Audio0
Evaluating Self-Supervised Speech Representations for Indigenous American Languages0
Fast-HuBERT: An Efficient Training Framework for Self-Supervised Speech Representation LearningCode1
QS-TTS: Towards Semi-Supervised Text-to-Speech Synthesis via Vector-Quantized Self-Supervised Speech Representation LearningCode1
Speech representation learning: Learning bidirectional encoders with single-view, multi-view, and multi-task methods0
MASR: Multi-label Aware Speech Representation0
On-Device Constrained Self-Supervised Speech Representation Learning for Keyword Spotting via Knowledge Distillation0
Flowchase: a Mobile Application for Pronunciation Training0
Label Aware Speech Representation Learning For Language Identification0
Simultaneous or Sequential Training? How Speech Representations Cooperate in a Multi-Task Self-Supervised Learning System0
An empirical study on speech restoration guided by self supervised speech representation0
INTapt: Information-Theoretic Adversarial Prompt Tuning for Enhanced Non-Native Speech Recognition0
TranUSR: Phoneme-to-word Transcoder Based Unified Speech Representation Learning for Cross-lingual Speech Recognition0
DinoSR: Self-Distillation and Online Clustering for Self-supervised Speech Representation LearningCode1
A multimodal dynamical variational autoencoder for audiovisual speech representation learningCode0
Learning Cross-lingual Visual Speech Representations0
FaceXHuBERT: Text-less Speech-driven E(X)pressive 3D Facial Animation Synthesis Using Self-Supervised Speech Representation LearningCode1
Self-supervised speech representation learning for keyword-spotting with light-weight transformers0
Structured Pruning of Self-Supervised Pre-trained Models for Speech Recognition and UnderstandingCode1
A low latency attention module for streaming self-supervised speech representation learningCode0
Efficient Speech Representation Learning with Low-Bit Quantization0
Improved Self-Supervised Multilingual Speech Representation Learning Combined with Auxiliary Language Information0
Disentangled Feature Learning for Real-Time Neural Speech Coding0
Show:102550
← PrevPage 2 of 6Next →

No leaderboard results yet.