SOTAVerified

Speech Enhancement

Speech Enhancement is a signal processing task that involves improving the quality of speech signals captured under noisy or degraded conditions. The goal of speech enhancement is to make speech signals clearer, more intelligible, and more pleasant to listen to, which can be used for various applications such as voice recognition, teleconferencing, and hearing aids. A representative Github project with online demo : ClearerVoice-Studio.

( Image credit: A Fully Convolutional Neural Network For Speech Enhancement )

Papers

Showing 751800 of 982 papers

TitleStatusHype
Resource-Efficient Speech Mask Estimation for Multi-Channel Speech Enhancement0
Instantaneous PSD Estimation for Speech Enhancement based on Generalized Principal ComponentsCode1
CLC: Complex Linear Coding for the DNS 2020 ChallengeCode1
Unsupervised Sound Separation Using Mixture Invariant Training0
Real Time Speech Enhancement in the Waveform DomainCode2
Boosting Objective Scores of a Speech Enhancement Model by MetricGAN Post-processing0
An Iterative Graph Spectral Subtraction Method for Speech Enhancement0
SE-MelGAN -- Speaker Agnostic Rapid Speech Enhancement0
HiFi-GAN: High-Fidelity Denoising and Dereverberation Based on Speech Deep Features in Adversarial NetworksCode1
A fully recurrent feature extraction for single channel speech enhancementCode0
A non-causal FFTNet architecture for speech enhancementCode1
Dilated U-net based approach for multichannel speech enhancement from First-Order Ambisonics recordings0
Phase-aware Single-stage Speech Denoising and Dereverberation with U-NetCode0
Similarity-and-Independence-Aware Beamformer: Method for Target Source Extraction using Magnitude Spectrogram as Reference0
SNR-Based Teachers-Student Technique for Speech Enhancement0
Sub-Band Knowledge Distillation Framework for Speech Enhancement0
Noise Robust TTS for Low Resource Speakers using Pre-trained Model and Speech Enhancement0
Lite Audio-Visual Speech EnhancementCode1
SERIL: Noise Adaptive Speech Enhancement using Regularization-based Incremental LearningCode1
TinyLSTMs: Efficient Neural Speech Enhancement for Hearing AidsCode1
The INTERSPEECH 2020 Deep Noise Suppression Challenge: Datasets, Subjective Testing Framework, and Challenge Results0
Sparse Mixture of Local Experts for Efficient Speech EnhancementCode0
Speaker Re-identification with Speaker Dependent Speech Enhancement0
Adversarial Feature Learning and Unsupervised Clustering based Speech Synthesis for Found Data with Acoustic and Textual Noise0
On the Role of Visual Cues in Audiovisual Speech Enhancement0
Towards a Competitive End-to-End Speech Recognition for CHiME-6 Dinner Party TranscriptionCode0
CHiME-6 Challenge:Tackling Multispeaker Speech Recognition for Unsegmented Recordings0
SNR-Based Features and Diverse Training Data for Robust DNN-Based Speech Enhancement0
WaveCRN: An Efficient Convolutional Recurrent Neural Network for End-to-end Speech EnhancementCode1
Characterizing Speech Adversarial Examples Using Self-Attention U-Net Enhancement0
Improving noise robust automatic speech recognition with single-channel time-domain enhancement network0
Tackling real noisy reverberant meetings with all-neural source separation, counting, and diarization system0
Phonetic Feedback for Speech Enhancement With and Without Parallel Speech DataCode1
Deep Residual-Dense Lattice Network for Speech EnhancementCode1
iSEGAN: Improved Speech Enhancement Generative Adversarial NetworksCode1
Efficient Trainable Front-Ends for Neural Speech Enhancement0
Consistency-aware multi-channel speech enhancement using deep neural networks0
Speech Enhancement using Self-Adaptation and Multi-Head Self-Attention0
Stable Training of DNN for Speech Enhancement based on Perceptually-Motivated Black-Box Cost Function0
Boosted Locality Sensitive Hashing: Discriminative Binary Codes for Source SeparationCode0
DNN-Based Distributed Multichannel Mask Estimation for Speech Enhancement in Microphone Arrays0
Weighted Speech Distortion Losses for Neural-network-based Real-time Speech EnhancementCode1
Robust Multi-channel Speech Recognition using Frequency Aligned Network0
Tensor-to-Vector Regression for Multi-channel Speech Enhancement based on Tensor-Train NetworkCode1
Single Channel Speech Enhancement Using Temporal Convolutional Recurrent Neural Networks0
Channel-Attention Dense U-Net for Multichannel Speech EnhancementCode1
Deep Xi as a Front-End for Robust Automatic Speech Recognition0
CLCNet: Deep learning-based Noise Reduction for Hearing Aids using Complex Linear Coding0
Noise dependent Super Gaussian-Coherence based dual microphone Speech Enhancement for hearing aid application using smartphone0
The INTERSPEECH 2020 Deep Noise Suppression Challenge: Datasets, Subjective Speech Quality and Testing Framework0
Show:102550
← PrevPage 16 of 20Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1ROSE-CD(PESQ)PESQ (wb)3.99Unverified
2PESQetarianPESQ (wb)3.82Unverified
3Mamba-SEUNet L (+PCS)PESQ (wb)3.73Unverified
4Schrödinger bridge (PESQ loss)PESQ (wb)3.7Unverified
5SEMamba (+PCS)PESQ (wb)3.69Unverified
6ZipEnhancer (S, \lamba_6 = 0)PESQ (wb)3.63Unverified
7PrimeK-NetPESQ (wb)3.61Unverified
8ZipEnhancer (S, \lamba_6 = 0.2)PESQ (wb)3.61Unverified
9MP-SENetPESQ (wb)3.6Unverified
10PCS_CS_WAVLMPESQ (wb)3.54Unverified
#ModelMetricClaimedVerifiedStatus
1BSRNN-S + MGDSI-SDR-WB21.4Unverified
2DTLNSI-SDR-WB16.34Unverified
3Non-Real-Time MultiScale+SI-SDR-WB16.22Unverified
4ZipEnhancer (M)PESQ-WB3.81Unverified
5TF-Locoformer (M)PESQ-WB3.72Unverified
6ZipEnhancer (S)PESQ-WB3.69Unverified
7MambAttentionPESQ-WB3.67Unverified
8MP-SENetPESQ-WB3.62Unverified
9xLSTM-SENetPESQ-WB3.59Unverified
10BSRNN-S + MRSDPESQ-WB3.53Unverified
#ModelMetricClaimedVerifiedStatus
1Inter-Channel Conv-TasNetSDR19.67Unverified
2CA Dense U-Net (Complex)SDR18.64Unverified
3Dense U-Net (Complex)SDR18.4Unverified
4Dense U-Net (Real)SDR16.86Unverified
5U-Net (Real)SDR15.97Unverified
6Noisy/unprocessedSDR6.5Unverified
#ModelMetricClaimedVerifiedStatus
1Schrödinger Bridge (PESQ loss)PESQ-WB3.09Unverified
2SGMSE+PESQ-WB2.5Unverified
3Demucs v4PESQ-WB2.37Unverified
4Schrödinger BridgePESQ-WB2.33Unverified
5Conv-TasNetPESQ-WB2.31Unverified
6CDiffuSEPESQ-WB1.6Unverified
#ModelMetricClaimedVerifiedStatus
1ReVISE (ch2)Audio Quality MOS4.19Unverified
2ReVISE (bf)Audio Quality MOS4.11Unverified
3Demucs (ch2)Audio Quality MOS2.95Unverified
4Demucs (bf)Audio Quality MOS2.39Unverified
5MaxDI (Baseline)PESQ1.17Unverified
6DAJA (MVDR,HMA,1000) (Overlapped Speech)SDR-4.76Unverified
#ModelMetricClaimedVerifiedStatus
1ZipEnhancer (M)PESQ-NB4.08Unverified
2DCCRN-MCPESQ-NB3.21Unverified
3DCCRN-MPESQ-NB3.15Unverified
4DCCRNPESQ-NB3.04Unverified
5RNN-ModulationPESQ-WB2.75Unverified
#ModelMetricClaimedVerifiedStatus
1MambAttentionESTOI0.8Unverified
2SEMambaESTOI0.8Unverified
3xLSTM-SENetESTOI0.8Unverified
4MP-SENetESTOI0.79Unverified
#ModelMetricClaimedVerifiedStatus
1SepFormerPESQ2.84Unverified
2DTLNPESQ2.23Unverified
3UnprocessedPESQ1.83Unverified
4Non-Real-Time MultiScale+PESQ1.52Unverified
#ModelMetricClaimedVerifiedStatus
1DCUNet-MCPESQ-NB3.44Unverified
2DCCRN-MPESQ-NB3.28Unverified
3DCUNetPESQ-NB3.25Unverified
#ModelMetricClaimedVerifiedStatus
1CleanMel-L-mapDNSMOS3.82Unverified
2SpatialNetDNSMOS BAK3.43Unverified
#ModelMetricClaimedVerifiedStatus
1rose_cd(PESQ )PESQ3.99Unverified
2ROSE-CDPESQ3.49Unverified
#ModelMetricClaimedVerifiedStatus
1Wave-U-NetCBAK3.24Unverified
#ModelMetricClaimedVerifiedStatus
1Audio-Visual concat-refPESQ2.7Unverified
#ModelMetricClaimedVerifiedStatus
1SE-MelGANAudio Quality MOS3.1Unverified
#ModelMetricClaimedVerifiedStatus
1DeFT-ANPESQ3.01Unverified
#ModelMetricClaimedVerifiedStatus
1Audio-Visual concat-refPESQ3.03Unverified
#ModelMetricClaimedVerifiedStatus
1SepFormerPESQ3.07Unverified