SOTAVerified

Speech Enhancement

Speech Enhancement is a signal processing task that involves improving the quality of speech signals captured under noisy or degraded conditions. The goal of speech enhancement is to make speech signals clearer, more intelligible, and more pleasant to listen to, which can be used for various applications such as voice recognition, teleconferencing, and hearing aids. A representative Github project with online demo : ClearerVoice-Studio.

( Image credit: A Fully Convolutional Neural Network For Speech Enhancement )

Papers

Showing 251300 of 982 papers

TitleStatusHype
Boosting Domain Incremental Learning: Selecting the Optimal Parameters is All You NeedCode0
Boosted Locality Sensitive Hashing: Discriminative Binary Codes for Source SeparationCode0
BLOOM-Net: Blockwise Optimization for Masking Networks Toward Scalable and Efficient Speech EnhancementCode0
Magnitude-Phase Dual-Path Speech Enhancement Network based on Self-Supervised Embedding and Perceptual Contrast Stretch BoostingCode0
A fully recurrent feature extraction for single channel speech enhancementCode0
A Fully Convolutional Neural Network for Speech EnhancementCode0
Learning with Learned Loss Function: Speech Enhancement with Quality-Net to Improve Perceptual Evaluation of Speech QualityCode0
Language and Noise Transfer in Speech Enhancement Generative Adversarial NetworkCode0
Lessons Learned from the URGENT 2024 Speech Enhancement ChallengeCode0
Investigating the effect of residual and highway connections in speech enhancement modelsCode0
An Investigation into the Effectiveness of Enhancement in ASR Training and Test for CHiME-5 Dinner Party TranscriptionCode0
Investigating Generative Adversarial Networks based Speech Dereverberation for Robust Speech RecognitionCode0
Investigating Training Objectives for Generative Speech EnhancementCode0
Let SSMs be ConvNets: State-space Modeling with Optimal Tensor ContractionsCode0
A variance modeling framework based on variational autoencoders for speech enhancementCode0
Direction of Arrival Correction through Speech Quality FeedbackCode0
Improving Design of Input Condition Invariant Speech EnhancementCode0
Improved Speech Enhancement with the Wave-U-NetCode0
How to train your ears: Auditory-model emulation for large-dynamic-range inputs and mild-to-severe hearing lossesCode0
Aura: Privacy-preserving Augmentation to Improve Test Set Diversity in Speech EnhancementCode0
Guided Source Separation Meets a Strong ASR Backend: Hitachi/Paderborn University Joint Investigation for Dinner Party ASRCode0
PlumberNet: Fixing interference leakage after GEV beamformingCode0
Exploiting Low-Rank Tensor-Train Deep Neural Networks Based on Riemannian Gradient Descent With Illustrations of Speech ProcessingCode0
Face Landmark-based Speaker-Independent Audio-Visual Speech Enhancement in Multi-Talker EnvironmentsCode0
Feature Enhancement with Deep Feature Losses for Speaker VerificationCode0
Estimation and Restoration of Unknown Nonlinear Distortion using DiffusionCode0
High-Resolution Speech Restoration with Latent Diffusion ModelCode0
End-to-End Multi-Task Denoising for joint SDR and PESQ OptimizationCode0
ESPnet-SE++: Speech Enhancement for Robust Speech Recognition, Translation, and UnderstandingCode0
Deep Xi as a Front-End for Robust Automatic Speech RecognitionCode0
Effective Noise-aware Data Simulation for Domain-adaptive Speech Enhancement Leveraging Dynamic Stochastic PerturbationCode0
Exploiting Hidden Representations from a DNN-based Speech Recogniser for Speech Intelligibility Prediction in Hearing-impaired ListenersCode0
ROSE: A Recognition-Oriented Speech Enhancement Framework in Air Traffic Control Using Multi-Objective LearningCode0
Deep Unfolding: Model-Based Inspiration of Novel Deep Architectures0
Deep Time Delay Neural Network for Speech Enhancement with Full Data Learning0
Audio-Visual Speech Enhancement Using Multimodal Deep Convolutional Neural Networks0
Deep Speech Enhancement for Reverberated and Noisy Signals using Wide Residual Networks0
Deep Residual Echo Suppression and Noise Reduction: A Multi-Input FCRN Approach in a Hybrid Speech Enhancement System0
Audio-Visual Speech Enhancement and Separation by Utilizing Multi-Modal Self-Supervised Embeddings0
A network of deep neural networks for distant speech recognition0
Deep Noise Suppression With Non-Intrusive PESQNet Supervision Enabling the Use of Real Training Data0
Deep Noise Suppression Maximizing Non-Differentiable PESQ Mediated by a Non-Intrusive PESQNet0
Deep neural network techniques for monaural speech enhancement: state of the art analysis0
Audio-visual multi-channel speech separation, dereverberation and recognition0
An Ensemble SVM-based Approach for Voice Activity Detection0
Deep neural network Based Low-latency Speech Separation with Asymmetric analysis-Synthesis Window Pair0
Audio-visual End-to-end Multi-channel Speech Separation, Dereverberation and Recognition0
Deep low-latency joint speech transmission and enhancement over a gaussian channel0
Audio Recording Device Identification Based on Deep Learning0
An Empirical Study on the Impact of Positional Encoding in Transformer-based Monaural Speech Enhancement0
Show:102550
← PrevPage 6 of 20Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1ROSE-CD(PESQ)PESQ (wb)3.99Unverified
2PESQetarianPESQ (wb)3.82Unverified
3Mamba-SEUNet L (+PCS)PESQ (wb)3.73Unverified
4Schrödinger bridge (PESQ loss)PESQ (wb)3.7Unverified
5SEMamba (+PCS)PESQ (wb)3.69Unverified
6ZipEnhancer (S, \lamba_6 = 0)PESQ (wb)3.63Unverified
7PrimeK-NetPESQ (wb)3.61Unverified
8ZipEnhancer (S, \lamba_6 = 0.2)PESQ (wb)3.61Unverified
9MP-SENetPESQ (wb)3.6Unverified
10PCS_CS_WAVLMPESQ (wb)3.54Unverified
#ModelMetricClaimedVerifiedStatus
1BSRNN-S + MGDSI-SDR-WB21.4Unverified
2DTLNSI-SDR-WB16.34Unverified
3Non-Real-Time MultiScale+SI-SDR-WB16.22Unverified
4ZipEnhancer (M)PESQ-WB3.81Unverified
5TF-Locoformer (M)PESQ-WB3.72Unverified
6ZipEnhancer (S)PESQ-WB3.69Unverified
7MambAttentionPESQ-WB3.67Unverified
8MP-SENetPESQ-WB3.62Unverified
9xLSTM-SENetPESQ-WB3.59Unverified
10BSRNN-S + MRSDPESQ-WB3.53Unverified
#ModelMetricClaimedVerifiedStatus
1Inter-Channel Conv-TasNetSDR19.67Unverified
2CA Dense U-Net (Complex)SDR18.64Unverified
3Dense U-Net (Complex)SDR18.4Unverified
4Dense U-Net (Real)SDR16.86Unverified
5U-Net (Real)SDR15.97Unverified
6Noisy/unprocessedSDR6.5Unverified
#ModelMetricClaimedVerifiedStatus
1Schrödinger Bridge (PESQ loss)PESQ-WB3.09Unverified
2SGMSE+PESQ-WB2.5Unverified
3Demucs v4PESQ-WB2.37Unverified
4Schrödinger BridgePESQ-WB2.33Unverified
5Conv-TasNetPESQ-WB2.31Unverified
6CDiffuSEPESQ-WB1.6Unverified
#ModelMetricClaimedVerifiedStatus
1ReVISE (ch2)Audio Quality MOS4.19Unverified
2ReVISE (bf)Audio Quality MOS4.11Unverified
3Demucs (ch2)Audio Quality MOS2.95Unverified
4Demucs (bf)Audio Quality MOS2.39Unverified
5MaxDI (Baseline)PESQ1.17Unverified
6DAJA (MVDR,HMA,1000) (Overlapped Speech)SDR-4.76Unverified
#ModelMetricClaimedVerifiedStatus
1ZipEnhancer (M)PESQ-NB4.08Unverified
2DCCRN-MCPESQ-NB3.21Unverified
3DCCRN-MPESQ-NB3.15Unverified
4DCCRNPESQ-NB3.04Unverified
5RNN-ModulationPESQ-WB2.75Unverified
#ModelMetricClaimedVerifiedStatus
1MambAttentionESTOI0.8Unverified
2SEMambaESTOI0.8Unverified
3xLSTM-SENetESTOI0.8Unverified
4MP-SENetESTOI0.79Unverified
#ModelMetricClaimedVerifiedStatus
1SepFormerPESQ2.84Unverified
2DTLNPESQ2.23Unverified
3UnprocessedPESQ1.83Unverified
4Non-Real-Time MultiScale+PESQ1.52Unverified
#ModelMetricClaimedVerifiedStatus
1DCUNet-MCPESQ-NB3.44Unverified
2DCCRN-MPESQ-NB3.28Unverified
3DCUNetPESQ-NB3.25Unverified
#ModelMetricClaimedVerifiedStatus
1CleanMel-L-mapDNSMOS3.82Unverified
2SpatialNetDNSMOS BAK3.43Unverified
#ModelMetricClaimedVerifiedStatus
1rose_cd(PESQ )PESQ3.99Unverified
2ROSE-CDPESQ3.49Unverified
#ModelMetricClaimedVerifiedStatus
1Wave-U-NetCBAK3.24Unverified
#ModelMetricClaimedVerifiedStatus
1Audio-Visual concat-refPESQ2.7Unverified
#ModelMetricClaimedVerifiedStatus
1SE-MelGANAudio Quality MOS3.1Unverified
#ModelMetricClaimedVerifiedStatus
1DeFT-ANPESQ3.01Unverified
#ModelMetricClaimedVerifiedStatus
1Audio-Visual concat-refPESQ3.03Unverified
#ModelMetricClaimedVerifiedStatus
1SepFormerPESQ3.07Unverified