SOTAVerified

Speech Enhancement

Speech Enhancement is a signal processing task that involves improving the quality of speech signals captured under noisy or degraded conditions. The goal of speech enhancement is to make speech signals clearer, more intelligible, and more pleasant to listen to, which can be used for various applications such as voice recognition, teleconferencing, and hearing aids. A representative Github project with online demo : ClearerVoice-Studio.

( Image credit: A Fully Convolutional Neural Network For Speech Enhancement )

Papers

Showing 801850 of 982 papers

TitleStatusHype
Zero-Reference Deep Curve Estimation for Low-Light Image EnhancementCode1
Improving GANs for Speech EnhancementCode1
Robust Speaker Recognition Using Speech Enhancement And Attention Model0
A Differentiable Perceptual Audio Metric Learned from Just Noticeable DifferencesCode1
Speech Enhancement based on Denoising Autoencoder with Multi-branched EncodersCode0
Monaural Speech Enhancement Using a Multi-Branch Temporal Convolutional Network0
Mixture of Inference Networks for VAE-based Audio-visual Speech Enhancement0
High-quality Speech Synthesis Using Super-resolution Mel-Spectrogram0
Time-Domain Multi-modal Bone/air Conducted Speech Enhancement0
MMTM: Multimodal Transfer Module for CNN FusionCode0
Distributed Microphone Speech Enhancement based on Deep Learning0
Sequential Multi-Frame Neural Beamforming for Speech Separation and Enhancement0
Speaker independence of neural vocoders and their effect on parametric resynthesis speech enhancement0
Robust Unsupervised Audio-visual Speech Enhancement Using a Mixture of Variational Autoencoders0
The Speed Submission to DIHARD II: Contributions & Lessons Learned0
Spleeter: A Fast And State-of-the Art Music Source Separation Tool With Pre-trained ModelsCode1
What does a network layer hear? Analyzing hidden representations of end-to-end ASR through speech synthesisCode0
Memory Requirement Reduction of Deep Neural Networks Using Low-bit Quantization of Parameters0
Does Speech enhancement of publicly available data help build robust Speech Recognition Systems?0
Feature Enhancement with Deep Feature Losses for Speaker VerificationCode0
A Recurrent Variational Autoencoder for Speech Enhancement0
Word-level Embeddings for Cross-Task Transfer Learning in Speech ProcessingCode0
AeGAN: Time-Frequency Speech Denoising via Generative Adversarial Networks0
Comparative Study between Adversarial Networks and Classical Techniques for Speech Enhancement0
Multi-Talker MVDR Beamforming Based on Extended Complex Gaussian Mixture Model0
Semi-Supervised Multichannel Speech Enhancement With a Deep Speech PriorCode1
使用語者轉換技術於語音合成資料庫之音質改進(Speech Enhancement for TTS Speech Corpora by using Voice Conversion Technologies)0
Speech enhancement based on the integration of fully convolutional network, temporal lowpass filtering and spectrogram masking0
AV Speech Enhancement Challenge using a Real Noisy Corpus0
FaSNet: Low-latency Adaptive Beamforming for Multi-microphone Audio ProcessingCode1
Multichannel Speech Enhancement by Raw Waveform-mapping using Fully Convolutional Networks0
An Investigation into the Effectiveness of Enhancement in ASR Training and Test for CHiME-5 Dinner Party TranscriptionCode0
CochleaNet: A Robust Language-independent Audio-Visual Model for Speech Enhancement0
A scalable noisy speech dataset and online subjective test framework0
Spoken Speech Enhancement using EEG0
Generative Speech Enhancement Based on Cloned Networks0
On Loss Functions for Supervised Monaural Time-Domain Speech Enhancement0
Speech Enhancement using Adaptive Mean Median Deviation and EMD Technique0
Coarse-to-fine Optimization for Speech Enhancement0
A Dual-Staged Context Aggregation Method Towards Efficient End-To-End Speech Enhancement0
Audio-visual Speech Enhancement Using Conditional Variational Auto-Encoders0
Deep learning for minimum mean-square error approaches to speech enhancement0
My lips are concealed: Audio-visual speech enhancement through obstructions0
Convolutional Neural Network-based Speech Enhancement for Cochlear Implant Recipients0
A Monaural Speech Enhancement Method for Robust Small-Footprint Keyword Spotting0
The Second DIHARD Diarization Challenge: Dataset, task, and baselinesCode0
rVAD: An Unsupervised Segment-Based Robust Voice Activity Detection MethodCode0
Increasing Compactness Of Deep Learning Based Speech Enhancement Models With Parameter Pruning And Quantization Techniques0
Guided Source Separation Meets a Strong ASR Backend: Hitachi/Paderborn University Joint Investigation for Dinner Party ASRCode0
Deep-Learning-Based Audio-Visual Speech Enhancement in Presence of Lombard Effect0
Show:102550
← PrevPage 17 of 20Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1ROSE-CD(PESQ)PESQ (wb)3.99Unverified
2PESQetarianPESQ (wb)3.82Unverified
3Mamba-SEUNet L (+PCS)PESQ (wb)3.73Unverified
4Schrödinger bridge (PESQ loss)PESQ (wb)3.7Unverified
5SEMamba (+PCS)PESQ (wb)3.69Unverified
6ZipEnhancer (S, \lamba_6 = 0)PESQ (wb)3.63Unverified
7PrimeK-NetPESQ (wb)3.61Unverified
8ZipEnhancer (S, \lamba_6 = 0.2)PESQ (wb)3.61Unverified
9MP-SENetPESQ (wb)3.6Unverified
10PCS_CS_WAVLMPESQ (wb)3.54Unverified
#ModelMetricClaimedVerifiedStatus
1BSRNN-S + MGDSI-SDR-WB21.4Unverified
2DTLNSI-SDR-WB16.34Unverified
3Non-Real-Time MultiScale+SI-SDR-WB16.22Unverified
4ZipEnhancer (M)PESQ-WB3.81Unverified
5TF-Locoformer (M)PESQ-WB3.72Unverified
6ZipEnhancer (S)PESQ-WB3.69Unverified
7MambAttentionPESQ-WB3.67Unverified
8MP-SENetPESQ-WB3.62Unverified
9xLSTM-SENetPESQ-WB3.59Unverified
10BSRNN-S + MRSDPESQ-WB3.53Unverified
#ModelMetricClaimedVerifiedStatus
1Inter-Channel Conv-TasNetSDR19.67Unverified
2CA Dense U-Net (Complex)SDR18.64Unverified
3Dense U-Net (Complex)SDR18.4Unverified
4Dense U-Net (Real)SDR16.86Unverified
5U-Net (Real)SDR15.97Unverified
6Noisy/unprocessedSDR6.5Unverified
#ModelMetricClaimedVerifiedStatus
1Schrödinger Bridge (PESQ loss)PESQ-WB3.09Unverified
2SGMSE+PESQ-WB2.5Unverified
3Demucs v4PESQ-WB2.37Unverified
4Schrödinger BridgePESQ-WB2.33Unverified
5Conv-TasNetPESQ-WB2.31Unverified
6CDiffuSEPESQ-WB1.6Unverified
#ModelMetricClaimedVerifiedStatus
1ReVISE (ch2)Audio Quality MOS4.19Unverified
2ReVISE (bf)Audio Quality MOS4.11Unverified
3Demucs (ch2)Audio Quality MOS2.95Unverified
4Demucs (bf)Audio Quality MOS2.39Unverified
5MaxDI (Baseline)PESQ1.17Unverified
6DAJA (MVDR,HMA,1000) (Overlapped Speech)SDR-4.76Unverified
#ModelMetricClaimedVerifiedStatus
1ZipEnhancer (M)PESQ-NB4.08Unverified
2DCCRN-MCPESQ-NB3.21Unverified
3DCCRN-MPESQ-NB3.15Unverified
4DCCRNPESQ-NB3.04Unverified
5RNN-ModulationPESQ-WB2.75Unverified
#ModelMetricClaimedVerifiedStatus
1MambAttentionESTOI0.8Unverified
2SEMambaESTOI0.8Unverified
3xLSTM-SENetESTOI0.8Unverified
4MP-SENetESTOI0.79Unverified
#ModelMetricClaimedVerifiedStatus
1SepFormerPESQ2.84Unverified
2DTLNPESQ2.23Unverified
3UnprocessedPESQ1.83Unverified
4Non-Real-Time MultiScale+PESQ1.52Unverified
#ModelMetricClaimedVerifiedStatus
1DCUNet-MCPESQ-NB3.44Unverified
2DCCRN-MPESQ-NB3.28Unverified
3DCUNetPESQ-NB3.25Unverified
#ModelMetricClaimedVerifiedStatus
1CleanMel-L-mapDNSMOS3.82Unverified
2SpatialNetDNSMOS BAK3.43Unverified
#ModelMetricClaimedVerifiedStatus
1rose_cd(PESQ )PESQ3.99Unverified
2ROSE-CDPESQ3.49Unverified
#ModelMetricClaimedVerifiedStatus
1Wave-U-NetCBAK3.24Unverified
#ModelMetricClaimedVerifiedStatus
1Audio-Visual concat-refPESQ2.7Unverified
#ModelMetricClaimedVerifiedStatus
1SE-MelGANAudio Quality MOS3.1Unverified
#ModelMetricClaimedVerifiedStatus
1DeFT-ANPESQ3.01Unverified
#ModelMetricClaimedVerifiedStatus
1Audio-Visual concat-refPESQ3.03Unverified
#ModelMetricClaimedVerifiedStatus
1SepFormerPESQ3.07Unverified