SOTAVerified

Keyword Spotting

In speech processing, keyword spotting deals with the identification of keywords in utterances.

( Image credit: Simon Grest )

Papers

Showing 151–200 of 407 papers

TitleStatusHype
Fully Unsupervised Training of Few-shot Keyword Spotting—0
Frequency & Channel Attention Network for Small Footprint Noisy Spoken Keyword Spotting—0
Challenges and Opportunities in Multi-device Speech Processing—0
An In-Vehicle KWS System with Multi-Source Fusion for Vehicle Applications—0
Flexible Keyword Spotting based on Homogeneous Audio-Text Embedding—0
Global-Local Convolution with Spiking Neural Networks for Energy-efficient Keyword Spotting—0
GraphemeAug: A Systematic Approach to Synthesized Hard Negative Keyword Spotting Examples—0
An Optimized Recurrent Unit for Ultra-Low-Power Keyword Spotting—0
GTTS-EHU Systems for QUESST at MediaEval 2014—0
Hardware Aware Training for Efficient Keyword Spotting on General Purpose and Specialized Hardware—0
Hardware/Software Co-Design of RISC-V Extensions for Accelerating Sparse DNNs on FPGAs—0
HEiMDaL: Highly Efficient Method for Detection and Localization of wake-words—0
Contrastive Augmentation: An Unsupervised Learning Approach for Keyword Spotting in Speech Technology—0
Hierarchical Neural Network Architecture In Keyword Spotting—0
Contrastive Learning With Audio Discrimination For Customizable Keyword Spotting In Continuous Speech—0
A Fast Network Exploration Strategy to Profile Low Energy Consumption for Keyword Spotting—0
Fixed-point quantization aware training for on-device keyword-spotting—0
How Tiny Can Analog Filterbank Features Be Made for Ultra-low-power On-device Keyword Spotting?—0
A Multitask Training Approach to Enhance Whisper with Contextual Biasing and Open-Vocabulary Keyword Spotting—0
IIIT-H System for MediaEval 2014 QUESST—0
台語關鍵詞辨識之實作與比較 (Implementation and Comparison of Keyword Spotting for Taiwanese) [In Chinese]—0
Implicit Acoustic Echo Cancellation for Keyword Spotting and Device-Directed Speech Detection—0
CUHK System for QUESST Task of MediaEval 2014—0
Improved low-resource Somali speech recognition by semi-supervised acoustic and language model training—0
Finding Opinion Manipulation Trolls in News Community Forums—0
Filterbank Learning for Noise-Robust Small-Footprint Keyword Spotting—0
Improving Reverberant Speech Training Using Diffuse Acoustic Simulation—0
Improving Small Footprint Few-shot Keyword Spotting with Supervision on Auxiliary Data—0
BUT QUESST 2015 System Description—0
Improving vision-inspired keyword spotting using dynamic module skipping in streaming conformer encoder—0
An Integrated Framework for Two-pass Personalized Voice Trigger—0
DASB -- Discrete Audio and Speech Benchmark—0
BUT QUESST 2014 System Description—0
Data Augmentation for Robust Keyword Spotting under Playback Interference—0
Keyword-Guided Adaptation of Automatic Speech Recognition—0
DCCRN-KWS: an audio bias based model for noise robust small-footprint keyword spotting—0
A Channel-Pruned and Weight-Binarized Convolutional Neural Network for Keyword Spotting—0
Keyword spotting -- Detecting commands in speech using deep learning—0
Keyword spotting for audiovisual archival search in Uralic languages—0
Keyword Spotting for Hearing Assistive Devices Robust to External Speakers—0
An Exploration into the Performance of Unsupervised Cross-Task Speech Representations for "In the Wild'' Edge Applications—0
A 14uJ/Decision Keyword Spotting Accelerator with In-SRAM-Computing and On Chip Learning for Customization—0
Leveraging Large Language Models for Exploiting ASR Uncertainty—0
A Lightweight dynamic filter for keyword spotting—0
Latency Control for Keyword Spotting—0
Learnable Front Ends Based on Temporal Modulation for Music Tagging—0
FEL: High Capacity Learning for Recommendation and Ranking via Federated Ensemble Learning—0
Learning Decoupling Features Through Orthogonality Regularization—0
Feature exploration for almost zero-resource ASR-free keyword spotting using a multilingual bottleneck extractor and correspondence autoencoders—0
Feature learning for efficient ASR-free keyword spotting in low-resource languages—0
Show:102550
← PrevPage 4 of 9Next →

Benchmark Results

#ModelMetricClaimedVerifiedStatus
1NNI non-filtered(for the development set)Cnxe6.09—Unverified
2NNI Choi(for the development set)Cnxe5.89—Unverified
3NTU rnn (eval)Cnxe2.01—Unverified
4NTU dtw (eval)Cnxe2.01—Unverified
5NTU dtw (dev)Cnxe2.01—Unverified
6NTU rnn (dev)Cnxe2.01—Unverified
7ELiRF SDTW (eval)Cnxe1.19—Unverified
8ELiRF SDTW-avg (eval)Cnxe1.07—Unverified
9ELiRF SDTW (dev)Cnxe1.07—Unverified
10CUNY [Subseq+MFCC] (eval)Cnxe1.07—Unverified
#ModelMetricClaimedVerifiedStatus
1WaveFormerGoogle Speech Commands V2 1298.8—Unverified
2QNNGoogle Speech Commands V2 3598.6—Unverified
3TripletLoss-res15Google Speech Commands V1 1298.56—Unverified
4M2DGoogle Speech Commands V2 3598.5—Unverified
5EAT-SGoogle Speech Commands V2 3598.15—Unverified
6Audio Spectrogram TransformerGoogle Speech Commands V2 3598.11—Unverified
7EdgeCRNN 2.0×Google Speech Commands V2 1298.05—Unverified
8BC-ResNet-8Google Speech Commands V1 1298—Unverified
9HTS-ATGoogle Speech Commands V2 3598—Unverified
10Wav2KWSGoogle Speech Commands V1 1297.9—Unverified
#ModelMetricClaimedVerifiedStatus
1Stacked 1D CNNError Rate1.99—Unverified
2End-to-end DNN-HMMError Rate1.7—Unverified
3HEiMDaLError Rate0.45—Unverified
#ModelMetricClaimedVerifiedStatus
1Res26Accuracy95.88—Unverified
2EfficientNet-A0 + SA + TLAccuracy95.83—Unverified
#ModelMetricClaimedVerifiedStatus
1QuaternionNeuralNetworkAccuracy (10-fold)98.53—Unverified
2SSAMBAAccuracy (10-fold)97.4—Unverified
#ModelMetricClaimedVerifiedStatus
1TensorFlow's model version 2TFMA89.7—Unverified
2TensorFlow's model version 1TFMA85.4—Unverified
#ModelMetricClaimedVerifiedStatus
12D-ConvNetAccuracy (%)95.4—Unverified
21D-ConvNetAccuracy (%)93.7—Unverified
#ModelMetricClaimedVerifiedStatus
1Quaternion Neural NetworksAccuracy(10-fold)98.53—Unverified
#ModelMetricClaimedVerifiedStatus
1MicroNet-KWS-LAccuracy95.3—Unverified