← 返回资源分享
Speech Commands: A Dataset for Limited-Vocabulary Speech Recognition
论文
论文
发布时间2018-04-09
发表arXiv:1804.03209
作者:Pete Warden
详细介绍
Describes an audio dataset of spoken words designed to help train and
evaluate keyword spotting systems. Discusses why this task is an interesting
challenge, and why it requires a specialized dataset that is different from
conventional datasets used for automatic speech recognition of full sentences.
Suggests a methodology for reproducible and comparable accuracy metrics for
this task. Describes how the data was collected and verified, what it contains,
previous versions and properties. Concludes by reporting baseline results of
models trained on this dataset.
代码仓库 (33)
TomVeniat/SANAS官方PyTorch
ncsoft/phonmatchnet官方TensorFlow
ROBOTICSENGINEER/End_to_End_Learning_of_Speech_2D_Feature_Trajectory官方TensorFlow
bozliu/E2E-Keyword-SpottingPyTorch
TMarquet/speech_recognitionTensorFlow
MohFakih/AudioFoolPyTorch
lukesin/nas-for-kws-2PyTorch
nitinvwaran/UrbanSound8K-audio-classification-with-ResNetTensorFlow
pspratling/Speech-Recognition
widzemin/audio_projectTensorFlow
