The People’s Speech is a free-to-download 30,000-hour and growing supervised conversational English speech recognition dataset licensed for academic and commercial usage under CC-BY-SA (with a CC-BY subset). The data is collected via searching the Internet for appropriately licensed audio data with existing transcriptions. A model trained on this dataset achieves a 9.98% word error rate on Librispeech’s test-clean test set.
← 返回资源分享
The People’s Speech
语音小管家 · 数据
数据Speech Recognition
提供方语音小管家
许可协议Creative Commons
发布时间2022-05-27
获取方式
https://openreview.net/pdf?id=R8CwidgJ0yT
