科研速览继续刷下去 →
◆ International Journal of Electronics and Telecommunications2026-07-31· Computer science

Google Speech Commands benchmarks tests with new dataset extension

Mateusz Kucharski, Radosław Idzikowski, Wiktoria Sarnicka

原始摘要(原文)
The Google Speech Commands dataset remains one of the most widely used benchmarks for evaluating keywordspotting and limited-vocabulary speech recognition systems. However, the discontinuation and partial loss of functionality of the Papers with Code platform has created gaps in the accessibility, transparency, and reproducibility of previously published benchmark results. This paper addresses this issue by reconstructing, validating, and comparing the performance of state-of-the-art keyword-spotting models whose code repositories remain publicly available. Four major frameworks—ML-KWSfor- MCU, Keyword Transformer (KWT), Howl, and SpeechCmdRecognition— were tested using both the original Google Speech Commands dataset and a newly developed extension recorded at Wrocław University of Science and Technology. The extension provides 3,710 high-quality recordings from 106 non-nativeEnglish speakers, captured under controlled acoustic conditions. Experimental results show that most architectures generalize well across datasets, with KWT models achieving the highest robustness and accuracy, including perfect classification under certain conditions. Conversely, classical DNN-based models exhibit significant performance degradation, demonstrating their sensitivity to acoustic variability. The study underscores the need for reliable benchmark repositories and standardized evaluationenvironments, particularly as emerging paradigms—such as quantum machine learning—begin to influence research in speech recognition. The findings provide a consolidated, verifiable comparison of contemporary models and lay groundwork for future benchmarking efforts in both classical and quantumenhanced approaches to speech processing.
读原文 ↗

AI 追问PRO

登录后使用 AI 追问

讨论区

登录后参与讨论

相关论文

Google Speech Commands benchmarks tests with new dataset extension — 科研速览 Science Skim