This dataset contains augmented speech command samples in 15 languages, derived from multiple public datasets. Only commands that overlap with the Google Speech Commands (GSC) vocabulary are included, making the dataset suitable for multilingual keyword spotting tasks aligned with GSC-style classification. Audio samples have been augmented using standard audio techniques to improve model robustness (e.g., time-shifting, noise injection, pitch variation). The dataset is organized in folders per command label. Metadata files are included to facilitate training and evaluation. - One folder per speech command (e.g., yes/, no/, go/, stop/, etc.) - traininglist.txt - validationlist.txt…
Independent publisher
Artur Muratov
artur-muratov
NLP, Speech Recognition, Computer Vision
Models in Library0
Datasets in Library1
Models on Hugging Face21
Followers—