SAVRN
Search Contact SAVRN

Independent publisher

Artur Muratov

artur-muratov

NLP, Speech Recognition, Computer Vision

Models in Library0
Datasets in Library1
Models on Hugging Face21
Followers

Datasets

This dataset contains augmented speech command samples in 15 languages, derived from multiple public datasets. Only commands that overlap with the Google Speech Commands (GSC) vocabulary are included, making the dataset suitable for multilingual keyword spotting tasks aligned with GSC-style classification. Audio samples have been augmented using standard audio techniques to improve model robustness (e.g., time-shifting, noise injection, pitch variation). The dataset is organized in folders per command label. Metadata files are included to facilitate training and evaluation. - One folder per speech command (e.g., yes/, no/, go/, stop/, etc.) - traininglist.txt - validationlist.txt…

Publicly accessible cc-by-4.0