Dataset · Audio classification
spanish-accents-speech
by Silencio Voice AI SilencioNetwork/spanish-accents-speech
Spanish from 12 speakers born in 9 countries. 19 clips, 0.3 hours, each labelled with the speaker's country of birth, first language, self-reported accent or regional variety, declared Spanish proficiency, age band, gender and recording device.
Dataset Card
Spanish from 12 speakers born in 9 countries. 19 clips, 0.3 hours, each labelled with the speaker's country of birth, first language, self-reported accent or regional variety, declared Spanish proficiency, age band, gender and recording device. Audio and speaker metadata only. Human-validated transcription is available on request. Silencio's Spanish catalogue is 65% Latin American by recorded hours, led by Venezuela, with Spain at 9%. It also holds a substantial body of Spanish spoken as a second language: 18% of hours come from speakers born in Africa, notably Nigeria, Kenya and Egypt. This sample covers 9 countries of birth across Latin America, Spain, Africa, the Middle East and the…
Excerpt from the card by Silencio Voice AI, licensed cc-by-nc-4.0.
Structure
default 19 rows
| Split | Rows | Size |
|---|---|---|
| test | 19 | 149.7 MB |
Details
- Repository
- SilencioNetwork/spanish-accents-speech
- Publisher
- Silencio Voice AI
- Task category
- Audio classification
- Tags
- spanish, accented-speech, accent-classification
- Size category
- n<1K
- Languages
- es
- Revision
- 965553e7e25a6fed4ae5fdf88d6331bc0767359b
- Last updated
- 2026-09-22
Files
3 files, 121.5 MB in total.
Every file
| File | Type | Size | SHA-256 |
|---|---|---|---|
| data/test-00000-of-00001.parquet | Data | 121.5 MB | 622bdba0abc3 |
| README.md | Documentation | 12.9 KB | — |
| .gitattributes | Repository | 2.5 KB | — |
License and Download
- License
- cc-by-nc-4.0
- Access
- No access gate
Released by Silencio Voice AI through its official repository on Hugging Face. Read the license.