SAVRN
Search Contact SAVRN

Independent publisher

Binh Nguyen

nguyenvulebinh

NLP, Speech and more

Models in Library1
Datasets in Library0
Models on Hugging Face52
Followers103

Models

Model · Speech recognition

wav2vec2-base-vi-vlsp2020

Binh Nguyen

Our models use wav2vec2 architecture, pre-trained on 13k hours of Vietnamese youtube audio (un-label data) and fine-tuned on 250 hours labeled of VLSP ASR dataset on 16kHz sampled speech audio. You can find more description here The ASR model parameters are made available for non-commercial use only, under the terms of the Creative Commons Attribution-NonCommercial 4.0 International (CC BY-NC 4.0) license. You can find details at: https://creativecommons.org/licenses/by-nc/4.0/legalcode [email protected]

Open weights cc-by-nc-4.0 transformers