SAVRN
Search Contact SAVRN

Independent publisher

Anurag Singh

anuragshas

Machine Translation, Speech

Models in Library1
Datasets in Library0
Models on Hugging Face65
Followers13

Models

Model · Speech recognition

wav2vec2-large-xlsr-53-telugu

Anurag Singh

Fine-tuned facebook/wav2vec2-large-xlsr-53 on Telugu using the OpenSLR SLR66 dataset. When using this model, make sure that your speech input is sampled at 16kHz. The model can be used directly (without a language model) as follows: 70% of the OpenSLR Telugu dataset was used for training. Train Split of annotations is here Test Split of annotations is here Training Data Preparation notebook can be found here Training notebook can be foundhere Evaluation notebook is here

Open weights apache-2.0 transformers