SAVRN
Search Contact SAVRN

Independent publisher

Harveen Singh Chadha

Harveenchadha

Speech Recognition, NLP, NLU, Indic Languages

Models in Library1
Datasets in Library0
Models on Hugging Face30
Followers48

Models

Fine-tuned on Multilingual Pretrained Model CLSRIL-23. The original fairseq checkpoint is present here. When using this model, make sure that your speech input is sampled at 16kHz. Note: The result from this model is without a language model so you may witness a higher WER in some cases. This model was trained on 4200 hours of Hindi Labelled Data. The labelled data is not present in public domain as of now. Models were trained using experimental platform setup by Vakyansh team at Ekstep. Here is the training repository. In case you want to explore training logs on wandb they are here. The model can be used directly (without a language model) as follows: The model can be evaluated as follows…

Open weights mit transformers