SAVRN
Search Contact SAVRN

Independent publisher

Mahmoud Ashraf

MahmoudAshraf

Models in Library1
Datasets in Library0
Models on Hugging Face4
Followers38

Models

Model · Speech recognition

mms-300m-1130-forced-aligner

Mahmoud Ashraf

This Python package provides an efficient way to perform forced alignment between text and audio using Hugging Face's pretrained models. it also features an improved implementation to use much less memory than TorchAudio forced alignment API. The model checkpoint uploaded here is a conversion from torchaudio to HF Transformers for the MMS-300M checkpoint trained on forced alignment dataset

Open weights cc-by-nc-4.0 315M parameters transformers