SAVRN
Search Contact SAVRN

Independent publisher

Huh

JaesungHuh

audio visual learning

Models in Library1
Datasets in Library0
Models on Hugging Face2
Followers5

Models

Model · Audio classification

voice-gender-classifier

Huh

This repo contains the inference code to use pretrained human voice gender classifier. - You could also try Huggingface online demo. First, clone the original github repository and install the packages via pip. For those who need pretrained weights, please download it in here State-of-the-art speaker verification model already produces good representation of the speaker's gender. I used the pretrained ECAPA-TDNN from TaoRuijie's repository, added one linear layer to make two-class classifier, and finetuned the model with the VoxCeleb2 dev set. The model achieved 98.7% accuracy on the VoxCeleb1 identification test split. I would like to note the training dataset I've used for this model…

Open weights mit 15M parameters transformers