SAVRN
Search Contact SAVRN

Independent publisher

Gary Stafford

garystafford

image, video, and audio generation, object detection, fine-tuning, MLOps

Models in Library1
Datasets in Library0
Models on Hugging Face4
Followers5

Models

Model · Audio classification

wav2vec2-deepfake-voice-detector

Gary Stafford

Fine-tuned Wav2Vec2 model for detecting AI-generated speech. Determines if audio was spoken by a human or created by AI text-to-speech/voice cloning software. Fine-tuned Wav2Vec2 transformer for binary audio classification (real vs AI-generated speech). Trained to distinguish authentic human speech from synthetic audio generated by AI text-to-speech and voice cloning services including: Note: This model uses transfer learning from a base model already trained for deepfake detection. Fast convergence is expected due to task similarity and TTS engine overlap with the base model's training data. The model outputs logits (raw, unnormalized scores) for two classes: Apply softmax to convert raw…

Open weights apache-2.0 316M parameters transformers