SAVRN
Search Contact SAVRN

SAVRN Model Hub · Datasets by Task

Audio to Audio Datasets

4 open-weight audio to audio datasets in the SAVRN Model Hub, with Muhammad Zohaib Hassan, Rvc Model and Sungkyun Chang publishing the most.

4Datasets
4Publishers
3Licenses

Most Downloaded

DatasetPublisherLicenseMonthly downloads
Quranic-Translation-Audio-Data Muhammad Zohaib Hassan apache-2.0 4.6k
speech-to-speech-collections Meghana Bommadi other 221
spansynth-edit-gallery Sungkyun Chang Not stated —
gbs Rvc Model other —

Licenses

LicenseDatasetsCommercial use
other2Read the license
not stated1Not stated
apache-2.01Yes

Who Publishes Them

PublisherDatasets
Muhammad Zohaib Hassan1
Rvc Model1
Sungkyun Chang1
Meghana Bommadi1

All 4 Datasets

Quranic Translation Audio Data is a highly curated, standardized, and streaming-optimized multilingual audio dataset containing the complete recitation of translation audios and commentaries of the Holy Quran across 51 different translation directories. Every audio track has been meticulously converted from heavy.mp3 source files into the modern, high-fidelity Opus (.opus) format at a streaming-optimized bitrate of 32kbps. Alongside the audio, the dataset includes highly efficient Protocol Buffer (.pb) files containing word-by-word/verse-by-verse timing data for real-time syncing. This ensures crystal-clear vocal clarity while achieving maximum compression and immediate playback…

Publicly accessible apache-2.0

Dataset · Audio to audio

speech-to-speech-collections

Meghana Bommadi

1779 train / 43 validation calls from a production collections voice agent (Hindi/Hinglish), processed into the same schema as MeghanaKap/speech-to-speech-data. sides can be modelled without speaker separation. Real customers are heard in these recordings, so access is gated and reviewed. Kept: the customer's first name (the model must learn to say it), the amounts discussed, and the last 4 digits of the loan the agent reads aloud. Removed: full loan/account numbers, phone numbers (muted in the audio, tagged in the text, and stripped from the word timings), conversation and client ids, the source recording URL, and any digit run of 5 or more in the reference script. slotdict / systemprompt…

Access requested at publisher other

Dataset · Audio to audio

gbs

Rvc Model

This bundle prepares one RVC v2, 40 kHz, F0-guided multi-speaker base model from: - AISHELL/AISHELL-3 on Hugging Face; - badayvedat/VCTK on Hugging Face; - the official JVS Google Drive archive; - hfmwhisper110spk.zip in Rvcmodel/gbs. The HFM package contains 13,303 mono 48 kHz/16-bit WAV files from 110 speakers, totaling 19.889 hours. Its SHA-256 is cfc382e9b9b176a3c9d854bbcd0edaf8647d3beaba13ca6ea5f84abdb899a3db. The script pins upstream RVC commit 81eed5e8f68b6bed1789f682fe78cdd324495afc. That revision supports multi-speaker manifests but caps the helper at speaker ID 109, so the bootstrap raises only that manifest-validation limit. The model embedding size is generated from the actual…

Publicly accessible other

Dataset · Audio to audio

spansynth-edit-gallery

Sungkyun Chang

Finished MIDI-guided music edits made with SpanSynth-Edit. Open a work's shared link to compare the original and edited audio, explore its score, and make an editable copy. Each work has its own folder: - input.wav and output.wav: the original clip and saved generated audio, at 48 kHz mono. - original.mid and edited.mid: the source and edited score, aligned to the clip. - context.wav: the normalized audio used by the model, including the history before the clip. - project.json: the title, description, notes, editing region, generation settings, and piano-roll display settings. Shared links have the form https://mimbres-spansynth-edit.hf.space/?work=FOLDERNAME. An unlisted work is still…

Publicly accessible

Other Tasks

See all