Quranic Translation Audio Data is a highly curated, standardized, and streaming-optimized multilingual audio dataset containing the complete recitation of translation audios and commentaries of the Holy Quran across 51 different translation directories. Every audio track has been meticulously converted from heavy.mp3 source files into the modern, high-fidelity Opus (.opus) format at a streaming-optimized bitrate of 32kbps. Alongside the audio, the dataset includes highly efficient Protocol Buffer (.pb) files containing word-by-word/verse-by-verse timing data for real-time syncing. This ensures crystal-clear vocal clarity while achieving maximum compression and immediate playback…
Publicly accessible
apache-2.0
1779 train / 43 validation calls from a production collections voice agent (Hindi/Hinglish), processed into the same schema as MeghanaKap/speech-to-speech-data. sides can be modelled without speaker separation. Real customers are heard in these recordings, so access is gated and reviewed. Kept: the customer's first name (the model must learn to say it), the amounts discussed, and the last 4 digits of the loan the agent reads aloud. Removed: full loan/account numbers, phone numbers (muted in the audio, tagged in the text, and stripped from the word timings), conversation and client ids, the source recording URL, and any digit run of 5 or more in the reference script. slotdict / systemprompt…
Access requested at publisher
other
This bundle prepares one RVC v2, 40 kHz, F0-guided multi-speaker base model from: - AISHELL/AISHELL-3 on Hugging Face; - badayvedat/VCTK on Hugging Face; - the official JVS Google Drive archive; - hfmwhisper110spk.zip in Rvcmodel/gbs. The HFM package contains 13,303 mono 48 kHz/16-bit WAV files from 110 speakers, totaling 19.889 hours. Its SHA-256 is cfc382e9b9b176a3c9d854bbcd0edaf8647d3beaba13ca6ea5f84abdb899a3db. The script pins upstream RVC commit 81eed5e8f68b6bed1789f682fe78cdd324495afc. That revision supports multi-speaker manifests but caps the helper at speaker ID 109, so the bootstrap raises only that manifest-validation limit. The model embedding size is generated from the actual…
Publicly accessible
other
Finished MIDI-guided music edits made with SpanSynth-Edit. Open a work's shared link to compare the original and edited audio, explore its score, and make an editable copy. Each work has its own folder: - input.wav and output.wav: the original clip and saved generated audio, at 48 kHz mono. - original.mid and edited.mid: the source and edited score, aligned to the clip. - context.wav: the normalized audio used by the model, including the history before the clip. - project.json: the title, description, notes, editing region, generation settings, and piano-roll display settings. Shared links have the form https://mimbres-spansynth-edit.hf.space/?work=FOLDERNAME. An unlisted work is still…
Publicly accessible