SAVRN
Search Contact SAVRN

Independent publisher

Vu Van

TruongVuVan

Models in Library0
Datasets in Library1
Models on Hugging Face
Followers

Datasets

Dataset chứa 6 giọng tiếng Việt và công cụ speak.py để chạy trên Google Colab với OmniVoice. Mỗi giọng gồm profile.json, voice.pt (prompt cache), audio mẫu và reftext.txt. 1. Lồng tiếng SRT: mở colab/OmivoiceVIColab.ipynb 2. Clone giọng mới: mở colab/OmivoiceVICloneColab.ipynb 3. Đặt HFREPO = " " 4. Runtime → GPU → chạy từng cell - native (mặc định, khuyến nghị): native ≤cap → dùng gap tới cue sau → stretch ≤x1.1 → chồng nhẹ đuôi (crossfade). Không đẩy timestamp SRT — mỗi cue neo đúng start gốc. Rút gọn câu: tự sửa trên file SRT. - cascade: giữ tốc độ tự nhiên, cue dài đẩy cue sau (lệch timeline) - fit: kéo nén tín hiệu vừa khung SRT - strict: ghép theo timestamp (có thể cắt audio)…

Publicly accessible apache-2.0 n<1K