SAVRN
Search Contact SAVRN

Independent publisher

Pham Nguyen Ngoc Bao

pnnbao-ump

Speech AI Specialist | Medical NLP Researcher | Author of VieNeu-TTS

Models in Library3
Datasets in Library0
Models on Hugging Face17
Followers128

Models

Model · Text to speech

VieNeu-TTS-v3-Turbo

Pham Nguyen Ngoc Bao

VieNeu-TTS v3 Turbo is the next generation of Vietnamese TTS — 48 kHz high-fidelity speech, 23 built-in preset voices across three regions (North / Central / South), instant voice cloning, real-time streaming with an OpenAI-compatible API (16 concurrent streams on one RTX 3060), inline emotion cues, and seamless bilingual (En–Vi) code-switching. The reference implementation is the vieneu Python SDK (v3.7.1). Its minimal install is torch-free: on CPU everything runs on ONNX Runtime (PyTorch is never imported), and on a CUDA machine it auto-switches to the PyTorch engine with automatic batching and a continuous-batching stream scheduler — same API, no code change. The VieNeu-TTS v3 Turbo…

Open weights apache-2.0 131M parameters 1,024 tokens

VieNeu-TTS-0.3B-Q4-0-GGUF is a Q40 quantized version of VieNeu-TTS-0.3B. This model is specifically optimized to run directly on the CPU, providing extremely fast speech synthesis without the need for a dedicated GPU. Training high-quality TTS models requires significant GPU resources. If you find this model useful, please consider supporting the development: eSpeak NG is mandatory for phonemization. Use the source code from GitHub for the best experience with full text preprocessing support: In the UI, select Backbone: VieNeu-TTS-0.3B-q4-gguf and Device: CPU. Install the SDK to integrate VieNeu-TTS-0.3B into your research or applications: This model is released under the CC BY-NC 4.0…

Open weights cc-by-nc-4.0

Model · Text to speech

VieNeu-TTS-v2

Pham Nguyen Ngoc Bao

VieNeu-TTS-v2 is the next generation of Vietnamese TTS, designed for Natural Communication, Podcasts, and Bilingual (En-Vi) Code-switching. This project features the flagship VieNeu-TTS-v2 architecture: Tác giả: Phạm Nguyễn Ngọc Bảo Training high-quality TTS models requires significant GPU resources. If you find this model useful, please consider supporting the development: Install the SDK to integrate VieNeu-TTS-0.3B into your research or applications: Deploy VieNeu-TTS as a high-performance API Server (powered by LMDeploy) with a single command. Start the Server with a Public Tunnel (No port forwarding needed): Once the server is running, you can connect from anywhere (Colab, Web Apps…

Open weights apache-2.0 294M parameters 4,096 tokens