SAVRN
Search Contact SAVRN

Open-weight model · Audio to audio

MOSS-v2-12to32-Enhancer-60M

by LAION eV laion/MOSS-v2-12to32-Enhancer-60M

MOSS-v2-12to32-Enhancer-60M is an open-weight model for audio to audio from LAION eV, released under Creative Commons Attribution-NonCommercial 4.0. Its published files total 19.9 GB.

A small causal Qwen3 model that predicts 32 MOSS v2 audio-codebook indices from the 12 indices of the current frame and up to ten previously generated HQ frames. It is trained from scratch on German, English, Spanish and French speech.

Parameters—
Context—
Weights19.5 GB
Licensecc-by-nc-4.0
AccessOpen weights
Monthly Downloads—

Model Card

A small causal Qwen3 model that predicts 32 MOSS v2 audio-codebook indices from the 12 indices of the current frame and up to ten previously generated HQ frames. It is trained from scratch on German, English, Spanish and French speech. This is an experimental enhancement model. Validation cross entropy measures token prediction under teacher forcing. It does not establish an improvement in perceived audio quality. Use the generated-history listening evaluation to judge speech content, speaker identity, artifacts and long-utterance stability. Root inference weights: step 64,699, selected by held-out extra-20-codebook cross entropy 5.818123. - Step 12,940: 100 originals / 200 input cases…

Excerpt from the card by LAION eV, licensed cc-by-nc-4.0.

Identity and Version

Repository
laion/MOSS-v2-12to32-Enhancer-60M
Publisher
LAION eV
Task
Audio to audio
Modality
Audio
Library
pytorch
Parameters
Not stated by the source
Languages
de, en, es, fr
Revision
55580487eccb2d802a519e1c64f5f923dbb4e185
First published
2026-10-01
Last updated
2026-10-02

Files and Weights

171 files, 19.9 GB in total. The weights are 40 files totalling 19.5 GB in pt, safetensors.

Weights40 files · 19.5 GB
Configuration109 files · 1.7 MB
Documentation9 files · 69.2 KB
Other12 files · 406.5 MB
Repository1 file · 1.7 KB
Every file
FileTypeSizeSHA-256
checkpoints/step-00003235/model.safetensorsWeights243.5 MB 31420d1f7b65
checkpoints/step-00003235/training_state.ptWeights730.5 MB a49baf3c8aa7
checkpoints/step-00006470/model.safetensorsWeights243.5 MB 30fb0e7e0698
checkpoints/step-00006470/training_state.ptWeights730.5 MB 62cdf5d0732c
checkpoints/step-00009705/model.safetensorsWeights243.5 MB 6487dbed78a8
checkpoints/step-00009705/training_state.ptWeights730.5 MB d4170d5e4bb3
checkpoints/step-00012940/model.safetensorsWeights243.5 MB 4aad5ead4357
checkpoints/step-00012940/training_state.ptWeights730.5 MB b1806348dd5e
checkpoints/step-00016175/model.safetensorsWeights243.5 MB 7e2bc9898821
checkpoints/step-00016175/training_state.ptWeights730.5 MB b7b01b748fa1
checkpoints/step-00019410/model.safetensorsWeights243.5 MB 7108e826108c
checkpoints/step-00019410/training_state.ptWeights730.5 MB 10143cd6cf22
checkpoints/step-00022645/model.safetensorsWeights243.5 MB e045eed5c76b
checkpoints/step-00022645/training_state.ptWeights730.5 MB 15f3125852ab
checkpoints/step-00025880/model.safetensorsWeights243.5 MB b52adf269fbf
checkpoints/step-00025880/training_state.ptWeights730.5 MB d4ce08bc66e8
checkpoints/step-00029115/model.safetensorsWeights243.5 MB 392169109cf3
checkpoints/step-00029115/training_state.ptWeights730.5 MB a098b77499a5
checkpoints/step-00032350/model.safetensorsWeights243.5 MB f4fd87c995bb
checkpoints/step-00032350/training_state.ptWeights730.5 MB 6aa13c4436bf
checkpoints/step-00035585/model.safetensorsWeights243.5 MB e7a6283fc163
checkpoints/step-00035585/training_state.ptWeights730.5 MB bff214f1cd06
checkpoints/step-00038820/model.safetensorsWeights243.5 MB 13aeb74b57fc
checkpoints/step-00038820/training_state.ptWeights730.5 MB 04bda991bc59
checkpoints/step-00042055/model.safetensorsWeights243.5 MB 77551f2c991d
checkpoints/step-00042055/training_state.ptWeights730.5 MB 90a0b95c8ff6
checkpoints/step-00045290/model.safetensorsWeights243.5 MB 2008d88725ed
checkpoints/step-00045290/training_state.ptWeights730.5 MB 68884d01716f
checkpoints/step-00048525/model.safetensorsWeights243.5 MB 34eb2ca563b5
checkpoints/step-00048525/training_state.ptWeights730.5 MB 549620ee18b0
checkpoints/step-00051760/model.safetensorsWeights243.5 MB 483841ec5a2f
checkpoints/step-00051760/training_state.ptWeights730.5 MB 473248c226bf
checkpoints/step-00054995/model.safetensorsWeights243.5 MB 0fe5759e77df
checkpoints/step-00054995/training_state.ptWeights730.5 MB 0f73ca158a86
checkpoints/step-00058230/model.safetensorsWeights243.5 MB f45d159b6837
checkpoints/step-00058230/training_state.ptWeights730.5 MB 899660258a5a
checkpoints/step-00061465/model.safetensorsWeights243.5 MB fb6e7b7b3c4f
checkpoints/step-00061465/training_state.ptWeights730.5 MB bbdd2baabe7d
checkpoints/step-00064699/model.safetensorsWeights243.5 MB 3429482ada7b
checkpoints/step-00064699/training_state.ptWeights730.5 MB 49a6cbc6c8a7
best_checkpoint.jsonConfiguration652 B —
checkpoints/step-00003235/checkpoint.jsonConfiguration2.2 KB —
checkpoints/step-00003235/config.jsonConfiguration419 B —
checkpoints/step-00003235/validation.jsonConfiguration171 B —
checkpoints/step-00006470/checkpoint.jsonConfiguration2.2 KB —
checkpoints/step-00006470/config.jsonConfiguration419 B —
checkpoints/step-00006470/validation.jsonConfiguration170 B —
checkpoints/step-00009705/checkpoint.jsonConfiguration2.2 KB —
checkpoints/step-00009705/config.jsonConfiguration419 B —
checkpoints/step-00009705/validation.jsonConfiguration170 B —
checkpoints/step-00012940/checkpoint.jsonConfiguration2.2 KB —
checkpoints/step-00012940/config.jsonConfiguration419 B —
checkpoints/step-00012940/validation.jsonConfiguration169 B —
checkpoints/step-00016175/checkpoint.jsonConfiguration2.2 KB —
checkpoints/step-00016175/config.jsonConfiguration419 B —
checkpoints/step-00016175/validation.jsonConfiguration171 B —
checkpoints/step-00019410/checkpoint.jsonConfiguration2.2 KB —
checkpoints/step-00019410/config.jsonConfiguration419 B —
checkpoints/step-00019410/validation.jsonConfiguration169 B —
checkpoints/step-00022645/checkpoint.jsonConfiguration2.2 KB —
checkpoints/step-00022645/config.jsonConfiguration419 B —
checkpoints/step-00022645/validation.jsonConfiguration169 B —
checkpoints/step-00025880/checkpoint.jsonConfiguration2.2 KB —
checkpoints/step-00025880/config.jsonConfiguration419 B —
checkpoints/step-00025880/validation.jsonConfiguration170 B —
checkpoints/step-00029115/checkpoint.jsonConfiguration2.2 KB —
checkpoints/step-00029115/config.jsonConfiguration419 B —
checkpoints/step-00029115/validation.jsonConfiguration168 B —
checkpoints/step-00032350/checkpoint.jsonConfiguration2.2 KB —
checkpoints/step-00032350/config.jsonConfiguration419 B —
checkpoints/step-00032350/validation.jsonConfiguration168 B —
checkpoints/step-00035585/checkpoint.jsonConfiguration2.2 KB —
checkpoints/step-00035585/config.jsonConfiguration419 B —
checkpoints/step-00035585/validation.jsonConfiguration170 B —
checkpoints/step-00038820/checkpoint.jsonConfiguration2.2 KB —
checkpoints/step-00038820/config.jsonConfiguration419 B —
checkpoints/step-00038820/validation.jsonConfiguration169 B —
checkpoints/step-00042055/checkpoint.jsonConfiguration2.2 KB —
checkpoints/step-00042055/config.jsonConfiguration419 B —
checkpoints/step-00042055/validation.jsonConfiguration169 B —
checkpoints/step-00045290/checkpoint.jsonConfiguration2.2 KB —
checkpoints/step-00045290/config.jsonConfiguration419 B —
checkpoints/step-00045290/validation.jsonConfiguration169 B —
checkpoints/step-00048525/checkpoint.jsonConfiguration2.2 KB —
checkpoints/step-00048525/config.jsonConfiguration419 B —
checkpoints/step-00048525/validation.jsonConfiguration170 B —
checkpoints/step-00051760/checkpoint.jsonConfiguration2.2 KB —
checkpoints/step-00051760/config.jsonConfiguration419 B —
checkpoints/step-00051760/validation.jsonConfiguration170 B —
checkpoints/step-00054995/checkpoint.jsonConfiguration2.2 KB —
checkpoints/step-00054995/config.jsonConfiguration419 B —
checkpoints/step-00054995/validation.jsonConfiguration169 B —
checkpoints/step-00058230/checkpoint.jsonConfiguration2.2 KB —
checkpoints/step-00058230/config.jsonConfiguration419 B —
checkpoints/step-00058230/validation.jsonConfiguration169 B —
checkpoints/step-00061465/checkpoint.jsonConfiguration2.2 KB —
checkpoints/step-00061465/config.jsonConfiguration419 B —
checkpoints/step-00061465/validation.jsonConfiguration170 B —
checkpoints/step-00064699/checkpoint.jsonConfiguration2.2 KB —
checkpoints/step-00064699/config.jsonConfiguration419 B —
checkpoints/step-00064699/validation.jsonConfiguration170 B —
configs/data_sources.jsonConfiguration77.4 KB —
configs/model_60m.jsonConfiguration422 B —
configs/release_licenses.jsonConfiguration1.7 KB —
configs/training_single_view_test.jsonConfiguration1.2 KB —
eval/step-00012940-short-preview/samples.jsonConfiguration191.2 KB —
eval/step-00012940-short-preview/summary.jsonConfiguration283.3 KB —
eval/step-00012940/gpu_probe.jsonConfiguration1.5 KB —
eval/step-00012940/samples.jsonConfiguration179.2 KB —
eval/step-00012940/selection.jsonConfiguration32.5 KB —
eval/step-00012940/summary.jsonConfiguration190.6 KB —
eval/step-00022645-four-10s/samples.jsonConfiguration7.6 KB —
eval/step-00022645-four-10s/selection.jsonConfiguration2.5 KB —
eval/step-00022645-four-10s/summary.jsonConfiguration9.4 KB —
eval/step-00048525-four-10s/samples.jsonConfiguration7.6 KB —
eval/step-00048525-four-10s/selection.jsonConfiguration2.5 KB —
eval/step-00048525-four-10s/summary.jsonConfiguration9.4 KB —
eval/step-00064699-10s/gpu_probe.jsonConfiguration1.5 KB —
eval/step-00064699-10s/samples.jsonConfiguration186.9 KB —
eval/step-00064699-10s/selection.jsonConfiguration37.8 KB —
eval/step-00064699-10s/summary.jsonConfiguration198.8 KB —
manifest.jsonConfiguration15.8 KB —
provenance.jsonConfiguration947 B —
release_licenses.jsonConfiguration1.7 KB —
scripts/evaluate_audio.pyConfiguration15.5 KB —
scripts/start_model_release.pyConfiguration7.8 KB —
scripts/upload_model.pyConfiguration22.2 KB —
scripts/watch_audio_eval.pyConfiguration4.2 KB —
src/moss_enhancer/__init__.pyConfiguration76 B —
src/moss_enhancer/augment.pyConfiguration7.1 KB —
src/moss_enhancer/checkpoints.pyConfiguration3.1 KB —
src/moss_enhancer/data.pyConfiguration5.5 KB —
src/moss_enhancer/eval_support.pyConfiguration10.6 KB —
src/moss_enhancer/followup.pyConfiguration13.6 KB —
src/moss_enhancer/frame_dataset.pyConfiguration3.4 KB —
src/moss_enhancer/inference.pyConfiguration7.6 KB —
src/moss_enhancer/model.pyConfiguration12.8 KB —
src/moss_enhancer/model_config.pyConfiguration3.2 KB —
src/moss_enhancer/parallel_codec.pyConfiguration1.6 KB —
training/__init__.pyConfiguration84 B —
training/build_store.pyConfiguration4.9 KB —
training/epoch_data.pyConfiguration5.9 KB —
training/estimate_epoch.pyConfiguration3.9 KB —
training/followup_mix.pyConfiguration4.4 KB —
training/model_60m.jsonConfiguration422 B —
training/orchestrate.pyConfiguration4.2 KB —
training/progress.jsonConfiguration197 B —
training/train_distributed.pyConfiguration17.0 KB —
training/training_single_view_test.jsonConfiguration1.2 KB —
LICENSEDocumentation19.3 KB —
LICENSE-codeDocumentation1.1 KB —
LICENSES.mdDocumentation2.7 KB —
README.mdDocumentation19.9 KB —
docs/LICENSE-codeDocumentation1.1 KB —
docs/model_card.mdDocumentation18.1 KB —
docs/release_licenses.mdDocumentation2.7 KB —
docs/single_view_test.mdDocumentation2.2 KB —
training/single_view_test.mdDocumentation2.2 KB —
docs/licenses/EARS-CC-BY-NC-4.0.txtOther19.3 KB —
docs/licenses/MLS-CC-BY-4.0.txtOther18.7 KB —
docs/licenses/MLS-MIRROR-CC0-1.0.txtOther7.0 KB —
eval/step-00012940-short-preview/listening.htmlOther106.5 MB 5c0e44209db1
eval/step-00012940/listening.htmlOther184.3 MB c969860656f5
eval/step-00022645-four-10s/listening.htmlOther4.3 MB —
eval/step-00048525-four-10s/listening.htmlOther4.3 MB —
eval/step-00064699-10s/listening.htmlOther107.1 MB fd3a14ddeabe
licenses/EARS-CC-BY-NC-4.0.txtOther19.3 KB —
licenses/MLS-CC-BY-4.0.txtOther18.7 KB —
licenses/MLS-MIRROR-CC0-1.0.txtOther7.0 KB —
pyproject.tomlOther643 B —
.gitattributesRepository1.7 KB —

License and Download

License
cc-by-nc-4.0
Access
Open weights, no gate
Download size
19.5 GB
Download from LAION eV

Released by LAION eV through its official repository on Hugging Face. Read the license.

Built From

  • Trained on (disclosed) laion/MOSS-v2-12to32-Enhancement-Tokens

Memory Requirements

PrecisionWeights in memory
As published19.5 GB

Weights only, from the published parameter count; the key-value cache and runtime add to this.

Questions About MOSS-v2-12to32-Enhancer-60M

Can I use MOSS-v2-12to32-Enhancer-60M commercially?

Not without separate permission. MOSS-v2-12to32-Enhancer-60M is released under Creative Commons Attribution-NonCommercial 4.0. CC BY-NC 4.0 permits sharing and adapting with credit for non-commercial purposes only. Commercial use needs separate permission from the rights holder.

Similar Models

Model · Audio to audio

bigvgan_v2_22khz_80band_256x

NVIDIA

[[Paper]](https://arxiv.org/abs/2206.04658) - [[Code]](https://github.com/NVIDIA/BigVGAN) - [[Showcase]](https://bigvgan-demo.github.io/) - [[Project Page]](https://research.nvidia.com/labs/adlr/projects/bigvgan/) - [[Weights]](https://huggingface.co/collections/nvidia/bigvgan-66959df3d97fd7d98d97dc9a) - [[Demo]](https://huggingface.co/spaces/nvidia/BigVGAN) - General refactor and code improvements for improved readability. - Fully fused CUDA kernel of anti-alised activation (upsampling + activation + downsampling) with inference speed benchmark. - We provide pretrained checkpoints of BigVGAN-v2 using diverse audio configurations, supporting up to 44 kHz sampling rate and 512x upsampling…

Open weights mit PyTorch

Model · Audio to audio

bigvgan_v2_44khz_128band_512x

NVIDIA

[[Paper]](https://arxiv.org/abs/2206.04658) - [[Code]](https://github.com/NVIDIA/BigVGAN) - [[Showcase]](https://bigvgan-demo.github.io/) - [[Project Page]](https://research.nvidia.com/labs/adlr/projects/bigvgan/) - [[Weights]](https://huggingface.co/collections/nvidia/bigvgan-66959df3d97fd7d98d97dc9a) - [[Demo]](https://huggingface.co/spaces/nvidia/BigVGAN) - General refactor and code improvements for improved readability. - Fully fused CUDA kernel of anti-alised activation (upsampling + activation + downsampling) with inference speed benchmark. - We provide pretrained checkpoints of BigVGAN-v2 using diverse audio configurations, supporting up to 44 kHz sampling rate and 512x upsampling…

Open weights mit PyTorch

Descript Audio Codec running on-device on the LiteRT CompiledModel GPU (ML Drift). The convolutional encoder/decoder run on the GPU; the RVQ runs on CPU. 43:1 compression (1 s → 12×50 codes), RTF ≈ 0.82 (faster than real-time) on Pixel 8a. - dac16khzencoderfp16.tflite (43 MB) — audio[1,1,16000] → latent[1,1024,50], GPU. - dac16khzdeconlyzsfp16.tflite (105 MB) — latent[1,1024,50] → audio, GPU. - dacrvq.bin (1.2 MB) — RVQ weights (12 codebooks) for the CPU quantizer (float32 LE). encoder 367/367 + decoder 398/398 nodes on the LiteRT GPU delegate (LITERTCL, 1 partition, no CPU fallback); warm RTF ~0.82; reconstruction corr 1.0 vs PyTorch DAC. The decoder's ConvTranspose1d are rewritten to a…

Open weights mit litert

Model · Audio to audio

ArkEcho-RVC-M3-C-F

Ethernos

本仓库(Repository)所包含的所有人工神经网络(Artificial Neural Network)权重文件(.pth/.index)、训练日志及相关代码,均为计算声学(Computational Acoustics)与深度学习(Deep Learning)领域的技术研究实验产物。 这些文件本质上是高维张量(High-dimensional Tensors)的数值序列,通过随机梯度下降(SGD)与反向传播算法(Backpropagation)对公开可获取的音频数据进行统计建模(Statistical Modeling)得到。其技术形态与数字图像处理中的卷积核(Convolution Kernels)、自然语言处理中的词向量(Word Embeddings)并无本质差异。 本仓库不构成对任何第三方知识产权的故意侵犯,所有代码遵循 MIT License 开源协议,模型权重文件仅作为技术实现的副产品(By-products)存在。 本技术实验所使用的训练数据集(Training Dataset)包含以下角色的语音样本: - Mon3tr(モンスター)、Mon3tr 中文语音 等 - 声源版权归属:上海鹰角网络科技有限公司(Hypergryph Network Technology Co., Ltd.)及其关联公司 - 原始作品:《明日方舟》(Arknights) 上述角色的声音版权、肖像权、姓名权及相关知识产权均完全归属于鹰角网络及其合法授权方。本实验仅基于已公开发布的游戏内语音资源进行技术层面的信号处理(Signal Processing)与特征提取(Feature…

Open weights cc-by-nc-4.0

Model · Audio to audio

spansynth-edit

Sungkyun Chang

Synthesise music from MIDI, or add, remove, and modify notes in a recording by revising its MIDI. SpanSynth-Edit generates the selected region, using surrounding audio for timbre guidance. Try it in your browser: upload audio, transcribe with YourMT3+, edit the piano roll, and generate. See the local web setup to run the app yourself. - CPU inference is also supported. Measured peak VRAM was about 3.9 GiB on a GH200 for synthesis and both editing methods with default settings (20.48 s crop, 16 steps, CFG 2.0). Install PyTorch for your GPU first, then install the CLI without downloading the demo audio: Model and codec weights download automatically on first use. No token is required.…

Open weights apache-2.0 481M parameters diffusers