SAVRN
Search Contact SAVRN

Open-weight model · Text to video

MiniMax-H3-x-Z-Image-native

by Joey joeygambino/MiniMax-H3-x-Z-Image-native

The comfy-native cuts of the MiniMax-H3 × Z-Image graft: Z-Image's spatial-attention profile on H3's engine — richer sets and textures, same identity, no per-shot sharpening creep. Full story, demos and verification on the GGUF page.

Parameters
Context
Weights304.6 GB
Licenseother
AccessOpen weights
Monthly Downloads33.5k

Model Card

The comfy-native cuts of the MiniMax-H3 × Z-Image graft: Z-Image's spatial-attention profile on H3's engine — richer sets and textures, same identity, no per-shot sharpening creep. Full story, demos and verification on the GGUF page. Load with the plain Load Diffusion Model node, ComfyUI 0.32+. Files are the pruned H3 builds with the graft baked in (zs05 = late-block gains, dose 0.5): - bf16 — the master (ref2va) - comfy-fp8 / fp8e5m2 — fp8 scaled - comfy-int8 / int8convrot — the fast pick on RTX 50 - comfy-w4a8 / w4a4 / nvfp4 — 4-bit family for 16 GB cards (w4a8 is the quality pick; nvfp4 is Blackwell-native, emulated elsewhere) - comfy-mxfp8 — 8-bit microscaling, Blackwell-specialized…

Excerpt from the card by Joey, licensed other.

Identity and Version

Repository
joeygambino/MiniMax-H3-x-Z-Image-native
Publisher
Joey
Task
Text to video
Modality
Video
Library
minimax-h3
Parameters
Not stated by the source
Languages
Not stated by the source
Revision
113de7b82dfb34a9a8958077c94cd3787c9b782e
First published
2026-08-22
Last updated
2026-08-24

Files and Weights

17 files, 304.6 GB in total. The weights are 15 files totalling 304.6 GB in safetensors.

Weights15 files · 304.6 GB
Documentation1 file · 1.5 KB
Repository1 file · 1.5 KB
Every file
FileTypeSizeSHA-256
MiniMax-H3-fl2va-pruned-zs05-comfy-fp8.safetensorsWeights21.0 GB bcc1368b7b94
MiniMax-H3-fl2va-pruned-zs05-comfy-int8.safetensorsWeights21.0 GB 73ab00194045
MiniMax-H3-fl2va-pruned-zs05-comfy-mxfp8.safetensorsWeights21.6 GB 52e4376a1bae
MiniMax-H3-fl2va-pruned-zs05-comfy-nvfp4.safetensorsWeights12.5 GB cc0a5da65bd0
MiniMax-H3-fl2va-pruned-zs05-comfy-w4a8.safetensorsWeights12.5 GB 938189b37b30
MiniMax-H3-ref2va-pruned-zs05-comfy-fp8.safetensorsWeights21.0 GB 330fc9fe5b3f
MiniMax-H3-ref2va-pruned-zs05-comfy-fp8e5m2.safetensorsWeights21.0 GB 0834c5d2f79f
MiniMax-H3-ref2va-pruned-zs05-comfy-int8.safetensorsWeights21.0 GB 497c0ff6377e
MiniMax-H3-ref2va-pruned-zs05-comfy-mxfp8.safetensorsWeights21.6 GB 5e8b2c3a8e9d
MiniMax-H3-ref2va-pruned-zs05-comfy-nvfp4.safetensorsWeights12.5 GB a66fc5284e9b
MiniMax-H3-ref2va-pruned-zs05-comfy-w4a4.safetensorsWeights11.3 GB e66e77706b7a
MiniMax-H3-ref2va-pruned-zs05-comfy-w4a8.safetensorsWeights12.5 GB 1f48ec1090f5
minimax_h3_fl2va_zs05_int8_convrot.safetensorsWeights34.0 GB 4f09561efd59
minimax_h3_ref2va_pruned_zs05_bf16.safetensorsWeights40.2 GB d7cb04b66d79
minimax_h3_ref2va_pruned_zs05_int8_convrot.safetensorsWeights21.0 GB 71b8085ac422
README.mdDocumentation1.5 KB
.gitattributesRepository1.5 KB

License and Download

License
other
Access
Open weights, no gate
Download size
304.6 GB
Download from Joey

Released by Joey through its official repository on Hugging Face. Read the license.

Built From

Memory Requirements

PrecisionWeights in memory
As published304.6 GB

Weights only, from the published parameter count; the key-value cache and runtime add to this.

Questions About MiniMax-H3-x-Z-Image-native

What license is MiniMax-H3-x-Z-Image-native released under?

other, as its publisher declares it. Read the license text before commercial use.

Similar Models

Model · Text to video

Wan2.2-T2V-A14B-GGUF

QuantStack

This GGUF file is a direct conversion of Wan-AI/Wan2.2-T2V-A14B Since this is a quantized model, all original licensing terms and usage restrictions remain in effect. Usage The model can be used with the ComfyUI custom node ComfyUI-GGUF by city96 Place model files in ComfyUI/models/unet see the GitHub readme for further installation instructions.

Open weights apache-2.0 gguf

Model · Text to video

MiniMax-H3-Turbo-Lora

Larryvrh

A LoRA for MiniMax-H3 that renders joint video + synchronized stereo audio in as few as 4 sampling steps instead of the usual ~20 — a ~5× sampling speedup — and keeps getting better as you add steps. For most work, use minimaxh3turbov4step600ema.safetensors. It's the markedly better micro-detail (faces, fingers, fine texture), and the over-sharpening / plastic look of the earlier v1 (~850) line is fully resolved. v4 introduced a static-frame enhancement — a big win for static and small-motion content. The one trade-off shows up only at 4 steps with large, fast motion, where v4 can produce motion-smear / trailing ghosting (we're actively fixing this). Two things address it: - Use 6–8 steps.…

Open weights apache-2.0 minimax-h3

Model · Text to video

MiniMax-H3-Turbo-Lora-ComfyUI

DRBAPH

This repository contains MiniMax-H3 Turbo LoRAs converted and optimized for ComfyUI: These LoRAs accelerate MiniMax-H3 video and synchronized-audio generation by reducing the required number of sampling steps. Newly added LoRA, located in the experimental/ folder: Manual recommended sigmas: 3-step 1.0, 0.961165, 0.853333, 0.0 4-step 1.0, 0.970874, 0.907249, 0.640000, 0.0 Three LoRAs extracted from VDN-H3 8 step: The main 8-step LoRA works on both FL2VA and Ref2VA. If you are running a pruned base, choose the pruned version that corresponds to your base — the pruned versions need their own matching pruned base. Three dynamically resized BF16 LoRAs are now included. Their source weights were…

Open weights apache-2.0 minimax-h3

Model · Text to video

Sulphur-2-base

Sulphur

Sulphur 2 An uncensored video generation model based on LTX 2.3 supporting both t2v and i2v natively, as well as all of the other ltx 2.3 formats. Follow us on X Join our Discord Support the next version of the project, even just a few dollars would go a long way: Kofi To get started with the model, I recommend downloading either of the dev versions, (fp8mixed or bf16) and downloading the distill lora provided. By the way, I'm aware the workflows contain sulphurfinal right now, just use the lora or use the full models, don't use both at the same time. This model contains a prompt enhancer. The easiest way to get started with the prompt enhancer is by using it on lmstudio. The way to…

Open weights diffusers

Model · Text to video

Wan2.1-VACE-1.3B-GGUF

Sam

Wan2.1 is an open-source suite of video foundation models, compatible with consumer-grade GPUs, that excels in various video generation tasks like text-to-video, image-to-video, and video editing, even supporting visual text generation. Download models using huggingface-cli: You can also download directly from this page. This model is a derivative work of the original model licensed under the Apache 2.0 License, and is therefore distributed under the terms of the same license. Thanks to Patrick Gillespie for creating the ASCII text art tool used in this project https://patorjk.com/software/taag/ Wan-AI for the Wan model https://huggingface.co/Wan-AI/Wan2.1-VACE-1.3B…

Open weights apache-2.0 diffusers

Model · Text to video

Wan2.1-T2V-1.3B-GGUF

Sam

Wan2.1 is an open-source suite of video foundation models, compatible with consumer-grade GPUs, that excels in various video generation tasks like text-to-video, image-to-video, and video editing, even supporting visual text generation. Download models using huggingface-cli: You can also download directly from this page. This model is a derivative work of the original model licensed under the Apache 2.0 License, and is therefore distributed under the terms of the same license. Thanks to Patrick Gillespie for creating the ASCII text art tool used in this project https://patorjk.com/software/taag/ Wan-AI for the Wan model https://huggingface.co/Wan-AI/Wan2.1-T2V-1.3B…

Open weights apache-2.0 diffusers