SAVRN
Search Contact SAVRN

Open-weight model · Image text to video

MiniMax-H3-GGUF

by Unsloth AI unsloth/MiniMax-H3-GGUF

Instructions further below. GGUF for MiniMax-H3, compatible on most platforms including stablediffusion.cpp and Unsloth.

Parameters
Context
Weights212.2 GB
Licenseother
AccessOpen weights
Monthly Downloads865.2k

Model Card

Instructions further below. GGUF for MiniMax-H3, compatible on most platforms including stablediffusion.cpp and Unsloth. You can run MiniMax-H3 via Unsloth: https://github.com/unslothai/unsloth/ GGUF quantizations of MiniMaxAI/MiniMax-H3 MiniMax H3 is an omni-modal generative system that produces video with native stereo audio, up to 15 seconds at 24 FPS with 32 kHz stereo audio. Both halves of the runtime are in this repo: the denoisers and the Qwen3-VL text encoder they need. H3 ships two denoisers, and which one you load decides what the model can be given: - fl2vapruned, the H3-Base first-and-last-frame variant. Text, plus zero, one or two frames. - ref2vapruned, the reference variant.…

Excerpt from the card by Unsloth AI, licensed other.

Identity and Version

Repository
unsloth/MiniMax-H3-GGUF
Publisher
Unsloth AI
Task
Image text to video
Modality
Other
Library
gguf
Parameters
Not stated by the source
Languages
en, zh
Revision
d629413c2e5b51b38c453668b75ca3b06ca92703
First published
2026-08-07
Last updated
2026-08-14

Files and Weights

24 files, 212.2 GB in total. The weights are 18 files totalling 212.2 GB in gguf, safetensors.

Weights18 files · 212.2 GB
Documentation3 files · 24.3 KB
Other2 files · 6.0 MB
Repository1 file · 2.8 KB
Every file
FileTypeSizeSHA-256
minimax_h3_fl2va_pruned-Q2_K.ggufWeights6.7 GB 696866a3c98e
minimax_h3_fl2va_pruned-Q3_K.ggufWeights8.8 GB 2c3cadbcc87c
minimax_h3_fl2va_pruned-Q4_K.ggufWeights11.4 GB dd948e08ad0b
minimax_h3_fl2va_pruned-Q5_0.ggufWeights13.9 GB e7e7d16af6fe
minimax_h3_fl2va_pruned-Q6_K.ggufWeights16.6 GB cf9453a56548
minimax_h3_fl2va_pruned-Q8_0.ggufWeights21.4 GB 1c77759fd30e
minimax_h3_fl2va_pruned-UD-Q2_K_XL.ggufWeights8.1 GB cfe0795c00ab
minimax_h3_fl2va_pruned-UD-Q3_K_XL.ggufWeights9.6 GB aee15137d311
minimax_h3_ref2va_pruned-Q2_K.ggufWeights6.7 GB 12089d0a9935
minimax_h3_ref2va_pruned-Q3_K.ggufWeights8.7 GB 1e93dbb3b881
minimax_h3_ref2va_pruned-Q4_K.ggufWeights11.4 GB 2fa5840021cf
minimax_h3_ref2va_pruned-Q5_0.ggufWeights13.9 GB b91ca85cecff
minimax_h3_ref2va_pruned-Q6_K.ggufWeights16.6 GB 4a7a3b910384
minimax_h3_ref2va_pruned-Q8_0.ggufWeights21.4 GB 60f8a47434ec
qwen3vl_32b_minimax_h3-Q2_K_M.ggufWeights13.1 GB a8ccadccd57e
qwen3vl_32b_minimax_h3-Q4_K_M.ggufWeights18.2 GB 11e6efe70a57
vae/minimax_h3_audio_vae_fp32.safetensorsWeights605.3 MB 8e505d95dd15
vae/minimax_h3_video_vae_fp16.safetensorsWeights5.2 GB 7c1f131492e7
LICENSEDocumentation17.6 KB
NOTICEDocumentation1.3 KB
README.mdDocumentation5.3 KB
assets/h3_gguf_ud_q2_k_xl.gifOther2.8 MB 21c66004ddf1
assets/h3_gguf_ud_q2_k_xl.mp4Other3.3 MB da4b9480b964
.gitattributesRepository2.8 KB

License and Download

License
other
Access
Open weights, no gate
Download size
212.2 GB
Download from Unsloth AI

Released by Unsloth AI through its official repository on Hugging Face. Read the license.

Built From

Memory Requirements

PrecisionWeights in memory
As published212.2 GB

Weights only, from the published parameter count; the key-value cache and runtime add to this.

Questions About MiniMax-H3-GGUF

What license is MiniMax-H3-GGUF released under?

other, as its publisher declares it. Read the license text before commercial use.

Similar Models

Model · Image text to video

Minimax-H3-nvfp4-INT4-INT8-Convrot

Jay

This repository is a community-compiled collection of quantized and pruned weights for MiniMax H3 (Hailuo 3.0), optimized for local inference environments like ComfyUI. By unifying various quantization formats (INT4, INT8, Mixed, and NVFP4) into a single structured repository, this hub makes it easier for users with consumer GPUs (16GB - 24GB VRAM) to experiment with MiniMax H3's powerful omni-modal text/image/audio-to-video generation capabilities. If you are new to local generation and aren't sure what to download, use this guide based on your graphics card. Perfect for RTX 4070 Ti Super, RTX 4080, etc. Perfect for RTX 3090, RTX 4090, etc. Exclusively for RTX 5090, PRO 6000, and other…

Open weights other diffusers

Model · Image text to video

MiniMax-H3-Pruned-GGUF

Jay

This repository (Abiray/MiniMax-H3-Pruned-GGUF) provides pruned and quantized GGUF weights for the MiniMax H3 omni-modal generative model. MiniMax H3 is designed for unified multimodal context processing, capable of generating synchronized high-definition video and 32 kHz stereo audio from text, image, audio, and video inputs. 1. Download your desired.gguf variant from the table above. 2. Place the downloaded.gguf file into the ComfyUI/models/unet/ directory. 3. In your ComfyUI workflow, load the model using the UnetLoaderGGUF node. MiniMax H3 is released under the MiniMax H3 Community License Agreement. Please refer to the MiniMaxAI/MiniMax-H3 repository and the repository's LICENSE file…

Open weights other

Model · Image text to video

MiniMax-H3-Realism-People-LoRA

Fal

A LoRA adapter for MiniMax H3 specialized in realistic people: faces that hold up in close-up, natural skin texture, believable expressions and gestures, film-style lighting and documentary camera movement. Same prompt, same seed — base model on the left, this adapter on the right: 19 pairs, same prompt, same seed, adapter on vs off — the only variable is the LoRA. The trigger word is present on both sides, so it is not doing the work. Each pair plays the base model first, then freezes and dims while the adapted version plays beside it. Close-up talking faces, arguments, several people speaking at once, weathered skin, children, ritual and travel scenes. Nothing cherry-picked from a larger…

Open weights other minimax-h3

Model · Image text to video

DaSiWa-MiniMax-H3-Hybrid

Bruno Po

Unmodified 4-step and 8-step DaSiWa MiniMax H3 Hybrid SafeTensors checkpoints mirrored for BRP Canvas downloads. The hybrid checkpoint supports text-to-video, reference-to-video, and first/last-frame-to-video through the same ComfyUI workflow. This is not an official MiniMax or DaSiWa distribution. Review the original model page and the included upstream MiniMax license before use. BRP Canvas defaults to shift video 9 and shift audio 4 for both distilled checkpoints. - 4-step: 56c52c7890c105308d28fe9c25c25fdb80e6cd6a54e2604d8af71732ba4ed74f - 8-step: e0441d26414f6e0c28f43d580e6cc56fad424da0fa4d261b698ca73188aa6332

Open weights other minimax-h3

Model · Image text to video

MiniMax-H3

MiniMax

Offical skills to improve prompt writing: skills on github Use MiniMax\-H3 directly via API\. Use MiniMax\-H3 directly via App\. MiniMax H3 is a general-purpose, omni-modal generative system. It supports unified understanding of multimodal contexts composed of text, images, video, and audio, and can generate video with native stereo audio at resolutions up to 2K and durations of up to 15 seconds. Thanks to its task-generalization-oriented system design, H3 already possesses broad multimodal context understanding and generation capabilities at the pre-training stage, enabling outstanding performance in following complex multimodal instructions. H3 supports the following input and output…

Open weights other 33.1B parameters minimax-h3