SAVRN
Search Contact SAVRN

Open-weight model · Image text to video

MiniMax-H3-Pruned-GGUF

by Jay Abiray/MiniMax-H3-Pruned-GGUF

This repository (Abiray/MiniMax-H3-Pruned-GGUF) provides pruned and quantized GGUF weights for the MiniMax H3 omni-modal generative model.

Parameters
Context
Weights197.0 GB
Licenseother
AccessOpen weights
Monthly Downloads143.2k

Model Card

This repository (Abiray/MiniMax-H3-Pruned-GGUF) provides pruned and quantized GGUF weights for the MiniMax H3 omni-modal generative model. MiniMax H3 is designed for unified multimodal context processing, capable of generating synchronized high-definition video and 32 kHz stereo audio from text, image, audio, and video inputs. 1. Download your desired.gguf variant from the table above. 2. Place the downloaded.gguf file into the ComfyUI/models/unet/ directory. 3. In your ComfyUI workflow, load the model using the UnetLoaderGGUF node. MiniMax H3 is released under the MiniMax H3 Community License Agreement. Please refer to the MiniMaxAI/MiniMax-H3 repository and the repository's LICENSE file…

Excerpt from the card by Jay, licensed other.

Identity and Version

Repository
Abiray/MiniMax-H3-Pruned-GGUF
Publisher
Jay
Task
Image text to video
Modality
Other
Library
Not stated by the source
Parameters
Not stated by the source
Languages
Not stated by the source
Revision
76061e7fe4866dbc68903bfdf852430f0f1b2487
First published
2026-08-07
Last updated
2026-08-07

Files and Weights

19 files, 197.0 GB in total. The weights are 14 files totalling 197.0 GB in gguf.

Weights14 files · 197.0 GB
Documentation3 files · 23.1 KB
Other1 file · 1.3 MB
Repository1 file · 2.9 KB
Every file
FileTypeSizeSHA-256
MiniMax-H3-FL2VA-Pruned-Q3_K_M.ggufWeights8.9 GB c844d70dbb4d
MiniMax-H3-FL2VA-Pruned-Q4_K_M.ggufWeights11.6 GB d74e644906fe
MiniMax-H3-FL2VA-Pruned-Q4_K_S.ggufWeights11.6 GB 5f1c73f8dca3
MiniMax-H3-FL2VA-Pruned-Q5_K_M.ggufWeights14.1 GB abe739ac98bc
MiniMax-H3-FL2VA-Pruned-Q5_K_S.ggufWeights14.1 GB 16a773d7e906
MiniMax-H3-FL2VA-Pruned-Q6_K.ggufWeights16.7 GB a4f8efd15b24
MiniMax-H3-FL2VA-Pruned-Q8_0.ggufWeights21.6 GB 8dfe882b4dee
MiniMax-H3-Ref2VA-Pruned-Q3_K_M.ggufWeights8.9 GB 6e22b83e7c22
MiniMax-H3-Ref2VA-Pruned-Q4_K_M.ggufWeights11.6 GB a67ff28c56f0
MiniMax-H3-Ref2VA-Pruned-Q4_K_S.ggufWeights11.6 GB 8744d71ce13a
MiniMax-H3-Ref2VA-Pruned-Q5_K_M.ggufWeights14.1 GB dd220ed38883
MiniMax-H3-Ref2VA-Pruned-Q5_K_S.ggufWeights14.1 GB bcc259231840
MiniMax-H3-Ref2VA-Pruned-Q6_K.ggufWeights16.7 GB 78eb11832967
MiniMax-H3-Ref2VA-Pruned-Q8_0.ggufWeights21.6 GB d42b22f417a3
LICENSEDocumentation17.6 KB
NOTICEDocumentation120 B
README.mdDocumentation5.3 KB
video/MiniMax_H3_00002-audio.mp4Other1.3 MB 6cd366e1efbe
.gitattributesRepository2.9 KB

License and Download

License
other
Access
Open weights, no gate
Download size
197.0 GB
Download from Jay

Released by Jay through its official repository on Hugging Face. Read the license.

Built From

Memory Requirements

PrecisionWeights in memory
As published197.0 GB

Weights only, from the published parameter count; the key-value cache and runtime add to this.

Questions About MiniMax-H3-Pruned-GGUF

What license is MiniMax-H3-Pruned-GGUF released under?

other, as its publisher declares it. Read the license text before commercial use.

Similar Models

Model · Image text to video

MiniMax-H3-GGUF

Unsloth AI

Instructions further below. GGUF for MiniMax-H3, compatible on most platforms including stablediffusion.cpp and Unsloth. You can run MiniMax-H3 via Unsloth: https://github.com/unslothai/unsloth/ GGUF quantizations of MiniMaxAI/MiniMax-H3 MiniMax H3 is an omni-modal generative system that produces video with native stereo audio, up to 15 seconds at 24 FPS with 32 kHz stereo audio. Both halves of the runtime are in this repo: the denoisers and the Qwen3-VL text encoder they need. H3 ships two denoisers, and which one you load decides what the model can be given: - fl2vapruned, the H3-Base first-and-last-frame variant. Text, plus zero, one or two frames. - ref2vapruned, the reference variant.…

Open weights other gguf

Model · Image text to video

Minimax-H3-nvfp4-INT4-INT8-Convrot

Jay

This repository is a community-compiled collection of quantized and pruned weights for MiniMax H3 (Hailuo 3.0), optimized for local inference environments like ComfyUI. By unifying various quantization formats (INT4, INT8, Mixed, and NVFP4) into a single structured repository, this hub makes it easier for users with consumer GPUs (16GB - 24GB VRAM) to experiment with MiniMax H3's powerful omni-modal text/image/audio-to-video generation capabilities. If you are new to local generation and aren't sure what to download, use this guide based on your graphics card. Perfect for RTX 4070 Ti Super, RTX 4080, etc. Perfect for RTX 3090, RTX 4090, etc. Exclusively for RTX 5090, PRO 6000, and other…

Open weights other diffusers

Model · Image text to video

MiniMax-H3-Realism-People-LoRA

Fal

A LoRA adapter for MiniMax H3 specialized in realistic people: faces that hold up in close-up, natural skin texture, believable expressions and gestures, film-style lighting and documentary camera movement. Same prompt, same seed — base model on the left, this adapter on the right: 19 pairs, same prompt, same seed, adapter on vs off — the only variable is the LoRA. The trigger word is present on both sides, so it is not doing the work. Each pair plays the base model first, then freezes and dims while the adapted version plays beside it. Close-up talking faces, arguments, several people speaking at once, weathered skin, children, ritual and travel scenes. Nothing cherry-picked from a larger…

Open weights other minimax-h3

Model · Image text to video

DaSiWa-MiniMax-H3-Hybrid

Bruno Po

Unmodified 4-step and 8-step DaSiWa MiniMax H3 Hybrid SafeTensors checkpoints mirrored for BRP Canvas downloads. The hybrid checkpoint supports text-to-video, reference-to-video, and first/last-frame-to-video through the same ComfyUI workflow. This is not an official MiniMax or DaSiWa distribution. Review the original model page and the included upstream MiniMax license before use. BRP Canvas defaults to shift video 9 and shift audio 4 for both distilled checkpoints. - 4-step: 56c52c7890c105308d28fe9c25c25fdb80e6cd6a54e2604d8af71732ba4ed74f - 8-step: e0441d26414f6e0c28f43d580e6cc56fad424da0fa4d261b698ca73188aa6332

Open weights other minimax-h3

Model · Image text to video

MiniMax-H3

MiniMax

Offical skills to improve prompt writing: skills on github Use MiniMax\-H3 directly via API\. Use MiniMax\-H3 directly via App\. MiniMax H3 is a general-purpose, omni-modal generative system. It supports unified understanding of multimodal contexts composed of text, images, video, and audio, and can generate video with native stereo audio at resolutions up to 2K and durations of up to 15 seconds. Thanks to its task-generalization-oriented system design, H3 already possesses broad multimodal context understanding and generation capabilities at the pre-training stage, enabling outstanding performance in following complex multimodal instructions. H3 supports the following input and output…

Open weights other 33.1B parameters minimax-h3