SAVRN
Search Contact SAVRN

Open-weight model · Image to video

MiniMax-H3-GGUF

by Leejet leejet/MiniMax-H3-GGUF

The license of the quantized files follows the license of the original model: These files are converted using https://github.com/leejet/stable-diffusion.cpp This model can be used with stable-diffusion.cpp.

Parameters
Context
Weights98.4 GB
License
AccessOpen weights
Monthly Downloads107.4k

Model Card

The license of the quantized files follows the license of the original model: These files are converted using https://github.com/leejet/stable-diffusion.cpp This model can be used with stable-diffusion.cpp. For setup instructions and usage details, please refer to: To use this model in ComfyUI, first install the following custom node: An example ComfyUI workflow is available here: src="https://huggingface.co/leejet/MiniMax-H3-GGUF/resolve/main/assets/example.mp4" controls muted

Excerpt from the card by Leejet.

Identity and Version

Repository
leejet/MiniMax-H3-GGUF
Publisher
Leejet
Task
Image to video
Modality
Other
Library
Not stated by the source
Parameters
Not stated by the source
Languages
en, zh
Revision
d9c4c6312b4728a68a15a35626d84775a6523783
First published
2026-08-03
Last updated
2026-08-30

Files and Weights

11 files, 98.4 GB in total. The weights are 7 files totalling 98.4 GB in gguf.

Weights7 files · 98.4 GB
Configuration1 file · 17.2 KB
Documentation1 file · 1.4 KB
Other1 file · 1.4 MB
Repository1 file · 2.1 KB
Every file
FileTypeSizeSHA-256
minimax_h3_fl2va-Q4_K_M.ggufWeights18.8 GB a25ee3184dd5
minimax_h3_fl2va_pruned-Q4_K_M.ggufWeights11.4 GB dd948e08ad0b
minimax_h3_ref2va-Q4_K_M.ggufWeights18.8 GB fdaecd3e73cf
minimax_h3_ref2va_pruned-Q2_K_M.ggufWeights6.7 GB 792651930a5d
minimax_h3_ref2va_pruned-Q4_K_M.ggufWeights11.4 GB 9ca1aef9f460
qwen3vl_32b_minimax_h3-Q2_K_M.ggufWeights13.1 GB a8ccadccd57e
qwen3vl_32b_minimax_h3-Q4_K_M.ggufWeights18.2 GB 11e6efe70a57
minimax_h3_fl2v_gguf.jsonConfiguration17.2 KB
README.mdDocumentation1.4 KB
assets/example.mp4Other1.4 MB d4cfa66036e3
.gitattributesRepository2.1 KB

License and Download

License
Not stated by the source
Access
Open weights, no gate
Download size
98.4 GB
Download from Leejet

Released by Leejet through its official repository on Hugging Face.

Built From

Memory Requirements

PrecisionWeights in memory
As published98.4 GB

Weights only, from the published parameter count; the key-value cache and runtime add to this.

Similar Models

Model · Image to video

Wan_2.2_ComfyUI_Repackaged

Comfy Org

Repackaged model files for ComfyUI. - https://huggingface.co/nvidia/ChronoEdit-14B-Diffusers - https://huggingface.co/Wan-AI/Wan2.2-Animate-14B - https://huggingface.co/alibaba-pai/Wan2.2-Fun-A14B-Control-Camera - https://huggingface.co/alibaba-pai/Wan2.2-Fun-A14B-Control - https://huggingface.co/alibaba-pai/Wan2.2-Fun-5B-Control - https://huggingface.co/alibaba-pai/Wan2.2-Fun-5B-InP - https://huggingface.co/alibaba-pai/Wan2.2-Fun-A14B-InP - https://huggingface.co/alibaba-pai/Wan2.2-VACE-Fun-A14B - https://huggingface.co/Wan-AI/Wan2.2-I2V-A14B - https://huggingface.co/Wan-AI/Wan2.2-S2V-14B - https://huggingface.co/Wan-AI/Wan2.2-T2V-A14B - https://huggingface.co/Wan-AI/Wan2.2-TI2V-5B…

Open weights apache-2.0 diffusion-single-file

Model · Image to video

LTX-2.5

LTX.io

LTX-2.5 is an open world model with open weights, built for local execution and fine-tuning. Its established use is generating synchronized, high-fidelity video and audio from text, image, and video inputs; applicability to emerging domains such as robotics and physical AI is developing. Full control and customization — self-host on your own infrastructure. No per-generation billing, no per-seat lock-in, no forced API dependency. Revenue is measured across the whole entity, including subsidiaries and affiliates under common control. The full, binding terms live in LICENSE. - Native multishot generation — generate connected scenes in a single pass: multiple shots that hold character…

Access requested at publisher other diffusion-single-file

Model · Image to video

Minimax-h3-Turbo

Lightx2v

Please check our repository or the LightX2V MiniMax-H3 examples to reproduce the results. Please check the model specifications for more details. Try the MiniMax-H3 Turbo LoRA directly in LightX2V Studio: The Studio currently uses the FL2V 8-step v1.0 768p LoRA, which provides improved video and audio generation quality with 8-step inference. Integrate MiniMax-H3 Turbo into your application through the LightX2V API

Open weights apache-2.0 diffusers

Model · Image to video

LTX-2.3

LTX.io

This model card focuses on the LTX-2.3 model, which is a significant update to the LTX-2 model with improved audio and visual quality as well as enhanced prompt adherence. LTX-2 was presented in the paper LTX-2: Efficient Joint Audio-Visual Foundation Model. If you want to dive in right to the code - it is available here. LTX-2.3 is a DiT-based audio-video foundation model designed to generate synchronized video and audio within a single model. It brings together the core building blocks of modern video generation, with open weights and a focus on practical, local execution. LTX-2.3 is accessible right away via the API Playground. You can use the models - full, distilled, upscalers and any…

Open weights other diffusers

Model · Image to video

MiniMax-H3-GGUF

Jay

This repository (Abiray/MiniMax-H3-GGUF) provides GGUF quantized versions and necessary component files for the MiniMax H3 model. MiniMax H3 is a general-purpose, omni-modal generative system that supports unified understanding of multimodal contexts composed of text, images, video, and audio. It can generate video with native stereo audio at resolutions up to 2K and durations of up to 15 seconds. If you are looking for a smaller model with the same great quality that fits better on consumer-tier GPUs, please check out the MiniMax-H3-Pruned-GGUF repository. The pruned architecture is compressed down to 8.9 GB – 21.6 GB, bringing MiniMax H3 execution directly to consumer hardware. This…

Open weights other

Model · Image to video

LTX-2.3-fp8

LTX.io

This is the FP8 versions of the LTX-2.3 model. All information below is derived from the base model. This model card focuses on the LTX-2.3 model, which is a significant update to the LTX-2 model with improved audio and visual quality as well as enhanced prompt adherence. LTX-2 was presented in the paper LTX-2: Efficient Joint Audio-Visual Foundation Model. If you want to dive in right to the code - it is available here. LTX-2.3 is a DiT-based audio-video foundation model designed to generate synchronized video and audio within a single model. It brings together the core building blocks of modern video generation, with open weights and a focus on practical, local execution. LTX-2.3 is…

Open weights other diffusers