SAVRN
Search Contact SAVRN

Open-weight model · Image to video

Minimax-h3-Turbo

by Lightx2v lightx2v/Minimax-h3-Turbo

Please check our repository or the LightX2V MiniMax-H3 examples to reproduce the results. Please check the model specifications for more details.

Parameters
Context
Weights58.8 GB
Licenseapache-2.0
AccessOpen weights
Monthly Downloads1.5M

Model Card

By Lightx2v, published under apache-2.0, revision 3ec17a324ced.

Please check our repository or the LightX2V MiniMax-H3 examples to reproduce the results. Please check the model specifications for more details. Try the MiniMax-H3 Turbo LoRA directly in LightX2V Studio: The Studio currently uses the FL2V 8-step v1.0 768p LoRA, which provides improved video and audio generation quality with 8-step inference. Integrate MiniMax-H3 Turbo into your application through the LightX2V API

Read Lightx2v's full model card

Please check our repository or the LightX2V MiniMax-H3 examples to reproduce the results.

Please check the model specifications for more details.

Online App

Try the MiniMax-H3 Turbo LoRA directly in LightX2V Studio:

The Studio currently uses the FL2V 8-step v1.0 768p LoRA, which provides improved video and audio generation quality with 8-step inference.

The model version deployed in the Studio may be updated over time.

Studio Preview

Online API

Integrate MiniMax-H3 Turbo into your application through the LightX2V API:

Identity and Version

Repository
lightx2v/Minimax-h3-Turbo
Publisher
Lightx2v
Task
Image to video
Modality
Other
Library
diffusers
Parameters
Not stated by the source
Languages
en, zh
Revision
3ec17a324ced54151364f24f8b5fb6bf7e26414f
First published
2026-08-07
Last updated
2026-09-10

Files and Weights

18 files, 58.8 GB in total. The weights are 16 files totalling 58.8 GB in safetensors.

Weights16 files · 58.8 GB
Documentation1 file · 1.4 KB
Repository1 file · 1.5 KB
Every file
FileTypeSizeSHA-256
minimax_h3_fl2v_turbo_4step_v0.1.safetensorsWeights1.4 GB 5ff4a12c8b45
minimax_h3_fl2v_turbo_4step_v1.0_768p_bf16.safetensorsWeights1.4 GB 1bdabc2e9fce
minimax_h3_fl2v_turbo_4step_v1.0_768p_comfyui_bf16.safetensorsWeights2.0 GB c396a9a06f58
minimax_h3_fl2v_turbo_4step_v1.1_768p_bf16.safetensorsWeights1.4 GB b5e25a59292d
minimax_h3_fl2v_turbo_4step_v1.1_768p_comfyui_bf16.safetensorsWeights2.0 GB 449d80f301ac
minimax_h3_fl2v_turbo_4step_v1.1_768p_fp8.safetensorsWeights34.0 GB 87751b0ab719
minimax_h3_fl2v_turbo_4step_v1.2_768p_bf16.safetensorsWeights1.4 GB c3d4a2cf618e
minimax_h3_fl2v_turbo_4step_v1.2_768p_comfyui_bf16.safetensorsWeights2.0 GB c8168ebc17bb
minimax_h3_fl2v_turbo_8step_v1.0_768p_bf16.safetensorsWeights1.4 GB 9b0efe3613b4
minimax_h3_fl2v_turbo_8step_v1.0_768p_comfyui_bf16.safetensorsWeights2.0 GB 08cfe946033a
minimax_h3_fl2v_turbo_8step_v1.0_bf16.safetensorsWeights1.4 GB e16ac20824d6
minimax_h3_fl2v_turbo_8step_v1.0_comfyui_bf16.safetensorsWeights2.0 GB 2339acdf19bf
minimax_h3_ref2v_turbo_4step_v0.1_bf16.safetensorsWeights1.4 GB 9e642fc8749c
minimax_h3_ref2v_turbo_4step_v0.1_comfyui_bf16.safetensorsWeights2.0 GB 5b9ab5ade15d
minimax_h3_ref2v_turbo_8step_v1.0_768p_bf16.safetensorsWeights1.4 GB 9bac880b1a5d
minimax_h3_ref2v_turbo_8step_v1.0_768p_comfyui_bf16.safetensorsWeights2.0 GB 6a56f41ab422
README.mdDocumentation1.4 KB
.gitattributesRepository1.5 KB

License and Download

License
apache-2.0
Access
Open weights, no gate
Download size
58.8 GB
Download from Lightx2v

Released by Lightx2v through its official repository on Hugging Face. Read the license.

Built From

Memory Requirements

PrecisionWeights in memory
As published58.8 GB

Weights only, from the published parameter count; the key-value cache and runtime add to this.

Questions About Minimax-h3-Turbo

Can I use Minimax-h3-Turbo commercially?

Yes. Minimax-h3-Turbo is released under Apache License 2.0. The Apache License 2.0 is a permissive open-source license. It permits commercial use, modification and redistribution. It requires keeping the license and copyright notices and any NOTICE file, stating significant changes, and it includes an express patent grant from contributors.

Similar Models

Model · Image to video

Wan_2.2_ComfyUI_Repackaged

Comfy Org

Repackaged model files for ComfyUI. - https://huggingface.co/nvidia/ChronoEdit-14B-Diffusers - https://huggingface.co/Wan-AI/Wan2.2-Animate-14B - https://huggingface.co/alibaba-pai/Wan2.2-Fun-A14B-Control-Camera - https://huggingface.co/alibaba-pai/Wan2.2-Fun-A14B-Control - https://huggingface.co/alibaba-pai/Wan2.2-Fun-5B-Control - https://huggingface.co/alibaba-pai/Wan2.2-Fun-5B-InP - https://huggingface.co/alibaba-pai/Wan2.2-Fun-A14B-InP - https://huggingface.co/alibaba-pai/Wan2.2-VACE-Fun-A14B - https://huggingface.co/Wan-AI/Wan2.2-I2V-A14B - https://huggingface.co/Wan-AI/Wan2.2-S2V-14B - https://huggingface.co/Wan-AI/Wan2.2-T2V-A14B - https://huggingface.co/Wan-AI/Wan2.2-TI2V-5B…

Open weights apache-2.0 diffusion-single-file

Model · Image to video

LTX-2.5

LTX.io

LTX-2.5 is an open world model with open weights, built for local execution and fine-tuning. Its established use is generating synchronized, high-fidelity video and audio from text, image, and video inputs; applicability to emerging domains such as robotics and physical AI is developing. Full control and customization — self-host on your own infrastructure. No per-generation billing, no per-seat lock-in, no forced API dependency. Revenue is measured across the whole entity, including subsidiaries and affiliates under common control. The full, binding terms live in LICENSE. - Native multishot generation — generate connected scenes in a single pass: multiple shots that hold character…

Access requested at publisher other diffusion-single-file

Model · Image to video

LTX-2.3

LTX.io

This model card focuses on the LTX-2.3 model, which is a significant update to the LTX-2 model with improved audio and visual quality as well as enhanced prompt adherence. LTX-2 was presented in the paper LTX-2: Efficient Joint Audio-Visual Foundation Model. If you want to dive in right to the code - it is available here. LTX-2.3 is a DiT-based audio-video foundation model designed to generate synchronized video and audio within a single model. It brings together the core building blocks of modern video generation, with open weights and a focus on practical, local execution. LTX-2.3 is accessible right away via the API Playground. You can use the models - full, distilled, upscalers and any…

Open weights other diffusers

Model · Image to video

MiniMax-H3-GGUF

Jay

This repository (Abiray/MiniMax-H3-GGUF) provides GGUF quantized versions and necessary component files for the MiniMax H3 model. MiniMax H3 is a general-purpose, omni-modal generative system that supports unified understanding of multimodal contexts composed of text, images, video, and audio. It can generate video with native stereo audio at resolutions up to 2K and durations of up to 15 seconds. If you are looking for a smaller model with the same great quality that fits better on consumer-tier GPUs, please check out the MiniMax-H3-Pruned-GGUF repository. The pruned architecture is compressed down to 8.9 GB – 21.6 GB, bringing MiniMax H3 execution directly to consumer hardware. This…

Open weights other

Model · Image to video

LTX-2.3-fp8

LTX.io

This is the FP8 versions of the LTX-2.3 model. All information below is derived from the base model. This model card focuses on the LTX-2.3 model, which is a significant update to the LTX-2 model with improved audio and visual quality as well as enhanced prompt adherence. LTX-2 was presented in the paper LTX-2: Efficient Joint Audio-Visual Foundation Model. If you want to dive in right to the code - it is available here. LTX-2.3 is a DiT-based audio-video foundation model designed to generate synchronized video and audio within a single model. It brings together the core building blocks of modern video generation, with open weights and a focus on practical, local execution. LTX-2.3 is…

Open weights other diffusers

Model · Image to video

LTX-2.3-GGUF

Unsloth AI

This is a GGUF quantized version of LTX-2.3. unsloth/LTX-2.3-GGUF uses Unsloth Dynamic 2.0 methodology for SOTA performance. - Important layers are upcasted to higher precision. - Uses tooling from ComfyUI-GGUF by city96. There are two sets of GGUF's published. One for the dev model and one for the distilled. The distilled model is optimized for few step generation, think 4-8 steps. dev on the other hand needs more steps at least 20, but you get better outputs. The distilled variant is useful as a drafting model or a refining model. In fact the workflow published below, uses the distilled lora on top of the dev model to refine the intial output. Download the mp4 in the repo and open it with…

Open weights other ggml