SAVRN
Search Contact SAVRN

Open-weight model · Image text to video

MiniMax-H3

by MiniMax MiniMaxAI/MiniMax-H3

Offical skills to improve prompt writing: skills on github Use MiniMax\-H3 directly via API\. Use MiniMax\-H3 directly via App\. MiniMax H3 is a general-purpose, omni-modal generative system.

Parameters33.1B
Context
Weights498.3 GB
Licenseother
AccessOpen weights
Monthly Downloads4.4M

Runs On

What it takes to serve MiniMax-H3 (33.1B parameters): the memory its weights need at each precision, and the cheapest way to rent enough data-center GPUs to hold them.

PrecisionWeightsMemory neededCheapest setupPer hourAlso fits
16-bit 66.2 GB 79.5 GB 1x MI300X (192 GB)
Vultr
$1.85 1x H100 $1.99 · 1x MI325X $2.00
8-bit 33.1 GB 39.7 GB 1x MI300X (192 GB)
Vultr
$1.85 1x H100 $1.99 · 1x MI325X $2.00
4-bit 16.6 GB 19.9 GB 1x MI300X (192 GB)
Vultr
$1.85 1x H100 $1.99 · 1x MI325X $2.00

Memory is the weights at that precision plus 20% for the runtime and a short context; a long context needs more. Prices are the lowest on-demand hourly rates in the SAVRN Index, read Sep 18, 2026.

SAVRN's Notes on MiniMax-H3

Fifteen seconds of video at up to 2K, with stereo audio, from text, image, video or audio prompts: that is MiniMax-H3's job. At 33.1B parameters the 16-bit working set is 79.5 GB, which fits on one MI300X with 192 GB at $1.85 an hour. At 8-bit (39.7 GB) or 4-bit (19.9 GB) it lands on the same card, so quantizing buys headroom, not a cheaper box. Storage is the bigger commitment: 280 files and about 498 GB, well above the 66.2 GB of weights.

The license is listed as other with no summary, so read the publisher's terms before a commercial deployment. No context length is published, no benchmarks are recorded, and no SAVRN Index host prices exist yet, so there is no per-token rental price to weigh against your own $1.85 an hour. Open access since July 28, 2026, last updated August 13, 2026; test the August files.

Model Card

Offical skills to improve prompt writing: skills on github Use MiniMax\-H3 directly via API\. Use MiniMax\-H3 directly via App\. MiniMax H3 is a general-purpose, omni-modal generative system. It supports unified understanding of multimodal contexts composed of text, images, video, and audio, and can generate video with native stereo audio at resolutions up to 2K and durations of up to 15 seconds. Thanks to its task-generalization-oriented system design, H3 already possesses broad multimodal context understanding and generation capabilities at the pre-training stage, enabling outstanding performance in following complex multimodal instructions. H3 supports the following input and output…

Excerpt from the card by MiniMax, licensed other.

Identity and Version

Repository
MiniMaxAI/MiniMax-H3
Publisher
MiniMax
Task
Image text to video
Modality
Other
Library
minimax-h3
Parameters
33.1B parameters
Languages
Not stated by the source
Revision
42ed227ee7df40d41602854ae760620d6eb651fe
First published
2026-07-28
Last updated
2026-08-13

Files and Weights

280 files, 498.5 GB in total. The weights are 104 files totalling 498.3 GB in safetensors.

Weights104 files · 498.3 GB
Configuration99 files · 986.4 KB
Tokenizer36 files · 103.4 MB
Documentation5 files · 99.3 KB
Other35 files · 40.6 MB
Repository1 file · 1.6 KB
Every file
FileTypeSizeSHA-256
FL2VA/audio_vae/model.safetensorsWeights605.4 MB 37dddc2f3e6d
FL2VA/text_encoder/model-00001-of-00014.safetensorsWeights4.9 GB 6b9dfbc930e5
FL2VA/text_encoder/model-00002-of-00014.safetensorsWeights4.9 GB d8bb44b4ff30
FL2VA/text_encoder/model-00003-of-00014.safetensorsWeights4.9 GB 54f22e8b3168
FL2VA/text_encoder/model-00004-of-00014.safetensorsWeights4.9 GB ad09c74d3c13
FL2VA/text_encoder/model-00005-of-00014.safetensorsWeights4.9 GB fc993c8a0e2a
FL2VA/text_encoder/model-00006-of-00014.safetensorsWeights4.9 GB 82f05620d1f7
FL2VA/text_encoder/model-00007-of-00014.safetensorsWeights4.9 GB fb91da8cb01f
FL2VA/text_encoder/model-00008-of-00014.safetensorsWeights4.9 GB 431ca56535c8
FL2VA/text_encoder/model-00009-of-00014.safetensorsWeights4.9 GB 3825e3f4302f
FL2VA/text_encoder/model-00010-of-00014.safetensorsWeights4.9 GB aded5a4d1d5e
FL2VA/text_encoder/model-00011-of-00014.safetensorsWeights4.9 GB 3820ffe8d8d6
FL2VA/text_encoder/model-00012-of-00014.safetensorsWeights4.9 GB 05ad2d08ce71
FL2VA/text_encoder/model-00013-of-00014.safetensorsWeights4.9 GB b64f22898712
FL2VA/text_encoder/model-00014-of-00014.safetensorsWeights3.3 GB e45b6c9998c7
FL2VA/transformer/model-00001-of-00013.safetensorsWeights5.2 GB 0b3386565e47
FL2VA/transformer/model-00002-of-00013.safetensorsWeights5.2 GB 9c98fd4579c9
FL2VA/transformer/model-00003-of-00013.safetensorsWeights5.2 GB fa484c940d41
FL2VA/transformer/model-00004-of-00013.safetensorsWeights5.2 GB 4a851df37cce
FL2VA/transformer/model-00005-of-00013.safetensorsWeights5.2 GB 3fe6ff94dc9d
FL2VA/transformer/model-00006-of-00013.safetensorsWeights5.2 GB 79b47e1b9ff0
FL2VA/transformer/model-00007-of-00013.safetensorsWeights5.2 GB 6ac4d6ce6397
FL2VA/transformer/model-00008-of-00013.safetensorsWeights5.2 GB 03531812243f
FL2VA/transformer/model-00009-of-00013.safetensorsWeights5.2 GB 7e82a80b0d0d
FL2VA/transformer/model-00010-of-00013.safetensorsWeights5.2 GB ef0a1f6b6514
FL2VA/transformer/model-00011-of-00013.safetensorsWeights5.2 GB bf885b8f2f07
FL2VA/transformer/model-00012-of-00013.safetensorsWeights5.2 GB dc717572d014
FL2VA/transformer/model-00013-of-00013.safetensorsWeights4.2 GB 8bfd852d5817
FL2VA/video_vae/source/model.safetensorsWeights10.4 GB 5f0c2e161d89
Ref2VA/audio_vae/model.safetensorsWeights605.4 MB 37dddc2f3e6d
Ref2VA/text_encoder/model-00001-of-00014.safetensorsWeights4.9 GB 6b9dfbc930e5
Ref2VA/text_encoder/model-00002-of-00014.safetensorsWeights4.9 GB d8bb44b4ff30
Ref2VA/text_encoder/model-00003-of-00014.safetensorsWeights4.9 GB 54f22e8b3168
Ref2VA/text_encoder/model-00004-of-00014.safetensorsWeights4.9 GB ad09c74d3c13
Ref2VA/text_encoder/model-00005-of-00014.safetensorsWeights4.9 GB fc993c8a0e2a
Ref2VA/text_encoder/model-00006-of-00014.safetensorsWeights4.9 GB 82f05620d1f7
Ref2VA/text_encoder/model-00007-of-00014.safetensorsWeights4.9 GB fb91da8cb01f
Ref2VA/text_encoder/model-00008-of-00014.safetensorsWeights4.9 GB 431ca56535c8
Ref2VA/text_encoder/model-00009-of-00014.safetensorsWeights4.9 GB 3825e3f4302f
Ref2VA/text_encoder/model-00010-of-00014.safetensorsWeights4.9 GB aded5a4d1d5e
Ref2VA/text_encoder/model-00011-of-00014.safetensorsWeights4.9 GB 3820ffe8d8d6
Ref2VA/text_encoder/model-00012-of-00014.safetensorsWeights4.9 GB 05ad2d08ce71
Ref2VA/text_encoder/model-00013-of-00014.safetensorsWeights4.9 GB b64f22898712
Ref2VA/text_encoder/model-00014-of-00014.safetensorsWeights3.3 GB e45b6c9998c7
Ref2VA/transformer/model-00001-of-00013.safetensorsWeights5.2 GB 902d6ae787f6
Ref2VA/transformer/model-00002-of-00013.safetensorsWeights5.2 GB 97b907e4d9b3
Ref2VA/transformer/model-00003-of-00013.safetensorsWeights5.2 GB fbdf425e3516
Ref2VA/transformer/model-00004-of-00013.safetensorsWeights5.2 GB 3831e9ff2640
Ref2VA/transformer/model-00005-of-00013.safetensorsWeights5.2 GB ec7c6bdd2698
Ref2VA/transformer/model-00006-of-00013.safetensorsWeights5.2 GB 536c4fe01ffd
Ref2VA/transformer/model-00007-of-00013.safetensorsWeights5.2 GB c7b80189722b
Ref2VA/transformer/model-00008-of-00013.safetensorsWeights5.2 GB 1a7da8bb7477
Ref2VA/transformer/model-00009-of-00013.safetensorsWeights5.2 GB 33f44f03128b
Ref2VA/transformer/model-00010-of-00013.safetensorsWeights5.2 GB adb8df74dcff
Ref2VA/transformer/model-00011-of-00013.safetensorsWeights5.2 GB 7175cdde6953
Ref2VA/transformer/model-00012-of-00013.safetensorsWeights5.2 GB 1c5c0bf33ab6
Ref2VA/transformer/model-00013-of-00013.safetensorsWeights4.2 GB 572e060b9b72
Ref2VA/video_vae/source/model.safetensorsWeights10.4 GB 5f0c2e161d89
audio_vae/diffusion_pytorch_model.safetensorsWeights605.4 MB 52c59e67ba8d
text_encoder/model-00001-of-00014.safetensorsWeights4.9 GB 6b9dfbc930e5
text_encoder/model-00002-of-00014.safetensorsWeights4.9 GB d8bb44b4ff30
text_encoder/model-00003-of-00014.safetensorsWeights4.9 GB 54f22e8b3168
text_encoder/model-00004-of-00014.safetensorsWeights4.9 GB ad09c74d3c13
text_encoder/model-00005-of-00014.safetensorsWeights4.9 GB fc993c8a0e2a
text_encoder/model-00006-of-00014.safetensorsWeights4.9 GB 82f05620d1f7
text_encoder/model-00007-of-00014.safetensorsWeights4.9 GB fb91da8cb01f
text_encoder/model-00008-of-00014.safetensorsWeights4.9 GB 431ca56535c8
text_encoder/model-00009-of-00014.safetensorsWeights4.9 GB 3825e3f4302f
text_encoder/model-00010-of-00014.safetensorsWeights4.9 GB aded5a4d1d5e
text_encoder/model-00011-of-00014.safetensorsWeights4.9 GB 3820ffe8d8d6
text_encoder/model-00012-of-00014.safetensorsWeights4.9 GB 05ad2d08ce71
text_encoder/model-00013-of-00014.safetensorsWeights4.9 GB b64f22898712
text_encoder/model-00014-of-00014.safetensorsWeights3.3 GB e45b6c9998c7
transformer/diffusion_pytorch_model-00001-of-00014.safetensorsWeights4.8 GB 2d847200c45c
transformer/diffusion_pytorch_model-00002-of-00014.safetensorsWeights4.7 GB 2c4d362eddd2
transformer/diffusion_pytorch_model-00003-of-00014.safetensorsWeights4.9 GB 949c5aafbbfa
transformer/diffusion_pytorch_model-00004-of-00014.safetensorsWeights4.6 GB eef761679010
transformer/diffusion_pytorch_model-00005-of-00014.safetensorsWeights4.7 GB 43fdf42d638e
transformer/diffusion_pytorch_model-00006-of-00014.safetensorsWeights4.9 GB 6442510b34d1
transformer/diffusion_pytorch_model-00007-of-00014.safetensorsWeights4.6 GB 29f48f535c91
transformer/diffusion_pytorch_model-00008-of-00014.safetensorsWeights4.7 GB c711b096c764
transformer/diffusion_pytorch_model-00009-of-00014.safetensorsWeights4.9 GB 44428defe397
transformer/diffusion_pytorch_model-00010-of-00014.safetensorsWeights4.6 GB 3d44939c374c
transformer/diffusion_pytorch_model-00011-of-00014.safetensorsWeights4.7 GB 224d24430b58
transformer/diffusion_pytorch_model-00012-of-00014.safetensorsWeights4.9 GB 48fa2bd8fe13
transformer/diffusion_pytorch_model-00013-of-00014.safetensorsWeights4.6 GB be5b4b1809f9
transformer/diffusion_pytorch_model-00014-of-00014.safetensorsWeights4.6 GB 8fbd5e6c1fb1
transformer_ref/diffusion_pytorch_model-00001-of-00014.safetensorsWeights4.8 GB 7a3fcad885f5
transformer_ref/diffusion_pytorch_model-00002-of-00014.safetensorsWeights4.7 GB 1638ae1dc8ae
transformer_ref/diffusion_pytorch_model-00003-of-00014.safetensorsWeights4.9 GB 1ef3c4954ffe
transformer_ref/diffusion_pytorch_model-00004-of-00014.safetensorsWeights4.6 GB 12d92f2975cf
transformer_ref/diffusion_pytorch_model-00005-of-00014.safetensorsWeights4.7 GB 304d41ce03d5
transformer_ref/diffusion_pytorch_model-00006-of-00014.safetensorsWeights4.9 GB 12a134b7c76d
transformer_ref/diffusion_pytorch_model-00007-of-00014.safetensorsWeights4.6 GB b96395261359
transformer_ref/diffusion_pytorch_model-00008-of-00014.safetensorsWeights4.7 GB 1897a6bf3b4f
transformer_ref/diffusion_pytorch_model-00009-of-00014.safetensorsWeights4.9 GB edfb38235adc
transformer_ref/diffusion_pytorch_model-00010-of-00014.safetensorsWeights4.6 GB f8710775cf34
transformer_ref/diffusion_pytorch_model-00011-of-00014.safetensorsWeights4.7 GB 9e18acc09f84
transformer_ref/diffusion_pytorch_model-00012-of-00014.safetensorsWeights4.9 GB ea2e18228f8b
transformer_ref/diffusion_pytorch_model-00013-of-00014.safetensorsWeights4.6 GB 1e12083b1875
transformer_ref/diffusion_pytorch_model-00014-of-00014.safetensorsWeights4.6 GB b340f44b5690
vae/diffusion_pytorch_model-00001-of-00003.safetensorsWeights5.1 GB 72f4c6be84ac
vae/diffusion_pytorch_model-00002-of-00003.safetensorsWeights5.0 GB 2e05e8bc23fa
vae/diffusion_pytorch_model-00003-of-00003.safetensorsWeights398.5 MB c05d6ac4b1a3
FL2VA/audio_vae/config.jsonConfiguration2.0 KB
FL2VA/audio_vae/config.yamlConfiguration91 B
FL2VA/audio_vae/dac_activations.pyConfiguration2.2 KB
FL2VA/audio_vae/dac_alias_free_act.pyConfiguration835 B
FL2VA/audio_vae/dac_alias_free_filter.pyConfiguration3.3 KB
FL2VA/audio_vae/dac_alias_free_resample.pyConfiguration1.7 KB
FL2VA/audio_vae/dac_attn_proj.pyConfiguration3.3 KB
FL2VA/audio_vae/dac_audio_vae.pyConfiguration7.3 KB
FL2VA/audio_vae/dac_bigvgan.pyConfiguration7.1 KB
FL2VA/audio_vae/dac_utils.pyConfiguration362 B
FL2VA/audio_vae/metadata.jsonConfiguration440 B
FL2VA/audio_vae/minimax_h3_audio_vae.pyConfiguration3.5 KB
FL2VA/model_index.jsonConfiguration719 B
FL2VA/processor/chat_template.jsonConfiguration5.5 KB
FL2VA/processor/preprocessor_config.jsonConfiguration390 B
FL2VA/processor/video_preprocessor_config.jsonConfiguration385 B
FL2VA/text_encoder/chat_template.jsonConfiguration5.5 KB
FL2VA/text_encoder/config.jsonConfiguration1.5 KB
FL2VA/text_encoder/model.safetensors.index.jsonConfiguration97.8 KB
FL2VA/text_encoder/preprocessor_config.jsonConfiguration390 B
FL2VA/text_encoder/video_preprocessor_config.jsonConfiguration385 B
FL2VA/transformer/config.jsonConfiguration604 B
FL2VA/transformer/model.safetensors.index.jsonConfiguration38.3 KB
FL2VA/video_vae/attention.pyConfiguration5.8 KB
FL2VA/video_vae/base_module.pyConfiguration9.5 KB
FL2VA/video_vae/config.jsonConfiguration1.8 KB
FL2VA/video_vae/conv.pyConfiguration4.5 KB
FL2VA/video_vae/flash.pyConfiguration5.8 KB
FL2VA/video_vae/func.pyConfiguration5.8 KB
FL2VA/video_vae/klvae.pyConfiguration48.6 KB
FL2VA/video_vae/minimax_h3_video_vae.pyConfiguration5.1 KB
FL2VA/video_vae/norm.pyConfiguration10.7 KB
FL2VA/video_vae/normalize.pyConfiguration1.2 KB
FL2VA/video_vae/parallel.pyConfiguration13.0 KB
FL2VA/video_vae/source/config.jsonConfiguration1.2 KB
FL2VA/video_vae/utils.pyConfiguration764 B
FL2VA/video_vae/vae_cnn.pyConfiguration8.8 KB
FL2VA/video_vae/vae_module.pyConfiguration1.9 KB
FL2VA/video_vae/vae_processor.pyConfiguration8.4 KB
FL2VA/video_vae/vae_vit.pyConfiguration13.5 KB
Ref2VA/audio_vae/config.jsonConfiguration2.0 KB
Ref2VA/audio_vae/config.yamlConfiguration91 B
Ref2VA/audio_vae/dac_activations.pyConfiguration2.2 KB
Ref2VA/audio_vae/dac_alias_free_act.pyConfiguration835 B
Ref2VA/audio_vae/dac_alias_free_filter.pyConfiguration3.3 KB
Ref2VA/audio_vae/dac_alias_free_resample.pyConfiguration1.7 KB
Ref2VA/audio_vae/dac_attn_proj.pyConfiguration3.3 KB
Ref2VA/audio_vae/dac_audio_vae.pyConfiguration7.3 KB
Ref2VA/audio_vae/dac_bigvgan.pyConfiguration7.1 KB
Ref2VA/audio_vae/dac_utils.pyConfiguration362 B
Ref2VA/audio_vae/metadata.jsonConfiguration440 B
Ref2VA/audio_vae/minimax_h3_audio_vae.pyConfiguration3.5 KB
Ref2VA/model_index.jsonConfiguration707 B
Ref2VA/processor/chat_template.jsonConfiguration5.5 KB
Ref2VA/processor/preprocessor_config.jsonConfiguration390 B
Ref2VA/processor/video_preprocessor_config.jsonConfiguration385 B
Ref2VA/text_encoder/chat_template.jsonConfiguration5.5 KB
Ref2VA/text_encoder/config.jsonConfiguration1.5 KB
Ref2VA/text_encoder/model.safetensors.index.jsonConfiguration97.8 KB
Ref2VA/text_encoder/preprocessor_config.jsonConfiguration390 B
Ref2VA/text_encoder/video_preprocessor_config.jsonConfiguration385 B
Ref2VA/transformer/config.jsonConfiguration604 B
Ref2VA/transformer/model.safetensors.index.jsonConfiguration38.3 KB
Ref2VA/video_vae/attention.pyConfiguration5.8 KB
Ref2VA/video_vae/base_module.pyConfiguration9.5 KB
Ref2VA/video_vae/config.jsonConfiguration1.8 KB
Ref2VA/video_vae/conv.pyConfiguration4.5 KB
Ref2VA/video_vae/flash.pyConfiguration5.8 KB
Ref2VA/video_vae/func.pyConfiguration5.8 KB
Ref2VA/video_vae/klvae.pyConfiguration48.6 KB
Ref2VA/video_vae/minimax_h3_video_vae.pyConfiguration5.1 KB
Ref2VA/video_vae/norm.pyConfiguration10.7 KB
Ref2VA/video_vae/normalize.pyConfiguration1.2 KB
Ref2VA/video_vae/parallel.pyConfiguration13.0 KB
Ref2VA/video_vae/source/config.jsonConfiguration1.2 KB
Ref2VA/video_vae/utils.pyConfiguration764 B
Ref2VA/video_vae/vae_cnn.pyConfiguration8.8 KB
Ref2VA/video_vae/vae_module.pyConfiguration1.9 KB
Ref2VA/video_vae/vae_processor.pyConfiguration8.4 KB
Ref2VA/video_vae/vae_vit.pyConfiguration13.5 KB
audio_scheduler/scheduler_config.jsonConfiguration96 B
audio_vae/config.jsonConfiguration2.3 KB
model_index.jsonConfiguration2.9 KB
modular_model_index.jsonConfiguration2.9 KB
processor/chat_template.jsonConfiguration5.5 KB
processor/preprocessor_config.jsonConfiguration390 B
processor/video_preprocessor_config.jsonConfiguration385 B
scheduler/scheduler_config.jsonConfiguration97 B
text_encoder/chat_template.jsonConfiguration5.5 KB
text_encoder/config.jsonConfiguration1.5 KB
text_encoder/model.safetensors.index.jsonConfiguration97.8 KB
text_encoder/preprocessor_config.jsonConfiguration390 B
text_encoder/video_preprocessor_config.jsonConfiguration385 B
transformer/config.jsonConfiguration546 B
transformer/diffusion_pytorch_model.safetensors.index.jsonConfiguration64.5 KB
transformer_ref/config.jsonConfiguration546 B
transformer_ref/diffusion_pytorch_model.safetensors.index.jsonConfiguration64.5 KB
vae/config.jsonConfiguration2.0 KB
vae/diffusion_pytorch_model.safetensors.index.jsonConfiguration74.2 KB
LICENSEDocumentation17.6 KB
README.mdDocumentation38.4 KB
docs/QA-about-License.mdDocumentation3.9 KB
docs/VIDEO_PROMPT_WRITING_GUIDE_base_en.mdDocumentation15.8 KB
docs/VIDEO_PROMPT_WRITING_GUIDE_ref_en.mdDocumentation23.6 KB
assets/fl2va.mp4Other1.1 MB 5a5d6e0fcdda
assets/full-arch.pngOther420.6 KB 5bd8e4077a11
assets/h3_direct_2k.mp4Other8.9 MB cc928dc5f406
assets/h3_direct_768p.mp4Other2.4 MB 642e444e8cfa
assets/i2va.mp4Other1.1 MB 5a5d6e0fcdda
assets/i2va_2k.mp4Other3.3 MB dcfec002e432
assets/i2va_direct_2k.mp4Other4.6 MB c3c1eef07982
assets/i2va_direct_768p.mp4Other1.4 MB fc73d7633bed
assets/minimax-h3.pngOther705.1 KB ebf43dfeaca4
assets/overview.pngOther212.9 KB b32290cdabb1
assets/r2va.mp4Other850.7 KB 5c9d33dfc517
assets/r2va_2k.mp4Other2.8 MB c3c96ce2bcd2
assets/r2va_direct_2k.mp4Other3.1 MB 26f07ac79f87
assets/r2va_direct_768p.mp4Other1.1 MB 9b940b139d16
assets/ref2va.mp4Other850.7 KB 5c9d33dfc517
assets/t2va.mp4Other1.6 MB d66903241362
assets/t2va_2k.mp4Other6.0 MB 6624095cbda8
scripts/readme/full-2k-i2va-h3-base.shOther1.2 KB
scripts/readme/full-2k-i2va-h3-context-ir.shOther1.3 KB
scripts/readme/full-2k-i2va-h3-regenerate-2k.shOther1.6 KB
scripts/readme/full-2k-i2va-reference-2k-result-by-directly-calling-open-platform-api.shOther1.3 KB
scripts/readme/full-2k-i2va-reference-768p-result-by-directly-calling-open-platform-api.shOther1.4 KB
scripts/readme/full-2k-ref2va-h3-api-2k-in-open-platform-for-reference.shOther1.7 KB
scripts/readme/full-2k-ref2va-h3-base.shOther1.5 KB
scripts/readme/full-2k-ref2va-h3-context-ir.shOther1.7 KB
scripts/readme/full-2k-ref2va-reference-2k-result-by-directly-calling-open-platform-api.shOther1.9 KB
scripts/readme/full-2k-ref2va-reference-768p-result-by-directly-calling-open-platform-api.shOther1.8 KB
scripts/readme/full-2k-t2va-h3-base.shOther906 B
scripts/readme/full-2k-t2va-h3-context-ir.shOther1.1 KB
scripts/readme/full-2k-t2va-h3-regenerate-2k.shOther1.3 KB
scripts/readme/full-2k-t2va-reference-2k-result-by-directly-calling-open-platform-api.shOther1.2 KB
scripts/readme/full-2k-t2va-reference-768p-result-by-directly-calling-open-platform-api.shOther1.2 KB
scripts/readme/reproducible-768p-fl2va-request.shOther5.7 KB
scripts/readme/reproducible-768p-ref2va-request.shOther5.0 KB
scripts/readme/reproducible-768p-t2va-request.shOther3.4 KB
.gitattributesRepository1.6 KB
FL2VA/processor/merges.txtTokenizer1.7 MB
FL2VA/processor/tokenizer.jsonTokenizer7.0 MB
FL2VA/processor/tokenizer_config.jsonTokenizer11.0 KB
FL2VA/processor/vocab.jsonTokenizer2.8 MB
FL2VA/text_encoder/merges.txtTokenizer1.7 MB
FL2VA/text_encoder/tokenizer.jsonTokenizer7.0 MB
FL2VA/text_encoder/tokenizer_config.jsonTokenizer11.0 KB
FL2VA/text_encoder/vocab.jsonTokenizer2.8 MB
FL2VA/tokenizer/merges.txtTokenizer1.7 MB
FL2VA/tokenizer/tokenizer.jsonTokenizer7.0 MB
FL2VA/tokenizer/tokenizer_config.jsonTokenizer11.0 KB
FL2VA/tokenizer/vocab.jsonTokenizer2.8 MB
Ref2VA/processor/merges.txtTokenizer1.7 MB
Ref2VA/processor/tokenizer.jsonTokenizer7.0 MB
Ref2VA/processor/tokenizer_config.jsonTokenizer11.0 KB
Ref2VA/processor/vocab.jsonTokenizer2.8 MB
Ref2VA/text_encoder/merges.txtTokenizer1.7 MB
Ref2VA/text_encoder/tokenizer.jsonTokenizer7.0 MB
Ref2VA/text_encoder/tokenizer_config.jsonTokenizer11.0 KB
Ref2VA/text_encoder/vocab.jsonTokenizer2.8 MB
Ref2VA/tokenizer/merges.txtTokenizer1.7 MB
Ref2VA/tokenizer/tokenizer.jsonTokenizer7.0 MB
Ref2VA/tokenizer/tokenizer_config.jsonTokenizer11.0 KB
Ref2VA/tokenizer/vocab.jsonTokenizer2.8 MB
processor/merges.txtTokenizer1.7 MB
processor/tokenizer.jsonTokenizer7.0 MB
processor/tokenizer_config.jsonTokenizer11.0 KB
processor/vocab.jsonTokenizer2.8 MB
text_encoder/merges.txtTokenizer1.7 MB
text_encoder/tokenizer.jsonTokenizer7.0 MB
text_encoder/tokenizer_config.jsonTokenizer11.0 KB
text_encoder/vocab.jsonTokenizer2.8 MB
tokenizer/merges.txtTokenizer1.7 MB
tokenizer/tokenizer.jsonTokenizer7.0 MB
tokenizer/tokenizer_config.jsonTokenizer11.0 KB
tokenizer/vocab.jsonTokenizer2.8 MB

License and Download

License
other
Access
Open weights, no gate
Download size
498.3 GB
Download from MiniMax

Released by MiniMax through its official repository on Hugging Face.

Memory Requirements

PrecisionWeights in memory
As published498.3 GB
16-bit66.2 GB
8-bit33.1 GB
4-bit16.6 GB

Weights only, from the published parameter count; the key-value cache and runtime add to this.

Built on This Model

Questions About MiniMax-H3

How much GPU memory does MiniMax-H3 need?

About 79.5 GB at 16-bit and 19.9 GB at 4-bit: the weights (33.1B parameters) plus a working margin. A long context needs more.

What is the cheapest GPU to run MiniMax-H3 on?

At 16-bit, 1x MI300X from $1.85 an hour; at 4-bit, 1x MI300X from $1.85 an hour, at the lowest on-demand prices the SAVRN Index lists.

What license is MiniMax-H3 released under?

other, as its publisher declares it. Read the license text before commercial use.

Similar Models

Model · Image text to video

MiniMax-H3-GGUF

Unsloth AI

Instructions further below. GGUF for MiniMax-H3, compatible on most platforms including stablediffusion.cpp and Unsloth. You can run MiniMax-H3 via Unsloth: https://github.com/unslothai/unsloth/ GGUF quantizations of MiniMaxAI/MiniMax-H3 MiniMax H3 is an omni-modal generative system that produces video with native stereo audio, up to 15 seconds at 24 FPS with 32 kHz stereo audio. Both halves of the runtime are in this repo: the denoisers and the Qwen3-VL text encoder they need. H3 ships two denoisers, and which one you load decides what the model can be given: - fl2vapruned, the H3-Base first-and-last-frame variant. Text, plus zero, one or two frames. - ref2vapruned, the reference variant.…

Open weights other gguf

Model · Image text to video

Minimax-H3-nvfp4-INT4-INT8-Convrot

Jay

This repository is a community-compiled collection of quantized and pruned weights for MiniMax H3 (Hailuo 3.0), optimized for local inference environments like ComfyUI. By unifying various quantization formats (INT4, INT8, Mixed, and NVFP4) into a single structured repository, this hub makes it easier for users with consumer GPUs (16GB - 24GB VRAM) to experiment with MiniMax H3's powerful omni-modal text/image/audio-to-video generation capabilities. If you are new to local generation and aren't sure what to download, use this guide based on your graphics card. Perfect for RTX 4070 Ti Super, RTX 4080, etc. Perfect for RTX 3090, RTX 4090, etc. Exclusively for RTX 5090, PRO 6000, and other…

Open weights other diffusers

Model · Image text to video

MiniMax-H3-Pruned-GGUF

Jay

This repository (Abiray/MiniMax-H3-Pruned-GGUF) provides pruned and quantized GGUF weights for the MiniMax H3 omni-modal generative model. MiniMax H3 is designed for unified multimodal context processing, capable of generating synchronized high-definition video and 32 kHz stereo audio from text, image, audio, and video inputs. 1. Download your desired.gguf variant from the table above. 2. Place the downloaded.gguf file into the ComfyUI/models/unet/ directory. 3. In your ComfyUI workflow, load the model using the UnetLoaderGGUF node. MiniMax H3 is released under the MiniMax H3 Community License Agreement. Please refer to the MiniMaxAI/MiniMax-H3 repository and the repository's LICENSE file…

Open weights other

Model · Image text to video

MiniMax-H3-Realism-People-LoRA

Fal

A LoRA adapter for MiniMax H3 specialized in realistic people: faces that hold up in close-up, natural skin texture, believable expressions and gestures, film-style lighting and documentary camera movement. Same prompt, same seed — base model on the left, this adapter on the right: 19 pairs, same prompt, same seed, adapter on vs off — the only variable is the LoRA. The trigger word is present on both sides, so it is not doing the work. Each pair plays the base model first, then freezes and dims while the adapted version plays beside it. Close-up talking faces, arguments, several people speaking at once, weathered skin, children, ritual and travel scenes. Nothing cherry-picked from a larger…

Open weights other minimax-h3

Model · Image text to video

DaSiWa-MiniMax-H3-Hybrid

Bruno Po

Unmodified 4-step and 8-step DaSiWa MiniMax H3 Hybrid SafeTensors checkpoints mirrored for BRP Canvas downloads. The hybrid checkpoint supports text-to-video, reference-to-video, and first/last-frame-to-video through the same ComfyUI workflow. This is not an official MiniMax or DaSiWa distribution. Review the original model page and the included upstream MiniMax license before use. BRP Canvas defaults to shift video 9 and shift audio 4 for both distilled checkpoints. - 4-step: 56c52c7890c105308d28fe9c25c25fdb80e6cd6a54e2604d8af71732ba4ed74f - 8-step: e0441d26414f6e0c28f43d580e6cc56fad424da0fa4d261b698ca73188aa6332

Open weights other minimax-h3