SAVRN
Search Contact SAVRN

Open-weight model · Image text to video

MiniMax-H3-Realism-People-LoRA

by Fal fal/MiniMax-H3-Realism-People-LoRA

A LoRA adapter for MiniMax H3 specialized in realistic people: faces that hold up in close-up, natural skin texture, believable expressions and gestures, film-style lighting and documentary camera movement.

Parameters
Context
Weights131.2 MB
Licenseother
AccessOpen weights
Monthly Downloads61.3k

Model Card

A LoRA adapter for MiniMax H3 specialized in realistic people: faces that hold up in close-up, natural skin texture, believable expressions and gestures, film-style lighting and documentary camera movement. Same prompt, same seed — base model on the left, this adapter on the right: 19 pairs, same prompt, same seed, adapter on vs off — the only variable is the LoRA. The trigger word is present on both sides, so it is not doing the work. Each pair plays the base model first, then freezes and dims while the adapted version plays beside it. Close-up talking faces, arguments, several people speaking at once, weathered skin, children, ritual and travel scenes. Nothing cherry-picked from a larger…

Excerpt from the card by Fal, licensed other.

Identity and Version

Repository
fal/MiniMax-H3-Realism-People-LoRA
Publisher
Fal
Task
Image text to video
Modality
Other
Library
minimax-h3
Parameters
Not stated by the source
Languages
fal
Revision
039cc8579d7aa357a882d7f4111b25da4f72dccc
First published
2026-08-10
Last updated
2026-08-12

Files and Weights

4 files, 176.3 MB in total. The weights are 1 file totalling 131.2 MB in safetensors.

Weights1 file · 131.2 MB
Documentation1 file · 5.9 KB
Other1 file · 45.1 MB
Repository1 file · 1.6 KB
Every file
FileTypeSizeSHA-256
h3-realism-people-t2v-i2v-r2v.safetensorsWeights131.2 MB acc529601d2d
README.mdDocumentation5.9 KB
before-after-comparison.mp4Other45.1 MB e7879117d6b3
.gitattributesRepository1.6 KB

License and Download

License
other
Access
Open weights, no gate
Download size
131.2 MB
Download from Fal

Released by Fal through its official repository on Hugging Face. Read the license.

Built From

Memory Requirements

PrecisionWeights in memory
As published131.2 MB

Weights only, from the published parameter count; the key-value cache and runtime add to this.

Questions About MiniMax-H3-Realism-People-LoRA

What license is MiniMax-H3-Realism-People-LoRA released under?

other, as its publisher declares it. Read the license text before commercial use.

Similar Models

Model · Image text to video

MiniMax-H3-GGUF

Unsloth AI

Instructions further below. GGUF for MiniMax-H3, compatible on most platforms including stablediffusion.cpp and Unsloth. You can run MiniMax-H3 via Unsloth: https://github.com/unslothai/unsloth/ GGUF quantizations of MiniMaxAI/MiniMax-H3 MiniMax H3 is an omni-modal generative system that produces video with native stereo audio, up to 15 seconds at 24 FPS with 32 kHz stereo audio. Both halves of the runtime are in this repo: the denoisers and the Qwen3-VL text encoder they need. H3 ships two denoisers, and which one you load decides what the model can be given: - fl2vapruned, the H3-Base first-and-last-frame variant. Text, plus zero, one or two frames. - ref2vapruned, the reference variant.…

Open weights other gguf

Model · Image text to video

Minimax-H3-nvfp4-INT4-INT8-Convrot

Jay

This repository is a community-compiled collection of quantized and pruned weights for MiniMax H3 (Hailuo 3.0), optimized for local inference environments like ComfyUI. By unifying various quantization formats (INT4, INT8, Mixed, and NVFP4) into a single structured repository, this hub makes it easier for users with consumer GPUs (16GB - 24GB VRAM) to experiment with MiniMax H3's powerful omni-modal text/image/audio-to-video generation capabilities. If you are new to local generation and aren't sure what to download, use this guide based on your graphics card. Perfect for RTX 4070 Ti Super, RTX 4080, etc. Perfect for RTX 3090, RTX 4090, etc. Exclusively for RTX 5090, PRO 6000, and other…

Open weights other diffusers

Model · Image text to video

MiniMax-H3-Pruned-GGUF

Jay

This repository (Abiray/MiniMax-H3-Pruned-GGUF) provides pruned and quantized GGUF weights for the MiniMax H3 omni-modal generative model. MiniMax H3 is designed for unified multimodal context processing, capable of generating synchronized high-definition video and 32 kHz stereo audio from text, image, audio, and video inputs. 1. Download your desired.gguf variant from the table above. 2. Place the downloaded.gguf file into the ComfyUI/models/unet/ directory. 3. In your ComfyUI workflow, load the model using the UnetLoaderGGUF node. MiniMax H3 is released under the MiniMax H3 Community License Agreement. Please refer to the MiniMaxAI/MiniMax-H3 repository and the repository's LICENSE file…

Open weights other

Model · Image text to video

DaSiWa-MiniMax-H3-Hybrid

Bruno Po

Unmodified 4-step and 8-step DaSiWa MiniMax H3 Hybrid SafeTensors checkpoints mirrored for BRP Canvas downloads. The hybrid checkpoint supports text-to-video, reference-to-video, and first/last-frame-to-video through the same ComfyUI workflow. This is not an official MiniMax or DaSiWa distribution. Review the original model page and the included upstream MiniMax license before use. BRP Canvas defaults to shift video 9 and shift audio 4 for both distilled checkpoints. - 4-step: 56c52c7890c105308d28fe9c25c25fdb80e6cd6a54e2604d8af71732ba4ed74f - 8-step: e0441d26414f6e0c28f43d580e6cc56fad424da0fa4d261b698ca73188aa6332

Open weights other minimax-h3

Model · Image text to video

MiniMax-H3

MiniMax

Offical skills to improve prompt writing: skills on github Use MiniMax\-H3 directly via API\. Use MiniMax\-H3 directly via App\. MiniMax H3 is a general-purpose, omni-modal generative system. It supports unified understanding of multimodal contexts composed of text, images, video, and audio, and can generate video with native stereo audio at resolutions up to 2K and durations of up to 15 seconds. Thanks to its task-generalization-oriented system design, H3 already possesses broad multimodal context understanding and generation capabilities at the pre-training stage, enabling outstanding performance in following complex multimodal instructions. H3 supports the following input and output…

Open weights other 33.1B parameters minimax-h3