Quantized diffusion transformers for MiniMax H3, a 33B omni-modal video+audio generator, built from the ComfyUI repack at Comfy-Org/MiniMax-H3. Output is 768p / 24 fps / 4–15 s with synchronized 32 kHz stereo audio. (2K output requires the separate H3-Regenerate-2K module, which is not part of this or Comfy-Org's release.) These are ComfyUI single-file checkpoints, not diffusers models. Filenames follow minimaxh3.safetensors. - fl2va — first/last-frame mode. Zero images = text-to-video, one or two = frame-conditioned. - ref2va — omni-reference mode (up to 9 images / 3 video clips / 3 audio clips). Both get identical treatment; pick the one matching your workflow. The three INT4 variants…