Updated: Aug 9, 2026
base modelDownload
1 variant available
int8 SafeTensor
minimax_h3_fl2va_pruned_int8_convrot.safetensors
8-bit integer, smaller file • 19.53 GB
Verified: 10 days ago
You need these files to run this model. We'll show the best match for your preferences.
H3 (Comfy) • minimax_h3_audio_vae_fp32.safetensors
H3 (Comfy) • minimax_h3_video_vae_fp16.safetensors
H3 (Comfy) • qwen3vl_32b_minimax_h3_nvfp4_awq.safetensors
Downloads your preferred variants
8,7750 1 2 3 4 5 6 7 8 9,0 1 2 3 4 5 6 7 8 90 1 2 3 4 5 6 7 8 90 1 2 3 4 5 6 7 8 9
Aug 3, 2026
MiniMax H3

5900 1 2 3 4 5 6 7 8 90 1 2 3 4 5 6 7 8 90 1 2 3 4 5 6 7 8 9
License:
Minimax H3 INT8/INT4 Convrot
Required Components added to description as the files link to the official Civitai Minimax H3 model page which is set to generation-only, with no option to unlink.
text_encoders/qwen3vl_32b_minimax_h3_nvfp4_awq.safetensors
vae/minimax_h3_audio_vae_fp32.safetensors
vae/minimax_h3_video_vae_fp16.safetensors
Just uploaded both FL2VA & REF2VA w4a8_mixed models quantized by Kijai. (Makes my mixed int4 models obsolete) True int4 model size with int8 activations, near int8 quality, same speed as int8. For RAM/VRAM-constrained systems. Update your ComfyUI to the latest version for support of the new w4a8 (Should be included in Stable ComfyUI v0.31.0 https://github.com/Comfy-Org/ComfyUI/pull/15308
FL2VA - first last (frame) to video / audio
REF2VA - ref_images / ref_videos / ref_video_audios / ref_audios: up to 9 reference images, 3 reference videos (each may carry its own paired soundtrack), and 3 standalone reference audio clips
Both models can generate t2v (Text to video), i2v (Image to video), v2v (Video to Video), a2v (Audio to video), and multiple references (image/video/audio). But were further fine-tuned/trained for higher-quality outputs for the intended use.
