Updated: Sep 12, 2026
toolDownload
1 variant available
Archive Other
VDN_H3_ComfyUI_Text_To_Video_MiniMax_H3_VDN_8-Step_No_LoRA.zip
9.35 KB
Verified: 15 days ago

376
1.8K
MiniMax H3 VDN makes cinematic video with audio in just 8 steps.
Who it's for: creators who want this pipeline in ComfyUI without assembling nodes from scratch. Not for: one-click results with zero tuning - you still choose inputs, prompts, and settings.
Open preloaded workflow on RunComfy
Open preloaded workflow on RunComfy (browser)
Why RunComfy first
- Fewer missing-node surprises - run the graph in a managed environment before you mirror it locally.
- Quick GPU tryout - useful if your local VRAM or install time is the bottleneck.
- Matches the published JSON - the zip follows the same runnable workflow you can open on RunComfy.
When downloading for local ComfyUI makes sense - you want full control over models on disk, batch scripting, or offline runs.
How to use (local ComfyUI)
1. Load inputs (images/video/audio) in the marked loader nodes.
2. Set prompts, resolution, and seeds; start with a short test run.
3. Export from the Save / Write nodes shown in the graph.
Expectations - First run may pull large weights; cloud runs may require a free RunComfy account.
Overview
Create MiniMax H3 videos with audio from text prompts. You get 8-step sampling with Video DeltaNet hybrid attention. Build fantasy, driving, or combat scenes. Compare outputs with the FastVideo branch. Refine prompts for creative control.
Important nodes:
Key nodes in Comfyui VDN H3 ComfyUI Text To Video workflow
ApplyVDNH3Advanced(#6126). Enables Video DeltaNet on MiniMax H3 using the selected checkpoint from the VDN-H3 release. Adjust only when you want to change the VDN checkpoint or switch attention backends; keep the hybrid attention defaults for general use. Reference implementation: Saganaki22/ComfyUI-VDN-H3.MiniMaxH3ImageToVideo(#219). Central control point for story and timing, accepting the text prompt, computed width and height, and the derived frame count. For audio, append a singleAudio:line to the prompt to steer Foley and ambience while avoiding speech unless explicitly requested.MiniMaxH3SigmaShift(#6131, #6007). Rebalances the sigma schedule for video and audio streams, which can help reduce jitter and preserve detail at low step counts. Consider small adjustments only if you change model variants or notice motion instability.VHS_VideoCombine(#6170, #6104). Muxes decoded frames and audio into MP4 with your chosen frame rate and filename prefix. Useful options include trimming to audio length and selecting the desired H.264 pixel format. Project page: ComfyUI-VideoHelperSuite.
Notes
VDN H3 ComfyUI Text To Video | MiniMax H3 VDN 8-Step, No LoRA - see RunComfy page for the latest node requirements.

