Feedback
AI Ad Video Example
Loading...
comfyui minimax h3
See how the comfyui minimax h3 workflow brings MiniMax H3's open weights into ComfyUI for 2K clips with native stereo audio and full node control.
All Tools
Discover our comprehensive AI-powered animation toolkit
MiniMax H3
MiniMax H3 AI Video Generator
Seedance 2.5
The Future of AI Video Is Here.

Seedance 2.0
The Future of AI Video Is Here.

Kling 3.0
Next-Gen AI Video Generator
Grok Video Generator
Create Videos from Text or Images with AI
MiniMax H3 video generator
MiniMax H3 AI Video Generator

Seedance 2.0
The Future of AI Video Is Here.

Gemini Omni
Gemini Omni Video Generator
Why Creators Choose the comfyui minimax h3 Pipeline
The comfyui minimax h3 workflow packages MiniMax's omni-modal generation model as downloadable open weights inside ComfyUI. Because it reads text, images, video, and audio in one shared context, the graph can output footage with stereo dialogue, effects, and music written in the same pass. Clips run up to roughly fifteen seconds at 24fps and 2K, and every sampler, resolution, and duration setting stays exposed at the node level.
- Audio Built Into Every RenderVoice, effects, and music arrive already mixed with the picture, so each export lands as a single MP4 with stereo channels aligned.
- Runs Entirely on Your Own MachineDownload the weights once, then tune duration, resolution, and sampler values yourself — nothing gets throttled by a remote API.
- Mix Text, Stills, Clips, and SoundFeed several reference types at once to hold a face, a look, a camera move, or a specific voice steady across the comfyui minimax h3 generation.
Setting Up the comfyui minimax h3 Workflow in Three Moves
Three short steps carry you from a fresh ComfyUI install to open-weight video with synced audio.
Core Strengths of the comfyui minimax h3 Pipeline
Three ready-made templates, omni-modal understanding, stereo audio in the same pass, reference-locked control, and optional Sage Attention acceleration — the comfyui minimax h3 pipeline covers a complete local production loop.
Three Ready-Made Templates
The comfyui minimax h3 library includes text-to-video, image-to-video, and reference-to-video examples, so each generation mode works the moment you open it.
One Shared Context for Every Modality
Text, stills, footage, and audio are read together by the comfyui minimax h3 model, letting you blend reference types inside a single generation.
Lock Identity, Style, and Motion
Pin a face, a look, a movement, a camera path, or a voice using up to 9 images, 3 videos, and 3 audio clips through the comfyui minimax h3 R2V node.
Clean On-Screen Text and Brand Marks
Words and brand elements come out legible with the comfyui minimax h3 model, and plain-language instructions can describe how each reference relates to the others.
Faster Runs With Sage Attention
Drop a Patch Sage Attention KJ node into the comfyui minimax h3 workflow and generation time roughly halves, with only a slight quality trade-off.
Smart Resolution and Duration Grid
The comfyui minimax h3 Resolution Selector derives width and height from aspect ratio and megapixels, snapped to the 32-pixel grid with 17-frame blocks at 24fps.
comfyui minimax h3: Questions Answered
Answers to the questions people ask most often about running MiniMax H3 as a ComfyUI workflow.
What does the comfyui minimax h3 workflow actually do?
It wires MiniMax H3 — MiniMax's omni-modal generation model, released as open weights — straight into ComfyUI. A single forward pass turns text, images, video, and audio references into footage that already carries stereo sound.
How sharp and how long can the output be?
Expect up to 2K resolution at 24fps across roughly fifteen seconds. The native canvas keeps a 768px short edge, tops out at 768x1344 pixels, and rounds dimensions to multiples of 32.
Which generation modes ship with it?
Three examples arrive in the template library: text-to-video, image-to-video with optional first and last frame control, and reference-to-video, which holds a character, style, motion, camera path, or voice steady.
Is audio generated as well?
It is. The comfyui minimax h3 model writes stereo voice, sound effects, and music alongside the picture in the same pass, and everything is saved together inside one MP4 file.
What is the fastest way to begin?
Upgrade ComfyUI to 0.30.0 or newer, open Template Library > Video, pick a comfyui minimax h3 workflow, and accept the download prompt for models in the Hugging Face Comfy-Org/MiniMax-H3 repository.
Is there a way to make generation faster?
Yes. Install SageAttention plus the KJNodes pack, then place a Patch Sage Attention KJ node between the UNETLoader and BasicGuider in the comfyui minimax h3 workflow to roughly double throughput.
Put the comfyui minimax h3 Workflow to Work
Bring MiniMax H3 onto your own hardware inside ComfyUI — stereo audio, open weights, and every parameter in reach, with text, image, and reference video modes ready to run.
