Feedback
AI Ad Video Example
Loading...
comfyui minimax h3
Use the comfyui minimax h3 workflow to run open-weight MiniMax H3 in ComfyUI and create 2K/24fps clips with native stereo audio for T2V, I2V, and R2V.
All Tools
Discover our comprehensive AI-powered animation toolkit
MiniMax H3
MiniMax H3 AI Video Generator
Seedance 2.5
The Future of AI Video Is Here.

Seedance 2.0
The Future of AI Video Is Here.
FLUX 3 Video Generator

Kling 3.0
Next-Gen AI Video Generator
Grok Video Generator
Create Videos from Text or Images with AI
MiniMax H3 video generator
MiniMax H3 AI Video Generator

Seedance 2.0
The Future of AI Video Is Here.
What Makes the comfyui minimax h3 Workflow Stand Out
The comfyui minimax h3 workflow brings MiniMax's open-weight, omni-modal model into ComfyUI. It processes text, images, video, and audio together, then outputs clips with native stereo audio — dialogue, effects, and music appear in a single pass. Expect up to 2K resolution at 24fps and about 15 seconds of footage, with every parameter exposed for node-level tuning.
- Stereo Sound Built InDialogue, effects, and music are rendered alongside the picture in a single MP4, synced automatically through the comfyui minimax h3 workflow.
- Complete Local ControlExecute the comfyui minimax h3 model on your own machine, tweaking resolution, duration, and all diffusion settings without hitting API quotas.
- Multiple Media ReferencesCombine text, images, clips, and audio cues in a single run, using comfyui minimax h3 nodes to lock down character, style, motion, camera movement, or voice.
Running the comfyui minimax h3 Workflow in Three Steps
Three steps are all you need to create open-weight videos with native sound through the comfyui minimax h3 pipeline.
What the comfyui minimax h3 Workflow Offers
A local video studio in one workflow: three ComfyUI presets, open-weight multimodal generation, stereo audio, reference-driven control, and optional Sage Attention acceleration with the comfyui minimax h3 nodes.
Three Built-In Template Options
The comfyui minimax h3 template set provides separate examples for text-to-video, image-to-video, and reference-to-video, so each generation mode is ready out of the box.
Unified Understanding Across Media
The model behind comfyui minimax h3 processes text, visuals, footage, and sound in a single shared context, letting you combine every reference type in one generation.
Reference-Based Video Control
Through the comfyui minimax h3 R2V node, you can pin down a character, style, movement, camera angle, or voice using up to 9 images, 3 footage clips, and 3 audio files.
Precise Text and Logo Output
The comfyui minimax h3 model renders spelled-out text and brand elements cleanly, and follows natural-language instructions that describe relationships between references.
Accelerate with Sage Attention
Add the Patch Sage Attention KJ node to the comfyui minimax h3 pipeline and get roughly 2x faster generation with only minimal quality drop.
Precise Resolution and Length Grid
The resolution selector calculates dimensions from aspect ratio and megapixels, snapping to the model's 32-pixel grid and 17-frame block duration at 24fps — all accessible in the comfyui minimax h3 workflow.
comfyui minimax h3 — Common Questions
Find answers to typical questions about using the MiniMax H3 model within ComfyUI and the comfyui minimax h3 workflow.
What does the comfyui minimax h3 workflow involve?
It’s ComfyUI’s built-in integration for MiniMax H3, an open-weight omni-modal generator. The workflow turns text, images, footage, and audio references into video with native stereo audio in a single inference step.
What resolution and frame rate can you expect?
You can generate up to 2K resolution at 24fps for about 15 seconds. The native canvas has a 768px short edge, tops out at 768x1344 pixels, and rounds to multiples of 32 within the comfyui minimax h3 workflow.
What generation modes are available?
The comfyui minimax h3 template library offers text-to-video, image-to-video (with optional first/last-frame control), and reference-to-video presets for locking in character, style, motion, camera, or voice.
Is audio actually generated?
Yes — the model creates native stereo audio covering speech, sound effects, and music, all generated together with the video and synced into one MP4 via the comfyui minimax h3 workflow.
What do I need to get started?
Update ComfyUI to 0.30.0 or later, open Template Library > Video, pick a comfyui minimax h3 workflow, and use the pop-up to fetch models from the Comfy-Org/MiniMax-H3 repository on Hugging Face.
Is it possible to make generation faster?
Yes — install SageAttention and KJNodes, then insert a Patch Sage Attention KJ node between the UNETLoader and BasicGuider in the comfyui minimax h3 workflow to roughly double speed.
Jump Into Video Creation with the comfyui minimax h3 Workflow
Get local access to MiniMax H3 in ComfyUI — open weights, native stereo sound, and complete parameter control make the comfyui minimax h3 workflow perfect for T2V, I2V, and R2V projects.
