Feedback
AI Ad Video Example
Loading...
comfyui minimax h3
Turn prompts, stills, or footage into 2K open-weight video in ComfyUI, complete with synced stereo sound, using the comfyui minimax h3 workflow.
All Tools
Discover our comprehensive AI-powered animation toolkit
MiniMax H3
MiniMax H3 AI Video Generator
Seedance 2.5
The Future of AI Video Is Here.

Seedance 2.0
The Future of AI Video Is Here.

Veo3.1
Create Stunning Videos with Veo3.1
Grok Video Generator
Create Videos from Text or Images with AI
MiniMax H3 video generator
MiniMax H3 AI Video Generator

Seedance 2.0
The Future of AI Video Is Here.
Sora2 AI
Advanced AI Video Generator for High-Quality Videos
Key Advantages of the comfyui minimax h3 Workflow
Packaging MiniMax's versatile all-in-one multimodal model as open weights, this ComfyUI integration lets you drive text, images, video, and audio from a single graph. The comfyui minimax h3 workflow creates dialogue, sound effects, and music in the same pass as the visuals, outputs a synchronized MP4 up to 2K at 24fps for about 15 seconds, and leaves every parameter adjustable via nodes.
- Built-in Stereo SoundSpeech, effects, and music are produced together with the footage and combined into one MP4, with everything aligned inside the comfyui minimax h3 workflow.
- Open-Weight Local FreedomHost the comfyui minimax h3 model on your own hardware, with complete control over resolution, clip length, and all diffusion settings — no API limits.
- One Reference, Many Input TypesCombine prompts, pictures, clips, and audio in a single run, using the comfyui minimax h3 nodes to hold character, style, movement, camera motion, or voice.
A Simple Guide to Running the comfyui minimax h3 Workflow
Follow these three steps to produce open-weight videos with built-in sound using the comfyui minimax h3 workflow.
Core Features of the comfyui minimax h3 Workflow
The comfyui minimax h3 workflow ships with three ready-made ComfyUI templates, open-weight multimodal generation, built-in stereo audio, reference-driven control, and optional Sage Attention acceleration — a complete local video toolkit.
Three Built-In Templates
The comfyui minimax h3 library includes text-to-video, image-to-video, and reference-to-video templates, each covering one generation mode out of the box.
All-in-One Multimodal Context
The comfyui minimax h3 model reads text, photos, footage, and sound in one shared context, allowing you to blend several reference types in a single generation.
Reference-Driven Production
Lock identity, style, motion, camera, or voice with references — up to 9 images, 3 videos, and 3 audio clips routed through the comfyui minimax h3 R2V node.
Accurate Text and Brand Rendering
The comfyui minimax h3 model renders spelled-out copy and brand assets cleanly, and follows natural-language instructions about how references relate.
Sage Attention Boost
Put a Patch Sage Attention KJ node into the comfyui minimax h3 workflow and you can roughly double generation speed with only a slight quality trade-off.
Resolution and Duration Grid
The comfyui minimax h3 Resolution Selector calculates width and height from aspect ratio and megapixels, matching the 32-multiple grid and a 17-frame block at 24fps.
Frequently Asked Questions: comfyui minimax h3 in ComfyUI
Find quick answers on using the MiniMax H3 model with the comfyui minimax h3 workflow.
What does the comfyui minimax h3 workflow do?
It gives ComfyUI built-in support for MiniMax H3, the open-weight omni-modal model from MiniMax. In one pass, the comfyui minimax h3 workflow turns prompts, images, video, and audio into a video with synchronized stereo sound.
What resolution and frame rate can I expect?
The comfyui minimax h3 workflow can output up to 2K resolution at 24fps for about 15 seconds. Its canvas uses a 768px short edge, caps at 768x1344, and rounds to multiples of 32.
Which generation modes come with the workflow?
The comfyui minimax h3 template collection includes text-to-video (T2V), image-to-video (I2V) with optional first/last-frame control, and reference-to-video (R2V) that preserves identity, style, motion, camera, or voice.
Does the workflow generate audio?
Yes. The comfyui minimax h3 model creates stereo audio — speech, sound effects, and music — together with the visuals in one pass and syncs it into a single MP4.
What is the fastest way to begin?
Update ComfyUI to 0.30.0 or later, open Template Library > Video, choose a comfyui minimax h3 workflow, and use the pop-up to download the open-weight models from Hugging Face Comfy-Org/MiniMax-H3.
Can I make the comfyui minimax h3 workflow faster?
Yes. Install SageAttention and KJNodes, then insert a Patch Sage Attention KJ node between the UNETLoader and BasicGuider in the comfyui minimax h3 workflow to approximately halve the generation time.
Begin Creating with the comfyui minimax h3 Workflow Now
Run the comfyui minimax h3 workflow locally in ComfyUI with open weights, stereo sound, and full parameter control — text-to-video, image-to-video, and reference-to-video templates are ready to use.
