Generate with the comfyui minimax h3 workflow
Create videos with built-in stereo audio using the comfyui minimax h3 workflow in ComfyUI.
AI Video Prompt Generator
10s

Feedback

AI Ad Video Example

Loading...

comfyui minimax h3

Turn prompts, stills, or footage into 2K open-weight video in ComfyUI, complete with synced stereo sound, using the comfyui minimax h3 workflow.

All Tools

Discover our comprehensive AI-powered animation toolkit

Key Advantages of the comfyui minimax h3 Workflow

Packaging MiniMax's versatile all-in-one multimodal model as open weights, this ComfyUI integration lets you drive text, images, video, and audio from a single graph. The comfyui minimax h3 workflow creates dialogue, sound effects, and music in the same pass as the visuals, outputs a synchronized MP4 up to 2K at 24fps for about 15 seconds, and leaves every parameter adjustable via nodes.

  • Built-in Stereo Sound
    Speech, effects, and music are produced together with the footage and combined into one MP4, with everything aligned inside the comfyui minimax h3 workflow.
  • Open-Weight Local Freedom
    Host the comfyui minimax h3 model on your own hardware, with complete control over resolution, clip length, and all diffusion settings — no API limits.
  • One Reference, Many Input Types
    Combine prompts, pictures, clips, and audio in a single run, using the comfyui minimax h3 nodes to hold character, style, movement, camera motion, or voice.

A Simple Guide to Running the comfyui minimax h3 Workflow

Follow these three steps to produce open-weight videos with built-in sound using the comfyui minimax h3 workflow.

Core Features of the comfyui minimax h3 Workflow

The comfyui minimax h3 workflow ships with three ready-made ComfyUI templates, open-weight multimodal generation, built-in stereo audio, reference-driven control, and optional Sage Attention acceleration — a complete local video toolkit.

Three Built-In Templates

The comfyui minimax h3 library includes text-to-video, image-to-video, and reference-to-video templates, each covering one generation mode out of the box.

All-in-One Multimodal Context

The comfyui minimax h3 model reads text, photos, footage, and sound in one shared context, allowing you to blend several reference types in a single generation.

Reference-Driven Production

Lock identity, style, motion, camera, or voice with references — up to 9 images, 3 videos, and 3 audio clips routed through the comfyui minimax h3 R2V node.

Accurate Text and Brand Rendering

The comfyui minimax h3 model renders spelled-out copy and brand assets cleanly, and follows natural-language instructions about how references relate.

Sage Attention Boost

Put a Patch Sage Attention KJ node into the comfyui minimax h3 workflow and you can roughly double generation speed with only a slight quality trade-off.

Resolution and Duration Grid

The comfyui minimax h3 Resolution Selector calculates width and height from aspect ratio and megapixels, matching the 32-multiple grid and a 17-frame block at 24fps.

FAQ

Frequently Asked Questions: comfyui minimax h3 in ComfyUI

Find quick answers on using the MiniMax H3 model with the comfyui minimax h3 workflow.

1

What does the comfyui minimax h3 workflow do?

It gives ComfyUI built-in support for MiniMax H3, the open-weight omni-modal model from MiniMax. In one pass, the comfyui minimax h3 workflow turns prompts, images, video, and audio into a video with synchronized stereo sound.

2

What resolution and frame rate can I expect?

The comfyui minimax h3 workflow can output up to 2K resolution at 24fps for about 15 seconds. Its canvas uses a 768px short edge, caps at 768x1344, and rounds to multiples of 32.

3

Which generation modes come with the workflow?

The comfyui minimax h3 template collection includes text-to-video (T2V), image-to-video (I2V) with optional first/last-frame control, and reference-to-video (R2V) that preserves identity, style, motion, camera, or voice.

4

Does the workflow generate audio?

Yes. The comfyui minimax h3 model creates stereo audio — speech, sound effects, and music — together with the visuals in one pass and syncs it into a single MP4.

5

What is the fastest way to begin?

Update ComfyUI to 0.30.0 or later, open Template Library > Video, choose a comfyui minimax h3 workflow, and use the pop-up to download the open-weight models from Hugging Face Comfy-Org/MiniMax-H3.

6

Can I make the comfyui minimax h3 workflow faster?

Yes. Install SageAttention and KJNodes, then insert a Patch Sage Attention KJ node between the UNETLoader and BasicGuider in the comfyui minimax h3 workflow to approximately halve the generation time.

Begin Creating with the comfyui minimax h3 Workflow Now

Run the comfyui minimax h3 workflow locally in ComfyUI with open weights, stereo sound, and full parameter control — text-to-video, image-to-video, and reference-to-video templates are ready to use.