comfyui minimax h3
Shape AI clips with matching stereo audio straight from the comfyui minimax h3 node graph
AI Video Prompt Generator
10s

Feedback

AI Ad Video Example

Loading...

comfyui minimax h3

Try the comfyui minimax h3 node setup free — turn text, stills, or clips into 2K video with synced stereo sound, running locally with no API caps.

All Tools

Discover our comprehensive AI-powered animation toolkit

What the comfyui minimax h3 Workflow Brings to Your Machine

Packed into ComfyUI's template library, the comfyui minimax h3 setup loads MiniMax's omni-modal generation model as downloadable open weights. Text, stills, footage, and audio all feed one shared context, so a single forward pass yields clips with built-in stereo sound — dialogue, effects, and score included. You get up to 2K at 24fps for roughly 15 seconds, with every parameter exposed on the node graph.

  • Built-In Stereo Sound
    Speech, effects, and music arrive bundled with the picture in a single MP4, aligned automatically as the comfyui minimax h3 graph completes its pass.
  • Open Weights, Total Control
    Load the comfyui minimax h3 checkpoint on your own GPU and tune duration, resolution, and every diffusion setting without hitting API quotas.
  • Blended Reference Inputs
    Mix text, stills, footage, and audio in one run, pinning a character, look, motion, camera angle, or voice through the comfyui minimax h3 nodes.

Running the comfyui minimax h3 Workflow in Three Moves

From install to finished render: three short moves put the comfyui minimax h3 workflow to work on open-weight video with built-in audio.

Core Capabilities of the comfyui minimax h3 Workflow

Native templates, open-weight multimodal generation, synchronized stereo sound, reference-based control, and optional Sage Attention acceleration — the comfyui minimax h3 workflow covers the whole local video pipeline.

Three Ready-Made Templates

Text-to-video, image-to-video, and reference-to-video examples ship inside the comfyui minimax h3 library, one covering each generation mode.

One Shared Omni-Modal Context

Text, stills, footage, and audio are interpreted together by the comfyui minimax h3 model, letting every reference type feed a single generation.

Guided by Reference Material

Hold a face, a look, a movement, a camera path, or a voice steady using up to 9 images, 3 videos, and 3 audio clips through the comfyui minimax h3 R2V node.

Crisp Text & Logo Rendering

Words and brand marks come out legible with the comfyui minimax h3 model, and natural-language instructions describe how each reference relates to the others.

Sage Attention Acceleration

Adding the Patch Sage Attention KJ node to your comfyui minimax h3 graph roughly doubles throughput while barely touching output quality.

Resolution & Duration Grid

The comfyui minimax h3 Resolution Selector derives width and height from ratio and megapixels, snapped to a 32-pixel multiple and 17-frame blocks at 24fps.

FAQ

Questions About the comfyui minimax h3 Workflow

Everything people ask before loading MiniMax H3 open weights into their own ComfyUI setup.

1

What does the comfyui minimax h3 workflow actually do?

It wraps MiniMax H3 — MiniMax's omni-modal generation model, published as open weights — into a native ComfyUI graph. From text, stills, footage, or audio references, one forward pass produces video paired with its own stereo soundtrack.

2

How high can the output quality go?

Renders top out around 2K at 24fps across roughly 15 seconds. The native canvas keeps a 768px short edge, never exceeds 768x1344 pixels, and rounds dimensions to multiples of 32.

3

Which generation modes ship with it?

Three examples come in the comfyui minimax h3 template library: text-to-video, image-to-video with optional first and last frame control, and reference-to-video for locking character, style, motion, camera, or voice.

4

Is audio part of the output?

It is. Speech, sound effects, and music are generated natively in stereo alongside the picture — one pass, one MP4, everything in sync.

5

How do I begin?

Move ComfyUI to 0.30.0 or higher, open Template Library > Video, pick a comfyui minimax h3 template, and accept the prompt to download models from the Hugging Face Comfy-Org/MiniMax-H3 repository.

6

Can rendering be made faster?

Yes. Install SageAttention plus the KJNodes custom nodes, then insert a Patch Sage Attention KJ node between UNETLoader and BasicGuider in the comfyui minimax h3 graph — expect roughly twice the speed.

Put the comfyui minimax h3 Workflow to Work

Load MiniMax H3 open weights on your own machine, keep native stereo audio, and control every parameter — text-to-video, image-to-video, and reference-to-video graphs are ready to run.