comfyui minimax h3
Turn an idea into a finished clip — the comfyui minimax h3 workflow adds native stereo audio to every render
AI Video Prompt Generator
10s

Feedback

AI Ad Video Example

Loading...

comfyui minimax h3

Build open-weight AI video inside ComfyUI with the comfyui minimax h3 workflow — native stereo audio, 2K/24fps output, and node-level control.

All Tools

Discover our comprehensive AI-powered animation toolkit

What Makes the comfyui minimax h3 Workflow Different

Inside ComfyUI, the comfyui minimax h3 workflow loads MiniMax's omni-modal model as open weights. A single context handles text, images, video, and audio at once, so dialogue, effects, and music are rendered alongside the picture in one pass. Clips reach roughly 15 seconds at 2K/24fps, and every parameter stays editable at the node level.

  • Stereo Sound, Baked In
    Voice, effects, and music arrive together with the picture in a single MP4 — synchronized during the same pass through the comfyui minimax h3 workflow.
  • Runs on Open Weights
    Load the comfyui minimax h3 model on your own machine and tune resolution, clip length, and diffusion settings freely — no API caps.
  • Mixed Reference Inputs
    Feed text, stills, video, and audio references into one job, pinning a face, look, movement, camera angle, or voice through the comfyui minimax h3 nodes.

Running the comfyui minimax h3 Workflow: Three Steps

Three quick steps take you from setup to a finished open-weight clip with audio via the comfyui minimax h3 workflow.

Inside the comfyui minimax h3 Workflow: Key Capabilities

Three ready-made ComfyUI templates, open-weight multimodal generation, built-in stereo audio, reference-driven control, and optional Sage Attention acceleration — the comfyui minimax h3 workflow covers a full local video pipeline.

Three Built-In Templates

The comfyui minimax h3 template library includes text-to-video, image-to-video, and reference-to-video examples, each ready to run for one generation mode.

One Omni-Modal Context

Text, images, video, and audio are all interpreted together by the comfyui minimax h3 model, so every reference type can feed a single generation.

Generation Anchored by References

Pin a face, style, motion, camera path, or voice from source material — the comfyui minimax h3 R2V node accepts up to 9 images, 3 videos, and 3 audio clips.

Clean Text and Brand Marks

On-screen words and logos come out sharp with the comfyui minimax h3 model, and natural-language instructions can describe how references relate.

Faster Runs with Sage Attention

Add the Patch Sage Attention KJ node to the comfyui minimax h3 workflow and generation speeds up roughly twofold with little quality loss.

Resolution and Duration Grid

The comfyui minimax h3 Resolution Selector derives width and height from aspect ratio and megapixels, aligned to the model's 32-multiple grid and 17-frame blocks at 24fps.

FAQ

comfyui minimax h3: Questions Answered

Straight answers about running the MiniMax H3 model locally in ComfyUI.

1

What exactly is the comfyui minimax h3 workflow?

It is ComfyUI's built-in integration of MiniMax H3, an omni-modal generation model MiniMax released with open weights. From text, images, video, and audio references it renders video plus stereo audio in a single forward pass.

2

How high does the output quality go?

The comfyui minimax h3 workflow can reach 2K at 24fps for roughly 15 seconds. Its native canvas keeps a 768px short edge, tops out at 768x1344, and rounds dimensions to multiples of 32.

3

Which generation modes ship with it?

Three examples come in the comfyui minimax h3 template library: text-to-video (T2V), image-to-video (I2V) with optional first and last frame control, and reference-to-video (R2V) for locking character, style, motion, camera, or voice.

4

Does the output include audio?

It does. The comfyui minimax h3 model renders native stereo audio — speech, effects, and music — together with the visuals in one pass, delivered as a synced MP4.

5

What is the fastest way to get started?

Upgrade ComfyUI to 0.30.0 or newer, open Template Library > Video, pick a comfyui minimax h3 workflow, and let the pop-up guide you through downloading models from the Hugging Face Comfy-Org/MiniMax-H3 repository.

6

Is there a way to speed things up?

Yes — install SageAttention plus the KJNodes custom nodes, then insert a Patch Sage Attention KJ node between the UNETLoader and BasicGuider in the comfyui minimax h3 workflow to roughly double throughput.

Put the comfyui minimax h3 Workflow to Work

Load MiniMax H3 on your own machine through ComfyUI — open weights, stereo audio in every render, and full parameter access across T2V, I2V, and R2V workflows.