comfyui minimax h3 Video Generator
Describe a scene, add a reference, and let the comfyui minimax h3 workflow render picture and stereo sound together.
AI Video Prompt Generator
10s

Feedback

AI Ad Video Example

Loading...

comfyui minimax h3

Test the comfyui minimax h3 node pack: turn prompts, stills, or clips into 2K video with synced stereo sound, all inside a browser workflow.

All Tools

Discover our comprehensive AI-powered animation toolkit

What the comfyui minimax h3 Node Pack Delivers

Built on MiniMax's omni-modal generation model, the comfyui minimax h3 workflow ships as open weights you can run inside ComfyUI. A single context absorbs text, stills, clips, and audio at once, then returns a video whose stereo soundtrack — speech, effects, and music — is modeled in the same forward pass. Clips reach roughly 15 seconds at 24fps in up to 2K, and every sampler, step, and dimension stays editable at node level.

  • Stereo Audio, Generated Inline
    Speech, effects, and music land in the same MP4 as the picture — one pass through the comfyui minimax h3 graph keeps them perfectly in sync.
  • Runs on Your Own Hardware
    Because the comfyui minimax h3 weights are public, you set resolution, length, and every diffusion knob yourself, with no API quota in the way.
  • Mix Text, Stills, Clips & Voice
    Feed several reference types into one run and pin down a face, a look, a motion, a camera path, or a voice using the comfyui minimax h3 nodes.

Three Steps to Run the comfyui minimax h3 Workflow

From a fresh setup to your first open-weight clip with native audio, the comfyui minimax h3 workflow takes only three moves.

What You Get from the comfyui minimax h3 Workflow

Six strengths turn the comfyui minimax h3 setup into a full local production line: three ready-made templates, omni-modal understanding, reference locking, crisp on-screen text, Sage Attention acceleration, and a resolution grid that always keeps dimensions valid.

Three Ready-Made Templates

Text-to-video, image-to-video, and reference-to-video examples ship in the comfyui minimax h3 template library, one per generation mode, ready to run as-is.

One Context, Every Modality

Text, stills, clips, and audio are read together by the comfyui minimax h3 model, so different reference kinds can shape a single generation.

Lock Onto a Reference

Hold a face, a visual style, a movement, a camera path, or a voice steady with reference material — as many as 9 images, 3 videos, and 3 audio clips via the R2V node.

Clean On-Screen Text and Logos

Lettering and brand marks come out legible with the comfyui minimax h3 model, and plain-language instructions can spell out how references relate.

Roughly 2x Faster with Sage Attention

Drop a Patch Sage Attention KJ node into the comfyui minimax h3 graph to nearly double throughput while barely touching visual quality.

Smart Resolution and Duration Grid

The comfyui minimax h3 Resolution Selector derives width and height from ratio and megapixels, snapping to the model's 32-pixel multiples and 17-frame blocks at 24fps.

FAQ

comfyui minimax h3: Questions Answered

Straight answers about running the comfyui minimax h3 model locally inside ComfyUI.

1

What exactly is the comfyui minimax h3 workflow?

It is ComfyUI's built-in support for MiniMax H3, an omni-modal generation model from MiniMax released with public weights. From a single forward pass it turns text, images, video, and audio references into video that already carries stereo sound.

2

What resolution and frame rate can it reach?

Clips run up to 2K at 24fps for around 15 seconds. The native canvas keeps the short edge at 768px, tops out at 768x1344, and rounds dimensions to a multiple of 32.

3

Which generation modes ship with it?

Three examples arrive in the template library: text-to-video (T2V), image-to-video (I2V) with optional first- and last-frame control, and reference-to-video (R2V) that pins a character, style, motion, camera, or voice.

4

Will it produce sound as well as picture?

Yes. Stereo audio — speech, effects, and music — is modeled alongside the visuals by the comfyui minimax h3 model and delivered together in one MP4.

5

How do I get it running?

Upgrade ComfyUI to 0.30.0 or newer, go to Template Library > Video, select a comfyui minimax h3 workflow, and accept the pop-up that downloads models from the Hugging Face Comfy-Org/MiniMax-H3 repository.

6

Is there a way to make it faster?

Install SageAttention plus the KJNodes custom nodes, then place a Patch Sage Attention KJ node between the UNETLoader and BasicGuider in the comfyui minimax h3 graph for roughly double the speed.

Put the comfyui minimax h3 Workflow to Work

Take MiniMax H3 for a spin right here — open weights, stereo sound, and every parameter exposed, with T2V, I2V, and R2V templates waiting for your first prompt.