Try the comfyui minimax h3 Generator
Describe your idea and the comfyui minimax h3 workflow turns it into video with synced sound.
AI Video Prompt Generator
10s

Feedback

AI Ad Video Example

Loading...

comfyui minimax h3

Turn a sentence, a photo, or a short clip into a polished 2K video with synced audio — all through the comfyui minimax h3 workflow in ComfyUI.

All Tools

Discover our comprehensive AI-powered animation toolkit

Key Advantages of the comfyui minimax h3 Workflow for Video

With the comfyui minimax h3 workflow, you can run MiniMax's open-weight omni-modal engine directly inside ComfyUI. It processes text, visuals, motion, and sound as one unified context, so voice, effects, and music render alongside the picture in a single pass. You get clips up to 2K at 24fps, lasting around 15 seconds, with every node exposed for full parameter tuning.

  • Synchronized Stereo Sound
    Voices, sound effects, and music are produced in the same MP4 as the visuals, all aligned by the pipeline in a single generation step.
  • Full Local Customization
    Keep the comfyui minimax h3 model on your own hardware and adjust resolution, length, and every sampling detail — no API quota to worry about.
  • Multi-Format Reference Blending
    Feed prompts, still images, clips, and audio together so the node graph locks in identity, visual style, motion, camera movement, or vocal tone.

Three Simple Steps to Run the comfyui minimax h3 Workflow

Follow this quick guide to produce open-weight clips with synchronized sound through the comfyui minimax h3 workflow.

Feature Highlights of the comfyui minimax h3 Workflow

From ready-made ComfyUI templates and open-weight generation to synced audio, reference-based control, and Sage Attention acceleration, the comfyui minimax h3 workflow gives you a full offline video toolkit.

Three Built-In Generation Templates

The workflow package includes ready-made examples for text-to-video, image-to-video, and reference-to-video, each handling one mode instantly.

Unified Multimodal Understanding

Text, visuals, motion, and sound are processed together in one context, letting you mix reference types in a single render.

Guided Generation from References

The comfyui minimax h3 reference-to-video node lets you control identity, style, movement, perspective, or voice using up to 9 photos, 3 clips, and 3 sound files.

Clean Text and Logo Output

Written words and brand assets render sharply, and the model follows natural-language instructions for relationships between references.

Sage Attention Acceleration

Drop the Patch Sage Attention KJ node into the graph to nearly double rendering speed while keeping image quality close to original.

Flexible Resolution and Timing

The resolution tool calculates dimensions from aspect ratio and megapixels, snapping to the model's 32-pixel grid and block-based duration at 24fps.

FAQ

Q&A: Running the comfyui minimax h3 Workflow in ComfyUI

Find quick answers about installing, using, and optimizing the comfyui minimax h3 workflow in ComfyUI.

1

What is the MiniMax H3 ComfyUI workflow?

It's ComfyUI's built-in integration for MiniMax H3, an open-weight omni-modal generation model. The comfyui minimax h3 workflow turns text, images, video, and audio references into video with synchronized stereo sound in a single pass.

2

What resolution and frame rate can I expect?

Expect up to 2K resolution at 24fps, for roughly 15 seconds of video from the workflow. The canvas starts at a 768px short edge, maxes out at 768x1344 pixels, and rounds dimensions to multiples of 32.

3

What generation modes come with the workflow?

You get three ready-made templates in the library: text-to-video, image-to-video with optional start/end frame control, and reference-to-video for locking character, look, motion, camera, or voice.

4

Is audio generated along with the video?

Yes. The model creates stereo audio — dialogue, sound effects, and music — together with the visuals in one pass, all delivered in a synchronized MP4.

5

How can I start using this workflow?

Update ComfyUI to 0.30.0 or newer, open Template Library > Video and select a template, then follow the prompt to download models from the Comfy-Org/MiniMax-H3 repo on Hugging Face.

6

Is there a way to make generation faster?

Yes. Install SageAttention and KJNodes, then insert a Patch Sage Attention KJ node after the UNETLoader in the graph. This can nearly double your generation speed.

Jump Into Video Creation with the comfyui minimax h3 Workflow

Set up the comfyui minimax h3 workflow in ComfyUI locally, keep full control over weights and settings, and produce text-to-video, image-to-video, or reference-based clips with audio locked to the action. Begin now.