Feedback
AI Ad Video Example
Loading...
comfyui minimax h3
Discover a smarter way to produce video with sound: the comfyui minimax h3 workflow in ComfyUI renders text, images, or footage into clips with native stereo audio — up to 2K at 24fps.
All Tools
Discover our comprehensive AI-powered animation toolkit
MiniMax H3
MiniMax H3 AI Video Generator
Seedance 2.5
The Future of AI Video Is Here.

Seedance 2.0
The Future of AI Video Is Here.

Veo3.1
Create Stunning Videos with Veo3.1

Kling 3.0
Next-Gen AI Video Generator
Grok Video Generator
Create Videos from Text or Images with AI
MiniMax H3 video generator
MiniMax H3 AI Video Generator

Seedance 2.0
The Future of AI Video Is Here.
Open-Weight Video + Audio: comfyui minimax h3 in ComfyUI
The comfyui minimax h3 integration brings MiniMax H3 into ComfyUI as an open-weight model that understands text, images, video, and audio in one context. It generates clips with native stereo sound, striking a balance between resolution (up to 2K), frame rate (24fps), and duration (about 15 seconds). Every node is exposed for customization.
- One File, Full StereoThe model generates dialogue, effects, and music together with the video track, so everything arrives in a single MP4 with perfect sync.
- Local Control, No LimitsWith open-weight access, you can adjust resolution, length, and every diffusion parameter on your own machine — no third-party limits.
- Blend Any Media TypeThe nodes let you mix text, images, video, and audio references in a single run to lock in a character, style, motion, camera angle, or voice.
A Quick Start Guide for the comfyui minimax h3 Pipeline
Follow these three steps to create open-weight video with native audio through the comfyui minimax h3 integration.
Capabilities of the comfyui minimax h3 Node Suite
The comfyui minimax h3 package includes three native ComfyUI templates, open-weight multimodal generation, synchronized stereo audio, reference-locked control, and optional Sage Attention acceleration — everything you need for a full local video production pipeline.
Ready-Made Graph Templates
You get text-to-video, image-to-video, and reference-to-video examples, each preconfigured for one particular generation mode.
Unified Multimodal Context
The model processes text, images, video, and audio in a single shared context, allowing you to combine multiple reference types in one generation.
Reference-Locked Rendering
With the R2V node, you can lock a character's identity, style, movement, camera move, or voice using up to 9 images, 3 videos, and 3 audio clips.
Accurate Text and Brand Rendering
The model renders spelled-out text and brand elements cleanly, and follows natural-language instructions that specify how references relate.
Sage Attention Acceleration
Insert the Patch Sage Attention KJ node into the workflow to roughly double speed with minimal quality loss.
Resolution and Duration Grid
The Resolution Selector derives width and height from your chosen aspect ratio and megapixels, snapping to the model's 32-multiple grid and 17-frame duration blocks at 24fps.
comfyui minimax h3: Frequently Asked Questions
Get straightforward answers about using MiniMax H3 inside ComfyUI with the comfyui minimax h3 integration — covering setup, output, audio, and speed.
What exactly is the comfyui minimax h3 integration?
It's ComfyUI's native bridge to MiniMax H3, an open-weight omni-modal model. The workflow generates video with built-in stereo audio from text, images, footage, or sound references in one go.
What render quality does it support?
You can expect up to 2K resolution at 24fps for about 15 seconds. The native canvas starts at a 768px short edge, limits to 768x1344 pixels, and snaps to a multiple of 32.
Which video modes are included?
The presets include three modes: text-to-video (T2V), image-to-video (I2V) with optional first/last-frame control, and reference-to-video (R2V) that can lock character, style, motion, camera, or voice.
Will the output include audio?
Yes. The model generates native stereo audio — voice, effects, and music — together with the visuals and syncs it all into a single MP4.
What's the first step?
Update ComfyUI to 0.30.0 or later, open Template Library > Video, select the workflow you want, and follow the prompt to download models from the Comfy-Org/MiniMax-H3 Hugging Face repository.
Can I make comfyui minimax h3 render faster?
Yes — install SageAttention and KJNodes, then put a Patch Sage Attention KJ node between the UNETLoader and BasicGuider in the graph. This can about double your render speed.
Launch Your comfyui minimax h3 Video Projects Today
Take control of MiniMax H3 in ComfyUI with the comfyui minimax h3 integration — open weights, native stereo audio, and T2V/I2V/R2V presets are all ready for you. Start now.
