Feedback
AI Ad Video Example
Loading...
comfyui minimax h3
The comfyui minimax h3 integration turns text, images, or references into 2K/24fps open-weight clips with embedded stereo audio — all from inside ComfyUI.
All Tools
Discover our comprehensive AI-powered animation toolkit
MiniMax H3
MiniMax H3 AI Video Generator

Seedance 2.0
The Future of AI Video Is Here.
FLUX 3 Video Generator

Kling 3.0
Next-Gen AI Video Generator
Grok Video Generator
Create Videos from Text or Images with AI
MiniMax H3 video generator
MiniMax H3 AI Video Generator

Seedance 2.0
The Future of AI Video Is Here.

Gemini Omni
Gemini Omni Video Generator
Why choose the comfyui minimax h3 pipeline in ComfyUI
The comfyui minimax h3 integration packages MiniMax H3 as an open-weight, omni-modal generator. It processes text, images, footage, and audio in one contextual pass, then outputs an MP4 with stereo audio — dialogue, effects, and music — already embedded. In about 15 seconds of footage at 2K/24fps, you get fine-grained control over every diffusion parameter.
- Audio Rendered in the Same PassSpeech, SFX, and score are created together with the visuals in a single MP4, so the audio is already time-locked to the footage when rendering through comfyui minimax h3.
- Local Open-Weight ControlThe comfyui minimax h3 model runs on your own machine, letting you tune resolution, duration, and every sampling parameter with no API limits or queue waits.
- Multi-Reference Input SupportFeed the comfyui minimax h3 nodes with text alongside images, video, or audio to lock a character, aesthetic, movement, shot angle, or voice.
Getting Started with comfyui minimax h3 in ComfyUI
Three steps to start producing local open-weight video with embedded audio using the comfyui minimax h3 pipeline.
Capabilities of the comfyui minimax h3 Integration
This comfyui minimax h3 integration bundles three ready-made ComfyUI templates, open-weight processing, embedded stereo audio, reference-driven control, and optional Sage Attention acceleration — a complete local video production setup.
Three Ready-Made Templates
The comfyui minimax h3 presets cover text-to-video, image-to-video, and reference-to-video, so each mode runs immediately with no extra wiring.
Unified Multimodal Context
Text, images, footage, and audio are interpreted together by the comfyui minimax h3 model, allowing every reference to steer the same render.
Reference-Locked Generation
The comfyui minimax h3 R2V node preserves identity, look, movement, camera motion, or voice from up to 9 pictures, 3 clips, and 3 audio files.
Clean Text and Brand Rendering
Legible text and brand elements are reproduced cleanly by the comfyui minimax h3 model, with natural-language instruction following that explains reference relationships.
Sage Attention Acceleration
Insert the Patch Sage Attention KJ node into the comfyui minimax h3 graph to roughly double render speed with only a minor quality impact.
Smart Resolution and Duration Grid
The comfyui minimax h3 selector derives width and height from aspect ratio and megapixels, snapped to 32-multiple dimensions and 17-frame blocks at 24fps.
Frequently Asked Questions About comfyui minimax h3
Quick answers on setting up, using, and tuning comfyui minimax h3 inside ComfyUI.
How does comfyui minimax h3 work in ComfyUI?
ComfyUI now includes native support for MiniMax H3, an open-weight omni-modal model. The comfyui minimax h3 integration lets you generate footage with embedded stereo audio from text, image, video, and audio references in one pass.
What resolution and frame rate can I get?
The comfyui minimax h3 export goes up to 2K at 24fps for roughly 15 seconds. It uses a 768px short edge, capped at 768x1344, and rounds to a multiple of 32.
Which generation modes are bundled with comfyui minimax h3?
Three templates are included: text-to-video, image-to-video with optional start/end frame control, and reference-to-video for locking character, style, motion, camera, or voice.
Does the output include audio?
Yes. The comfyui minimax h3 model produces stereo audio — voice, effects, and music — alongside the video in a single pass, with everything time-aligned in one MP4.
What do I need to do to start?
Update to ComfyUI 0.30.0+, open Template Library > Video, pick a comfyui minimax h3 template, and follow the prompt to download models from the Hugging Face Comfy-Org/MiniMax-H3 repository.
Can I make the comfyui minimax h3 generation faster?
Yes. With SageAttention and KJNodes installed, add a Patch Sage Attention KJ node between UNETLoader and BasicGuider in the comfyui minimax h3 graph to cut render time nearly in half.
Kick Off Your comfyui minimax h3 Video Creations
Launch the comfyui minimax h3 integration to generate open-weight clips with embedded stereo audio and full control — T2V, I2V & R2V are ready.
