Feedback
AI Ad Video Example
Loading...
comfyui minimax h3
From text, images, or reference clips, the comfyui minimax h3 workflow renders 2K videos in ComfyUI with synchronized audio.
All Tools
Discover our comprehensive AI-powered animation toolkit
MiniMax H3
MiniMax H3 AI Video Generator
Seedance 2.5
The Future of AI Video Is Here.

Seedance 2.0
The Future of AI Video Is Here.

Kling 3.0
Next-Gen AI Video Generator
Grok Video Generator
Create Videos from Text or Images with AI
MiniMax H3 video generator
MiniMax H3 AI Video Generator

Seedance 2.0
The Future of AI Video Is Here.

Kling Motion Control
Turn reference images into amazing motion videos in minutes
What Makes the comfyui minimax h3 Workflow a Strong Choice
Through this ComfyUI port of MiniMax H3, the comfyui minimax h3 workflow brings omni-modal generation to your local setup. It handles text, visuals, and sound together, creating videos up to 2K at 24fps for about 15 seconds — with native stereo audio and adjustable node parameters.
- Audio Built into Every RenderDialogue, effects, and music are produced alongside the visuals in one MP4, perfectly synced by the comfyui minimax h3 workflow without extra editing steps.
- Full Control via Open WeightsBecause the comfyui minimax h3 model runs locally, you adjust resolution, duration, and every diffusion parameter directly — no API quotas or hidden restrictions.
- Mix Multiple Reference TypesCombining text, images, video, and audio references in a single generation lets you lock in a character, look, motion, camera movement, or voice using the comfyui minimax h3 nodes.
Getting Started with the comfyui minimax h3 Workflow
A simple three-step process to produce open-weight videos with built-in audio via the comfyui minimax h3 workflow.
Key Capabilities of the comfyui minimax h3 Workflow
A full local video production stack comes together in the comfyui minimax h3 workflow, with native ComfyUI templates, omni-modal open-weight generation, reference-driven control, and Sage Attention acceleration.
Native ComfyUI Templates
The comfyui minimax h3 template library includes text-to-video, image-to-video, and reference-to-video examples, each ready to run in its own mode.
Unified Omni-Modal Understanding
Text, images, video, and audio are processed together inside the comfyui minimax h3 model, so you can blend every reference type into one generation.
Precise Reference Locking
Lock identity, style, motion, camera move, or voice from references — up to 9 images, 3 videos, and 3 audio clips through the comfyui minimax h3 R2V node.
Crisp Text and Brand Rendering
The comfyui minimax h3 model renders spelled-out text and logos cleanly, and follows instructions that describe reference relationships in natural language.
Faster with Sage Attention
Adding a Patch Sage Attention KJ node to the comfyui minimax h3 workflow nearly doubles speed while keeping quality largely intact.
Flexible Resolution and Duration Grid
The comfyui minimax h3 Resolution Selector calculates dimensions from aspect ratio and megapixels, snapped to the model's 32-multiple grid and 17-frame-per-block duration at 24fps.
Frequently Asked Questions About comfyui minimax h3
Find quick answers about running the comfyui minimax h3 workflow in ComfyUI, including setup, output, and audio questions.
What does the comfyui minimax h3 workflow do?
The comfyui minimax h3 workflow is ComfyUI's official way to run MiniMax H3, an open-weight omni-modal model. It can take text, images, clips, or sound and return a video with synchronized stereo audio in one pass.
What quality can I expect from comfyui minimax h3?
The comfyui minimax h3 workflow produces up to 2K resolution at 24fps for about 15 seconds, using a native canvas with a 768px short edge capped at 768x1344 pixels and rounded to a multiple of 32.
Which modes are available in the comfyui minimax h3 template library?
You get three ready-made templates: text-to-video, image-to-video with optional first/last-frame control, and reference-to-video that locks character, style, motion, camera, or voice.
Does the comfyui minimax h3 model really generate audio?
Yes. Voice, sound effects, and music are synthesized together with the visuals in one pass, then synced into a single MP4 file by the comfyui minimax h3 workflow.
How do I start using comfyui minimax h3?
Update ComfyUI to 0.30.0 or later, open Template Library > Video, select a comfyui minimax h3 workflow, and follow the pop-up to download models from the Hugging Face Comfy-Org/MiniMax-H3 repository.
Can I accelerate comfyui minimax h3 generation?
Install SageAttention and KJNodes, then insert a Patch Sage Attention KJ node between the UNETLoader and BasicGuider in the comfyui minimax h3 workflow to roughly double speed.
Begin Generating with the comfyui minimax h3 Workflow Today
Run MiniMax H3 locally in ComfyUI with native stereo audio, open weights, and full parameter control — text-to-video, image-to-video, and reference-to-video workflows are ready when you are.
