Feedback
AI Ad Video Example
Loading...
comfyui minimax h3
Create 2K/24fps open-weight videos in ComfyUI with MiniMax H3 and native stereo audio, using prompts, images, or clips through the comfyui minimax h3 workflow.
All Tools
Discover our comprehensive AI-powered animation toolkit
MiniMax H3
MiniMax H3 AI Video Generator

Seedance 2.0
The Future of AI Video Is Here.

Kling 3.0
Next-Gen AI Video Generator
Grok Video Generator
Create Videos from Text or Images with AI
MiniMax H3 video generator
MiniMax H3 AI Video Generator

Seedance 2.0
The Future of AI Video Is Here.

Kling Motion Control
Turn reference images into amazing motion videos in minutes

Suno AI Music Generator
Create Professional Music with AI
Why the comfyui minimax h3 Workflow Is a Must for Video Creators
By loading MiniMax's open-weight omni-modal model inside ComfyUI, the comfyui minimax h3 workflow lets a single context understand text, visuals, motion, and sound at once. It renders video together with stereo audio — dialogue, SFX, and music in one pass — and supports outputs up to 2K at 24 fps for roughly 15 seconds, with each node fully adjustable.
- Stereo Audio in the Same PassSpeech, effects, and music are produced in the same MP4 as the footage, so the sound and picture remain aligned without extra steps.
- Local Control Without API WallsDeploy the MiniMax H3 model on your own machine and tune resolution, duration, and diffusion settings directly — no external service limits.
- Rich Multi-Input Reference SupportFeed text, pictures, clips, or audio into a single run, and the MiniMax H3 nodes lock down a character's look, a visual style, movement, camera motion, or vocal tone.
Three Quick Steps to Run the comfyui minimax h3 Workflow
Follow this straightforward guide to produce open-weight video with built-in audio using the comfyui minimax h3 workflow in just three steps.
What the comfyui minimax h3 Workflow Brings to ComfyUI Video Creation
The comfyui minimax h3 workflow combines three bundled ComfyUI templates, open-weight omni-modal generation, native stereo audio, reference-based control, and optional Sage Attention acceleration into one local video production suite.
Ready-Made Template Trio
The comfyui minimax h3 library includes three prebuilt examples — text-to-video, image-to-video, and reference-to-video — so every generation mode is available immediately.
Unified Multimodal Understanding
Within one context, the MiniMax H3 model processes text, pictures, footage, and sound together, so you can blend all reference types in a single run.
Reference-Based Creative Control
Anchor a character's look, a visual style, motion, camera movement, or voice from reference input — up to nine images, three videos, and three audio files through the MiniMax H3 R2V node.
Sharp Text and Logo Reproduction
The MiniMax H3 model renders legible on-screen text and brand marks reliably, while its instruction following lets you describe reference relationships in plain language.
Sage Attention-Accelerated Rendering
Attach the Patch Sage Attention KJ node to the workflow and you can nearly double rendering speed with only minimal quality impact.
Flexible Output Dimensions and Length
The MiniMax H3 resolution selector derives width and height from aspect ratio and megapixels, respecting a 32-multiple layout and the model's 17-frame-per-block duration at 24 fps.
Answers to Common comfyui minimax h3 Workflow Questions
Straightforward answers about running MiniMax H3 in ComfyUI, covering quality, modes, audio, speed, and setup.
What exactly is the comfyui minimax h3 workflow?
It's ComfyUI's built-in implementation of MiniMax H3, an open-weight omni-modal model from MiniMax. In one forward pass, the workflow creates video together with stereo audio from text, images, video, or audio references.
What resolution and frame rate should I expect from MiniMax H3?
It generates up to 2K footage at 24 fps for around 15 seconds. The default canvas starts with a 768px short edge, maxes out at 768x1344, and stays aligned to the 32-pixel multiple.
What generation modes are included in the comfyui minimax h3 workflow?
It comes with three template types: T2V for text-to-video, I2V for image-to-video (with optional first/last frame control), and R2V for reference-to-video, which pins down character, style, motion, camera, or voice.
Does the MiniMax H3 model create sound along with video?
Absolutely — the workflow generates stereo audio, including dialogue, effects, and music, all modeled at the same time as the footage and delivered as one synchronized MP4.
What is the quickest way to start using the comfyui minimax h3 workflow?
Upgrade ComfyUI to 0.30.0 or newer, go to Template Library > Video, choose the workflow, and follow the prompt to download the model from the Comfy-Org/MiniMax-H3 Hugging Face repo.
Can the comfyui minimax h3 workflow be sped up?
Yes. Install SageAttention and KJNodes, then place a Patch Sage Attention KJ node between the UNETLoader and BasicGuider in the workflow. That can roughly double the generation speed.
Begin Your Video AI Journey with the comfyui minimax h3 Workflow
Fire up the comfyui minimax h3 workflow on your own ComfyUI setup, enjoy stereo audio, open weights, and total parameter control — with T2V, I2V, and R2V workflows ready to use.
