Feedback
AI Ad Video Example
Loading...
comfyui minimax h3
Use the comfyui minimax h3 workflow to run MiniMax H3 with open weights and native stereo audio, delivering up to 2K at 24fps.
All Tools
Discover our comprehensive AI-powered animation toolkit
MiniMax H3
MiniMax H3 AI Video Generator
Seedance 2.5
The Future of AI Video Is Here.

Seedance 2.0
The Future of AI Video Is Here.

Kling 3.0
Next-Gen AI Video Generator
MiniMax H3 video generator
MiniMax H3 AI Video Generator

Seedance 2.0
The Future of AI Video Is Here.

Kling Motion Control
Turn reference images into amazing motion videos in minutes
Sora2 AI
Advanced AI Video Generator for High-Quality Videos
Key Perks of the comfyui minimax h3 Workflow
The comfyui minimax h3 workflow bundles MiniMax's omni-modal open-weight model for ComfyUI. It processes text, images, video, and audio in a unified context, producing clips with native stereo sound — speech, effects, and music all at once. You get up to 2K resolution at 24fps for around 15 seconds, plus granular node-level tweaking.
- Integrated Stereo SoundtrackVoice, effects, and music are produced together with your footage in a single MP4, perfectly timed by the comfyui minimax h3 process.
- Complete Local CustomizationExecute the comfyui minimax h3 model on your own machine, adjusting resolution, length, and each diffusion setting freely — without API restrictions.
- Multi-Modal Reference IntegrationFeed in text, images, clips, and audio simultaneously to lock a character, aesthetic, movement, camera path, or voice through the comfyui minimax h3 nodes.
Quick Start Guide for the comfyui minimax h3 Workflow
Create open-weight videos with built-in audio in just three steps by following the comfyui minimax h3 workflow.
Notable Features of the comfyui minimax h3 Workflow
The comfyui minimax h3 workflow bundles three built-in ComfyUI templates, open-weight multi-modal generation, native stereo sound, reference-based control, and an optional Sage Attention boost for a full local video production suite.
Preconfigured Templates for Three Modes
The comfyui minimax h3 library comes with ready-made examples for text-to-video, image-to-video, and reference-to-video, each addressing a different creation mode immediately.
Unified Multimodal Understanding
With the comfyui minimax h3 model, text, images, clips, and audio are interpreted within one shared context, merging every reference type in a single generation run.
Reference-Guided Creation
Fix a character's look, a visual style, a movement, a camera angle, or a voice from reference sources — up to 9 pictures, 3 video clips, and 3 audio files using the comfyui minimax h3 R2V node.
Crisp Text and Logo Rendering
Clear text and logo details are rendered faithfully by the comfyui minimax h3 model, along with natural-language instructions that define how references relate.
Sage Attention Performance Boost
Approximately double the render speed with negligible quality impact by inserting the Patch Sage Attention KJ node into the comfyui minimax h3 workflow.
Resolution and Length Controls
The comfyui minimax h3 Resolution Selector calculates dimensions from aspect ratio and megapixels, aligning to the model's 32-pixel grid and 17-frame block duration at 24fps.
Frequently Asked Questions About the comfyui minimax h3 Workflow
Find answers to common questions about using the MiniMax H3 model within ComfyUI.
What does the comfyui minimax h3 workflow do?
This is ComfyUI's built-in support for MiniMax H3, an open-weight omni-modal generation model from MiniMax. It enables video creation with native stereo audio from text, images, clips, and audio references in one forward pass.
What video quality can this workflow reach?
The comfyui minimax h3 workflow can render clips up to 2K resolution at 24fps lasting roughly 15 seconds. The default canvas is 768px on the short side, limited to 768x1344 pixels and rounded to multiples of 32.
What creation modes come with this workflow?
The comfyui minimax h3 template set offers three variants: text-to-video (T2V), image-to-video (I2V) with optional first and last frame settings, and reference-to-video (R2V) which can fix a character, style, movement, camera, or voice.
Is audio generation included?
Absolutely — the comfyui minimax h3 model creates native stereo audio covering speech, sound effects, and music, all generated together with the video and synced into one MP4.
How can I start using this workflow?
Upgrade ComfyUI to 0.30.0 or newer, go to Template Library > Video, select a comfyui minimax h3 workflow, and use the pop-up to download models from the Comfy-Org/MiniMax-H3 repo on Hugging Face.
Is there a way to accelerate rendering?
Yes — install SageAttention and the KJNodes custom pack, then insert a Patch Sage Attention KJ node between UNETLoader and BasicGuider in the comfyui minimax h3 workflow to get around twice the render speed.
Launch Your Video Creation with comfyui minimax h3
Use MiniMax H3 directly in ComfyUI with open weights, native stereo sound, and full control over every parameter — T2V, I2V, and R2V setups are ready whenever you are.
