Feedback
AI Ad Video Example
Loading...
comfyui minimax h3
Within ComfyUI, MiniMax H3 lets you generate open-weight footage alongside its soundtrack, starting from prompts, pictures, or existing video, at up to 2K resolution and 24fps.
All Tools
Discover our comprehensive AI-powered animation toolkit
MiniMax H3
MiniMax H3 AI Video Generator
Seedance 2.5
The Future of AI Video Is Here.

Seedance 2.0
The Future of AI Video Is Here.

Kling 3.0
Next-Gen AI Video Generator
Grok Video Generator
Create Videos from Text or Images with AI
MiniMax H3 video generator
MiniMax H3 AI Video Generator

Seedance 2.0
The Future of AI Video Is Here.

Kling Motion Control
Turn reference images into amazing motion videos in minutes
Unlocking Video and Audio Generation with the comfyui minimax h3 Workflow
With the comfyui minimax h3 workflow, ComfyUI users can tap into MiniMax's open-weight, omni-modal model. In a single pass, it handles text, images, video, and audio and outputs synchronized audio built from voice, effects, and music. The result reaches 2K at 24fps for roughly 15 seconds, and every parameter is exposed through nodes.
- Built-in Stereo SoundVoice, effects, and music are created together with the visuals in a single MP4, all aligned through the comfyui minimax h3 node graph.
- Unrestricted Local ExecutionExecute the comfyui minimax h3 model on your machine, adjusting resolution, duration, and diffusion settings without hitting API quotas.
- Rich Reference IntegrationMerge text, stills, footage, and audio into a single run, preserving characters, artistic styles, movement, camera motion, or voices through the comfyui minimax h3 nodes.
3 Simple Steps with the comfyui minimax h3 Workflow
Kick off your video generation journey with the comfyui minimax h3 workflow — three easy steps to get open-weight clips with sound.
Core Features of the comfyui minimax h3 Workflow
The comfyui minimax h3 workflow combines three ComfyUI templates, open-weight multimodal generation, stereo audio, reference-driven controls, and Sage Attention acceleration to form a complete on-premises video studio.
Triple Template Collection
The comfyui minimax h3 library contains example graphs for text-to-video, image-to-video, and reference-to-video, with each mode ready to use immediately.
Unified Understanding
With a single unified context, the comfyui minimax h3 model processes text, images, video, and audio simultaneously, letting you combine any references in one output.
Reference-Powered Creation
Preserve a character, visual style, movement, camera angle, or voice from source materials — supporting up to 9 images, 3 videos, and 3 audio files through the comfyui minimax h3 R2V node.
Sharp Text and Brand Accuracy
The comfyui minimax h3 model renders legible text and logos, and follows natural language instructions to describe how references relate to each other.
Faster Rendering with Sage Attention
Add the Patch Sage Attention KJ node within the comfyui minimax h3 workflow to nearly double rendering speed while preserving quality.
Smart Resolution and Duration Controls
The Resolution Selector in the comfyui minimax h3 workflow calculates dimensions from aspect ratio and megapixels, snapping to the model's 32-multiple grid and 17-frame-per-block timing at 24fps.
Frequently Asked Questions: comfyui minimax h3
Quick answers about using MiniMax H3 through the comfyui minimax h3 workflow in ComfyUI.
What is the comfyui minimax h3 workflow used for?
It's ComfyUI's built-in support for MiniMax H3, an open-weight, omni-modal model from MiniMax. This workflow generates video that includes stereo audio from text, images, video, and audio references in a single forward pass.
What resolution and frame rate are available?
With the comfyui minimax h3 workflow, you can get up to 2K resolution at 24fps with a maximum duration of roughly 15 seconds. The native canvas has a short edge of 768px, caps at 768×1344, and snaps to multiples of 32.
What generation modes can I use?
The comfyui minimax h3 library offers three ready-made workflows: text-to-video, image-to-video with optional first and last frame settings, and reference-to-video that preserves character, style, movement, camera, or voice.
Is audio included in the video?
Absolutely. The comfyui minimax h3 model creates stereo sound — dialogue, effects, and music — simultaneously with the visual content, and the final MP4 has the audio embedded.
What steps do I need to follow?
Start by updating ComfyUI to 0.30.0 or higher, open Template Library > Video, select any comfyui minimax h3 workflow, then follow the prompt to grab models from the Comfy-Org/MiniMax-H3 repo on Hugging Face.
Are there performance optimizations?
Definitely! Install SageAttention and KJNodes, then insert a Patch Sage Attention KJ node between the UNETLoader and BasicGuider within the comfyui minimax h3 workflow to nearly double inference speed.
Ready to Generate with the comfyui minimax h3 Workflow?
Launch MiniMax H3 on your own machine, complete with stereo audio, open weights, and adjustable parameters. The comfyui minimax h3 workflow includes T2V, I2V, and R2V templates, so you can start right away.
