comfyui minimax h3
Produce open-weight clips with built-in stereo sound via the comfyui minimax h3 setup.
AI Video Prompt Generator
10s

Feedback

AI Ad Video Example

Loading...

comfyui minimax h3

Within ComfyUI, MiniMax H3 lets you generate open-weight footage alongside its soundtrack, starting from prompts, pictures, or existing video, at up to 2K resolution and 24fps.

All Tools

Discover our comprehensive AI-powered animation toolkit

Unlocking Video and Audio Generation with the comfyui minimax h3 Workflow

With the comfyui minimax h3 workflow, ComfyUI users can tap into MiniMax's open-weight, omni-modal model. In a single pass, it handles text, images, video, and audio and outputs synchronized audio built from voice, effects, and music. The result reaches 2K at 24fps for roughly 15 seconds, and every parameter is exposed through nodes.

  • Built-in Stereo Sound
    Voice, effects, and music are created together with the visuals in a single MP4, all aligned through the comfyui minimax h3 node graph.
  • Unrestricted Local Execution
    Execute the comfyui minimax h3 model on your machine, adjusting resolution, duration, and diffusion settings without hitting API quotas.
  • Rich Reference Integration
    Merge text, stills, footage, and audio into a single run, preserving characters, artistic styles, movement, camera motion, or voices through the comfyui minimax h3 nodes.

3 Simple Steps with the comfyui minimax h3 Workflow

Kick off your video generation journey with the comfyui minimax h3 workflow — three easy steps to get open-weight clips with sound.

Core Features of the comfyui minimax h3 Workflow

The comfyui minimax h3 workflow combines three ComfyUI templates, open-weight multimodal generation, stereo audio, reference-driven controls, and Sage Attention acceleration to form a complete on-premises video studio.

Triple Template Collection

The comfyui minimax h3 library contains example graphs for text-to-video, image-to-video, and reference-to-video, with each mode ready to use immediately.

Unified Understanding

With a single unified context, the comfyui minimax h3 model processes text, images, video, and audio simultaneously, letting you combine any references in one output.

Reference-Powered Creation

Preserve a character, visual style, movement, camera angle, or voice from source materials — supporting up to 9 images, 3 videos, and 3 audio files through the comfyui minimax h3 R2V node.

Sharp Text and Brand Accuracy

The comfyui minimax h3 model renders legible text and logos, and follows natural language instructions to describe how references relate to each other.

Faster Rendering with Sage Attention

Add the Patch Sage Attention KJ node within the comfyui minimax h3 workflow to nearly double rendering speed while preserving quality.

Smart Resolution and Duration Controls

The Resolution Selector in the comfyui minimax h3 workflow calculates dimensions from aspect ratio and megapixels, snapping to the model's 32-multiple grid and 17-frame-per-block timing at 24fps.

FAQ

Frequently Asked Questions: comfyui minimax h3

Quick answers about using MiniMax H3 through the comfyui minimax h3 workflow in ComfyUI.

1

What is the comfyui minimax h3 workflow used for?

It's ComfyUI's built-in support for MiniMax H3, an open-weight, omni-modal model from MiniMax. This workflow generates video that includes stereo audio from text, images, video, and audio references in a single forward pass.

2

What resolution and frame rate are available?

With the comfyui minimax h3 workflow, you can get up to 2K resolution at 24fps with a maximum duration of roughly 15 seconds. The native canvas has a short edge of 768px, caps at 768×1344, and snaps to multiples of 32.

3

What generation modes can I use?

The comfyui minimax h3 library offers three ready-made workflows: text-to-video, image-to-video with optional first and last frame settings, and reference-to-video that preserves character, style, movement, camera, or voice.

4

Is audio included in the video?

Absolutely. The comfyui minimax h3 model creates stereo sound — dialogue, effects, and music — simultaneously with the visual content, and the final MP4 has the audio embedded.

5

What steps do I need to follow?

Start by updating ComfyUI to 0.30.0 or higher, open Template Library > Video, select any comfyui minimax h3 workflow, then follow the prompt to grab models from the Comfy-Org/MiniMax-H3 repo on Hugging Face.

6

Are there performance optimizations?

Definitely! Install SageAttention and KJNodes, then insert a Patch Sage Attention KJ node between the UNETLoader and BasicGuider within the comfyui minimax h3 workflow to nearly double inference speed.

Ready to Generate with the comfyui minimax h3 Workflow?

Launch MiniMax H3 on your own machine, complete with stereo audio, open weights, and adjustable parameters. The comfyui minimax h3 workflow includes T2V, I2V, and R2V templates, so you can start right away.