← all repositories
T8mars/comfyui-minimax-h3-audio-T8

MiniMax H3, But Make It a ComfyUI Post-Production Suite

This node pack turns MiniMax H3 video generation into a full ComfyUI workflow suite with lip sync, long-video stitching, and surgical face repair.

1k stars Python Image · Video · Audio
comfyui-minimax-h3-audio-T8
Collecting fresh signals — velocity needs a few days of history.
collecting data…
star history

What it does

This is a ComfyUI node collection for MiniMax H3 that bundles text-to-video, image-to-video, audio-driven generation, and reference-based controls into ready-made workflows. It handles the entire lifecycle: generation, acceleration, long-video segmentation with breakpoint recovery, video outpainting, face-refinement windows, and final upscaling or H.264 export. Think of it as a post-production suite that happens to live inside your inference UI.

The interesting bit

The project treats video generation less like a one-shot prompt and more like an editable film strip. Features such as Vocal Lock isolate a voice track to drive animation while mixing the full song only at export, and Face Refine Window lets you surgically regenerate just the bad frames without discarding the original audio. Even the video outpainting workflow forces a human-in-the-loop preview before it commits to rendering the rest of the clip.

Key highlights

  • 327 nodes covering generation (T2VA, I2VA, FL2VA), audio control, lip sync, and H3-World camera/character steering via WASD keys.
  • Long-video support with multi-keyframe segmentation, serial queue generation, and optional Color Match to dampen seam jumps between segments.
  • Video outpainting with two decode modes—joint_decode rebuilds the whole canvas to avoid hard edges, while preserve_source keeps original pixels at the cost of potential seams.
  • Acceleration experiments including OpenVDN 8-step, progressive sampling, and a TensorRT VAE backend that the authors note is not guaranteed to speed up the full pipeline.
  • Face Refine workflows that accept or reject per-window repairs via a recoverable manifest, preserving the original audio track throughout.

Caveats

  • Many advanced features (OpenVDN, DLSS frame interpolation, TRT VAE, progressive sampling) are marked as experimental and explicitly cannot be stacked together.
  • The README repeatedly warns that acceleration tricks do not always improve end-to-end time; for example, TRT VAE was slightly slower than native in short-film tests when load and save are included.
  • Video outpainting and progressive sampling can introduce banding, repeated textures, or audio discrepancies, and the authors make no blanket promise of seamless results across all materials.

Verdict

Worth a look if you are already running MiniMax H3 inside ComfyUI and want production-grade controls like segmented long-video editing and audio-preserving face repair. Skip it if you are after a simple, single-click video generator—this expects you to manage models, workflows, and manual review steps.

Frequently asked

What is T8mars/comfyui-minimax-h3-audio-T8?
This node pack turns MiniMax H3 video generation into a full ComfyUI workflow suite with lip sync, long-video stitching, and surgical face repair.
Is comfyui-minimax-h3-audio-T8 open source?
Yes — T8mars/comfyui-minimax-h3-audio-T8 is an open-source project tracked on heatdrop.
What language is comfyui-minimax-h3-audio-T8 written in?
T8mars/comfyui-minimax-h3-audio-T8 is primarily written in Python.
How popular is comfyui-minimax-h3-audio-T8?
T8mars/comfyui-minimax-h3-audio-T8 has 1k stars on GitHub.
Where can I find comfyui-minimax-h3-audio-T8?
T8mars/comfyui-minimax-h3-audio-T8 is on GitHub at https://github.com/T8mars/comfyui-minimax-h3-audio-T8.

heatdrop uses Google Analytics to see which pages get read — nothing else. Your call. How we handle data.