← all repositories

smthemex/ComfyUI_Sonic

A ComfyUI custom node for audio-driven portrait animation using the Sonic method with Whisper and Stable Video Diffusion models.

1.1k stars Python Image · Video · Audio
ComfyUI_Sonic
Not currently ranked — collecting fresh signals.
star history

This repository provides a ComfyUI integration for Sonic, a portrait animation technique that animates a person’s face driven by audio input. It leverages OpenAI’s Whisper-tiny model for audio processing and Stability AI’s Stable Video Diffusion (SVD) model for video generation. Users can input a portrait image and audio to generate synchronized talking-head animations.

Frequently asked

What is smthemex/ComfyUI_Sonic?
A ComfyUI custom node for audio-driven portrait animation using the Sonic method with Whisper and Stable Video Diffusion models.
Is ComfyUI_Sonic open source?
Yes — smthemex/ComfyUI_Sonic is open source, released under the MIT license.
What language is ComfyUI_Sonic written in?
smthemex/ComfyUI_Sonic is primarily written in Python.
How popular is ComfyUI_Sonic?
smthemex/ComfyUI_Sonic has 1.1k stars on GitHub.
Where can I find ComfyUI_Sonic?
smthemex/ComfyUI_Sonic is on GitHub at https://github.com/smthemex/ComfyUI_Sonic.

heatdrop uses Google Analytics to see which pages get read — nothing else. Your call. How we handle data.