← all repositories

harlanhong/ACTalker

An ICCV 2025 end-to-end video diffusion framework for talking head synthesis supporting audio and expression control.

459 stars Python Image · Video · Audio
ACTalker
Not currently ranked — collecting fresh signals.
star history

ACTalker is a video diffusion model for generating realistic talking head videos from audio and expression signals. It employs masked selective state spaces modeling to achieve audio-visual synchronization and natural motion in avatar animation. The framework supports both single and multi-signal control for digital human synthesis and is intended for high-quality video generation of talking avatars.

Frequently asked

What is harlanhong/ACTalker?
An ICCV 2025 end-to-end video diffusion framework for talking head synthesis supporting audio and expression control.
Is ACTalker open source?
Yes — harlanhong/ACTalker is an open-source project tracked on heatdrop.
What language is ACTalker written in?
harlanhong/ACTalker is primarily written in Python.
How popular is ACTalker?
harlanhong/ACTalker has 459 stars on GitHub.
Where can I find ACTalker?
harlanhong/ACTalker is on GitHub at https://github.com/harlanhong/ACTalker.

heatdrop uses Google Analytics to see which pages get read — nothing else. Your call. How we handle data.