#128 · Primary category: Video & Animation

echomimic

aaai2025 audio-driven-portrait-animations audio-driven-talking-face human-animation talking-face-generation talking-head

[AAAI 2025] EchoMimic: Lifelike Audio-Driven Portrait Animations through Editable Landmark Conditioning

Project last updated:04/07/26

GitHub Stars

4.3K

Forks

467

Contributors

6

License

Apache-2.0

Why we included this project

EchoMimic turns a still portrait and a voice track into a video of the person speaking and moving their head. What sets it apart from plain lip-sync tools is the editable landmark-conditioning stage: on top of matching the audio, you can steer the expression, the pose, or individual facial movements. That flexibility means one codebase covers straightforward audio-to-video talking heads and the trickier audio-plus-landmark hybrids people hit in production. The repo includes inference scripts, accelerated models, pretrained weights for English and Mandarin, and Gradio and ComfyUI interfaces, so a small team can get it running instead of just reading about it. If you are building talking-avatar pipelines, virtual presenters, or dubbing work, this is a practical reference.

Articles for this project

No articles for this project yet.

To suggest a topic or contribute an article, contact us.

Related projects in this category