#128 · Primary category: Video & Animation
echomimic
[AAAI 2025] EchoMimic: Lifelike Audio-Driven Portrait Animations through Editable Landmark Conditioning
Project last updated:04/07/26
GitHub Stars
4.3K
Forks
467
Contributors
6
License
Apache-2.0
Why we included this project
EchoMimic turns a still portrait and a voice track into a video of the person speaking and moving their head. What sets it apart from plain lip-sync tools is the editable landmark-conditioning stage: on top of matching the audio, you can steer the expression, the pose, or individual facial movements. That flexibility means one codebase covers straightforward audio-to-video talking heads and the trickier audio-plus-landmark hybrids people hit in production. The repo includes inference scripts, accelerated models, pretrained weights for English and Mandarin, and Gradio and ComfyUI interfaces, so a small team can get it running instead of just reading about it. If you are building talking-avatar pipelines, virtual presenters, or dubbing work, this is a practical reference.
Articles for this project
No articles for this project yet.
To suggest a topic or contribute an article, contact us.
Related projects in this category
yt-dlp
A feature-rich command-line audio/video downloader
MoneyPrinterTurbo
Generate HD short videos from a topic or keyword with an automated AI workflow.
Deep-Live-Cam
real time face swap and one-click video deepfake with only a single image
manim
Animation engine for explanatory math videos
anime
JavaScript animation engine