#196 · Primary category: Video & Animation

hallo

face-animation image-animation video-animation

Hallo: Hierarchical Audio-Driven Visual Synthesis for Portrait Image Animation

Project last updated:09/14/24

GitHub Stars

8.7K

Forks

1.1K

Contributors

15

License

MIT

Why we included this project

For believable talking-head video, this is one of the more practical open implementations to start from. Hallo takes a single portrait photo and an audio track, then produces a short clip where the face lip-syncs and moves with the speech, a task that usually sits behind expensive commercial APIs. It uses a diffusion-based architecture with hierarchical audio conditioning to keep the face identity stable while generating natural head motion and mouth shapes. Creators building avatars, dubbing pipelines, virtual presenters, or video-localization tools will find the codebase a reasonable reference for training and inference. The project ships pretrained checkpoints, a Hugging Face Space demo, and a paper describing the method, so you can check output quality before committing to deeper integration.

Articles for this project

No articles for this project yet.

To suggest a topic or contribute an article, contact us.

Related projects in this category