#234 · Primary category: Video & Animation

MuseV

diffusion human-video-generation image2video infinite-length musev video-generation

MuseV: Infinite-length and High Fidelity Virtual Human Video Generation with Visual Conditioned Parallel Denoising

Project last updated:06/28/24

GitHub Stars

2.8K

Forks

303

Contributors

10

License

Other

Why we included this project

Teams building virtual humans and character-driven video are the audience for MuseV, a diffusion-based framework that tackles one of the harder problems in this niche: keeping a consistent subject on screen for long stretches. Its Visual Conditioned Parallel Denoising scheme removes the usual length limit, which matters when you need extended talking-head or performance clips rather than short loops. Because it sits inside the Stable Diffusion ecosystem, you can bring your own base models, LoRA weights, and ControlNet setups, and it handles image-to-video, text-to-image-to-video, and video-to-video inputs. Multi-reference conditioning through IPAdapter and ReferenceNet-style methods helps hold a character's identity across frames. MuseV also connects with the companion MuseTalk and MusePose tools, so it can anchor a fuller virtual human pipeline instead of working as an isolated generator.

Articles for this project

No articles for this project yet.

To suggest a topic or contribute an article, contact us.

Related projects in this category