#234 · Primary category: Video & Animation
MuseV
MuseV: Infinite-length and High Fidelity Virtual Human Video Generation with Visual Conditioned Parallel Denoising
Project last updated:06/28/24
GitHub Stars
2.8K
Forks
303
Contributors
10
License
Other
Why we included this project
Teams building virtual humans and character-driven video are the audience for MuseV, a diffusion-based framework that tackles one of the harder problems in this niche: keeping a consistent subject on screen for long stretches. Its Visual Conditioned Parallel Denoising scheme removes the usual length limit, which matters when you need extended talking-head or performance clips rather than short loops. Because it sits inside the Stable Diffusion ecosystem, you can bring your own base models, LoRA weights, and ControlNet setups, and it handles image-to-video, text-to-image-to-video, and video-to-video inputs. Multi-reference conditioning through IPAdapter and ReferenceNet-style methods helps hold a character's identity across frames. MuseV also connects with the companion MuseTalk and MusePose tools, so it can anchor a fuller virtual human pipeline instead of working as an isolated generator.
Articles for this project
No articles for this project yet.
To suggest a topic or contribute an article, contact us.
Related projects in this category
yt-dlp
A feature-rich command-line audio/video downloader
MoneyPrinterTurbo
Generate HD short videos from a topic or keyword with an automated AI workflow.
Deep-Live-Cam
real time face swap and one-click video deepfake with only a single image
manim
Animation engine for explanatory math videos
anime
JavaScript animation engine