#25 · Primary category: 3D Generation & Asset Creation

Diffuman4D

3d-computer-vision 3d-vision avatar-generation avatar-generator computer-vision

[ICCV 2025] Diffuman4D: 4D Consistent Human View Synthesis from Sparse-View Videos with Spatio-Temporal Diffusion Models

Project last updated:04/10/26

GitHub Stars

632

Forks

36

Contributors

2

License

Apache-2.0

Why we included this project

Diffuman4D tackles a genuinely hard problem in human view synthesis: keeping a reconstructed person consistent across time when you only have a handful of synchronized cameras. It pairs 3D Gaussian splatting with a spatio-temporal diffusion model, so rendered frames stay stable instead of flickering as the viewpoint moves. The repo is practical to try, with a quick-start install, example data from the DNA-Rendering dataset, and an interactive browser demo, so you can test it on your own sparse-view captures without much setup. Researchers following recent human view synthesis work will also find the EasyVolcap-based pipeline clearly documented, and the re-annotated multi-view labels are a useful addition for training and evaluation.

Articles for this project

No articles for this project yet.

To suggest a topic or contribute an article, contact us.

Related projects in this category