#15 · Primary category: Image Generation
Sana
SANA: Efficient High-Resolution Image Synthesis with Linear Diffusion Transformer
Project last updated:08/27/26
GitHub Stars
8.9K
Forks
715
Contributors
23
License
Apache-2.0
Why we included this project
Sana is NVIDIA's text-to-image project built around a linear diffusion transformer, a design that keeps image quality high while cutting memory and compute well below typical diffusion models. A 0.6B model can generate detailed images on a single consumer GPU, and the 4-bit variant runs in under 8GB, so a small team can actually run and fine-tune it without renting a cluster of H100s. The repository covers both sides of the work: training and inference guides for getting models up and running, plus companion models for controllable generation, video, streaming, and reinforcement-learning alignment. For teams that want production-grade image synthesis on hardware they already own, this is a practical place to start.
Articles for this project
No articles for this project yet.
To suggest a topic or contribute an article, contact us.
Related projects in this category
stable-diffusion-webui
Stable Diffusion web UI
ComfyUI
The most powerful and modular diffusion model GUI, api and backend with a graph/nodes interface.
diffusers
🤗 Diffusers: State-of-the-art diffusion models for image, video, and audio generation in PyTorch.
upscayl
🆙 Upscayl - #1 Free and Open Source AI Image Upscaler for Linux, MacOS and Windows.
InvokeAI
InvokeAI is a leading open-source creative engine for Stable Diffusion, offering an industry-leading web UI for generating and editing visual media.