#128 · Primary category: Image Generation
ComfyUI-Gemini
Using Gemini in ComfyUI
Project last updated:05/22/24
GitHub Stars
788
Forks
66
Contributors
1
License
GPL-3.0
Why we included this project
Google's Gemini models plug into ComfyUI through this custom node pack, which adds nodes for generating prompts, describing images, and running multi-turn chat. It covers the text-only Gemini Pro, the vision model, and the newer 1.5 Pro, which adds system instructions and single-file uploads for audio, PDFs, and text. The batch-labeling workflow is useful if you train LoRAs, since Gemini Pro Vision can tag images automatically. Because the API key can be set as an environment variable, shared workflows don't leak credentials. It's a clean way to bring a capable external model into an existing image-generation setup.
Articles for this project
No articles for this project yet.
To suggest a topic or contribute an article, contact us.
Related projects in this category
stable-diffusion-webui
Stable Diffusion web UI
ComfyUI
The most powerful and modular diffusion model GUI, api and backend with a graph/nodes interface.
diffusers
🤗 Diffusers: State-of-the-art diffusion models for image, video, and audio generation in PyTorch.
upscayl
🆙 Upscayl - #1 Free and Open Source AI Image Upscaler for Linux, MacOS and Windows.
InvokeAI
InvokeAI is a leading open-source creative engine for Stable Diffusion, offering an industry-leading web UI for generating and editing visual media.