#811 · Primary category: AI Agents & Automation
gemini-skill
gemini drawing MCP & skill through browser, can be used in openclaw or any agent that supports MCP. Gemini画图 MCP和sill,支持龙虾或任何agent使用٩(๑>◡<๑)۶
Project last updated:08/01/26
GitHub Stars
830
Forks
119
Contributors
4
License
MIT
Why we included this project
Agents that can only talk are half useful, and this project gives them Gemini's image generation and chat through the browser, without paying for a separate API tier. It runs a standard MCP server plus a skill, so any MCP-capable client can call tools that drive a real logged-in browser session over CDP. The browser daemon handles the tedious parts, like keeping the session alive, evading automation detection, downloading full-resolution images, and stripping watermarks. That makes it a drop-in way to add Gemini drawing and image extraction to an existing agent setup, and the OpenAI-compatible Atlas Cloud provider offers a pure API route if you want to compare browser automation against it. Just know it needs a locally installed browser and a logged-in Google account, so it suits environments where that setup is acceptable.
Articles for this project
No articles for this project yet.
To suggest a topic or contribute an article, contact us.
Related projects in this category
openclaw
Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞
hermes-agent
The agent that grows with you
n8n
Fair-code workflow automation platform with native AI capabilities. Combine visual building with custom code, self-host or cloud, 400+ integrations.
deepseek-harness
DeepSeek Harness: Everything is a Plugin.
AutoGPT
AutoGPT is the vision of accessible AI for everyone, to use and to build on. Our mission is to provide the tools, so that you can focus on what matters.