Generate high-quality images from text prompts using Google's Gemini model through the MCP protocol.
Generate high-quality images from text prompts using Google's Gemini model through the MCP protocol. Exposed over MCP by the gemini image generator mcp server, that capability becomes something an assistant can invoke while it works, not something you go and do afterwards.
This MCP server allows any AI assistant to generate images using Google's Gemini AI model. The server handles prompt engineering, text-to-image conversion, filename generation, and local image storage, making it easy to create and manage AI-generated images through any MCP client.
Everything the assistant can do here goes through one of these:
Prerequisites — The Prerequisites tool exposed by this serverInstallation — The Installation tool exposed by this serverYou will need 2 environment variables: GEMINI_API_KEY, OUTPUT_IMAGE_PATH. The server will not start without them, which is usually why the tools fail to appear on a first run. Keep credentials in your client's env block or a secrets manager rather than in a file you might commit.
Installation goes through your MCP client rather than a global install: point it at @smithery/cli on npm and it is fetched when the client starts. The copy-paste blocks for Claude Desktop, Claude Code and Cursor are further down this page.
Among the AI and media services options, the useful question is rarely "what can it do" but "what does it cost you to run" — permissions, credentials, and how much of your context its toolset consumes. Gemini Image Generator's toolset — Prerequisites, Installation — is a fair guide to whether it matches your workflow. It is maintained by qhdrl12; worth a glance at recent repository activity before you build anything load-bearing on it.
This entry was verified against Gemini Image Generator's own documentation before publication; SyncDev keeps the directory reviewed rather than auto-generated.
| Tool | What it does |
|---|---|
| Prerequisites | The Prerequisites tool exposed by this server. |
| Installation | The Installation tool exposed by this server. |
{
"mcpServers": {
"server-gemini-image-generator": {
"command": "npx",
"args": ["-y", "@smithery/cli"],
"env": {
"GEMINI_API_KEY": "your-value",
"OUTPUT_IMAGE_PATH": "your-value"
}
}
}
}Add to claude_desktop_config.json, then restart Claude Desktop.
| Variable | Description | Required |
|---|---|---|
| GEMINI_API_KEY | Credential the server authenticates with. | Yes |
| OUTPUT_IMAGE_PATH | Filesystem location the server is allowed to use. | Optional |
Build a programmable telecommunications stack for connecting telephony services with the Internet via a cloud-based utility.
Search built for AI, not humans — semantic web search that returns model-ready content, plus code context.
Answers, not links — delegate questions to Perplexity's search-grounded models and get cited responses back.
Give your assistant a voice — text-to-speech, voice cloning and audio tools from the ElevenLabs API.
Give your assistant a real code sandbox — isolated cloud VMs for actually running the code it writes.
The ML hub in your context window — search models, datasets, papers and run Spaces from the official server.