STDIO MCP server for OpenAI and Gemini image generation
If you already use Image, the image mcp server is the piece that lets your assistant work with it directly. STDIO MCP server for OpenAI and Gemini image generation.
Setup follows the usual MCP pattern — install or clone the server, register it in your client's configuration file, restart the client. The configuration blocks on this page cover the common clients.
The toolset is worth reading before you wire it up, because it tells you what the integration is really for:
prompt — What the image should containfilename — Base filename without an extensiongenerate_image_openai — Generate one or more images with the OpenAI Image APIgenerate_image_gemini — Generate one or more images with the Gemini Interactions APIConfiguration is passed through the environment: OPENAI_API_KEY, GEMINI_API_KEY. Treat anything key-shaped as a real credential — scope it to the minimum the server needs, and rotate it if it ever lands in a shared config.
This sits in the AI and media services group, where several servers overlap in what they claim to do but differ sharply once you actually set them up. Image's toolset — prompt, filename, generate_image_openai and 1 more — is a fair guide to whether it matches your workflow. It is maintained by s4shibam; worth a glance at recent repository activity before you build anything load-bearing on it.
This entry was verified against Image's own documentation before publication; SyncDev keeps the directory reviewed rather than auto-generated.
| Tool | What it does |
|---|---|
| prompt | What the image should contain |
| filename | Base filename without an extension |
| generate_image_openai | Generate one or more images with the OpenAI Image API |
| generate_image_gemini | Generate one or more images with the Gemini Interactions API |
{
"mcpServers": {
"image-gen-mcp": {
"command": "npx",
"args": ["-y", "@s4shibam/image-gen-mcp"],
"env": {
"OPENAI_API_KEY": "your-openai-api-key",
"GEMINI_API_KEY": "your-gemini-api-key"
}
}
}
}Configuration as documented by the project. Restart the client after saving.
| Variable | Description | Required |
|---|---|---|
| OPENAI_API_KEY | Credential the server authenticates with. | Yes |
| GEMINI_API_KEY | Credential the server authenticates with. | Yes |
Build a programmable telecommunications stack for connecting telephony services with the Internet via a cloud-based utility.
Search built for AI, not humans — semantic web search that returns model-ready content, plus code context.
Answers, not links — delegate questions to Perplexity's search-grounded models and get cited responses back.
Give your assistant a voice — text-to-speech, voice cloning and audio tools from the ElevenLabs API.
Give your assistant a real code sandbox — isolated cloud VMs for actually running the code it writes.
The ML hub in your context window — search models, datasets, papers and run Spaces from the official server.