Gemini Image GeneratorSAFE
MCP server for AI image generation and editing using Google's Gemini Flash models. Create images from text prompts with intelligent filename generation and strict text exclusion. Supports text-to-image generation with future expansion to image editing capabilities.
Overview
From the repository's own README, as read at the audited commit. Badges and raw HTML are left out.
[](https://mseep.ai/app/qhdrl12-mcp-server-gemini-image-generator) [](https://smithery.ai/server/@qhdrl12/mcp-server-gemini-image-gen)
Generate high-quality images from text prompts using Google's Gemini model through the MCP protocol.
Overview
This MCP server allows any AI assistant to generate images using Google's Gemini AI model. The server handles prompt engineering, text-to-image conversion, filename generation, and local image storage, making it easy to create and manage AI-generated images through any MCP client.
Features
- Text-to-image generation using Gemini 2.0 Flash
- Image-to-image transformation based on text prompts
- Support for both file-based and base64-encoded images
- Automatic intelligent filename generation based on prompts
- Automatic translation of non-English prompts
- Local image storage with configurable output path
- Strict text exclusion from generated images
- High-resolution image output
- Direct access to both image data and file path
Available MCP Tools
The server provides the following MCP tools for AI assistants:
1. generate_image_from_text
Creates a new image from a text prompt description.
generate_image_from_text(prompt: str) -> Tuple[bytes, str]
Parameters:
prompt: Text description of the image you want to generate
Returns:
- A tuple containing:
- Raw image data (bytes)
- Path to the saved image file (str)
This dual return format allows AI assistants to either work with the image data directly or reference the saved file pa
f4a38489855bOBSERVED · 2026-10-08Connect
Built from this server's own package name, version and transport as found in its source — not copied from anyone's documentation, so it cannot drift against a page we do not control. Replace the environment placeholders with a token scoped to the least it needs.
claude mcp add mcp-server-gemini-image-generator --env GEMINI_API_KEY=${GEMINI_API_KEY} -- uvx mcp-server-gemini-image-generator{
"mcpServers": {
"mcp-server-gemini-image-generator": {
"command": "uvx",
"args": [
"mcp-server-gemini-image-generator"
],
"env": {
"GEMINI_API_KEY": "${GEMINI_API_KEY}"
}
}
}
}Exposed tools (3)
3 read · 0 write · 0 destructive.
| Tool | Risk | Description |
|---|---|---|
generate_image_from_text | read | Generate an image based on the given text prompt using Google |
transform_image_from_encoded | read | Transform an existing image based on the given text prompt using Google |
transform_image_from_file | read | Transform an existing image file based on the given text prompt using Google |
Trust audit
SAFEgrade B · trust 89/100 Nothing in the source contradicts what it says it does. Grade A is reserved for packages that have also passed the behavioural sandbox.
| Layer | What it checks | Result |
|---|---|---|
| L0 | Provenance & inventory | PASS |
| L1 | Static analysis of the code | PASS |
| L2 | Instruction surface (what it tells the agent) | PASS |
| L3 | Class-specific surface | PASS |
| L4 | Behavioural (sandbox) | SKIPPED |
What the source does
- Filesystem
- none-observed
- Network
- none-observed
- Shell
- none-observed
- Dependencies
- pinned
- Secrets in source
- none-found
Findings (5)
image_bytes = base64.b64decode(image_data)
image_data = base64.b64decode(base64_string)
examples/flying_pig_scifi_city.png
examples/pig_cute_baby_whale.png
cat > .env << 'EOF'
Gates applied: no_behavioural_pass.
f4a38489855bfull audit observations/trust-audit/mcp-server/qhdrl12__gemini-image-generator.json · Report an issue / request a re-scanAudit history
Every audit this server has had. A grade with a past is a grade somebody is still checking.
| Date | Source | Verdict | Grade | Score | Change |
|---|---|---|---|---|---|
| 2026-10-08 | f4a38489855b | SAFE | B | 89 | first audit |
Questions
What is the Gemini Image Generator MCP server?
MCP server for AI image generation and editing using Google's Gemini Flash models. Create images from text prompts with intelligent filename generation and strict text exclusion. Supports text-to-image generation with future expansion to image editing capabilities.
What tools does Gemini Image Generator expose?
3 in total: 3 read-only, 0 that write, and 0 that can delete or overwrite. Every one is listed on this page with its risk.
Is Gemini Image Generator safe to connect to an agent?
The audit found nothing in the source that contradicts what it says it does, and graded it B (89/100). Grade A is held back for packages that have also passed a sandboxed behavioural run, which is why a clean server reads B.
What credentials does Gemini Image Generator need?
It reads GEMINI_API_KEY from the environment. Give it a token scoped to the least it needs — an agent that can be talked into calling a tool can be talked into calling it with your credentials.
How does Gemini Image Generator run?
It speaks stdio, so it runs as a local process your client starts. It is published on PyPI as mcp-server-gemini-image-generator.
How current is this page?
The grade is for one exact copy of the source (f4a38489855b), read on 2026-10-08. The repository is watched and re-audited when it changes.