No project description provided

These details have not been verified by PyPI

Project description

Gemini Image Generator Server MCP server

Gemini Image Generator MCP Server

Generate high-quality images from text prompts using Google's Gemini model through the MCP protocol.

Overview

This MCP server allows any AI assistant to generate images using Google's Gemini AI model. The server handles prompt engineering, text-to-image conversion, filename generation, and local image storage, making it easy to create and manage AI-generated images through any MCP client.

Features

Text-to-image generation using Gemini 2.0 Flash
Image-to-image transformation based on text prompts
Support for both file-based and base64-encoded images
Automatic intelligent filename generation based on prompts
Automatic translation of non-English prompts
Local image storage with configurable output path
Strict text exclusion from generated images
High-resolution image output
Direct access to both image data and file path

Available MCP Tools

The server provides the following MCP tools for AI assistants:

1. `generate_image_from_text`

Creates a new image from a text prompt description.

generate_image_from_text(prompt: str) -> Tuple[bytes, str]

Parameters:

prompt: Text description of the image you want to generate

Returns:

A tuple containing:
- Raw image data (bytes)
- Path to the saved image file (str)

This dual return format allows AI assistants to either work with the image data directly or reference the saved file path.

Examples:

"Generate an image of a sunset over mountains"
"Create a photorealistic flying pig in a sci-fi city"

Example Output

This image was generated using the prompt:

"Hi, can you create a 3d rendered image of a pig with wings and a top hat flying over a happy futuristic scifi city with lots of greenery?"

Flying pig over sci-fi city

A 3D rendered pig with wings and a top hat flying over a futuristic sci-fi city filled with greenery

Known Issues

When using this MCP server with Claude Desktop Host:

Performance Issues: Using transform_image_from_encoded may take significantly longer to process compared to other methods. This is due to the overhead of transferring large base64-encoded image data through the MCP protocol.
Path Resolution Problems: There may be issues with correctly resolving image paths when using Claude Desktop Host. The host application might not properly interpret the returned file paths, making it difficult to access the generated images.

For the best experience, consider using alternative MCP clients or the transform_image_from_file method when possible.

2. `transform_image_from_encoded`

Transforms an existing image based on a text prompt using base64-encoded image data.

transform_image_from_encoded(encoded_image: str, prompt: str) -> Tuple[bytes, str]

Parameters:

encoded_image: Base64 encoded image data with format header (must be in format: "data:image/[format];base64,[data]")
prompt: Text description of how you want to transform the image

Returns:

A tuple containing:
- Raw transformed image data (bytes)
- Path to the saved transformed image file (str)

Example:

"Add snow to this landscape"
"Change the background to a beach"

3. `transform_image_from_file`

Transforms an existing image file based on a text prompt.

transform_image_from_file(image_file_path: str, prompt: str) -> Tuple[bytes, str]

Parameters:

image_file_path: Path to the image file to be transformed
prompt: Text description of how you want to transform the image

Returns:

A tuple containing:
- Raw transformed image data (bytes)
- Path to the saved transformed image file (str)

Examples:

"Add a llama next to the person in this image"
"Make this daytime scene look like night time"

Example Transformation

Using the flying pig image created above, we applied a transformation with the following prompt:

"Add a cute baby whale flying alongside the pig"

Before: Flying pig over sci-fi city

After: Flying pig with baby whale

The original flying pig image with a cute baby whale added flying alongside it

Setup

Prerequisites

Python 3.11+
Google AI API key (Gemini)
MCP host application (Claude Desktop App, Cursor, or other MCP-compatible clients)

Getting a Gemini API Key

Visit Google AI Studio API Keys page
Sign in with your Google account
Click "Create API Key"
Copy your new API key for use in the configuration
Note: The API key provides a certain quota of free usage per month. You can check your usage in the Google AI Studio

Installation

Installing via Smithery

To install Gemini Image Generator MCP for Claude Desktop automatically via Smithery:

npx -y @smithery/cli install @qhdrl12/mcp-server-gemini-image-gen --client claude

Manual Installation

Clone the repository:

git clone https://github.com/your-username/mcp-server-gemini-image-generator.git
cd mcp-server-gemini-image-generator

Create a virtual environment and install dependencies:

# Using uv (recommended)
uv venv
source .venv/bin/activate  # On Windows: .venv\Scripts\activate
uv pip install -e .

# Or using regular venv
python -m venv .venv
source .venv/bin/activate  # On Windows: .venv\Scripts\activate
pip install -e .

Set up environment variables (choose one method):

Method A: Using .env file (optional)

# Create .env file in the project root
cat > .env << 'EOF'
GEMINI_API_KEY=your-gemini-api-key-here
OUTPUT_IMAGE_PATH=/path/to/save/images
EOF

Method B: Set directly in Claude Desktop config (recommended)

Set environment variables directly in the claude_desktop_config.json (shown in configuration section below)

Configure Claude Desktop

Add the following to your claude_desktop_config.json:

macOS: ~/Library/Application Support/Claude/claude_desktop_config.json
Windows: %APPDATA%\Claude\claude_desktop_config.json

{
    "mcpServers": {
        "gemini-image-generator": {
            "command": "uv",
            "args": [
                "--directory",
                "/absolute/path/to/mcp-server-gemini-image-generator",
                "run",
                "mcp-server-gemini-image-generator"
            ],
            "env": {
                "GEMINI_API_KEY": "your-actual-gemini-api-key-here",
                "OUTPUT_IMAGE_PATH": "/absolute/path/to/your/images/directory"
            }
        }
    }
}

Important Configuration Notes:

Replace paths with your actual paths:
- Change /absolute/path/to/mcp-server-gemini-image-generator to the actual location where you cloned this repository
- Change /absolute/path/to/your/images/directory to where you want generated images to be saved
Environment Variables:
- Replace your-actual-gemini-api-key-here with your real Gemini API key from Google AI Studio
- Use absolute paths for OUTPUT_IMAGE_PATH to ensure images are saved correctly
Example with real paths:

{
    "mcpServers": {
        "gemini-image-generator": {
            "command": "uv",
            "args": [
                "--directory",
                "/Users/username/Projects/mcp-server-gemini-image-generator",
                "run",
                "mcp-server-gemini-image-generator"
            ],
            "env": {
                "GEMINI_API_KEY": "GEMINI_API_KEY",
                "OUTPUT_IMAGE_PATH": "OUTPUT_IMAGE_PATH"
            }
        }
    }
}

Usage

Once installed and configured, you can ask Claude to generate or transform images using prompts like:

Generating New Images

"Generate an image of a sunset over mountains"
"Create an illustration of a futuristic cityscape"
"Make a picture of a cat wearing sunglasses"

Transforming Existing Images

"Transform this image by adding snow to the scene"
"Edit this photo to make it look like it was taken at night"
"Add a dragon flying in the background of this picture"

The generated/transformed images will be saved to your configured output path and displayed in Claude. With the updated return types, AI assistants can also work directly with the image data without needing to access the saved files.

Testing

You can test the application by running the FastMCP development server:

fastmcp dev server.py

This command starts a local development server and makes the MCP Inspector available at http://localhost:5173/. The MCP Inspector provides a convenient web interface where you can directly test the image generation tool without needing to use Claude or another MCP client. You can enter text prompts, execute the tool, and see the results immediately, which is helpful for development and debugging.

License

MIT License

Project details

These details have not been verified by PyPI

Release history Release notifications | RSS feed

This version

0.1.0

Nov 19, 2025

Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

iflow_mcp_mcp_server_gemini_image_generator-0.1.0.tar.gz (3.1 MB view details)

Uploaded Nov 19, 2025 Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

The dropdown lists show the available interpreters, ABIs, and platforms. Enable javascript to be able to filter the list of wheel files.

iflow_mcp_mcp_server_gemini_image_generator-0.1.0-py3-none-any.whl (13.0 kB view details)

Uploaded Nov 19, 2025 Python 3

File details

Details for the file iflow_mcp_mcp_server_gemini_image_generator-0.1.0.tar.gz.

File metadata

Download URL: iflow_mcp_mcp_server_gemini_image_generator-0.1.0.tar.gz
Upload date: Nov 19, 2025
Size: 3.1 MB
Tags: Source
Uploaded using Trusted Publishing? No
Uploaded via: uv/0.6.9

File hashes

Hashes for iflow_mcp_mcp_server_gemini_image_generator-0.1.0.tar.gz
Algorithm	Hash digest
SHA256	`d98ce69362ed0ad99608505b689b6a6d4a64fdc6027c87f7f9c9cdb7bd8e38be`
MD5	`ff64e7319771ec212be3c57933711e24`
BLAKE2b-256	`c8462b2c5fe21713afeaf23744ab8eb66f2704a360139ebac12a8ffbb33f80cc`

See more details on using hashes here.

File details

Details for the file iflow_mcp_mcp_server_gemini_image_generator-0.1.0-py3-none-any.whl.

File metadata

Download URL: iflow_mcp_mcp_server_gemini_image_generator-0.1.0-py3-none-any.whl
Upload date: Nov 19, 2025
Size: 13.0 kB
Tags: Python 3
Uploaded using Trusted Publishing? No
Uploaded via: uv/0.6.9

File hashes

Hashes for iflow_mcp_mcp_server_gemini_image_generator-0.1.0-py3-none-any.whl
Algorithm	Hash digest
SHA256	`1dcf3bfa8ef82ba96dc6517c4aa171cff7630b0e5e1894045698b16313b98e3e`
MD5	`e878ddfb6f4c4fefe17c7e9564355480`
BLAKE2b-256	`2d27fc03f630975ee8045086a6cad68644059814bf1b6ccc2e2b076f509488b5`

See more details on using hashes here.

iflow-mcp_mcp-server-gemini-image-generator 0.1.0

Navigation

Verified details

Maintainers

Unverified details

Meta

Project description

Gemini Image Generator MCP Server

Overview

Features

Available MCP Tools

1. generate_image_from_text

Example Output

Known Issues

2. transform_image_from_encoded

3. transform_image_from_file

Example Transformation

Setup

Prerequisites

Getting a Gemini API Key

Installation

Installing via Smithery

Manual Installation

Configure Claude Desktop

Usage

Generating New Images

Transforming Existing Images

Testing

License

Project details

Verified details

Maintainers

Unverified details

Meta

Release history Release notifications | RSS feed

Download files

Source Distribution

Built Distribution

File details

File metadata

File hashes

File details

File metadata

File hashes

1. `generate_image_from_text`

2. `transform_image_from_encoded`

3. `transform_image_from_file`