Skip to main content
Pre-release

This release is a pre-release and may not be stable for production use.

Gemini Research MCP Server

PyPI version CI Python 3.12+ License: MIT

MCP server for AI-powered research using Gemini. Fast grounded search, URL extraction, comprehensive Deep Research, and session management.

Built on FastMCP 4.0.0b5 (beta, exact-pinned) with the modern sessionless MCP protocol, Gemini 3.7 Flash, MCP Tasks (SEP-1732), the guard-pattern elicitation flow, a BM25-compacted tool catalog, and pluggable Disk/Redis storage with a zero-configuration local default.

Architecture

Architecture

Mermaid source
flowchart TB
    subgraph Client["MCP Client"]
        Claude["Claude / Copilot"]
    end

    subgraph Server["gemini-research-mcp"]
        direction TB
        FastMCP["FastMCP 4 Server<br/>@mcp.tool()<br/>BM25SearchTransform"]
        
        subgraph Tools["Tools"]
            RW["research_web<br/>Quick lookup 5-30s"]
            RD["research_deep<br/>Autonomous 3-20min"]
            RF["research_followup<br/>Continue session"]
            RR["resume_research<br/>Recover interrupted"]
            FW["fetch_webpage<br/>Content extraction"]
            EX["export_research_session<br/>MD/JSON/DOCX"]
            LS["list_research_sessions"]
            LT["list_format_templates"]
        end

        subgraph Modules["Core Modules"]
            Quick["quick.py<br/>Web grounding"]
            Deep["deep.py<br/>Deep research agent"]
            Content["content.py<br/>SSRF protection"]
            StorageMod["storage.py<br/>Session + artifact store"]
            Templates["templates.py<br/>Format templates"]
        end
    end

    subgraph External["External Services"]
        Gemini["Google Gemini API"]
        Web["Web Sources<br/>via trafilatura"]
    end

    subgraph Storage["Persistence"]
        Disk["DiskStore<br/>XDG data directory"]
        Redis["Redis/Valkey<br/>shared multi-worker storage"]
    end

    Claude -->|"MCP Protocol"| FastMCP
    FastMCP --> Tools
    
    RW --> Quick
    RD --> Deep
    RF --> StorageMod
    RR --> StorageMod
    FW --> Content
    LT --> Templates
    
    Quick -->|"grounding"| Gemini
    Deep -->|"agentic"| Gemini
    Content -->|"httpx"| Web
    StorageMod --> Disk
    StorageMod -.-> Redis

Tools

The server exposes a BM25-compacted catalog: only the 5 tools below plus the synthetic search_tools/call_tool pair are listed by default (fastmcp.server.transforms.search.BM25SearchTransform). Utility tools (fetch_webpage, research_followup, list_research_sessions, list_format_templates, refine_research_plan, inspect_mcp_server_for_gemini) are hidden from the default listing to keep the catalog small for LLM tool-selection, but remain fully callable directly by name or via the call_tool proxy, and are discoverable by relevance through search_tools.

Tool Description Latency Visible by default
research_web Fast web search with citations 5-30 sec
research_deep Multi-step autonomous research (MCP Tasks) 3-20 min
research_deep_max Maximum-comprehensiveness Deep Research for exhaustive/high-stakes work longer-running
resume_research Resume interrupted/in-progress sessions instant
export_research_session Disk-first export to persistent Markdown, JSON, or DOCX artifacts instant
search_tools Discover hidden utility tools by relevance (BM25) instant
call_tool Proxy to invoke any hidden tool by name varies
research_followup Continue conversation after research 5-30 sec discoverable
list_research_sessions List saved research sessions instant discoverable
list_format_templates Browse report format templates instant discoverable
refine_research_plan Iterate on or approve a collaborative_planning=True plan instant-3min discoverable
fetch_webpage Extract article content from a specific URL (SSRF-protected, chunkable) 0.5-2 sec discoverable
inspect_mcp_server_for_gemini Inspect remote MCP reachability and schemas (diagnostic only) varies discoverable

research_deep / research_deep_max Deep Research parameters

Parameter Type Default Description
visualization "off" | "auto" "off" Let the agent produce and persist supporting images/charts. Images are persisted as MCP resource artifacts (research://exports/{id}), never inlined as text.
collaborative_planning boolean false Return the drafted research plan and an interaction ID instead of running the full report. Approve or iterate on the plan with `refine_research_plan(previous_interaction_id=..., decision="approve"
mcp_servers array | null null Disabled. Any non-empty value fails before network or Gemini API access because provider-side Deep Research remote MCP is not reliable.

fetch_webpage Parameters

fetch_webpage is discoverable through search_tools in the default server listing.

The fetch_webpage tool supports chunked reading for large pages and optional proxy routing:

Parameter Type Default Description
url string required HTTP/HTTPS URL to fetch
max_length integer | null null Maximum characters to return (chunk size)
start_index integer 0 Character offset for pagination
proxy_url string | null null Optional HTTP(S) proxy URL for the request

Notes:

  • SSRF protection is always applied (private/internal hosts are blocked).
  • robots.txt is checked before fetch when protego is installed.
  • When output is truncated, the response includes a continuation hint with next start_index.
  • If proxy_url is omitted, the server falls back to FETCH_PROXY_URL when set.
  • proxy_url must be a public HTTP(S) host (private/internal proxy hosts are blocked).

Install the web extra for the highest-quality fetch_webpage experience:

pip install 'gemini-research-mcp[web]'
# or
uv add 'gemini-research-mcp[web]'

Without [web], fetch_webpage still works using the built-in HTML fallback, but trafilatura extraction and protego-based robots.txt checks are unavailable.

Power User Workflow

Power User Workflow

Key insight: Gemini Deep Research runs asynchronously on Google's servers. Even if VS Code disconnects, your research continues. The resume_research tool retrieves completed work.

Features

  • Auto-Clarification: research_deep asks clarifying questions for vague queries. On the modern sessionless MCP protocol this uses a stateless guard pattern (InputRequiredResult, two independent tool calls, no server-held connection); legacy handshake clients still use MCP Elicitation (ctx.elicit())
  • Deep Research Max: research_deep_max exposes Google's Max agent for exhaustive, high-stakes, and offline research workflows
  • Collaborative Planning: research_deep(..., collaborative_planning=True) returns the drafted plan for approval before running the full report; refine or approve it with refine_research_plan
  • Visualization: visualization="auto" lets Deep Research produce supporting images, persisted as downloadable MCP resource artifacts
  • MCP Tasks: Real-time progress with streaming updates
  • Session Persistence: Research sessions are automatically saved and can be resumed later; shareable across instances with Redis (see Storage backends)
  • Persistent, Disk-First Exports: Export to Markdown, JSON, or professional DOCX with Table of Contents; artifacts survive restarts and can be shared through Redis while files are still written to disk by default
  • File Search: Search your own data alongside web using file_search_store_names
  • Fail-closed remote MCP: Deep Research rejects mcp_servers before network/provider access until Google exposes a reliable structured result contract
  • Format Instructions: Control report structure (sections, tables, tone)
  • LangChain-ready: verified consumable via langchain.mcp.MCPAdapter (LangChain 1.4.0a2) over both stdio and streamable-http - see scripts/langchain_interop_smoke.py. LangChain is never a dependency of this package.

Installation

PyPI (recommended)

pip install gemini-research-mcp
# or
uv add gemini-research-mcp

Claude Desktop (MCPB Bundle)

Download the .mcpb bundle from GitHub Releases and open it in Claude Desktop for single-click installation.

The bundle uses UV runtime - dependencies are installed automatically, no Python required.

Configuration

Variable Required Default Description
GEMINI_API_KEY Yes Google AI Studio API key
GEMINI_MODEL No gemini-3.7-flash Model for research_web
GEMINI_SUMMARY_MODEL No gemini-3.7-flash Model for session summaries, titles, and clarification (thinking level low)
DEEP_RESEARCH_AGENT No deep-research-preview-04-2026 Default agent for research_deep; accepts fast, standard, deep-research, max, deep-research-max, or exact agent IDs
FETCH_PROXY_URL No Default HTTP(S) proxy for fetch_webpage
GEMINI_RESEARCH_STORAGE_URL No redis://... URL to share sessions, exports, and (unless overridden) Tasks across multiple instances/workers. Local DiskStore + memory Tasks remain the default
FASTMCP_DOCKET_URL No Advanced Tasks backend override. Takes priority over GEMINI_RESEARCH_STORAGE_URL; unset preserves memory Tasks locally
GEMINI_RESEARCH_STORAGE_PATH No XDG data dir Custom directory for the local DiskStore
GEMINI_RESEARCH_TTL_SECONDS No backend default Override session/export TTL
GEMINI_RESEARCH_EXPORT_DIR No ~/.gemini-research/exports/ Disk-first destination when export_research_session has no output_path
GEMINI_RESEARCH_TRANSPORT No stdio stdio (default, historical) or streamable-http (opt-in, see Transports)
GEMINI_RESEARCH_HTTP_HOST No 127.0.0.1 Bind host for streamable-http. Non-loopback requires GEMINI_RESEARCH_HTTP_BEARER_TOKEN
GEMINI_RESEARCH_HTTP_PORT No 8000 Bind port for streamable-http
GEMINI_RESEARCH_HTTP_PATH No /mcp URL path for streamable-http
GEMINI_RESEARCH_HTTP_BEARER_TOKEN No Static bearer token required to call streamable-http. Never reuse GEMINI_API_KEY for this
cp .env.example .env
# Edit .env with your API key

Transports

The server defaults to stdio, matching every existing VS Code/Claude Desktop configuration - no changes required for local, single-client use.

Streamable HTTP is opt-in, for remote or multi-client/multi-worker deployments, and is sessionless (no sticky session required across calls):

# Local-only (no auth required, loopback binding):
uv run gemini-research-mcp --transport streamable-http

# Remote-accessible (bearer token required - refuses to start otherwise):
GEMINI_RESEARCH_HTTP_BEARER_TOKEN=$(openssl rand -hex 32) \
  uv run gemini-research-mcp --transport streamable-http --host 0.0.0.0 --port 8000

Binding to any non-loopback host (0.0.0.0, ::, a LAN/public IP, etc.) without GEMINI_RESEARCH_HTTP_BEARER_TOKEN set causes the server to refuse to start - this prevents accidentally exposing your Gemini API quota to the public internet. 127.0.0.1/localhost/::1 never require a token.

--transport, --host, --port, and --path CLI flags mirror the GEMINI_RESEARCH_TRANSPORT/GEMINI_RESEARCH_HTTP_HOST/GEMINI_RESEARCH_HTTP_PORT/GEMINI_RESEARCH_HTTP_PATH environment variables (CLI flags take precedence).

Storage backends

Research sessions and export artifacts are stored through a single backend-agnostic layer:

  • Local (default, no Redis required): sessions and exports use DiskStore under the XDG data directory, while FastMCP Tasks use in-process memory (GEMINI_RESEARCH_STORAGE_PATH to override) - zero configuration, single process/single machine.
  • Distributed (Redis/Valkey): set GEMINI_RESEARCH_STORAGE_URL=redis://host:6379/0 to share sessions, exports, and Tasks across multiple server instances or workers. Set FASTMCP_DOCKET_URL only when Tasks must use a different backend. Requires the distributed extra:
uv add 'gemini-research-mcp[distributed]'

Deep Research vs Deep Research Max

Google exposes Deep Research variants through the Gemini Interactions API agent field, not the regular Gemini model field:

  • research_deep uses deep-research-preview-04-2026 by default. Use it for interactive research, comparisons, investigations, and latency/cost-sensitive synthesis.
  • research_deep_max uses deep-research-max-preview-04-2026. Use it when the user explicitly asks for Max, exhaustive/comprehensive due diligence, market maps, literature reviews, board-ready reports, offline/nightly research, or maximum completeness over speed.

For Copilot and other LLM clients, the two tools are intentionally separate so Max can be selected from the tool name and description. There is no public model parameter for Deep Research, because follow-up and quick research use Gemini models while Deep Research uses Interactions agents.

Remote MCP servers for Deep Research

Disabled in v0.16.0b2. Any non-empty mcp_servers value is rejected for both research_deep and research_deep_max before remote inspection, network access, or Gemini API consumption.

The request shape is valid and Google documents MCP as a Deep Research tool, but repeated paid E2E runs completed without any mcp_server_tool_call/mcp_server_tool_result steps. One run also invented substitute evidence after failing to obtain the fixture data. See googleapis/python-genai#2126.

inspect_mcp_server_for_gemini remains available to inspect endpoint reachability, tool names, and schema compatibility. It does not enable the disabled Deep Research integration.

Remote MCP will only be reconsidered after an upstream correction and repeated E2E runs that retain non-empty structured tool-call and tool-result steps.

Usage

VS Code MCP

Add to .vscode/mcp.json:

{
  "servers": {
    "gemini-research": {
      "command": "uvx",
      "args": ["gemini-research-mcp"],
      "env": {
        "GEMINI_API_KEY": "your-api-key"
      }
    }
  }
}

Or run from source:

{
  "servers": {
    "gemini-research": {
      "command": "uv",
      "args": ["run", "--directory", "path/to/gemini-research-mcp", "gemini-research-mcp"],
      "envFile": "${workspaceFolder}/path/to/gemini-research-mcp/.env"
    }
  }
}

Command Line

uv run gemini-research-mcp
# or
uvx gemini-research-mcp

DOCX Export

Export research sessions to professional Word documents with:

  • Cover page with title, date, and research metadata
  • Clickable Table of Contents with navigation to sections
  • Professional typography: Calibri fonts, 1-inch margins, 1.5x line spacing
  • Executive summary with elegant formatting
  • Full research report with proper heading hierarchy
  • Sources section with full clickable URLs
  • Metadata table with session details

VS Code Setup

To enable DOCX export, install with the [docx] extra:

{
  "servers": {
    "gemini-research": {
      "command": "uvx",
      "args": ["--from", "gemini-research-mcp[docx]", "gemini-research-mcp"],
      "env": {
        "GEMINI_API_KEY": "your-api-key"
      }
    }
  }
}

Downloading Files

export_research_session is disk-first: the file is always written to disk and the absolute path is returned on the first line of the response text (e.g. ✅ **Saved to:** /…/report.docx). This means any MCP client — GUI or headless — gets a usable file path back.

By default exports are written to GEMINI_RESEARCH_EXPORT_DIR (defaults to ~/.gemini-research/exports/; falls back to the system temp dir if that location isn't writable). Override per-call with the output_path argument:

{
  "name": "export_research_session",
  "arguments": {
    "interaction_id": "v1_...",
    "format": "docx",
    "output_path": "/absolute/or/relative/path/report.docx"
  }
}

When output_path is supplied, the parent directory must already exist (no silent mkdir). GUI hosts (e.g. VS Code Copilot Chat) also receive an EmbeddedResource attachment backed by the persistent research://exports/{id} resource store for native "Save As" — clients that can't render it can safely ignore it.

Client compatibility

research_deep requires MCP Tasks support (SEP-1732) on the client. Clients that do not advertise the tasks capability will receive a -32600 error.

Known client status:

  • VS Code Copilot Chat / MCP Inspector / Claude Desktop — supported.
  • GitHub Copilot CLI — tracked upstream at github/copilot-cli#2538; until that lands, use research_web from the CLI.

Installation (pip/uv)

# Install with DOCX support
pip install 'gemini-research-mcp[docx]'
# or
uv add 'gemini-research-mcp[docx]'

Features

Feature Description
Cover Page Title, date, duration, tokens, AI agent
Clickable TOC Internal hyperlinks navigate to sections
Syntax Highlighting Pygments-powered code blocks with GitHub colors
Professional Styling Calibri fonts, proper heading hierarchy (H1-H4)
Page Margins Standard 1-inch (2.54cm) margins
Heading Spacing keep_with_next prevents orphan headings
Sources Full URLs as clickable hyperlinks
Pure Python No external binaries (Pandoc not required)

Resources

MCP Resources provide read-only data that clients can access:

Resource Description
research://models Available models and their capabilities
research://exports List cached exports ready for download
research://exports/{id} Download an exported file (Markdown, JSON, or DOCX)

File Downloads

The export_research_session tool creates exports and returns a resource URI. Clients (like VS Code) can then fetch the resource to download the file with proper MIME type handling.

Development

uv sync --extra dev
uv run pytest
uv run mypy src/
uv run ruff check src/

Tests

uv run pytest                    # Unit tests
uv run pytest -m e2e             # E2E tests (requires GEMINI_API_KEY)
uv run pytest --cov=src/gemini_research_mcp  # With coverage

Pricing

Tool Typical Cost
research_web ~$0.01-0.05 per query
research_deep ~$2-5 per task

Deep Research uses ~80-160 searches and ~250k-900k tokens per task.

License

MIT

Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

gemini_research_mcp-0.16.0b2.tar.gz (462.5 kB view details)

Uploaded Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

gemini_research_mcp-0.16.0b2-py3-none-any.whl (97.1 kB view details)

Uploaded Python 3

File details

Details for the file gemini_research_mcp-0.16.0b2.tar.gz.

File metadata

  • Download URL: gemini_research_mcp-0.16.0b2.tar.gz
  • Upload date:
  • Size: 462.5 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? Yes
  • Uploaded via: twine/7.0.0 CPython/3.13.14

File hashes

Hashes for gemini_research_mcp-0.16.0b2.tar.gz
Algorithm Hash digest
SHA256 5d9718bcb8b3cff501e35a73436990ea9b2f0fddebce4c94b65a4c51c1a3d03b
MD5 bdafd991f602b183aa8924560b652555
BLAKE2b-256 4082f02fa98d25629c4709cdae1efe98ecbd2d50a9e59c74970d01ab48bbace5

See more details on using hashes here.

Provenance

The following attestation bundles were made for gemini_research_mcp-0.16.0b2.tar.gz:

Publisher: publish.yml on machinemates-ai/gemini-research-mcp

Attestations: Values shown here reflect the state when the release was signed and may no longer be current.

File details

Details for the file gemini_research_mcp-0.16.0b2-py3-none-any.whl.

File metadata

File hashes

Hashes for gemini_research_mcp-0.16.0b2-py3-none-any.whl
Algorithm Hash digest
SHA256 53806a9d52da47a6868ef0628bdfce146337d0d144d49f97d866f7957d1a018f
MD5 3fad7eadb523909a7fe8335767cea47f
BLAKE2b-256 ee07f5829a25074771c3ef79db4e6e99e9f4a70a7a44d6857b306c3d6f83e0f3

See more details on using hashes here.

Provenance

The following attestation bundles were made for gemini_research_mcp-0.16.0b2-py3-none-any.whl:

Publisher: publish.yml on machinemates-ai/gemini-research-mcp

Attestations: Values shown here reflect the state when the release was signed and may no longer be current.

Release history Release notifications | RSS feed

This release

0.16.0b2 This release

2 files

0.15.1

2 files

0.15.0

2 files

0.13.4

2 files

0.13.3

2 files

0.13.2

2 files

0.13.1

2 files

0.13.0

2 files

0.12.1

2 files

0.12.0

2 files

0.11.2

2 files

0.11.1

2 files

0.11.0

2 files

0.10.4

2 files

0.10.3

2 files

0.10.2

2 files

0.10.0

2 files

0.7.0

2 files

0.6.5

2 files

0.6.4

2 files

0.6.3

2 files

0.6.2

2 files

0.6.1

2 files

0.6.0

2 files

0.5.1

2 files

0.5.0

2 files

0.4.0

2 files

0.3.1

2 files

0.3.0

2 files

0.2.4

2 files

0.2.3

2 files

0.2.2

2 files

0.2.1

2 files

0.2.0

2 files

0.1.4

2 files

0.1.3

2 files

0.1.2

2 files

0.1.1

2 files

0.1.0

2 files

Anthropic, PBC Visionary sponsor Bloomberg Visionary sponsor Hudson River Trading Visionary sponsor Meta Visionary sponsor NVIDIA Visionary sponsor Microsoft Sustainability sponsor Depot Continuous Integration AWS Cloud computing and Security Sponsor Datadog Monitoring Fastly CDN Google Download Analytics Sentry Error logging StatusPage Status page