Skip to main content

gemini-mcp

MCP gateway for massive-payload Gemini analysis on the AI Studio free tier. Runs on a small VPS; a local Freebuff Desktop agent sends huge bodies to it and gets back a compact consolidated result — the agent's own context stays clean.

  • Cascading Flash chain 3.8 → 3.7 → 3.6 → 3.5 → 3 via gemini-router (per key×model RPD/RPM/TPM ledger shared with tldr-digest).
  • Over-limit bodies: split into overlapping chunks, each marked [CHUNK N/M], analyzed separately, then consolidated (recursively if needed).
  • thinking_level default high, per-request override; opt-in overflow (flash-lite → gemma) when the daily quota is gone; hard stop otherwise; per-request wait policy for minute-limit resets (default: wait, countdown reported in progress).
  • Deep progress to the caller (MCP progress + logging notifications + embedded trace), rotating disk logs, gemini-mcp quota|logs|calls|doctor in the terminal.
  • Transports: SSH-stdio (primary) and streamable HTTP daemon (127.0.0.1, bearer).
  • Dual backend (0.8.0): router (fleet: shared SQLite ledger on the VPS) or direct (plain Google AI Studio: same engine, in-memory counters, no server at all). backend: auto picks router only when persistent state exists; env GEMINI_MCP_BACKEND overrides. See docs/08-dual-mode.md.

Quickstart — no server (direct mode)

uvx --from gemini-mcp-gateway gemini-mcp setup
# wizard: fleet-vs-simple question → masked key paste (aistudio.google.com/app/apikey)
# → ~/.config/gemini-mcp/{config.yaml,.env 0600} → live listModels probe → client snippet

Paste the printed snippet into your MCP client (Freebuff ~/.agents/mcp.json, Claude Desktop, Cursor — any stdio client) and ask the agent to call quota_status. No env vars needed: the gateway auto-discovers ~/.config/gemini-mcp/. Manual alternative: export GEMINI_API_KEYS="AIza…,AIza…" + uvx --from gemini-mcp-gateway gemini-mcp stdio. gemini-mcp doctor re-checks keys/chain/limits; gemini-mcp uninstall --yes removes everything the wizard wrote.

Direct mode needs nothing but API keys: quota counters are per-process approximations (reset on restart) and every response says backend=direct in its report line. Fleet deploys (VPS daemon + shared ledger) pin backend: router in config.yaml. The npm shim @vernikr/gemini-mcp (pnpm-first install) ships the same core at the same version.

Status

v0.8.0 — published consumer package. Fleet side (VPS daemon 0.6.x, SSH-stdio + HTTP tunnel) runs in production; consumer side ships as PyPI gemini-mcp-gateway (+ npm shim @vernikr/gemini-mcp) with the dual-mode backend (router/direct, docs/08) and the setup wizard (M7.2). Router core: gemini-router 0.1.1 on PyPI. 80 tests green, no network needed (respx/fakes only). Next: M7.3 dual-publish CI, M7.4 clean-machine beta.

Development

uv sync --extra dev    # dev pin: gemini-router from git tag (tool.uv.sources);
                       # published metadata uses the PyPI version range — no auth needed
uv run pytest -q       # 80 tests: chunker, pipeline, tools/server, logging, direct
                       # backend (respx), setup wizard — all faked, no network
uv run ruff check src tests
GEMINI_API_KEYS=... uv run gemini-mcp setup|quota|calls|logs|prune|doctor
Stage Artifact
1 · Unpacking/reframing + Q&A docs/01-unpacking-reframing.md
2 · Functional/business requirements docs/02-requirements.md
3 · Architecture & stack docs/03-architecture.md
4 · Plan (M0–M6) docs/04-plan.md
Risks docs/blockers/

Planned usage (Freebuff, ~/.agents/mcp.json)

{ "mcpServers": {
    "gemini": {                       // A) SSH-stdio — primary, encrypted, no open ports
      "command": "ssh",
      "args": ["root@38.244.152.2", "/opt/apps/gemini-mcp/.venv/bin/gemini-mcp", "stdio"] },
    "gemini-http": {                  // B) HTTP via `ssh -L 8790:127.0.0.1:8790 root@38.244.152.2`
      "type": "http", "url": "http://127.0.0.1:8790/mcp",
      "headers": { "Authorization": "Bearer $GEMINI_MCP_TOKEN" } } } }

Tools: analyze_large (chunking pipeline), gemini_generate (single shot), quota_status, router_logs.

Deployment (M4)

From the workstation (sibling checkout of gemini-router required next to this repo):

bash scripts/run_remote.sh setup    # rsync both repos → venv, .env, systemd unit, doctor
bash scripts/run_remote.sh smoke    # first real-key run on the VPS (spends ≤1 RPD)
bash scripts/run_remote.sh status|logs

Freebuff Desktop wiring (SSH-stdio and HTTP-via-tunnel snippets, verification steps): docs/freebuff-wiring.md. The daemon never leaves 127.0.0.1:8790; no firewall changes are made (server-spec compliant).

Ops (after M4)

ssh root@38.244.152.2
sudo systemctl status gemini-mcp          # daemon (HTTP)
/opt/apps/gemini-mcp/.venv/bin/gemini-mcp quota    # what's left, per key×model
/opt/apps/gemini-mcp/.venv/bin/gemini-mcp logs -f  # live app log

Conventions: English in files, Conventional Commits, SemVer, docs updated with every functional change (see AGENTS.md).

Metadata

Release files for gemini-mcp-gateway 0.8.0

For a detailed explanation of source distributions (sdists) and built distributions (wheels), please see the package formats documentation.

Source distribution (sdist)

Source distribution for gemini-mcp-gateway 0.8.0
File Size Uploaded
gemini_mcp_gateway-0.8.0.tar.gz 144.0 kB Details

Built distribution (wheel)

Table of built distributions (wheels) for gemini-mcp-gateway 0.8.0
File Interpreter ABI Platform
gemini_mcp_gateway-0.8.0-py3-none-any.whl Python 3 none any Details

Total release size: 188.9 kB

Release files / gemini_mcp_gateway-0.8.0.tar.gz

Download URL gemini_mcp_gateway-0.8.0.tar.gz
Size 144.0 kB
Tags Source
SHA-256 checksum
How to use checksums
2e2eee301cb971cd95a1887b00420122daa0920f8f3169a2f52fd121f9ce9bc0
BLAKE2b-256 checksum
How to use checksums
7d1ad90494a5a8f46e113ffca9d0870b40ae13a88699d074b6c782d29109251e
Upload date
Uploaded using Trusted Publishing?
What is trusted publishing?
No
Uploaded via uv/0.12.13 {"installer":{"name":"uv","version":"0.12.13","subcommand":["publish"]},"python":null,"implementation":{"name":null,"version":null},"distro":{"name":"macOS","version":null,"id":null,"libc":null},"system":{"name":null,"release":null},"cpu":null,"openssl_version":null,"setuptools_version":null,"rustc_version":null,"ci":null}

Release files / gemini_mcp_gateway-0.8.0-py3-none-any.whl

Download URL gemini_mcp_gateway-0.8.0-py3-none-any.whl
Size 44.9 kB
Tags Python 3
SHA-256 checksum
How to use checksums
ff1625fd756a193a26b24797e535c9a16fbc679006bf3d8145298ad5fb8c17e1
BLAKE2b-256 checksum
How to use checksums
0467ec9872f4cd3bc2d56fce073cce189a5da303f7b169e02e705243a3d51fcb
Upload date
Uploaded using Trusted Publishing?
What is trusted publishing?
No
Uploaded via uv/0.12.13 {"installer":{"name":"uv","version":"0.12.13","subcommand":["publish"]},"python":null,"implementation":{"name":null,"version":null},"distro":{"name":"macOS","version":null,"id":null,"libc":null},"system":{"name":null,"release":null},"cpu":null,"openssl_version":null,"setuptools_version":null,"rustc_version":null,"ci":null}

Release history Release notifications | RSS feed

0.12.0

2 release files

0.11.0

2 release files

0.10.1

2 release files

0.10.0

2 release files

0.9.3

2 release files

0.9.2

2 release files

0.9.1

2 release files

0.9.0

2 release files

0.8.2

2 release files

0.8.1

2 release files

This release

0.8.0 This release

2 release files

0.7.1

2 release files

Anthropic, PBC Visionary sponsor Bloomberg Visionary sponsor Hudson River Trading Visionary sponsor Meta Visionary sponsor NVIDIA Visionary sponsor Microsoft Sustainability sponsor Depot Continuous Integration AWS Cloud computing and Security Sponsor Datadog Monitoring Fastly CDN Google Download Analytics Sentry Error logging StatusPage Status page