Skip to main content

PyPI version Python 3.11+ License: MIT CI

EggPool

A lightweight, LAN-hosted proxy that aggregates multiple AI provider accounts behind one OpenAI/Anthropic-compatible endpoint.

Features

  • Proxies model requests across multiple providers and accounts behind a single endpoint
  • OpenAI- and Anthropic-compatible upstream request paths, with transparent bidirectional protocol transcoding
  • Dynamically discovers available models; routes by quota utilization (load-based, never cost-based)
  • Per-account outbound proxy support (pproxy — SOCKS5, HTTP, Shadowsocks)
  • Tracks requests, tokens, latency, errors, and cost provenance in SQLite (provider_reported, trusted local derived/partial, bounded estimated; reservation is advisory, not a floor)
  • Multi-page dashboard with 50+ themes, reliability, routing, and runtime views
  • Model metadata enrichment from provider catalogs, OpenRouter, Artificial Analysis, and Hugging Face
  • Provider-neutral request shaping: cache reporting, safe suffix compression, policy-scoped overrides, optional synthetic cache controls, and advisory threshold tuning
  • Thinking/reasoning capability-aware routing with configurable budget mapping
  • Designed for lightweight deployments (Raspberry Pi, SBCs)

Quick Start

# Install (one-shot)
curl -fsSL https://raw.githubusercontent.com/eggstack/eggpool/main/scripts/install.sh | bash

# Interactive onboarding — connect providers, validate, start
eggpool onboard

# Install as a systemd service
sudo env "PATH=$PATH" "$(command -v eggpool)" deploy systemd --install

See Deployment for alternative install methods (pipx, manual, production) and the full deployment guide.

CLI Reference

Command Description
eggpool serve Start the proxy server (--daemon to detach)
eggpool stop Stop the running server
eggpool restart Fully restart the server (stop then start)
eggpool rehash Restart to apply config changes
eggpool onboard Interactive onboarding wizard
eggpool connect Add a provider account interactively
eggpool connect list List supported providers
eggpool logout Remove a configured provider account
eggpool check-config Validate configuration
eggpool migrate Run database migrations
eggpool models refresh Refresh the model catalog
eggpool accounts list List configured provider accounts
eggpool accounts status Show account status (provider, priority, weight, enabled)
eggpool accounts explain Show per-account routing eligibility for a model
eggpool stats transcoding Show protocol transcoding statistics
eggpool stats repair-costs Dry-run/apply repair for suspicious historical request costs (incl. reservation-fallback rows where canonical cost equals the inflated reservation while a smaller local estimate exists)
eggpool stats recompute-costs Recompute cost_microdollars on historical requests
eggpool runtime-status Print runtime health summary
eggpool backup Create a timestamped backup
eggpool recover Restore from a backup archive
eggpool deploy systemd Print/install systemd unit
eggpool deploy cron Install watchdog cron (non-systemd)
eggpool deploy backup-cron Install daily backup cron job
eggpool deploy logrotate Print/install logrotate config
eggpool deploy all Print every deployment snippet in sequence
eggpool update Check for and install updates
eggpool uninstall Uninstall EggPool from this machine

All commands accept --config /path/to/config.toml. Config resolution: --config > $EGGPOOL_CONFIG > ~/.config/eggpool/config.toml > ./config.toml.

Full command reference: docs/deployment.md

Configuration

Configuration lives in a single TOML file. API keys are loaded from environment variables or .env.

# Example provider configuration
[providers.opencode-go]
id = "opencode-go"
base_url = "https://opencode.ai/zen/go/v1"
protocols = ["openai", "anthropic"]

[[providers.opencode-go.accounts]]
name = "personal"
api_key = "sk-your-opencode-go-key"

Use eggpool connect for interactive provider setup. See docs/providers.md for the full provider catalog, configuration details, and troubleshooting.

Key Config Sections

Section Purpose
[server] Bind address, port (default 11300), API key, logging, threads
[upstream] Upstream API base URL, timeouts, connection pool
[database] SQLite path, WAL mode
[models] Catalog refresh, exposure mode, model collapse, withdrawal policy
[routing] Routing strategy, retry limits, quota mode, same-tier fairness
[dashboard] Dashboard toggle, theme, refresh interval
[providers.*] Provider configs with accounts and routing priority
[network] Outbound transport, DNS cache
[model_info] Optional model metadata refresh, aliases, overrides, and external source settings
[transcoder] Protocol transcoding between OpenAI and Anthropic formats
[compression] Request shaping: observe/safe, stable thresholds, transform toggles, advanced policy overrides
[cache] Synthetic cache controls (post-route, disabled by default, dry-run first)

The catalog refresh is non-destructive by default: failed, empty, or partial upstream responses never silently de-pool a healthy account. Set [models].catalog_withdrawal_policy (preserve_until_health default, confirmed_once, confirmed_twice) to opt into destructive behavior on authoritative refreshes. See architecture/README.md § Catalog Refresh Semantics.

Full config reference: config.example.toml | docs/providers.md

Protocol transcoding

When [transcoder] enabled = true, EggPool bridges OpenAI Chat Completions and Anthropic Messages bidirectionally so a single client ecosystem (e.g. OpenCode, which speaks only OpenAI) can reach Anthropic-only upstreams and vice versa.

What gets translated:

  • Request bodies (text + tool-use + vision + thinking + structured outputs)
  • Streaming SSE events (including tool-call deltas and thinking deltas)
  • Non-retryable error envelopes
  • Usage and cost fields (preserved exactly as the upstream reported them)

What is dropped with a structured warning log:

  • OpenAI fields with no Anthropic equivalent (logit_bias, presence_penalty, top_logprobs, etc.)
  • Anthropic fields with no OpenAI equivalent (top_k, cache_control)

Phase 6 feature flags ([transcoder.features]) — all off by default:

  • tools — bidirectional tool calling translation
  • vision — image/document content parts
  • thinking — extended thinking ↔ reasoning_content
  • structured_outputsresponse_format / json_schema coercion
  • anthropic_primitivestop_k, cache_control, context_management, container, mcp_servers

See docs/transcoding.md for the full translation table and known limitations.

Request shaping

EggPool’s request-shaping stack is opt-in and cache-preserving by default. With the shipped config, the public surface stays in reporting mode: no route changes, no stable-prefix mutation, and no synthetic cache annotations on the wire. Routing remains load-based; cache, compression, synthetic-cache, and tuning fields never enter QuotaFairScorer.

Operator surfaces

Surface Default Mutates requests? Primary stats/API
Cache reporting on no /api/stats/cache-observability
Request segmentation on no /api/stats/canonical-request-segmentation
Native cache preservation on when transcoding no /api/stats/cache-stability
Compression opportunities off unless [compression].enabled = true no /api/stats/compression-observability
Safe compression off (mode = "observe") volatile suffix only, fail-closed /api/stats/compression-runtime
Policy overrides off ([[compression.policies]] = []) scoped overlay only /api/stats/compression-policies
Synthetic cache controls off (enabled = false, dry_run = true) provider-bound stable prefix only /api/stats/synthetic-cache-observability
Advisory tuning off (enabled = false) no /api/stats/compression-tuning
Request-shaping summary on with dashboard/runtime no /api/stats/request-shaping
Routing guardrails always on no /api/stats/runtime

Stable config knobs

The default example now exposes only the normal operator knobs:

[compression]
enabled = false
mode = "observe"
min_candidate_tokens = 2048
min_savings_tokens = 1024
max_compression_latency_ms = 25.0

[compression.transforms]
fold_repeated_lines = true
compact_logs = true
compact_search_results = true
elide_base64_blobs = true
minify_machine_json = true
compact_stack_traces = true

[cache.synthetic_cache_controls]
enabled = false
dry_run = true
min_stable_tokens = 1024

Advanced knobs still exist for policy-scoped rollouts and advisory tuning, but they are intentionally pushed into the docs instead of the main example:

  • [[compression.policies]] for scoped overrides
  • compression.placement, respect_cache_boundaries, header_* for advanced routing-adjacent operator workflows
  • cache.synthetic_cache_controls.provider_kinds, ttl, max_breakpoints, placements, require_policy
  • compression.tuning.* for recommendation-only threshold guidance

Safety rules

  • Safe compression mutates only eligible volatile_suffix string leaves.
  • Stable-prefix preservation is verified with stable_prefix_content_hash; any mismatch falls back to the original payload.
  • Native provider cache annotations are preserved byte-for-byte.
  • Synthetic cache controls are disabled by default and dry-run first when enabled.
  • Context-limit checks happen before compression; compression never rescues an over-limit request.
  • Routing stays load-based and reporting-only metrics never influence scorer inputs.

Recommended rollout

  1. Enable cache reporting and request segmentation only.
  2. Run compression in observe mode for 24-48 hours.
  3. Turn on safe compression for one client or policy.
  4. Expand safe compression if fallbacks stay at zero.
  5. Try synthetic cache controls in dry-run for Anthropic-compatible upstreams.
  6. Move synthetic cache to apply mode only after the dry-run is clean.
  7. Enable advisory tuning only if you want threshold suggestions.

Further reading

API Endpoints

Method Path Description
GET /v1/models List available models
POST /v1/chat/completions OpenAI-compatible chat completions
POST /v1/messages Anthropic-compatible messages
GET /v1/healthz Liveness check
GET /v1/readyz Readiness check
GET /api/backoffs Active upstream-derived account backoffs (?now=<epoch> for reproducible snapshots)
GET /api/model-info Enriched model metadata summaries
GET /api/model-info/{model_id} Enriched metadata detail for one model
GET /api/model-info/{model_id}/aliases Source-keyed alias rows for one model
GET /api/model-info/sources Model-info source health
POST /api/model-info/refresh Trigger model-info refresh — ?model_id=<id>&source=<provider_catalog|openrouter|artificial_analysis|huggingface>&force=1 for a single-model force refresh (auth-gated). model_id accepts provider-suffixed IDs (gpt-4o/openai); unknown source values return HTTP 400
GET /api/stats/cache-observability Cache counter status coverage
GET /api/stats/canonical-request-segmentation Segmentation status and per-region token estimates
GET /api/stats/cache-stability Transcoder cache boundary tracker counters
GET /api/stats/compression-observability Observe-mode opportunity, per-policy roll-ups
GET /api/stats/compression-runtime Safe-mode applied/fallback counts and latency
GET /api/stats/compression-policies Per-policy roll-up table
GET /api/stats/synthetic-cache-observability Synthetic cache candidate / applied / native-preserved counts
GET /api/stats/compression-tuning Threshold tuning recommendations
GET /api/stats/request-shaping Operator-facing request-shaping summary
GET /api/stats/runtime Runtime metrics + hardcoded routing guardrails

When [dashboard].enabled = true, a multi-page dashboard is served at / with request stats, latency metrics, provider health, model-info detail pages, and more. Stats API available under /api/stats/*.

Model-info observability

  • The dashboard /models page renders a degraded-state notice above the table if the model-info service is unattached (no app.state.model_info) or if get_summary_map() raises an exception. The exception's full traceback is logged under eggpool.dashboard.routes — the page never embeds the traceback text in HTML.
  • The dashboard /models/{model_id:path} detail page distinguishes "no canonical row exists" (empty-state copy) from "lookup failed" (degraded-state notice) when the service throws.
  • /api/stats/runtime includes a model_info section with enabled, canonical_count, catalog_model_count, provider_model_count, due_count, and a source_health dict (no raw source payloads). Failures surface as *_error keys and probe_errors, never as raised exceptions.

Documentation

Topic Link
Deployment (install, systemd, production) docs/deployment.md
Provider catalog & configuration docs/providers.md
Backup & restore docs/backup-restore.md
Per-account outbound proxy docs/proxy.md
Model context limits docs/model-limits.md
Raspberry Pi setup docs/raspberry-pi.md
Firewall configuration docs/firewall.md
Filesystem layout docs/filesystem-layout.md
Network & DNS diagnostics docs/network-diagnostics.md
Protocol transcoding docs/transcoding.md
Cache & compression operator guide docs/cache-compression.md
Cache & compression profiles docs/cache-compression-profiles.md
Cache & compression troubleshooting docs/cache-compression-troubleshooting.md
Thinking & reasoning docs/thinking.md
Architecture overview architecture/README.md

Development

uv sync --extra dev
uv run ruff check src/ tests/ scripts/
uv run ruff format src/ tests/ scripts/
uv run pyright src/ scripts/
uv run pytest

Agent Configuration

eggpool configsetup generates configuration snippets for popular coding agents:

Target Command Output --write default Model
OpenCode eggpool configsetup opencode JSON provider config N/A (clipboard) auto
Claude Code eggpool configsetup claude-code JSON snippet N/A (clipboard) N/A
Aider eggpool configsetup aider Shell env exports .env.eggpool recommended
Codex eggpool configsetup codex TOML provider block N/A (printed) recommended
Qwen Code eggpool configsetup qwen-code JSON provider block N/A (printed) optional
Kilo eggpool configsetup kilo JSON provider block N/A (printed) optional
Continue eggpool configsetup continue YAML model block ~/.continue/eggpool.yaml usually yes
Cline eggpool configsetup cline JSON profile cline-eggpool.json recommended
Roo Code eggpool configsetup roo-code JSON profile roo-eggpool.json recommended
Goose eggpool configsetup goose Shell env exports N/A (printed) recommended
OpenHands eggpool configsetup openhands Shell env exports N/A (printed) recommended

Shared options: --host, --base-url, --model, --write, --output, --force, --no-clipboard, --print-secret. Generated JSON, TOML, YAML, and shell snippets escape catalog/config values for the target format, including provider-suffixed model IDs.

Examples:

eggpool configsetup aider --model openai/gpt-4 --write
eggpool configsetup continue --model claude-sonnet-4 --output ~/.continue/eggpool.yaml
eggpool configsetup cline --no-clipboard

License

MIT

Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

eggpool-0.5.2.tar.gz (973.8 kB view details)

Uploaded Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

eggpool-0.5.2-py3-none-any.whl (883.3 kB view details)

Uploaded Python 3

File details

Details for the file eggpool-0.5.2.tar.gz.

File metadata

  • Download URL: eggpool-0.5.2.tar.gz
  • Upload date:
  • Size: 973.8 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/6.2.0 CPython/3.14.2

File hashes

Hashes for eggpool-0.5.2.tar.gz
Algorithm Hash digest
SHA256 44ed9ec06133eee72051dc72738691851ba4aa0cb1dcca5c4e56e23f52061fad
MD5 f2960ff805b4f104256dd883ed444c8c
BLAKE2b-256 0a40594a470b47582299dbbb86f6bfc70d2e2fbc24fee65c8138076ee674e90c

See more details on using hashes here.

File details

Details for the file eggpool-0.5.2-py3-none-any.whl.

File metadata

  • Download URL: eggpool-0.5.2-py3-none-any.whl
  • Upload date:
  • Size: 883.3 kB
  • Tags: Python 3
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/6.2.0 CPython/3.14.2

File hashes

Hashes for eggpool-0.5.2-py3-none-any.whl
Algorithm Hash digest
SHA256 8b94bd91359b897d73c76551a2deffe9b90812a376ba328b45c989142ef5b088
MD5 25ee583dec3fa04e1a72d39a76a36217
BLAKE2b-256 6445d79beeb419cdabd1077db36b81717a302dcd92b3c9e2f2f92b608873214f

See more details on using hashes here.

Release history Release notifications | RSS feed

0.8.0

3 files

0.7.4

2 files

0.7.3

2 files

0.7.2

2 files

0.7.1

2 files

0.7.0

2 files

0.6.9

2 files

0.6.8

2 files

0.6.7

2 files

0.6.6

2 files

0.6.5

2 files

0.6.4

2 files

0.6.3

2 files

0.6.2

2 files

0.6.1

2 files

0.6.0

2 files

0.5.9

2 files

0.5.8

2 files

0.5.7

2 files

0.5.6

2 files

0.5.5

2 files

0.5.4

2 files

0.5.3

2 files

This release

0.5.2 This release

2 files

0.5.1

2 files

0.5.0

2 files

0.4.9

2 files

0.4.8

2 files

0.4.7

2 files

0.4.6

2 files

0.4.5

2 files

0.4.4

2 files

0.4.3

2 files

0.4.2

2 files

0.4.1

2 files

0.4.0

2 files

0.3.9

2 files

0.3.8

2 files

0.3.7

2 files

0.3.6

2 files

0.3.5

2 files

0.3.4

2 files

0.3.3

2 files

0.3.2

2 files

0.3.1

2 files

0.3.0

2 files

0.2.2

2 files

0.2.1

2 files

0.2.0

2 files

0.1.7

2 files

0.1.6

2 files

0.1.5

2 files

0.1.4

2 files

0.1.3

2 files

0.1.2

2 files

0.1.1

2 files

0.1.0

2 files

Anthropic, PBC Visionary sponsor Bloomberg Visionary sponsor Hudson River Trading Visionary sponsor Meta Visionary sponsor NVIDIA Visionary sponsor Microsoft Sustainability sponsor Depot Continuous Integration AWS Cloud computing and Security Sponsor Datadog Monitoring Fastly CDN Google Download Analytics Sentry Error logging StatusPage Status page