Skip to main content

Official Python SDK for TokenBee LLM inference gateway and observability.

Project description

TokenBee Python SDK

Official Python SDK for TokenBee - The Intelligent LLM Inference Gateway with Observability, Compression, and Privacy.

Features

  • Unified API: Access multiple LLM providers (OpenAI, Anthropic, Google, Mistral, etc.) through a single interface.
  • Intelligent Compression: Reduce token usage and latency with context-aware compression.
  • Privacy Guard: Automatic PII masking and privacy-preserving inference.
  • Built-in Observability: Automatic tracking of latency, costs, and token usage.

Links

Installation

pip install tokenbee-sdk

Quick Start

from tokenbee import TokenBee, TokenBeeModel, CompressionRate

# Initialize the client with your TokenBee API key 
# AND your LLM provider key (Bring Your Own Key - BYOK)
client = TokenBee(
    api_key="your_tokenbee_api_key",
    llm_key="your_llm_provider_key", # e.g. OpenAI or Anthropic key
    compression="auto",
    rate=CompressionRate.MEDIUM
)

# Send a request
response = client.send(
    model=TokenBeeModel.OPENAI_GPT_4O,
    input={
        "messages": [
            {"role": "user", "content": "Explain quantum entanglement in simple terms."}
        ]
    }
)

print(response["choices"][0]["message"]["content"])

Bring Your Own Key (BYOK)

TokenBee is a stateless gateway. We do not store your LLM provider API keys in our database. You pass your provider key (OpenAI, Anthropic, etc.) through the SDK's llm_key parameter. The SDK sends this in the X-LLM-Key header, allowing the TokenBee proxy to forward requests to the provider on your behalf while you maintain full control over your billing and security.

Advanced Usage

Compression Control

You can specify the compression rate and method per request. TokenBee uses an intelligent semantic engine to reduce token usage while preserving meaning.

response = client.send(
    model=TokenBeeModel.ANTHROPIC_CLAUDE_3_5_SONNET,
    input={
        "messages": [...],
        "compression": "auto",      # "auto" (default), "on", or "off"
        "rate": CompressionRate.HIGH, # MEDIUM (0.5), HIGH (0.33), etc.
        "privacy": True
    }
)
  • compression: Set to "auto" to let TokenBee decide when to compress, or "off" to bypass the compression engine entirely for high-precision tasks.
  • rate: Controls the aggressiveness of compression. HIGH aims for ~67% token reduction.
  • sessionId: (Optional) String ID to group multiple requests into a single replayable session in the dashboard.
  • userId: (Optional) String ID to track usage and costs per unique end-user.
  • privacy: (Optional) Set to True to disable payload logging and session replays for this request. Metadata (latency, tokens) will still be recorded for observability.

Supported Models

The SDK provides a TokenBeeModel enum with popular models:

  • TokenBeeModel.OPENAI_GPT_4O
  • TokenBeeModel.ANTHROPIC_CLAUDE_3_5_SONNET
  • TokenBeeModel.GEMINI_2_0_FLASH
  • ... and many others.

License

MIT

Project details


Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

tokenbee_sdk-1.3.0.tar.gz (4.6 kB view details)

Uploaded Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

tokenbee_sdk-1.3.0-py3-none-any.whl (4.4 kB view details)

Uploaded Python 3

File details

Details for the file tokenbee_sdk-1.3.0.tar.gz.

File metadata

  • Download URL: tokenbee_sdk-1.3.0.tar.gz
  • Upload date:
  • Size: 4.6 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/6.2.0 CPython/3.14.0

File hashes

Hashes for tokenbee_sdk-1.3.0.tar.gz
Algorithm Hash digest
SHA256 30fc2935758cdebeaedcc7549d9a1b53f3f58617222e70a79ef79fa0efb564ab
MD5 fccd102ef18d11f55fc426f156f9c9c5
BLAKE2b-256 38793de277b23e0e0a46c61cb38c73990f2cd6003101fb24b5ac487116c18d91

See more details on using hashes here.

File details

Details for the file tokenbee_sdk-1.3.0-py3-none-any.whl.

File metadata

  • Download URL: tokenbee_sdk-1.3.0-py3-none-any.whl
  • Upload date:
  • Size: 4.4 kB
  • Tags: Python 3
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/6.2.0 CPython/3.14.0

File hashes

Hashes for tokenbee_sdk-1.3.0-py3-none-any.whl
Algorithm Hash digest
SHA256 6b9bde41efb5b5c32262a83907e451fd727c36fabbbcd652b28ab7be16fde56d
MD5 7e72a197b1d334f7ef2e53576006ccbc
BLAKE2b-256 afe370e1ef5a500d81fd2db9d9deee2520421e9d33ca5e6da3d1531a1c36ff09

See more details on using hashes here.

Supported by

AWS Cloud computing and Security Sponsor Datadog Monitoring Depot Continuous Integration Fastly CDN Google Download Analytics Pingdom Monitoring Sentry Error logging StatusPage Status page