Lightweight multi-provider (OpenAI, Anthropic, Gemini) analytics tracking wrapper
Project description
tokvera
tokvera is a lightweight Python SDK that wraps OpenAI, Anthropic, and Gemini clients and emits usage analytics in a fire-and-forget way.
What's New in v0.2.2
- Added Trace Context v1 tags.
- New optional tags:
trace_id,run_id,conversation_id,span_id,parent_span_id,step_name. - Added Evaluation Signals v1 fields:
outcome,retry_reason,fallback_reason,quality_label,feedback_score. - Added FastAPI middleware integration helpers.
- Added LangChain callback integration helpers.
- Added LlamaIndex callback integration helpers.
- Auto-generates
trace_idandspan_idwhen you do not provide them.
Installation
pip install tokvera
For development:
pip install -e .[dev]
Environment Variable Setup
Set your ingestion endpoint:
# Linux/macOS
export TOKVERA_INGEST_URL="https://api.tokvera.com/v1/events"
# Windows PowerShell
$env:TOKVERA_INGEST_URL = "https://api.tokvera.com/v1/events"
If TOKVERA_INGEST_URL is not set, analytics are skipped automatically.
Trace Context v1
Use trace tags to reconstruct request chains without sending prompt payloads.
Recommended semantics:
trace_id: one end-to-end workflow/request.run_id: one execution run of an agent/workflow.conversation_id: one user conversation/session.span_id: one model call.parent_span_id: parent model call when nested.step_name: readable stage label (retrieve_context,draft_reply,quality_retry).
Example:
client = track_openai(
openai_client,
api_key="tokvera_project_key",
feature="support_bot",
tenant_id="acme",
trace_id="trace_req_20260304_001",
run_id="run_agent_20260304_001",
conversation_id="conv_9832",
span_id="span_root_1",
parent_span_id=None,
step_name="draft_reply",
)
FastAPI Middleware Integration
Use middleware to create request-level trace context and pass it into SDK calls.
from fastapi import FastAPI, Request
from openai import OpenAI
from tokvera import (
create_fastapi_tracking_middleware,
get_fastapi_track_kwargs,
track_openai,
)
app = FastAPI()
openai_client = OpenAI(api_key="sk-...")
middleware = create_fastapi_tracking_middleware(
defaults={"feature": "support_bot", "environment": "production"},
context_resolver=lambda request: {"tenant_id": request.headers.get("x-tenant-id")},
)
@app.middleware("http")
async def tokvera_context(request: Request, call_next):
return await middleware(request, call_next)
@app.post("/reply")
async def reply():
tracked = track_openai(
openai_client,
api_key="tokvera_project_key",
**get_fastapi_track_kwargs(step_name="draft_reply"),
)
return tracked.chat.completions.create(
model="gpt-4o-mini",
messages=[{"role": "user", "content": "Hello"}],
)
LangChain Callback Integration
Use a callback handler to emit Tokvera events from LangChain LLM runs.
from langchain_openai import ChatOpenAI
from tokvera import create_langchain_callback_handler
callback = create_langchain_callback_handler(
api_key="tokvera_project_key",
feature="agent_support",
tenant_id="acme",
environment="production",
)
model = ChatOpenAI(
model="gpt-4o-mini",
callbacks=[callback],
)
result = model.invoke("Hello")
LlamaIndex Callback Integration
Use a callback handler to emit Tokvera events from LlamaIndex workflows.
from llama_index.core.callbacks import CallbackManager
from tokvera import create_llamaindex_callback_handler
tokvera_handler = create_llamaindex_callback_handler(
api_key="tokvera_project_key",
feature="agent_support",
tenant_id="acme",
environment="production",
)
callback_manager = CallbackManager([tokvera_handler])
Quick Start
OpenAI
from openai import OpenAI
from tokvera import track_openai
openai_client = OpenAI(api_key="sk-...")
client = track_openai(
openai_client,
api_key="tokvera_project_key",
feature="support_bot",
tenant_id="acme",
trace_id="trace_support_001",
run_id="run_support_001",
conversation_id="conv_42",
step_name="draft_reply",
outcome="success",
quality_label="good",
feedback_score=5,
plan="pro",
environment="production",
template_id="support_v3",
capture_content=False,
)
response = client.chat.completions.create(
model="gpt-4o-mini",
messages=[{"role": "user", "content": "Hello"}],
)
Anthropic
from anthropic import Anthropic
from tokvera import track_anthropic
anthropic_client = Anthropic(api_key="sk-ant-...")
client = track_anthropic(
anthropic_client,
api_key="tokvera_project_key",
feature="support_bot",
tenant_id="acme",
environment="production",
)
client.messages.create(
model="claude-3-5-sonnet-latest",
max_tokens=256,
messages=[{"role": "user", "content": "Hello"}],
)
Gemini
from google import genai
from tokvera import track_gemini
gemini_client = genai.Client(api_key="AIza...")
client = track_gemini(
gemini_client,
api_key="tokvera_project_key",
feature="assistant",
tenant_id="acme",
environment="production",
)
client.models.generate_content(
model="gemini-2.0-flash",
contents="Hello",
)
Event Schema
Canonical specification: tokvera-api/docs/EVENT_SCHEMA.md
Events include:
schema_version:2026-02-16event_type:openai.request,anthropic.request, orgemini.requestprovider:openai,anthropic, orgeminiendpoint:chat.completions.create,responses.create,messages.create,models.generate_contentstatus:successorfailurelatency_msmodelusage:prompt_tokens,completion_tokens,total_tokenstags:feature,tenant_id,customer_id,attempt_type,plan,environment,template_id,trace_id,run_id,conversation_id,span_id,parent_span_id,step_name- Evaluation signals (optional):
outcome,retry_reason,fallback_reason,quality_label,feedback_score(emitted intagsand top-levelevaluation) erroron failure events
trace_id and span_id are auto-generated per request if not provided.
Privacy
By default, prompt/response content is not sent.
If capture_content=True, content is hashed (SHA-256) before ingestion. Raw content is never sent by this SDK.
Disable Tracking
You can disable tracking by either:
- Using the original OpenAI client directly (do not wrap it), or
- Unsetting
TOKVERA_INGEST_URLso no events are emitted.
Project details
Release history Release notifications | RSS feed
Download files
Download the file for your platform. If you're not sure which to choose, learn more about installing packages.
Source Distribution
Built Distribution
Filter files by name, interpreter, ABI, and platform.
If you're not sure about the file name format, learn more about wheel file names.
Copy a direct link to the current filters
File details
Details for the file tokvera-0.2.2.tar.gz.
File metadata
- Download URL: tokvera-0.2.2.tar.gz
- Upload date:
- Size: 51.7 kB
- Tags: Source
- Uploaded using Trusted Publishing? No
- Uploaded via: twine/6.2.0 CPython/3.12.10
File hashes
| Algorithm | Hash digest | |
|---|---|---|
| SHA256 |
ab1b82bf3f0119682dc9144002d3c27ad8dfcefd94e7665839f1a734d43e9f85
|
|
| MD5 |
33fd709c09e35d7749402512f098d2ae
|
|
| BLAKE2b-256 |
cf777db571f5763ab3e2963d4d54500a8c78bb55693b6e8d05f80ee3fc6dfa15
|
File details
Details for the file tokvera-0.2.2-py3-none-any.whl.
File metadata
- Download URL: tokvera-0.2.2-py3-none-any.whl
- Upload date:
- Size: 17.7 kB
- Tags: Python 3
- Uploaded using Trusted Publishing? No
- Uploaded via: twine/6.2.0 CPython/3.12.10
File hashes
| Algorithm | Hash digest | |
|---|---|---|
| SHA256 |
a12db4d1da72145eea2d5ae02d4cc8dd665f08c7056b1174baf57dc63722fc0f
|
|
| MD5 |
5000de860efdfb35721792cb7c216368
|
|
| BLAKE2b-256 |
2c5cfc22a87771c49a82f1c35e85dbd89db28cf29f9468c11951d415f82d4fed
|