Skip to main content

AgentGuard

Stop runaway agents before they burn money.

Zero-dependency Python kill switch for AI agents. Hard budget caps. Loop detection. Local traces. MIT.

PyPI Downloads Python CI License: MIT

pip install agentguard47

Getting started

1. Install and verify

pip install agentguard47
agentguard doctor   # package ok?
agentguard demo     # offline proof (no API keys)

2. Guard an OpenAI client

from agentguard import BudgetGuard, LoopGuard, Tracer, patch_openai

budget = BudgetGuard(max_cost_usd=5.00, warn_at_pct=0.8)
loop = LoopGuard(max_repeats=3)
tracer = Tracer(service="my-agent", guards=[loop])

patch_openai(tracer, budget_guard=budget)
# every OpenAI call is now traced + budget-enforced

When spend crosses the hard limit, BudgetExceeded is raised and the run stops.

3. Cap a single task

Session budget can still have headroom. One goal can still be killed:

with budget.goal("refund", max_cost_usd=0.50, warn_at_pct=0.8) as g:
    g.attempt()
    budget.consume(cost_usd=0.12)
    # BudgetExceeded names the goal when it crosses

4. Read the local proof

agentguard report .agentguard/traces.jsonl
agentguard incident .agentguard/traces.jsonl

Or scaffold a starter file:

agentguard quickstart --framework raw --write
python agentguard_raw_quickstart.py

What it stops

Problem Guard Exception
Spend blowup BudgetGuard BudgetExceeded
Same tool forever LoopGuard LoopDetected
Fuzzy / A-B-A-B loops FuzzyLoopGuard LoopDetected
Retry storms RetryGuard RetryLimitExceeded
Hung runs TimeoutGuard TimeoutExceeded
Spam calls RateLimitGuard —
Wallet drain (x402/USDC) X402SpendGuard BudgetExceeded

Not a dashboard. Not a model router. An in-process exception that kills the bad run mid-flight.

Cap your agent's x402 wallet spend

Agents that pay per-call via x402 (USDC micropayments) can drain a wallet in a silent loop. X402SpendGuard wraps the payment step and refuses before paying:

from agentguard import X402SpendGuard

guard = X402SpendGuard(
    max_total_usd=5.00,        # wallet cap, add period="day" for a daily reset
    max_per_endpoint_usd=1.00, # cap per resource URL
    max_per_call_usd=0.10,     # refuse any single payment above this
)
guard.charge(0.001, "https://api.example.com/search", my_x402_pay_step)

AgentGuard meters and refuses; it never signs or settles. Amounts come from your x402 client. No crypto dependencies.

Features

  • Hard stops — exceptions inside your process, not after-the-fact alerts
  • Task-level budgets — BudgetGuard.goal(...) for sub-task caps + warn hooks
  • Local traces — JSONL by default; no network unless you opt in
  • Zero deps — stdlib only; Python 3.9+
  • Provider patches — patch_openai / patch_anthropic
  • Framework hooks — LangChain, LangGraph, CrewAI (optional extras)

Local by default

  • No API key required for local proof
  • No network unless you configure HttpSink
  • MIT licensed

The SDK is the free local proof path. Start local. Add hosted ingest later only if you want retained history, alerts, team visibility, spend trends, hosted decision history, or dashboard-managed remote kill signals. Local guards remain authoritative. HttpSink mirrors trace and decision events; it does not execute remote kill signals by itself.

Integrations

OpenAI · Anthropic · LangChain · LangGraph · CrewAI · raw agent loops

pip install "agentguard47[langchain]"   # optional extras as needed

Security

The base install declares zero runtime dependencies. pip install agentguard47 pulls nothing, so a default install adds no third-party exposure.

Extras install third-party packages and need a separate audit. The LangChain and LangGraph extras now require Python 3.10+ and raise their minimum versions to the tested September 2026 releases. OpenTelemetry requires 1.44.0 or newer. The base SDK remains compatible with Python 3.9+.

The optional [crewai] extra requires CrewAI 1.15.21 or newer. Its current dependency tree still installs ChromaDB 1.1.1, with four distinct unresolved advisories: CVE-2026-45829, CVE-2026-45830, CVE-2026-45831, and CVE-2026-45833. The audit database provides no fixed version. Avoid this extra unless you have reviewed that upstream exposure. Installing AgentGuard alone does not install ChromaDB or start a server. See the upstream advisory and the release audit.

HttpSink validates the address it actually connects to, retains TLS hostname verification, and refuses cross-origin redirects. It connects directly and does not use environment proxy settings. A local guard stops instrumented work in your Python process; it does not cancel an agent loop running on a provider's server. Cost estimates are not invoices; supply provider-reported cost or use strict cost resolution when an estimate is insufficient.

Docs

Links

The hosted page is an optional next step, not a requirement. The SDK stays free, local, and MIT, and the local guards stay authoritative. Nothing in this package phones home. The only network egress is a sink or exporter you configure yourself, such as HttpSink or an OpenTelemetry exporter.


MIT · Built for people who ship agents and hate surprise bills.

Latest Release Notes (1.3.0)

(2026-09-12)

This release includes the accumulated, unpublished 1.2.14 candidate work below.

Security and enforcement fixes

  • LangChain now propagates guard exceptions through its real callback manager. A zero-call budget stops the tool before its body runs; previously LangChain could log the exception and continue. Sync and async dispatch run inline.
  • Budget and timeout caps reject invalid, negative, boolean, and non-finite values. Corrupt stored budget counters fail closed without rewriting state. Warning callbacks run outside budget locks and zero limits do not divide by zero.
  • Failed x402 payment callbacks refund only their original budget generation, so a reset or day rollover cannot reduce a later period's spending.
  • HTTP trace delivery rejects credential-bearing URLs, cross-origin redirects, mapped private IPv6 addresses, and private/reserved DNS answers at connection time. Connections use the validated address while TLS retains hostname checks. This transport deliberately does not use environment proxies.
  • Retry-After delays are finite, non-negative, and capped at 30 seconds.
  • MCP dependency updates resolve the npm audit findings in the committed lockfile.

Optional dependency compatibility

  • LangChain requires 1.6.3+, LangGraph 1.2.11+ with checkpoint 4.2.0+ and SDK 0.4.4+, OpenTelemetry 1.44.0+, and CrewAI 1.15.21+.
  • LangChain and LangGraph extras require Python 3.10+. The dependency-free base package remains compatible with Python 3.9+.
  • The optional CrewAI tree still installs ChromaDB with four distinct unresolved advisories (CVE-2026-45829, CVE-2026-45830, CVE-2026-45831, CVE-2026-45833). No fixed upstream version was available in the audit. Avoid this extra unless its exposure has been reviewed. Base installs do not include ChromaDB.
  • Audit scope, regression results, dependency resolutions, and limitations: September audit.

Reliability

  • Added the file-backed JsonFileStateStore integration for BudgetGuard(store=...), so configured budget usage can persist across processes and scheduled tasks. This is local persistence, not distributed coordination or a fairness guarantee.
  • Hardened the cross-process state lock (JsonFileStateStore, used by BudgetGuard(store=...)) against two Windows races that crashed concurrent processes under contention: an exclusive lock create that fails with PermissionError instead of FileExistsError during a concurrent release ("delete pending"), and an os.replace that transiently fails with access-denied when an antivirus/indexer holds the destination. Both now retry safely, so cross-process budget enforcement holds on Windows scheduled tasks.

Budget Goals

  • Added BudgetGuard.goal(...) for scoped per-goal caps on tokens, calls, and cost, with an optional warn_at_pct threshold and on_warning callback. Goal warnings are emitted once per goal while hard caps still refuse excess spend.

Payment Guardrails

  • Added X402SpendGuard for local caps on total, per-endpoint, and per-call x402/USDC spend. It checks and reserves configured spend before payment and rolls the reservation back if the payment callback raises. It does not settle x402 payments or add a crypto dependency.

Cost Accounting

  • Added maximum-precision billable-cost resolution with explicit source labels for provider-reported values, caller prices, estimates, zero-cost tool/local work, and unknown cost. Unknown usage stays conservative or fails in strict mode; the result is not a provider invoice.

Usage Accounting

  • Anthropic usage normalization now preserves thinking/reasoning tokens and separates them from answer tokens when the provider payload exposes that detail, alongside cache-read and cache-write fields.

Hardening

  • Rejected NaN, infinite, and negative budget inputs before state mutation so non-finite values cannot bypass a cost ceiling.
  • Made LoopGuard argument fingerprinting tolerate non-JSON-serializable tool arguments instead of crashing the guard while it checks for repeats.

Public Docs

  • Made the reader-facing surface fully model-agnostic to match the already-vendor-neutral code path: the README/PyPI "As a skill" heading now leads with Codex alongside Claude Code, and the budget-aware escalation example notes the escalate target can be any provider's model, not just Claude.

Onboarding

  • Bare agentguard now prints a friendly first-run welcome with the 60-second local path and the star call to action instead of an argparse help dump.
  • Added python -m agentguard as an entry point so the CLI works even when the agentguard script is not on PATH.
  • Added agentguard welcome and agentguard badge. badge prints a paste-able "Guarded by AgentGuard" README badge (markdown, rST, or HTML) so adopters can advertise the SDK and drive new installs.

Distribution

  • Added an opt-in bridge to the hosted AgentGuard page (bmdpat.com/tools/agentguard) from the README/PyPI page, the agentguard --help footer, and the first-run welcome. These are static links only: the SDK still makes no network calls unless you configure HttpSink, and nothing in the package phones home. The links carry UTM parameters so the site can measure click-through; no identifier is sent from your machine.

Full changelog: CHANGELOG.md

Release files for agentguard47 1.3.0

For a detailed explanation of source distributions (sdists) and built distributions (wheels), please see the package formats documentation.

Source distribution (sdist)

Source distribution for agentguard47 1.3.0
File Size Uploaded
agentguard47-1.3.0.tar.gz 215.4 kB Details

Built distribution (wheel)

Table of built distributions (wheels) for agentguard47 1.3.0
File Interpreter ABI Platform
agentguard47-1.3.0-py3-none-any.whl Python 3 none any Details

Total release size: 331.4 kB

Release files / agentguard47-1.3.0.tar.gz

Download URL agentguard47-1.3.0.tar.gz
Size 215.4 kB
Tags Source
SHA-256 checksum
How to use checksums
b32192ecd5c85272ff4a60065fe294873192af55f34f09a68aee4f015a653b90
BLAKE2b-256 checksum
How to use checksums
ddec0a2b59dcd755945e419427b3e68ec03fe559b775679977120cf977c62d78
Upload date
Uploaded using Trusted Publishing?
What is trusted publishing?
Yes
Uploaded via twine/7.0.0 CPython/3.13.14

Provenance

Provenance describes where a file came from. On PyPI, provenance is shared via attestations, which provide a verifiable record of the build or publishing details. View details, limitations and caveats.

PyPI Publish Attestation

PyPI verified that this artifact, at this checksum, originated from the publisher listed below.

Signed by GitHub Actions, verified by PyPI on Sep 12, 2026.

Transparency log

Release files / agentguard47-1.3.0-py3-none-any.whl

Download URL agentguard47-1.3.0-py3-none-any.whl
Size 116.0 kB
Tags Python 3
SHA-256 checksum
How to use checksums
ee791ed4ff92faca056766314b489dbe981fdee432d54dd5b64aa8d1438a8bc9
BLAKE2b-256 checksum
How to use checksums
4eae24810941e684a43161b6936850832a1d7d864135c1f02644e1e403f31536
Upload date
Uploaded using Trusted Publishing?
What is trusted publishing?
Yes
Uploaded via twine/7.0.0 CPython/3.13.14

Provenance

Provenance describes where a file came from. On PyPI, provenance is shared via attestations, which provide a verifiable record of the build or publishing details. View details, limitations and caveats.

PyPI Publish Attestation

PyPI verified that this artifact, at this checksum, originated from the publisher listed below.

Signed by GitHub Actions, verified by PyPI on Sep 12, 2026.

Transparency log

Release history Release notifications | RSS feed

1.4.0

2 release files

1.3.2

2 release files

1.3.1

2 release files

This release

1.3.0 This release

2 release files

1.2.13

2 release files

1.2.9

2 release files

1.2.8

2 release files

1.2.6

2 release files

1.2.5

2 release files

1.2.4

2 release files

1.2.3

2 release files

1.2.2

2 release files

1.2.1

2 release files

1.2.0

2 release files

1.0.0

2 release files

0.8.0

2 release files

0.7.0

2 release files

0.6.0

2 release files

0.5.0

2 release files

0.4.0

2 release files

0.3.0

2 release files

0.2.0

2 release files

Anthropic, PBC Visionary sponsor Bloomberg Visionary sponsor Hudson River Trading Visionary sponsor Meta Visionary sponsor NVIDIA Visionary sponsor Microsoft Sustainability sponsor Depot Continuous Integration AWS Cloud computing and Security Sponsor Datadog Monitoring Fastly CDN Google Download Analytics Sentry Error logging StatusPage Status page