Skip to main content

CTRLRun

CTRLRun verified

Transaction safety for AI-agent actions.

Agents can retry. Your refund shouldn't.

CTRLRun is open-source infrastructure for controlling consequential AI-agent actions.

You decide, per action, what an agent can do autonomously, what requires human approval, and what is blocked.

CTRLRun binds approvals to the exact action, blocks duplicate execution attempts for the same logical effect, stops blind retries when an execution outcome is uncertain, and records what actually happened.

Autonomy belongs to the action, not the agent.


The problem

Agents are getting write access to the real world: refunds, emails, deploys, permission grants, record changes. Frameworks already let you approve or deny a tool call. That is not the hard part.

The hard part is what happens at the boundary between intention and effect:

  • A refund commits at Stripe, the response times out, the agent retries. Did you just refund twice?
  • A human approves a €500 refund. The agent changes it to €5,000 before executing. Should that approval still count?
  • Two agents pick up the same task and issue the same refund. Which one wins?
  • A call timed out. Your framework marks it failed and retries. It wasn't failed. It was unknown.

CTRLRun owns that boundary.

Quick start

pip install ctrlrun
ctrlrun demo

ctrlrun demo runs the five failure scenarios below in well under a second, with no external services.

Protect your first action:

import ctrlrun


@ctrlrun.protect("stripe.refund", effect="refund:{payment_id}")
def refund(payment_id: str, amount: int, currency: str = "EUR"):
    return stripe.refunds.create(payment_intent=payment_id, amount=amount)

Configure autonomy per action in ctrlrun.yaml:

schema: ctrlrun.policy/v1

actions:
  customer.read:
    decision: allow

  stripe.refund:
    rules:
      - when: { amount_gte: 0, amount_lte: 50000 }      # up to €500.00
        decision: allow
      - when: { amount_gte: 0, amount_lte: 500000 }     # up to €5,000.00
        decision: approve
      - decision: deny

Amounts are integer minor units — cents, not euros. Floats are rejected outright, because 0.1 and 0.10 are the same money and different hashes.

Now the same agent can refund €100 on its own, must get a human to approve €2,000, and cannot refund €20,000 at all. Neither can it refund a negative amount, which is a charge wearing a refund's name — an upper bound alone is not a range. Unknown actions are denied. CTRLRun fails closed.

Protect an existing MCP server

No agent changes. Point the client at the gateway instead of at the tool server:

pip install "ctrlrun[gateway]"
ctrlrun gateway --upstream http://localhost:8000/mcp --alias acme --principal refund-agent

The gateway prints, on the line that starts it, every action in your policy that has no effect: template — because a write with no effect key is exactly the configuration this exists to prevent, and it should not be discovered in a receipt three weeks later:

1 action(s) have no effect: template and get no reservation:
  mcp.acme.list_payments
That is right for a read, and wrong for anything that changes the world.

Tools become actions named mcp.<alias>.<tool>, decided by the same ctrlrun.yaml. Declare their effect and resource templates there, since a tool call has no decorator to carry them:

schema: ctrlrun.policy/v2

actions:
  mcp.acme.create_refund:
    effect: "refund:{payment_id}"
    resource: "payment:{payment_id}"
    decision: approve

Everything but tools/call is relayed untouched. A lost response over the wire blocks the retry exactly as it does in-process — that is the whole point of putting it here.

Who is acting, and what are they entitled to?

A policy decides how much autonomy an action has. It cannot see who is asking — deliberately, since v0.1. That answers "may a €50,000 refund run without a human" and not "may this agent propose a €50,000 refund at all", and a system that can only ask the first will eventually answer the second by accident.

authority: is the second axis. It is opt-in, and then fail-closed: a policy with no authority: section behaves exactly as it did before, and the moment one exists every principal needs a grant and no grant means denied.

schema: ctrlrun.policy/v3

authority:
  grants:
    - id: head-of-support
      subject: { agent: "head-of-support", user: "dana@example.com" }
      actions: ["stripe.refund"]
      constraints: { amount_lte: 10000000 }   # €100,000.00
      delegable: true
      expires_at: "2027-01-01T00:00:00Z"

A grant carries no decision:. How much autonomy stripe.refund has is the same for everybody; what differs is whether they may ask. The two axes are evaluated separately, authority first, and combine as the stricter of the two — so neither can loosen the other.

A principal holding a delegable grant can narrow it at runtime with ctrlrun delegate, and a delegated grant is valid only if it is provably a subset of its parent on every dimension, at creation and again at every evaluation. A check performed only at creation would leave every delegation exactly as wide as the file used to be, which is the shape of every stale-permission incident there has ever been. Omitting a dimension the parent constrains is rejected, not inherited — a child that drops resources: would authorize resources its parent never could. ctrlrun revoke cuts a chain of any depth with one write.

Roll it out with mode: observe first. One top-level line runs every real decision against real traffic and records what would have been blocked, without blocking anything — then ctrlrun stats gives you the numbers before you enforce them. It is not a dry run: it executes, and effects land at remotes. It is a way to learn what enforcement will cost.

Identity is consumed, never invented. pip install "ctrlrun[identity]" adds a JWTIdentityProvider that verifies a bearer token against a JWKS or a pinned key and maps the verified claims onto a principal. CTRLRun issues no credential and defines no identity format.

Does it hold in your setup?

Everything above is proven by this repository's tests against this repository's configurations. That is the right place to start and the wrong place to stop, because what you deploy is your policy, your grants and your store.

$ ctrlrun verify
CTRLRun verify — ctrlrun 0.4.0, catalogue ctrlrun.guarantees/v1
policy     examples/authority/payments.yaml (ctrlrun.policy/v3, mode: enforce)
authority  same document, 3 grants
store      sqlite, scratch (created and destroyed for this run)

G1   mutated approval refused         PASS  stripe.refund
G2   replayed approval refused        PASS  stripe.refund
G3   duplicate effect refused         PASS  stripe.refund
G4   one winner under concurrency     PASS  stripe.refund (8 processes)
G5   ambiguous blocks a blind retry   PASS  stripe.refund
G6   unknown action refused           PASS
G7   no principal refused             PASS  stripe.refund
G8   expired authority refused        PASS  head-of-support
G9   delegation cannot escalate       PASS  head-of-support (6 of 6 dimensions)
G10  unknown exception is ambiguous   PASS  stripe.refund

10/10 declared guarantees pass. 0 not applicable.

It runs the kernel's own failure scenarios against the configuration in front of it, in a scratch store, with fake executors, and no network. Your .ctrlrun/state.db is byte-identical before and after.

Not applicable is not a pass. A policy with no approve rule cannot exercise the approval-binding guarantees, so they are reported N/A with the reason, excluded from the denominator and listed separately — 5/5 (5 not applicable), never 10/10. There is no flag that folds one into the count.

The badge means the declared guarantees pass — every guarantee this configuration can exercise was exercised, and none of them failed. It does not mean secure, safe, compliant, certified or audited, and docs/verify.md says on the same screen what verify cannot see: your executors, your reconcile hooks, where you put the decorator, your deployment, and whether your policy is the right policy.

There is a GitHub Action:

      - uses: CTRLRun/ctrlrun@main
        with:
          policy: ctrlrun.yaml

Two more things

Resolving an unknown outcome without a human. @protect(..., reconcile=...) takes a function that asks the remote what happened to an effect, and it is the only thing besides a human permitted to move a record out of AMBIGUOUS — and only in the direction its answer points.

Exporting to your tracing backend. pip install "ctrlrun[otel]" adds an OTelEventSink: one OpenTelemetry span per action, one span event per step. Argument values stay out of it unless you ask for them.

What ctrlrun demo shows

$ ctrlrun demo
CTRLRun demo — five ways an agent action goes wrong, and what stops it.
Policy: refunds up to €1,000 are autonomous, up to €10,000 need a human, above that are denied.

1. Duplicate effect after a lost response

   refund €500  →  remote commits  →  response lost  →  effect: AMBIGUOUS
   agent retries the same refund
   ✗ BLOCKED — effect may already have committed; blind retry refused
   remote refund calls: 1
   only a human moves it on:  ctrlrun resolve refund:txn_1 --committed|--failed

2. Approval mutation

   agent proposes refund €2,000  →  human approves apr_0aa78e0380ba55d77a601dc782f57095 (bound to the action hash)
   agent executes refund €5,000  →
   ✗ BLOCKED — approved action ≠ requested action (mismatch)

3. Concurrent agents, same effect

   Agent A  reserve refund:txn_123  →  ACQUIRED  →  executes
   Agent B  reserve refund:txn_123  →
   ✗ BLOCKED — already reserved (in_progress)

4. Approval replay

   approval apr_dbc8bc6f06690cdf2e2c55a4e591ef3b used once  →  consumed
   same approval presented again                            →
   ✗ BLOCKED — single-use approval already consumed

5. Authority escalation

   human €100,000 delegable  →  finance agent €25,000  →  support agent €2,000
   support agent's grant: dlg_5f8d41938a3f29972d5489d676cd9edb
   support agent requests €50,000  →
   ✗ BLOCKED — outside the delegated grant (authority_constraint)
   remote refund calls: 0
   finance agent tries to delegate €50,000 under its own €25,000  →  refused (containment: constraints)
   support agent requests €1,500  →  authority permits it, and the policy asks a human (apr_f86eca24dd80206ab5189ccb1b62aa55)
   two axes, and an action needs both: the stricter of the pair wins

Receipts (8): .ctrlrun/demo/receipts.jsonl
Events:       .ctrlrun/demo/events.jsonl

Read them:    CTRLRUN_STATE=.ctrlrun/demo/state.db ctrlrun receipts

Approval ids are generated per run; everything else is byte-for-byte what the demo prints.

Note scenario 1: remote refund calls: 1. The refund committed at the remote, the response was lost, and the retry was refused — so the customer was refunded once, not twice. Nothing but a human resolving the effect moves it on.

Every executed action produces a portable JSON receipt: who, what, arguments, decision, approval, effect key, and result (committed, failed, or ambiguous).

What CTRLRun is not

CTRLRun does not host models, plan, prompt, retrieve, route, remember, or orchestrate. It is not a guardrail library, an IAM system, a workflow engine, or a compliance product. It issues no credential: authority: decides what an identity somebody else vouched for is entitled to do, and CTRLRun runs no authorization server, mints no token, and performs no OAuth flow. If an agent only reads and answers, you don't need CTRLRun. The moment it can send, pay, refund, delete, deploy, grant, revoke, approve, submit, purchase, or cancel, you do.

CTRLRun cannot guarantee exactly-once execution against external systems it doesn't control. It guarantees that it will not knowingly execute the same logical effect twice, and that it will never treat an unknown outcome as a failure.

Documentation

Doc Purpose
docs/SPEC-v0.1.md The v0.1 contract: models, invariants, acceptance tests
docs/SPEC-v0.2.md The v0.2 delta: gateway, sinks, reconciliation, webhooks
docs/SPEC-v0.3.md The v0.3 delta: identity, authority, delegation, observe mode
docs/SPEC-v0.4.md The v0.4 delta: the guarantee catalogue, the scenario engine, the badge
docs/verify.md ctrlrun verify, the guarantees, the N/A rule, and what the badge means
docs/OWASP-AGENTIC-TOP10.md A reading of the OWASP Top 10 for Agentic Applications against the guarantees
docs/authority.md Grants, delegation and the omission rule, in plain language
docs/ACS.md The OWASP Agent Control Standard: what maps, and where it is silent
docs/ARCHITECTURE.md Kernel design and key decisions
docs/ROADMAP.md v0.1 → v1.0
docs/THREAT_MODEL.md What CTRLRun defends against and what it doesn't
docs/CLAIMS.md Every claim above, mapped to the code and the test that proves it
SECURITY.md Reporting a vulnerability
VISION.md Where this can go — not a build spec

License

Apache-2.0. The enforcement kernel is and will remain fully open source.

Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

ctrlrun-0.4.0.tar.gz (633.7 kB view details)

Uploaded Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

ctrlrun-0.4.0-py3-none-any.whl (215.8 kB view details)

Uploaded Python 3

File details

Details for the file ctrlrun-0.4.0.tar.gz.

File metadata

  • Download URL: ctrlrun-0.4.0.tar.gz
  • Upload date:
  • Size: 633.7 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? Yes
  • Uploaded via: twine/7.0.0 CPython/3.13.14

File hashes

Hashes for ctrlrun-0.4.0.tar.gz
Algorithm Hash digest
SHA256 f34072b5fd7a42a705d8fd939956d776d87499b3e2252bcf63e5cd569276657e
MD5 1da4502048581701ee7135ed11dd19e9
BLAKE2b-256 7f742565fa381bb072f33ca36f3de8e2d89ad9b9e98a6a2c1413373c6f47e149

See more details on using hashes here.

Provenance

The following attestation bundles were made for ctrlrun-0.4.0.tar.gz:

Publisher: publish.yml on CTRLRun/ctrlrun

Attestations: Values shown here reflect the state when the release was signed and may no longer be current.

File details

Details for the file ctrlrun-0.4.0-py3-none-any.whl.

File metadata

  • Download URL: ctrlrun-0.4.0-py3-none-any.whl
  • Upload date:
  • Size: 215.8 kB
  • Tags: Python 3
  • Uploaded using Trusted Publishing? Yes
  • Uploaded via: twine/7.0.0 CPython/3.13.14

File hashes

Hashes for ctrlrun-0.4.0-py3-none-any.whl
Algorithm Hash digest
SHA256 5bb869906293c0f73df43b82905506c19083a1890480d75074270ee5e4e9cf61
MD5 afefd832a1801599bcfb336ae410f950
BLAKE2b-256 e9d06da13e4ccc36e13fd894a4736c561b93de768dc52f67a19ed71cff5506e3

See more details on using hashes here.

Provenance

The following attestation bundles were made for ctrlrun-0.4.0-py3-none-any.whl:

Publisher: publish.yml on CTRLRun/ctrlrun

Attestations: Values shown here reflect the state when the release was signed and may no longer be current.

Release history Release notifications | RSS feed

0.5.0

2 files

This release

0.4.0 This release

2 files

0.2.0

2 files

0.1.0

2 files

Anthropic, PBC Visionary sponsor Bloomberg Visionary sponsor Hudson River Trading Visionary sponsor Meta Visionary sponsor NVIDIA Visionary sponsor Microsoft Sustainability sponsor Depot Continuous Integration AWS Cloud computing and Security Sponsor Datadog Monitoring Fastly CDN Google Download Analytics Sentry Error logging StatusPage Status page