hiddenlayer-openai-guardrails

Guardrails for the OpenAI Agents SDK

These details have not been verified by PyPI

Project description

HiddenLayer Guardrails for OpenAI Agents (Beta)

Drop-in replacement for the Agents SDK Agent that wires HiddenLayer guardrails into agent and tool execution. Agent input/output and every tool call are sent through HiddenLayer's analyze endpoint so prompt-injection and policy violations are caught automatically.

Note: OpenAI's native guardrails only support blocking content—they do not support redacting sensitive information from input or output. This library provides redaction capabilities through HiddenLayer's REDACT action, allowing you to sanitize content while still allowing the request to proceed.

Installation

pip install hiddenlayer-openai-guardrails

Configuration

Environment Variables

The following environment variables must be set for authentication:

HIDDENLAYER_CLIENT_ID - HiddenLayer API client ID (required)
HIDDENLAYER_CLIENT_SECRET - HiddenLayer API client secret (required)

Optional environment variables:

HIDDENLAYER_PROJECT_ID - HiddenLayer project ID for policy routing
HIDDENLAYER_REQUESTER_ID - Identifier for tracking requests (default: "hiddenlayer-openai-integration")

# Required
export HIDDENLAYER_CLIENT_ID="your-client-id"
export HIDDENLAYER_CLIENT_SECRET="your-client-secret"

# Optional
export HIDDENLAYER_PROJECT_ID="your-project-id"
export HIDDENLAYER_REQUESTER_ID="your-app-name"

HiddenLayerParams

Configure HiddenLayer behavior using the HiddenLayerParams object:

from hiddenlayer_openai_guardrails import HiddenLayerParams

params = HiddenLayerParams(
    project_id="my-project",       # Optional: HiddenLayer project ID for policy routing
    model="gpt-4o-mini",            # Optional: Model name for tracking (auto-detected from agent if not set)
    requester_id="my-app-v1",      # Optional: Identifier for tracking requests (default: "hiddenlayer-openai-integration")
)

All fields are optional. If model is not provided, it will be automatically detected from the agent's model configuration.

Usage

Basic Agent with Guardrails

The Agent class mirrors agents.Agent but adds HiddenLayer guardrails to the agent and all tools. Guardrails automatically block malicious content:

from agents import Runner, function_tool
from agents.run import RunConfig
from hiddenlayer_openai_guardrails import Agent, HiddenLayerParams


@function_tool
def get_weather(city: str) -> str:
    """returns weather info for the specified city."""
    return f"The weather in {city} is sunny"


# Configure HiddenLayer parameters
params = HiddenLayerParams(
    project_id="my-project",  # optional: for policy routing
)

agent = Agent(
    name="Haiku agent",
    instructions="Always respond in haiku form",
    model="gpt-4o-mini",
    tools=[get_weather],  # tool input/output are screened by HiddenLayer
    hiddenlayer_params=params,  # optional: defaults will be used if not provided
)

result = Runner.run_sync(
    agent,
    "What's the weather in Toronto",
    run_config=RunConfig(tracing_disabled=True),
)
print(result.final_output)

Redacting Input and Output

Since OpenAI's guardrails can only block (not redact), this library provides helper functions for content redaction:

from agents import Runner
from hiddenlayer_openai_guardrails import (
    Agent,
    HiddenLayerParams,
    redact_input,
    redact_output,
    InputBlockedError,
    OutputBlockedError,
)

# Configure HiddenLayer parameters
params = HiddenLayerParams(project_id="my-project")

agent = Agent(
    name="Assistant",
    instructions="You are a helpful assistant.",
    hiddenlayer_params=params,
)

try:
    # Redact sensitive info from user input before processing
    safe_input = await redact_input(
        user_input,
        hiddenlayer_params=params,
    )

    # Run agent (guardrails will block malicious content)
    result = await Runner.run(agent, safe_input)

    # Redact sensitive info from output before showing to user
    safe_output = await redact_output(
        result.final_output,
        hiddenlayer_params=params,
    )
    print(safe_output)

except InputBlockedError:
    print("Input was blocked by HiddenLayer")
except OutputBlockedError:
    print("Output was blocked by HiddenLayer")

Redacting Streamed Output

For streaming responses, use redact_streamed_output to buffer, scan, and replay content:

from agents import Runner
from hiddenlayer_openai_guardrails import Agent, HiddenLayerParams, redact_streamed_output

# Configure HiddenLayer parameters
params = HiddenLayerParams(project_id="my-project")

agent = Agent(
    name="Assistant",
    instructions="Help users",
    hiddenlayer_params=params,
)
result = Runner.run_streamed(agent, user_input)

async for chunk in redact_streamed_output(result, hiddenlayer_params=params):
    print(chunk, end="", flush=True)

How it works

hiddenlayer_openai_guardrails.agents.Agent returns a regular agents.Agent configured with:
- Agent-level input/output guardrails that analyze user and assistant messages.
- Tool-level guardrails that inspect arguments before execution and outputs afterward.
Guardrails rely on AsyncHiddenLayer.interactions.analyze and will raise when HiddenLayer signals a blocking action.

Development

Run tests after installing dev deps (pytest and pytest-asyncio): pytest tests
Code lives in src/hiddenlayer_openai_guardrails/agents.py; tests are in tests/test_agents.py.

Project details

These details have not been verified by PyPI

Release history Release notifications | RSS feed

0.4.0

Apr 8, 2026

0.3.0

Feb 27, 2026

This version

0.2.0

Feb 10, 2026

0.1.0

Jan 13, 2026

Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

hiddenlayer_openai_guardrails-0.2.0.tar.gz (8.1 kB view details)

Uploaded Feb 10, 2026 Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

The dropdown lists show the available interpreters, ABIs, and platforms. Enable javascript to be able to filter the list of wheel files.

hiddenlayer_openai_guardrails-0.2.0-py3-none-any.whl (9.6 kB view details)

Uploaded Feb 10, 2026 Python 3

File details

Details for the file hiddenlayer_openai_guardrails-0.2.0.tar.gz.

File metadata

Download URL: hiddenlayer_openai_guardrails-0.2.0.tar.gz
Upload date: Feb 10, 2026
Size: 8.1 kB
Tags: Source
Uploaded using Trusted Publishing? Yes
Uploaded via: twine/6.1.0 CPython/3.13.7

File hashes

Hashes for hiddenlayer_openai_guardrails-0.2.0.tar.gz
Algorithm	Hash digest
SHA256	`01f0d59e1ded7c11759d74d46352ff8fb60247afcdc79b54e5dec0e07f99898d`
MD5	`6bc5aa053ee33e2af32d0d32bfc85be1`
BLAKE2b-256	`358ebcadf389349f40ef6683f4c8868ae32c533cad613883c94fb48d14425ae1`

See more details on using hashes here.

Provenance

The following attestation bundles were made for hiddenlayer_openai_guardrails-0.2.0.tar.gz:

Publisher: publish.yml on hiddenlayerai/hiddenlayer-openai-guardrails

Attestations: Values shown here reflect the state when the release was signed and may no longer be current.

Statement:
- Statement type: https://in-toto.io/Statement/v1
- Predicate type: https://docs.pypi.org/attestations/publish/v1
- Subject name: hiddenlayer_openai_guardrails-0.2.0.tar.gz
- Subject digest: 01f0d59e1ded7c11759d74d46352ff8fb60247afcdc79b54e5dec0e07f99898d
- Sigstore transparency entry: 938445549
- Sigstore integration time: Feb 10, 2026
Source repository:
- Permalink: hiddenlayerai/hiddenlayer-openai-guardrails@d926275b0c582e19e88243a55cd729ad2a807f92
- Branch / Tag: refs/heads/main
- Owner: https://github.com/hiddenlayerai
- Access: public
Publication detail:
- Token Issuer: https://token.actions.githubusercontent.com
- Runner Environment: github-hosted
- Publication workflow: publish.yml@d926275b0c582e19e88243a55cd729ad2a807f92
- Trigger Event: workflow_dispatch

File details

Details for the file hiddenlayer_openai_guardrails-0.2.0-py3-none-any.whl.

File metadata

Download URL: hiddenlayer_openai_guardrails-0.2.0-py3-none-any.whl
Upload date: Feb 10, 2026
Size: 9.6 kB
Tags: Python 3
Uploaded using Trusted Publishing? Yes
Uploaded via: twine/6.1.0 CPython/3.13.7

File hashes

Hashes for hiddenlayer_openai_guardrails-0.2.0-py3-none-any.whl
Algorithm	Hash digest
SHA256	`763495bc13ed3da803a6d0a75157652d03a5791c56eb3531fe34afecbfd8e75b`
MD5	`f4394aaa1a9204cf079643da653b2721`
BLAKE2b-256	`f4ea9ce415dd24bac6ca48ec8dccea54e35fb527a10489d5a175bc9b3413f4f6`

See more details on using hashes here.

Provenance

The following attestation bundles were made for hiddenlayer_openai_guardrails-0.2.0-py3-none-any.whl:

Publisher: publish.yml on hiddenlayerai/hiddenlayer-openai-guardrails

Attestations: Values shown here reflect the state when the release was signed and may no longer be current.

Statement:
- Statement type: https://in-toto.io/Statement/v1
- Predicate type: https://docs.pypi.org/attestations/publish/v1
- Subject name: hiddenlayer_openai_guardrails-0.2.0-py3-none-any.whl
- Subject digest: 763495bc13ed3da803a6d0a75157652d03a5791c56eb3531fe34afecbfd8e75b
- Sigstore transparency entry: 938445558
- Sigstore integration time: Feb 10, 2026
Source repository:
- Permalink: hiddenlayerai/hiddenlayer-openai-guardrails@d926275b0c582e19e88243a55cd729ad2a807f92
- Branch / Tag: refs/heads/main
- Owner: https://github.com/hiddenlayerai
- Access: public
Publication detail:
- Token Issuer: https://token.actions.githubusercontent.com
- Runner Environment: github-hosted
- Publication workflow: publish.yml@d926275b0c582e19e88243a55cd729ad2a807f92
- Trigger Event: workflow_dispatch

hiddenlayer-openai-guardrails 0.2.0

Navigation

Verified details

Maintainers

Unverified details

Meta

Classifiers

Project description

HiddenLayer Guardrails for OpenAI Agents (Beta)

Installation

Configuration

Environment Variables

HiddenLayerParams

Usage

Basic Agent with Guardrails

Redacting Input and Output

Redacting Streamed Output

How it works

Development

Project details

Verified details

Maintainers

Unverified details

Meta

Classifiers

Release history Release notifications | RSS feed

Download files

Source Distribution

Built Distribution

File details

File metadata

File hashes

Provenance

File details

File metadata

File hashes

Provenance