Skip to main content
Pre-release

This release is a pre-release and may not be stable for production use.

Dream-RSI SDK · Alpha

Bring your agent. Record its search. Evolve executable exploration policies.

An independent, unofficial Python SDK inspired by Dream-RSI research. Maintained by TheAstrayDev, who is not a Google or Google DeepMind employee. This is a personal research initiative, not a commercial Google development, official SDK or endorsed product.

Python 3.11+ · zero core third-party dependencies · Apache-2.0 · version 0.2.0a2.

Install

python -m pip install --upgrade dreamrsi==0.2.0a2

The package includes the Python API and the dreamrsi command. Publishing policy packages to GitHub additionally requires GitHub CLI (gh).

Third-party policy and replay packages

Version 0.2.0a2 adds portable JSON bundles and GitHub-hosted package sharing. Community developers can distribute policy versions and recorded discovery trees; recipients import them into local SQLite memory with their own agent, quality metric, task-family contract and acceptance rules. These replay datasets contain observed search transitions, not model weights. Saved policies can avoid repeated development when they meet the recipient's quality floor; new tasks still run the recipient's agent.

dreamrsi list
dreamrsi install OWNER/REPOSITORY
dreamrsi save my-policy-pack --all
dreamrsi publish .dreamrsi/packages/exports/my-policy-pack.dreamrsi.json

Use a real package reference for OWNER/REPOSITORY. Export requires previously saved memory; publication requires GitHub CLI and creates a public repository. See the complete walkthrough for the diagram, executable integration example, and compatibility requirements.

Basic agent integration

from dreamrsi import Budget, DreamRSI

rsi = DreamRSI(
    agent=lambda task: task.upper(),
    evaluator=lambda answer: float(len(answer)),
    budget=Budget(model_calls=4),
)
result = rsi.run_sync("hello")
print(result.best)  # HELLO

This example demonstrates integration, not quality improvement. Stateful adapters support generate/evaluate/refine tasks. Replay uses recorded outcomes without new discovery calls. LLMPolicyDeveloper writes and iteratively rewrites executable source from measured feedback. Version 0.2.0a1 adds opt-in persistent champion reuse and more conservative replay cost accounting. It can replay already recorded runs without a new training call, delays holdout collection until a replay-improving candidate exists, and stops repeated policy revisions. The default promotion gate requires paired raw-quality and probe evidence; applications needing score-only decisions can explicitly use ReplayOnlyGate. An end-to-end cost advantage for an LLM discovery agent has not been proven. The SDK includes configurable Docker-free policy interpreters, optional process execution, held-out validation, SQLite recovery and reported token/USD accounting.

Two measured experiments

In the Bonsai Q2 experiment, a real local model wrote executable policy code while a deterministic agent solved the tasks. Across 64 fresh fixture tasks, counted operations fell from 256 to 248 with raw quality 0.9 throughout. This is a logical-operation proxy, not a measured token or dollar saving for an LLM agent.

In the GPT-6 Luna xhigh experiment, the real model both solved tasks and developed policies. Deployment fell from 24 to six model requests across six held-out tasks, but 94 preparation requests made the full path 100 versus 24. Mean reported score was slightly lower. This demonstrates learning and reuse, not an all-in economic win.

Alpha APIs may change. Generated source runs in a bounded Python-syntax subset, not arbitrary Python. External state isolation and remote cancellation require adapter support. Usage caps depend on accurate provider reports and ceilings. Research-scale performance and broad generalization remain unverified.

Metadata

Release files for dreamrsi 0.2.0a2

For a detailed explanation of source distributions (sdists) and built distributions (wheels), please see the package formats documentation.

Source distribution (sdist)

Source distribution for dreamrsi 0.2.0a2
File Size Uploaded
dreamrsi-0.2.0a2.tar.gz 2.3 MB Details

Built distribution (wheel)

Table of built distributions (wheels) for dreamrsi 0.2.0a2
File Interpreter ABI Platform
dreamrsi-0.2.0a2-py3-none-any.whl Python 3 none any Details

Total release size: 2.5 MB

Release files / dreamrsi-0.2.0a2.tar.gz

Download URL dreamrsi-0.2.0a2.tar.gz
Size 2.3 MB
Tags Source
SHA-256 checksum
How to use checksums
edc25bf58815cf0e1df3689bf53aad3423f12dedacb1bad106eaa1ddd8e9daed
BLAKE2b-256 checksum
How to use checksums
6bfc1fdec571ef566bb91e9b38299a1a2a0f6913bf167b09f248dc1e6aa8c973
Upload date
Uploaded using Trusted Publishing?
What is trusted publishing?
Yes
Uploaded via twine/7.0.0 CPython/3.13.14

Provenance

Provenance describes where a file came from. On PyPI, provenance is shared via attestations, which provide a verifiable record of the build or publishing details. View details, limitations and caveats.

PyPI Publish Attestation

PyPI verified that this artifact, at this checksum, originated from the publisher listed below.

Signed by GitHub Actions, verified by PyPI on Sep 26, 2026.

Transparency log

Release files / dreamrsi-0.2.0a2-py3-none-any.whl

Download URL dreamrsi-0.2.0a2-py3-none-any.whl
Size 137.1 kB
Tags Python 3
SHA-256 checksum
How to use checksums
84f3e9840992bfc438e26884046161590863456931d6c1248546b90d4a3fae3c
BLAKE2b-256 checksum
How to use checksums
352a0cec39f9e7dd67ce2e63c5a8d739a45cd9992c4d5cfc23ea0712e4084da6
Upload date
Uploaded using Trusted Publishing?
What is trusted publishing?
Yes
Uploaded via twine/7.0.0 CPython/3.13.14

Provenance

Provenance describes where a file came from. On PyPI, provenance is shared via attestations, which provide a verifiable record of the build or publishing details. View details, limitations and caveats.

PyPI Publish Attestation

PyPI verified that this artifact, at this checksum, originated from the publisher listed below.

Signed by GitHub Actions, verified by PyPI on Sep 26, 2026.

Transparency log
Anthropic, PBC Visionary sponsor Bloomberg Visionary sponsor Hudson River Trading Visionary sponsor Meta Visionary sponsor NVIDIA Visionary sponsor Microsoft Sustainability sponsor Depot Continuous Integration AWS Cloud computing and Security Sponsor Datadog Monitoring Fastly CDN Google Download Analytics Sentry Error logging StatusPage Status page