rnow CLI - Reinforcement Learning platform command-line interface

Project description

Documentation

See the documentation for a technical overview of the platform and train your first agent

Quick Start

1. Install uv (Python package manager)

# macOS/Linux:
$ curl -LsSf https://astral.sh/uv/install.sh | sh

# Windows:
PS> powershell -c "irm https://astral.sh/uv/install.ps1 | iex"

2. Install ReinforceNow

uv init && uv venv --python 3.11
source .venv/bin/activate  # Windows: .\.venv\Scripts\Activate.ps1
uv pip install rnow

3. Authenticate

rnow login

4. Create & Run Your First Project

rnow init --template sft
rnow run

That's it! Your training run will start on ReinforceNow's infrastructure. Monitor progress in the dashboard.

ReinforceNow Graph

Core Concepts

Go from raw data to a reliable AI agent in production. ReinforceNow gives you the flexibility to define:

1. Reward Functions

Define how your model should be evaluated using the @reward decorator:

from rnow.core import reward, RewardArgs

@reward
async def accuracy(args: RewardArgs, messages: list) -> float:
    """Check if the model's answer matches ground truth."""
    response = messages[-1]["content"]
    expected = args.metadata["answer"]
    return 1.0 if expected in response else 0.0

→ Write your first reward function

2. Tools (for Agents)

Give your model the ability to call functions during training:

from rnow.core import tool

@tool
def search(query: str, max_results: int = 5) -> dict:
    """Search the web for information."""
    # Your implementation here
    return {"results": [...]}

→ Train an agent with custom tools

3. Training Data

Create a train.jsonl file with your prompts and reward assignments:

{"messages": [{"role": "user", "content": "Balance the equation: Fe + O2 → Fe2O3"}], "rewards": ["accuracy"], "metadata": {"answer": "4Fe + 3O2 → 2Fe2O3"}}
{"messages": [{"role": "user", "content": "Balance the equation: H2 + O2 → H2O"}], "rewards": ["accuracy"], "metadata": {"answer": "2H2 + O2 → 2H2O"}}
{"messages": [{"role": "user", "content": "Balance the equation: N2 + H2 → NH3"}], "rewards": ["accuracy"], "metadata": {"answer": "N2 + 3H2 → 2NH3"}}

→ Learn about training data format

Contributing

We welcome contributions! ❤️ Please open an issue to discuss your ideas before submitting a PR

Project details

Release history Release notifications | RSS feed

0.4.37

Feb 9, 2026

This version

0.4.36

Feb 9, 2026

0.4.35

Feb 8, 2026

0.4.34

Feb 8, 2026

0.4.33

Feb 7, 2026

0.4.32

Feb 5, 2026

0.4.31

Feb 5, 2026

0.4.30

Feb 5, 2026

0.4.29

Feb 5, 2026

0.4.28

Feb 4, 2026

0.4.27

Feb 4, 2026

0.4.26

Feb 4, 2026

0.4.25

Feb 4, 2026

0.4.24

Feb 4, 2026

0.4.23

Feb 4, 2026

0.4.22

Feb 3, 2026

0.4.21

Feb 3, 2026

0.4.20

Feb 3, 2026

0.4.17

Jan 27, 2026

0.4.16

Jan 27, 2026

0.4.15

Jan 26, 2026

0.4.14

Jan 26, 2026

0.4.13

Jan 25, 2026

0.4.12

Jan 24, 2026

0.4.11

Jan 24, 2026

0.4.10

Jan 22, 2026

0.4.9

Jan 22, 2026

0.4.8

Jan 20, 2026

0.4.7

Jan 20, 2026

0.4.6

Jan 20, 2026

0.4.5

Jan 20, 2026

0.4.4

Jan 20, 2026

0.4.3

Jan 20, 2026

0.4.2

Jan 20, 2026

0.4.1

Jan 20, 2026

0.4.0

Jan 20, 2026

0.3.25

Jan 20, 2026

0.3.24

Jan 20, 2026

0.3.23

Jan 20, 2026

0.3.22

Jan 20, 2026

0.3.21

Jan 20, 2026

0.3.20

Jan 19, 2026

0.3.19

Jan 19, 2026

0.3.18

Jan 17, 2026

0.3.17

Jan 15, 2026

0.3.16

Jan 15, 2026

0.3.15

Jan 15, 2026

0.3.14

Jan 14, 2026

0.3.13

Jan 14, 2026

0.3.12

Jan 14, 2026

0.3.11

Jan 13, 2026

0.3.10

Jan 13, 2026

0.3.9

Jan 13, 2026

0.3.8

Jan 12, 2026

0.3.7

Jan 12, 2026

0.3.6

Jan 12, 2026

0.3.5

Jan 12, 2026

0.3.4

Jan 12, 2026

0.3.2

Jan 11, 2026

0.3.1

Jan 3, 2026

0.3.0

Jan 3, 2026

0.2.9

Dec 27, 2025

0.2.8

Dec 26, 2025

0.2.7

Dec 26, 2025

0.2.6

Dec 14, 2025

0.2.5

Dec 13, 2025

0.2.4

Dec 13, 2025

0.2.1

Dec 8, 2025

0.2.0

Dec 8, 2025

0.1.9

Dec 8, 2025

0.1.8

Dec 6, 2025

0.1.6

Dec 6, 2025

0.1.5

Dec 6, 2025

0.1.4

Dec 6, 2025

0.1.3

Dec 6, 2025

0.1.2

Dec 6, 2025

0.1.1

Dec 6, 2025

0.1.0

Dec 6, 2025

Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

rnow-0.4.36.tar.gz (1.4 MB view details)

Uploaded Feb 9, 2026 Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

The dropdown lists show the available interpreters, ABIs, and platforms. Enable javascript to be able to filter the list of wheel files.

rnow-0.4.36-py3-none-any.whl (1.5 MB view details)

Uploaded Feb 9, 2026 Python 3

File details

Details for the file rnow-0.4.36.tar.gz.

File metadata

Download URL: rnow-0.4.36.tar.gz
Upload date: Feb 9, 2026
Size: 1.4 MB
Tags: Source
Uploaded using Trusted Publishing? Yes
Uploaded via: twine/6.1.0 CPython/3.13.7

File hashes

Hashes for rnow-0.4.36.tar.gz
Algorithm	Hash digest
SHA256	`beeaeff5bf5a5b93f49ab06236d4c9e8ab8cdf695cd6663ef03a9748e9798a2b`
MD5	`0a8b398d226eabb61ff790f0f857cc9f`
BLAKE2b-256	`11ad7fe357df5da66c3d71af637a302c496d51caf8cc605c3bfacc1930c0e265`

See more details on using hashes here.

File details

Details for the file rnow-0.4.36-py3-none-any.whl.

File metadata

Download URL: rnow-0.4.36-py3-none-any.whl
Upload date: Feb 9, 2026
Size: 1.5 MB
Tags: Python 3
Uploaded using Trusted Publishing? Yes
Uploaded via: twine/6.1.0 CPython/3.13.7

File hashes

Hashes for rnow-0.4.36-py3-none-any.whl
Algorithm	Hash digest
SHA256	`8bc1ae76c3afd93d7d8764ab3d29ff1a360905dee3e75649fcf9e4cef7440e56`
MD5	`641f13532307e84eea710e8d3ae51a81`
BLAKE2b-256	`8b5aef02aa456be59880d183a247ef5e6a089f571d92cd452ff40e9a94460a24`

See more details on using hashes here.

rnow 0.4.36

Navigation

Verified details

Maintainers

Unverified details

Meta

Project description

Documentation

Quick Start

1. Install uv (Python package manager)

2. Install ReinforceNow

3. Authenticate

4. Create & Run Your First Project

Core Concepts

1. Reward Functions

2. Tools (for Agents)

3. Training Data

Contributing

Project details

Verified details

Maintainers

Unverified details

Meta

Release history Release notifications | RSS feed

Download files

Source Distribution

Built Distribution

File details

File metadata

File hashes

File details

File metadata

File hashes