Skip to main content

A lightweight utility library for reinforcement learning projects in JAX and Equinox.

Project description

JymKit: A Lightweight Utility Library for JAX-based RL Projects

JymKit lets you

  1. 🕹️ Import your favourite environments from various libraries with a single API and automatically wrap them to a common standard.
  2. 🚀 Bootstrap new JAX RL projects with a single CLI command and get started instantly with a complete codebase.
  3. 🤖 JymKit comes equiped with standard general RL implementations based on a near-single-file philosophy. You can either import these as off-the-shelf algorithms or copy over the code and tweak them for your problem. These algorithms follow the ideas of PureJaxRL for extremely fast end-to-end RL training in JAX.

📖 More details over at the Documentation

🚀 Getting started

JymKit lets you bootstrap your new reinforcement learning projects directly from the command line. As such, for new projects, the easiest way to get started is via uv:

uvx jymkit <projectname>
uv run example_train.py

# ... or via pipx
pipx run jymkit <projectname>
# activate a virtual environment in your preferred way, e.g. conda
python example_train.py

This will set up a Python project folder structure with (optionally) an environment template and (optionally) algorithm code for you to tailor to your problem.

For existing projects, you can simply install JymKit via pip and import the required functionality.

pip install jymkit
import jax
import jymkit as jym
from jymkit.algorithms import PPO

env = jym.make("CartPole")
env = jymkit.LogWrapper(env)
rng = jax.random.PRNGKey(0)
agent = PPO(total_timesteps=5e5, learning_rate=2.5e-3)
agent = agent.train(rng, env)

🏠 Environments

JymKit is not aimed at delivering a full environment suite. However, it does come equipped with a jym.make(...) command to import environments from existing suites (provided that these are installed) and wrap them appropriately to the JymKit API standard. For example, using environments from Gymnax:

import jymkit as jym
from jymkit.algorithms import PPO
import jax

env = jym.make("Breakout-MinAtar")
env = jym.FlattenObservationWrapper(env)
env = jym.LogWrapper(env)

agent = PPO(**some_good_hyperparameters)
agent = agent.train(jax.random.PRNGKey(0), env)

# > Using an environment from Gymnax via gymnax.make(Breakout-MinAtar).
# > Wrapping Gymnax environment with GymnaxWrapper
# >  Disable this behavior by passing wrapper=False
# > Wrapping environment in VecEnvWrapper
# > ... training results

For convenience, JymKit does include the 5 classic-control environments.

Currently, importing from external libraries is possible for Gymnax and Brax. More are coming up!

Environment API

The JymKit API stays close to the somewhat established Gymnax API for the reset() and step() functions, but allows for truncated episodes in a manner closer to Gymnasium.

env = jym.make(...)

obs, env_state = env.reset(key) # <-- Mirroring Gymnax

# env.step(): Gymnasium Timestep tuple with state information
(obs, reward, terminated, truncated, info), env_state = env.step(key, state, action)

🤖 Algorithms

Algorithms in jymkit.algorithms are built following a near-single-file implementation philosophy in mind. In contrast to implementations in CleanRL or PureJaxRL, JymKit algorithms are built in Equinox and follow a class-based design with a familiar Stable-Baselines API.

Each algorithm supports both discrete- and continuous action/observation space -- adjusting based on the provided environment observation_space and action_space. Additionally, the implementations support multi-agent environments out of the box.

from jymkit.algorithms import PPO
import jax

env = ...
agent = PPO(**some_good_hyperparameters)
agent = agent.train(jax.random.PRNGKey(0), env)

Currently, only a PPO implementation is implemented. More will be included in the near future. However, the current goal is not to include as many algorithms as possible.

Project details


Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

jymkit-0.0.12.tar.gz (39.5 kB view details)

Uploaded Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

jymkit-0.0.12-py3-none-any.whl (56.0 kB view details)

Uploaded Python 3

File details

Details for the file jymkit-0.0.12.tar.gz.

File metadata

  • Download URL: jymkit-0.0.12.tar.gz
  • Upload date:
  • Size: 39.5 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/6.1.0 CPython/3.11.13

File hashes

Hashes for jymkit-0.0.12.tar.gz
Algorithm Hash digest
SHA256 b7038f3337ee11687af6ef78fb373e7a42db0df1ed643e52af499f6b571f73cb
MD5 e3659a608311c5b0a057eea0207a5dd4
BLAKE2b-256 05fbeca77a145b868821dd3143af77ae48289539921d9135249272758c2b8c90

See more details on using hashes here.

File details

Details for the file jymkit-0.0.12-py3-none-any.whl.

File metadata

  • Download URL: jymkit-0.0.12-py3-none-any.whl
  • Upload date:
  • Size: 56.0 kB
  • Tags: Python 3
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/6.1.0 CPython/3.11.13

File hashes

Hashes for jymkit-0.0.12-py3-none-any.whl
Algorithm Hash digest
SHA256 28ac20435497b65cbd18bd89c1ea062eac126c1b4b6b2bbf113767825143f3e7
MD5 dd8297daf8a821b19ae2dcc2ce8f57c5
BLAKE2b-256 0840b6410f4e8fbb97ef77f256327993937df89700eeb14708cbfc2a9b58bcf4

See more details on using hashes here.

Supported by

AWS Cloud computing and Security Sponsor Datadog Monitoring Depot Continuous Integration Fastly CDN Google Download Analytics Pingdom Monitoring Sentry Error logging StatusPage Status page