Skip to main content

Unified API for training and inference

Project description

SkyRL: A Modular Full-stack RL Library for LLMs

| Documentation | Twitter/X | Huggingface | Slack Workspace |


Overview

SkyRL is a full-stack RL library that provides the following components:

  • skyrl: Our new unified library for RL on your own hardware, with support for the Tinker API. skyrl combines our previous work:

    • skyrl-train: A modular, performant training framework for RL.
    • skyrl-tx: A cross-platform library implementing a backend for the Tinker API, with a unified engine for training and inference.
  • skyrl-agent: Our agent layer for training long-horizon, real-world agents. For exact reproduction of SkyRL-v0 results, please checkout to commit a0d50c482436af7fac8caffa4533616a78431d66.

  • skyrl-gym: Our gymnasium of tool-use tasks, including a library of math, coding, search and SQL environments implemented in the Gymnasium API.

Getting Started

For a guide on developing with SkyRL, take at look at our Development Guide docs.

For model training, checkout skyrl to start using, modifying, or building on top of the SkyRL training stack. See our quickstart docs to ramp up!

For building environments, checkout skyrl-gym to integrate your task in the simple gymnasium interface.

For agentic pipelines, check out skyrl-agent for our work on optimizing and scaling pipelines for multi-turn tool use LLMs on long-horizon, real-environment tasks.

For a list of supported models, see our Supported Models docs.

News

  • [2026/02/17] 🎉 SkyRL is officially integrated with Harbor! Train your terminal-use agent! [Blog]
  • [2026/02/13] 🎉 SkyRL now implements the Tinker API! Run any training script written in the Tinker API on your local GPUs with SkyRL! [Blog]
  • [2025/11/26] 🎉 We released SkyRL-Agent: An agent layer for efficient, multi-turn, long-horizon agent training and evaluation. [Paper]
  • [2025/10/06] 🎉 We released SkyRL tx: An open implementation of a backend for the Tinker API to run a Tinker-like service on their own hardware. [Blog]
  • [2025/06/26] 🎉 We released SkyRL-v0.1: A highly-modular, performant RL training framework. [Blog]
  • [2025/06/26] 🎉 We released SkyRL-Gym: A library of RL environments for LLMs implemented with the Gymnasium API. [Blog]
  • [2025/05/20] 🎉 We released SkyRL-SQL: a multi-turn RL training pipeline for Text-to-SQL, along with SkyRL-SQL-7B — a model trained on just 653 samples that outperforms both GPT-4o and o4-mini!
  • [2025/05/06] 🎉 We released SkyRL-v0: our open RL training pipeline for multi-turn tool use LLMs, optimized for long-horizon, real-environment tasks like SWE-Bench!

Links

Projects using SkyRL

Acknowledgement

This work is done at Berkeley Sky Computing Lab in collaboration with Anyscale, with generous compute support from AnyscaleDatabricks, NVIDIA, Lambda Labs, AMD, AWS, Modal, and Daytona.

We adopt many lessons and code from several great projects such as veRL, OpenRLHF, Search-R1, OpenReasonerZero, and NeMo-RL. We appreciate each of these teams and their contributions to open-source research!

Citation

If you find the work in this repository helpful, please consider citing:

@misc{cao2025skyrl,
  title     = {SkyRL-v0: Train Real-World Long-Horizon Agents via Reinforcement Learning},
  author    = {Shiyi Cao and Sumanth Hegde and Dacheng Li and Tyler Griggs and Shu Liu and Eric Tang and Jiayi Pan and Xingyao Wang and Akshay Malik and Graham Neubig and Kourosh Hakhamaneshi and Richard Liaw and Philipp Moritz and Matei Zaharia and Joseph E. Gonzalez and Ion Stoica},
  year      = {2025},
}
@misc{liu2025skyrlsql,
      title={SkyRL-SQL: Matching GPT-4o and o4-mini on Text2SQL with Multi-Turn RL},
      author={Shu Liu and Sumanth Hegde and Shiyi Cao and Alan Zhu and Dacheng Li and Tyler Griggs and Eric Tang and Akshay Malik and Kourosh Hakhamaneshi and Richard Liaw and Philipp Moritz and Matei Zaharia and Joseph E. Gonzalez and Ion Stoica},
      year={2025},
}
@misc{griggs2025skrylv01,
      title={Evolving SkyRL into a Highly-Modular RL Framework},
      author={Tyler Griggs and Sumanth Hegde and Eric Tang and Shu Liu and Shiyi Cao and Dacheng Li and Charlie Ruan and Philipp Moritz and Kourosh Hakhamaneshi and Richard Liaw and Akshay Malik and Matei Zaharia and Joseph E. Gonzalez and Ion Stoica},
      year={2025},
      note={Notion Blog}
}

Project details


Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

skyrl-0.3.0.tar.gz (797.9 kB view details)

Uploaded Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

skyrl-0.3.0-py3-none-any.whl (925.0 kB view details)

Uploaded Python 3

File details

Details for the file skyrl-0.3.0.tar.gz.

File metadata

  • Download URL: skyrl-0.3.0.tar.gz
  • Upload date:
  • Size: 797.9 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? No
  • Uploaded via: uv/0.9.4

File hashes

Hashes for skyrl-0.3.0.tar.gz
Algorithm Hash digest
SHA256 aab4ba1fcc194c1f6158760bc6291f2ffcebc0a96a08749e8a1e5fc99d173930
MD5 88c9d885fc87c26f74b384b0c962d10d
BLAKE2b-256 ffd633c743adced5b4918002239804406f11529e05b4a6822fc876771c6d9b4b

See more details on using hashes here.

File details

Details for the file skyrl-0.3.0-py3-none-any.whl.

File metadata

  • Download URL: skyrl-0.3.0-py3-none-any.whl
  • Upload date:
  • Size: 925.0 kB
  • Tags: Python 3
  • Uploaded using Trusted Publishing? No
  • Uploaded via: uv/0.9.4

File hashes

Hashes for skyrl-0.3.0-py3-none-any.whl
Algorithm Hash digest
SHA256 0110e9434a58d6eb2aa533dd06e6346c8303757800af3c2b5a0d3f341cbad414
MD5 c352a89659e93a8a97233fa9a6c6c595
BLAKE2b-256 790fc89d23009539945c0e75ae428bf25363ca5b24a70e0aabe3377367de03a4

See more details on using hashes here.

Supported by

AWS Cloud computing and Security Sponsor Datadog Monitoring Depot Continuous Integration Fastly CDN Google Download Analytics Pingdom Monitoring Sentry Error logging StatusPage Status page