swekit

Tools for running a SWE agent using Composio platform

These details have not been verified by PyPI

Project links

Homepage

Project description

Production Ready Toolset for AI Agents

Build Software engineering Agents fast and easy!

SWEKIT Docs »

✨ Socials >> Discord | Youtube | Twitter | Linkedin

⛏️ Contribute >> Report Bugs | Request Feature | Contribute

Overview

swekit is a framework for building SWE agents on by utilising composio tooling ecosystem. SWE Kit allows you to

Scaffold agents which works out-of-the-box with choice of your agentic framework, crewai, llamaindex, etc...
Tools to add or optimise your agent's abilities
Benchmark your agents against SWE-bench

Dependencies

Before getting started, ensure you have the following set up:

Installation:
```
pip install swekit composio-core
```
Install agentic framework of your choice and the Composio plugin for the same: Here we're using crewai for the example:
```
pip install crewai composio-crewai
```
GitHub Access Token:

The agent requires a github access token to work with your repositories, You can create one at https://github.com/settings/tokens with necessary permissions and export it as an environment variable using export GITHUB_ACCESS_TOKEN=<your_token>
LLM Configuration: You also need to setup API key for the LLM provider you're planning to use. By default the agents scaffolded by swekit uses openai client, so export OPENAI_API_KEY before running your agent

Getting Started

Creating a new agent

Scaffold your agent using:
```
swekit scaffold crewai -o <path>
```
This creates a new agent in <path>/agent with four key files:
- main.py: Entry point to run the agent on your issue
- agent.py: Agent definition (edit this to customise behaviour)
- prompts.py: Agent prompts
- benchmark.py: SWE-Bench benchmark runner
Run the agent:
```
cd agent
python main.py
```
You'll be prompted for the repository name and issue.

Workspace Environment

The SWE-agent runs in Docker by default for security and isolation. This sandboxes the agent's operations, protecting against unintended consequences of arbitrary code execution.

The composio toolset has support for different types of workspaces.

Host - This will run on the host machine.

from composio import ComposioToolSet, WorkspaceType

toolset = ComposioToolSet(
    workspace_config=WorkspaceType.Host()
)

Docker - This will run inside a docker container

from composio import ComposioToolSet, WorkspaceType

toolset = ComposioToolSet(
    workspace_config=WorkspaceType.Docker()
)

On the docker container you can configure and expose the port for development as per your requirements. You can also use workspace.as_prompt() method to generate a workspace description for setting up your agent.

from composio import ComposioToolSet, WorkspaceType

toolset = ComposioToolSet(
    workspace_config=WorkspaceType.Docker(
        ports={
            8001: 8001,
        }
    )
)

You can read more about configuring docker ports here.

E2B - This will run inside a E2B Sandbox

from composio import ComposioToolSet, WorkspaceType

toolset = ComposioToolSet(
    workspace_config=WorkspaceType.E2B(),
)

FlyIO - This will run inside a FlyIO machine

from composio import ComposioToolSet, WorkspaceType

toolset = ComposioToolSet(
    workspace_config=WorkspaceType.FlyIO(),
)

FlyIO also allows for configuring ports for development/deployment.

from composio import ComposioToolSet, WorkspaceType

composio_toolset = ComposioToolSet(
    workspace_config=WorkspaceType.FlyIO(
        image="composio/composio",
        ports=[
            {
                "ports": [
                    {"port": 443, "handlers": ["tls", "http"]},
                ],
                "internal_port": 80,
                "protocol": "tcp",
            }
        ],
    )
)

You can read more abour configuring network ports on flyio machine here

Customising the workspace environment

The workspace environment contains following environment variables by default

COMPOSIO_API_KEY: The composio API key for interacting with composio API.
COMPOSIO_BASE_URL: Base URL for composio API server.
GITHUB_ACCESS_TOKEN: Github access token for the agent.
ACCESS_TOKEN: Access token for composio tooling server.

If you want to provide additional environment configuration you can use environment argument when creating a workspace configuration.

composio_toolset = ComposioToolSet(
    workspace_config=WorkspaceType.Docker(
        environment={
            "SOME_API_TOKEN": "<SOME_API_TOKEN>",
        }
    )
)

Running the Benchmark

SWE-Bench is a comprehensive benchmark designed to evaluate the performance of software engineering agents. It comprises a diverse collection of real-world issues from popular Python open-source projects, providing a robust testing environment.

To run the benchmark:

Ensure Docker is installed and running on your system.
Execute the following command:
```
cd agent
python benchmark.py --test-split=<test_split>
```
- By default, python benchmark.py runs only 1 test instance.
- Specify a test split ratio to run more tests, e.g., --test-split=1:300 runs 300 tests.

To run the benchmarks in E2B or FlyIO sandbox, you can set the workspace_env in the evaluate function call in benchmark.py

from composio import WorkspaceType

(...)

    evaluate(
        bench,
        dry_run=False,
        test_range=test_range,
        test_instance_ids=test_instance_ids_list,
        workspace_env=WorkspaceType.E2B
    )

To use E2B or FlyIO sandboxes you'll require API key for respective platforms, to use E2B export your API key as E2B_API_KEY and to use FlyIO export your API token as FLY_API_TOKEN.

Note: We utilize SWE-Bench-Docker to ensure each test instance runs in an isolated container with its specific environment and Python version.

To extend the functionality of the SWE agent by adding new tools or extending existing ones, refer to the Development Guide.

Project details

These details have not been verified by PyPI

Project links

Homepage

Release history Release notifications | RSS feed

0.2.42

Nov 21, 2024

0.2.41

Nov 20, 2024

0.2.40

Nov 14, 2024

0.2.39

Nov 11, 2024

0.2.38

Nov 6, 2024

0.2.37

Nov 6, 2024

0.2.36

Nov 4, 2024

0.2.35rc2 pre-release

Nov 2, 2024

0.2.34

Oct 28, 2024

This version

0.2.34rc1 pre-release

Nov 1, 2024

0.2.33

Oct 25, 2024

0.2.32

Oct 23, 2024

0.2.31

Oct 17, 2024

0.2.30

Oct 5, 2024

0.2.28

Sep 25, 2024

0.2.27

Sep 23, 2024

0.2.26

Sep 23, 2024

0.2.25

Sep 19, 2024

0.2.24

Sep 18, 2024

0.2.23

Sep 14, 2024

0.2.22

Sep 13, 2024

0.2.21

Sep 13, 2024

0.2.20

Sep 13, 2024

0.2.19

Sep 12, 2024

0.2.18

Sep 12, 2024

0.2.17

Sep 11, 2024

0.2.16

Sep 10, 2024

0.2.15

Sep 5, 2024

0.2.14

Sep 5, 2024

0.2.13

Sep 3, 2024

0.2.12

Sep 3, 2024

0.2.11

Aug 31, 2024

0.2.10

Aug 29, 2024

0.2.9

Aug 28, 2024

0.2.8

Aug 27, 2024

0.2.7

Aug 27, 2024

0.2.6

Aug 25, 2024

0.2.5

Aug 24, 2024

0.2.4

Aug 24, 2024

0.2.3

Aug 22, 2024

0.2.2

Aug 22, 2024

0.2.1

Aug 22, 2024

0.2.0

Aug 21, 2024

0.1.10

Aug 19, 2024

0.1.9

Aug 17, 2024

0.1.8

Aug 16, 2024

0.1.7

Aug 12, 2024

0.1.6

Aug 5, 2024

0.1.5

Jul 31, 2024

0.1.4

Jul 26, 2024

0.1.3

Jul 26, 2024

0.1.2

Jul 24, 2024

0.1.1

Jul 23, 2024

0.1.0

Jul 23, 2024

0.1.0rc4 pre-release

Jul 16, 2024

0.1.0rc3 pre-release

Jul 12, 2024

0.1.0rc2 pre-release

Jul 10, 2024

0.1.0rc1 pre-release

Jul 9, 2024

0.1.0rc0 pre-release

Jul 9, 2024

Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

swekit-0.2.34rc1.tar.gz (70.8 kB view details)

Uploaded Nov 1, 2024 Source

Built Distribution

swekit-0.2.34rc1-py3-none-any.whl (102.8 kB view details)

Uploaded Nov 1, 2024 Python 3

File details

Details for the file swekit-0.2.34rc1.tar.gz.

File metadata

Download URL: swekit-0.2.34rc1.tar.gz
Upload date: Nov 1, 2024
Size: 70.8 kB
Tags: Source
Uploaded using Trusted Publishing? No
Uploaded via: twine/5.1.1 CPython/3.12.7

File hashes

Hashes for swekit-0.2.34rc1.tar.gz
Algorithm	Hash digest
SHA256	`831c269e6a48966a71b1f89e8f57313f01120b99fe17ad899bc145c65a0abb6d`
MD5	`c5e2bba4a545c11d1313c0bf2db71bf2`
BLAKE2b-256	`82271e4f4e986712d6826b12c2ff774fb85355daa7edbc98fa3b5294678f9271`

See more details on using hashes here.

File details

Details for the file swekit-0.2.34rc1-py3-none-any.whl.

File metadata

Download URL: swekit-0.2.34rc1-py3-none-any.whl
Upload date: Nov 1, 2024
Size: 102.8 kB
Tags: Python 3
Uploaded using Trusted Publishing? No
Uploaded via: twine/5.1.1 CPython/3.12.7

File hashes

Hashes for swekit-0.2.34rc1-py3-none-any.whl
Algorithm	Hash digest
SHA256	`d8b71025dc456de93c1b3e30fb77e1f90db994e40937822dae517f9245dbe266`
MD5	`6a83510fc0dec55c815aa459423b28ae`
BLAKE2b-256	`caec4a92838da7ec6c1775745b8cd507b65ab3e3e9271cff1cb77ba670b28ceb`