Auto-GPT Benchmarks
Built for the purpose of benchmarking the performance of agents regardless of how they work.
Objectively know how well your agent is performing in categories like code, retrieval, memory, and safety.
Save time and money while doing it through smart dependencies. The best part? It's all automated.
Scores:
Ranking overall:
Detailed results:
Click here to see the results and the raw data!!
More agents coming soon !
Release files for agbenchmark 0.0.10
For a detailed explanation of source distributions (sdists) and built distributions (wheels), please see the package formats documentation.
Source distribution (sdist)
| File | Size | Uploaded | |
|---|---|---|---|
| agbenchmark-0.0.10.tar.gz | 101.5 kB | Details |
Built distribution (wheel)
| File | Interpreter | ABI | Platform | Reset |
|---|---|---|---|---|
| agbenchmark-0.0.10-py3-none-any.whl | Python 3 | none | any | Details |
Total release size: 313.7 kB
Release files / agbenchmark-0.0.10.tar.gz
| Download URL | agbenchmark-0.0.10.tar.gz |
|---|---|
| Size | 101.5 kB |
| Tags | Source |
|
SHA-256 checksum How to use checksums |
14793e98507f530eab7473a3dc7dfc218a49260de7b34f71939cb081877fb981
|
|
BLAKE2b-256 checksum How to use checksums |
a56989d87beadf1ab4d5835b0d601bbdf8be3b4cbb3081209541f33b2819e1e6
|
| Upload date | |
|
Uploaded using Trusted Publishing? What is trusted publishing? |
No |
| Uploaded via |
poetry/1.6.1 CPython/3.8.18 Linux/6.2.0-1011-azure
|
Release files / agbenchmark-0.0.10-py3-none-any.whl
| Download URL | agbenchmark-0.0.10-py3-none-any.whl |
|---|---|
| Size | 212.2 kB |
| Tags | Python 3 |
|
SHA-256 checksum How to use checksums |
e165be0d05eb84700057a0660fe05132b4d7857d1395466d0d59bb0592143200
|
|
BLAKE2b-256 checksum How to use checksums |
89cf984bbcea12511aff6437ffe0fbead4af1e7762ab27a5204cda339ef047d7
|
| Upload date | |
|
Uploaded using Trusted Publishing? What is trusted publishing? |
No |
| Uploaded via |
poetry/1.6.1 CPython/3.8.18 Linux/6.2.0-1011-azure
|