Benchmarking the performance of agents far and wide, regardless of how they are set up and how they work
Project description
Auto-GPT Benchmarks
Built for the purpose of benchmarking the performance of agents regardless of how they work.
Objectively know how well your agent is performing in categories like code, retrieval, memory, and safety.
Save time and money while doing it through smart dependencies. The best part? It's all automated.
Scores:
Ranking overall:
Detailed results:
Click here to see the results and the raw data!!
More agents coming soon !
Project details
Release history Release notifications | RSS feed
Download files
Download the file for your platform. If you're not sure which to choose, learn more about installing packages.
Source Distribution
agbenchmark-0.0.10.tar.gz
(101.5 kB
view hashes)
Built Distribution
agbenchmark-0.0.10-py3-none-any.whl
(212.2 kB
view hashes)
Close
Hashes for agbenchmark-0.0.10-py3-none-any.whl
Algorithm | Hash digest | |
---|---|---|
SHA256 | e165be0d05eb84700057a0660fe05132b4d7857d1395466d0d59bb0592143200 |
|
MD5 | b433851b7484772b18530fa4d6f02903 |
|
BLAKE2b-256 | 89cf984bbcea12511aff6437ffe0fbead4af1e7762ab27a5204cda339ef047d7 |