ZenGuard AI Benchmarks
This repository contains benchmarks for ZenGuard AI and information on how to run them.
There are two types of benchmarks that we run against ZenGuard AI:
- Hugging Face datasets based benchmarks
- ZenGuard AI Generated Benchmark - Zen Bench
Here you can find both benchmark results and how to run them yourself.
Public Datasets benchmarks
We are constantly monitoring Hugging Face for new datasets that relate to GenAI security. Then we run them against ZenGuard AI to find any potential security issues with our product.
ZenGuard AI Accuracy against Hugging Face datasets
| # | Dataset | Accuracy | Date Added |
|---|---|---|---|
| 1 | xTRam1/safe-guard-prompt-injection | 96% | 2024-07-01 |
| 2 | yanismiraoui/prompt_injections | 96.5% | 2024-08-01 |
| 3 | deepset/prompt-injections | 87% | 2024-05-15 |
| 4 | JasperLS/prompt-injections | 87% | 2024-05-15 |
| 4 | aporia-ai/prompt_injection | 87.68% | 2024-05-15 |
Check for yourself. Or run your own dataset.
We have developed the ZenGuard Benchmarks PyPi package to help test and benchmark ZenGuard AI better.
Here are the instructions on how to use the package.
Benchmarking Output
Here is an example of what the benchmarking output looks like:
Where:
Total Samples: The total number of prompts processed.Correct: The number of prompts that were classified correctly.False Positives: The number of prompts incorrectly identified as attacks.False Negatives: The number of actual prompt attacks that went undetected.Accuracy: The ratio of correctly classified prompts to the total number of samples.
Zen Bench
More information
A much more detailed documentation is available at docs.zenguard.ai.
Test the capabilities of ZenGuard AI in our ZenGuard Playground. It's available to start for free to understand how our guardrails can enhance your GenAI applications.
Check out our Client library to get started with integrating ZenGuard AI into your project.
Support
Book a Demo or just shoot us an email to hello@zenguard.ai
Topics we care about - LLM Security, LLM Guardrails, Prompt Injections, GenAI Security.
Developed with ❤️ by ZenGuard AI
Release files for zenguard-benchmarks 0.1.5
For a detailed explanation of source distributions (sdists) and built distributions (wheels), please see the package formats documentation.
Source distribution (sdist)
| File | Size | Uploaded | |
|---|---|---|---|
| zenguard_benchmarks-0.1.5.tar.gz | 4.7 kB | Details |
Built distribution (wheel)
| File | Interpreter | ABI | Platform | Reset |
|---|---|---|---|---|
| zenguard_benchmarks-0.1.5-py3-none-any.whl | Python 3 | none | any | Details |
Total release size: 10.2 kB
Release files / zenguard_benchmarks-0.1.5.tar.gz
| Download URL | zenguard_benchmarks-0.1.5.tar.gz |
|---|---|
| Size | 4.7 kB |
| Tags | Source |
|
SHA-256 checksum How to use checksums |
5aadb62748ac935137ebcfc651aa009856bd6027cccf4c392f8c255deea7fb25
|
|
BLAKE2b-256 checksum How to use checksums |
ddd5a225a80df0ef2574d503d8b1ffa8fa35340b0d766565b2a8e0c6ffeab9ba
|
| Upload date | |
|
Uploaded using Trusted Publishing? What is trusted publishing? |
No |
| Uploaded via |
poetry/1.7.1 CPython/3.11.9 Darwin/23.6.0
|
Release files / zenguard_benchmarks-0.1.5-py3-none-any.whl
| Download URL | zenguard_benchmarks-0.1.5-py3-none-any.whl |
|---|---|
| Size | 5.5 kB |
| Tags | Python 3 |
|
SHA-256 checksum How to use checksums |
75dc5aa753c24b02b1519b6fada17a0eb8cf6c99cab2ce57bd0cf954bd174deb
|
|
BLAKE2b-256 checksum How to use checksums |
1668ce3afecaa6a3eb4d15c3ef4f8e040f5b7032d976167439aa8914eee9d3eb
|
| Upload date | |
|
Uploaded using Trusted Publishing? What is trusted publishing? |
No |
| Uploaded via |
poetry/1.7.1 CPython/3.11.9 Darwin/23.6.0
|