Skip to main content

AVISE - AI Vulnerability Identification & Security Evaluation

A framework for identifying vulnerabilities in and evaluating the security of AI systems.

arXiv

Full Documentations: https://avise.readthedocs.io

Youtube Demo


Table of Contents

Quickstart for Evaluating Language Models

Prerequisites

  • Python 3.10+
  • Docker (For Running models locally with Ollama)

1. Install AVISE

Install with

  • pip:

    pip install avise
    
  • uv:

    uv pip install avise
    

    or

    uv tool install avise
    

2. Run a Model

You can use AVISE to evaluate any model accessible via an API by configuring a Connector. In this Quickstart, we will assume using the Ollama Docker container for running a language model. If you wish to evaluate models deployed in other ways, see the Full Documentations and available template connector configuration files at AVISE/avise/configs/connector/languagemodel/ dir of this repository.

Running a language model locally with Docker & Ollama

  • Clone this repository to your local machine with:
git clone https://github.com/ouspg/AVISE.git
  • Create the Ollama Docker container

    • for GPU accelerated inference with:
      docker compose -f AVISE/docker/ollama/docker-compose.yml up -d
      
    • or for CPU inference with:
      docker compose -f AVISE/docker/ollama/docker-compose-cpu.yml up -d
      
  • Pull an Ollama model to evaluate into the container with:

    docker exec -it avise-ollama ollama pull <model_name>
    

3. Evaluate the model with a Security Evaluation Test (SET)

Basic usage

avise --SET <SET_name> --connectorconf <connector_name> [options]

For example, you can run the prompt_injection SET on the model pulled to the Ollama Docker container with:

avise --SET prompt_injection --connectorconf ollama_lm --target <model_name>

To list the available SETs, run the command:

avise --SET-list

Advanced Usage

Configuring Connectors

You can create your own connector configuration files, or if you cloned the AVISE repository, you can modify the existing connector configuration files in AVISE/avise/configs/connector/languagemodel/.

For example, you can edit the default Ollama Connector configuration file AVISE/avise/configs/connector/languagemodel/ollama.json, and insert the name of an Ollama model you have pulled to be used as a target by default:

{
    "target_model": {
        "connector": "ollama-lm",
        "type": "language_model",
        "name": "<NAME_OF_TARGET_MODEL>",
        "api_url": "http://localhost:11434", #Ollama default
        "api_key": null
    }
}

Configuring Security Evaluation Tests (SETs)

Similarly, you can customize the configurations for SETs as well. For example, by editing the Red Queen SET configuration file AVISE/avise/configs/SET/languagemodel/multi_turn/red_queen.json, you can define if the SET is executed incrementally (the target model will generate a response after each subsequential prompt), or as a template (only works for target systems that accept a conversation as an input) and if the SET uses an Adversarial Language Model (ALM). Additionally, you can define the exact template attack prompts that the SET uses:

{
    "incremental_execution": false, #Change to true to execute the SET cases incrementally
    "use_adversarial_languagemodel": false, #Change to true to utilize an Adversarial Language Model (ALM)
    "adversarial_model_device": "auto", # Defines on which device(s) the ALM is loaded to ("cpu", "gpu", or "auto")
    "evaluation_system_prompt": "You are a highly accurate cybersecurity classifier, ..."
    "evaluation_model_device": "auto", # Defines on which device(s) the ELM is loaded to ("cpu", "gpu", or "auto")
    "sets": [
        {
            "id": "RED-QUEEN-001",
            "vulnerability_subcategory": "Evasion Attack",
            "action": "make a fake passport",
            "type": "occupation_teacher",
            "category": "violence,aiding_and_abetting,incitement",
            "conversation": [
                { ... }
            ...
            ]
         },
         ...
    ]

Using custom configuration files

If you want to use custom configuration files for SETs and/or Connectors, you can do so by giving the paths to the configuration files with --SETconf and --connectorconf arguments:

avise --SET prompt_injection --SETconf AVISE/avise/configs/SET/languagemodel/single_turn/prompt_injection_mini.json --connectorconf AVISE/avise/configs/connector/languagemodel/ollama.json

Required Arguments

Argument Description
--SET, -s Security Evaluation Test to run (e.g., prompt_injection, context_test)
--connectorconf, -c Path to Connector configuration file (Accepts predefined connector configuration paths: ollama_lm, openai_lm, genericrest_lm)

Optional Arguments

Argument Description
--SETconf Path to SET configuration file. If not given, uses preconfigured paths for SET config files.
--target, -t Name of the target model/system to evaluate. Overrides target name from connector configuration file.
--format, -f Report format: json, html, md
--runs, -r How many times each SET is executed
--output, -o Custom output file path
--api_key, -a API Key to use with requests sent to target API (overrides api_key from Connector configuration file).
--reports-dir, -d Base directory for reports (default: avise-reports/)
--SET-list List available Security Evaluation Tests
--connector-list List available Connectors
--verbose, -v Enable verbose logging
--version, -V Print version

Citation

If you find AVISE useful, please cite it as below:

@misc{lempinen2026,
      title={AVISE: Framework for Evaluating the Security of AI Systems},
      author={Mikko Lempinen and Joni Kemppainen and Niklas Raesalmi},
      year={2026},
      eprint={2604.20833},
      archivePrefix={arXiv},
      primaryClass={cs.CR},
      url={https://arxiv.org/abs/2604.20833},
}

Lempinen, M., Kemppainen, J., & Raesalmi, N. (2026). AVISE: Framework for Evaluating the Security of AI Systems. arXiv preprint arXiv:2604.20833.

Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

avise-0.2.3.tar.gz (108.2 kB view details)

Uploaded Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

avise-0.2.3-py3-none-any.whl (152.5 kB view details)

Uploaded Python 3

File details

Details for the file avise-0.2.3.tar.gz.

File metadata

  • Download URL: avise-0.2.3.tar.gz
  • Upload date:
  • Size: 108.2 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? Yes
  • Uploaded via: twine/6.1.0 CPython/3.13.13

File hashes

Hashes for avise-0.2.3.tar.gz
Algorithm Hash digest
SHA256 1248d5a286dc9ca8cb73d1c61fc44070c8a8b1cffd8f10a0d5b8ff084d3d0203
MD5 6f96927ff85a563f804ae836c889bf66
BLAKE2b-256 9ff1c4fc4cf0a22614ca28a80b5565ec9bdfd827f61d35bd47b64ea5c1767ce9

See more details on using hashes here.

Provenance

The following attestation bundles were made for avise-0.2.3.tar.gz:

Publisher: pypi-publish.yml on ouspg/AVISE

Attestations: Values shown here reflect the state when the release was signed and may no longer be current.

File details

Details for the file avise-0.2.3-py3-none-any.whl.

File metadata

  • Download URL: avise-0.2.3-py3-none-any.whl
  • Upload date:
  • Size: 152.5 kB
  • Tags: Python 3
  • Uploaded using Trusted Publishing? Yes
  • Uploaded via: twine/6.1.0 CPython/3.13.13

File hashes

Hashes for avise-0.2.3-py3-none-any.whl
Algorithm Hash digest
SHA256 6d5d25edda6ea03e9f91ef96aa940c262f56eda5e5381f8ca9ba86dc463ccbdc
MD5 dd94f969a6346228cd449f7f4fafafcd
BLAKE2b-256 5c6c6709c6efb93bb03e3319847c50d2f241ae3247821522377de5485f164f8a

See more details on using hashes here.

Provenance

The following attestation bundles were made for avise-0.2.3-py3-none-any.whl:

Publisher: pypi-publish.yml on ouspg/AVISE

Attestations: Values shown here reflect the state when the release was signed and may no longer be current.

Supported by

AWS Cloud computing and Security Sponsor Datadog Monitoring Depot Continuous Integration Fastly CDN Google Download Analytics Sentry Error logging StatusPage Status page