Skip to main content

CLI tool and Model Server for VirtueAI VirtueRed

Project description

VirtueRed

VirtueRed is a comprehensive package for the VirtueAI Redteaming system, providing both a CLI tool and a Model Server.

Installation

pip install virtuered

Components

1. Model Server

The Model Server allows you to serve your custom models for use with the VirtueRed system.

Usage

from virtuered.client import ModelServer

# Start the model server
server = ModelServer(
    port=4299                   # Optional, defaults to 4299
)
server.start()

Local Model Setup

  1. Create a models directory:
models/
├── model1.py
└── model2.py
  1. Create a new Python file in the ./models folder with a descriptive name for your model, e.g., model1.py. Create a function called chat in model1.py, the chat function will takes a list of chat messages and returns the response from the language model. For example, if we want to create a chat function with LLAMA-3-8b from huggingface:
import os
from transformers import pipeline, LlamaTokenizer, LlamaForCausalLM

# Step 1: Load the LLaMA model and tokenizer
model_name = "meta-llama/Meta-Llama-3-8B"  # Replace with the actual model name on Hugging Face Hub
tokenizer = LlamaTokenizer.from_pretrained(model_name)
model = LlamaForCausalLM.from_pretrained(model_name)

# Initialize the Hugging Face pipeline with the model and tokenizer
chatbot = pipeline("text-generation", model=model, tokenizer=tokenizer)

# Don't change the name of the function or the function signature
def chat(chats):
    """
    This function takes a list of chat messages and returns the response from the language model.
    
    Parameters:
    chats (list): A list of dictionaries. Each dictionary contains a "prompt" and optionally an "answer".
                  The last item in the list should have a "prompt" without an "answer".
    
    Returns:
    str: The response from the language model.
    
    Example of `chats` list with few-shot prompts:
    [
        {"prompt": "Hello, how are you?", "answer": "I'm an AI, so I don't have feelings, but thanks for asking!"},
        {"prompt": "What is the capital of France?", "answer": "The capital of France is Paris."},
        {"prompt": "Can you tell me a joke?"}
    ]
    
    Another example of `chats` list with one-shot prompt:
    [
        {"prompt": "What is the weather like today?"}
    ]
    """
    
    # Step 2: Prepare the chat history as a single string
    chat_history = []
    for c in chats:
        # Add the user prompt and assistant's answer to the chat history
        chat_history.append({"role": "user", "content": c["prompt"]})
        if "answer" in c.keys():
            chat_history.append({"role": "assistant", "content": c["answer"]})
        else:
            # If there is no answer, it means this is the prompt we need a response for
            break
    
    # Step 3: Generate the model's response
    response = chatbot(chat_history, max_length=1000, num_return_sequences=1)
    
    # Step 4: Extract and return the generated text
    generated_text = response[0]['generated_text']
    assistant_response = generated_text.split("Assistant:")[-1].strip()
    
    return assistant_response 

2. CLI Tool

The CLI tool provides a command-line interface for managing VirtueRed operations.

CLI Commands

# View all available commands
virtuered --help

# Configure a custom server address instead of the default http://localhost:4401(persistent)
virtuered config http://localhost:4403
virtuered show-config

# List all runs
virtuered list

# List all custom models
virtuered models

# Monitor ongoing scans
virtuered monitor

# Initialize a new scan
virtuered scan                # Uses default scan_config.json
virtuered scan myconfig.json  # Uses custom config file

# Get summary of a run
virtuered summary my_scan
virtuered summary 1

# Pause/Resume a scan
virtuered pause my_scan
virtuered resume 1

# Generate report
virtuered report my_scan

# Delete a run
virtuered delete 1

# For temporary custom server URL:
virtuered --server http://localhost:4403 list

Available Commands

  • --help: View all available commands
  • config: Set persistent default server URL
  • show-config: View current configuration
  • list: Show all runs
  • models: Show all custom models
  • scan: Initialize a new scan using configuration file
  • monitor: Monitor ongoing scans
  • summary: Get detailed summary of a run
  • report: Generate PDF report
  • pause: Pause a running scan
  • resume: Resume a paused scan
  • delete: Delete a run

Scan Configuration

Create a JSON configuration file (scan_config.json or custom name) before initiating a scan:

{
    "name": "my_scan",
    "model": "together_template",
    "datasets": [
        {
            "name": "EU Artificial Intelligence Act",
            "subcategories": [
                "Criminal justice/Predictive policing",
                "Persons (including murder)"
            ]
        },
        {
            "name": "White House AI Executive Order"
        }
    ],
    "extra_args": {
        "modelname": "model name",
        "apikey": "your-api-key"
    }
}

Configuration Fields:

  • Required:
    • name: Unique scan identifier (alphanumeric and underscores only)
    • model: Model specification (local or cloud-hosted template)
    • datasets: Array of at least one dataset object with name field
  • Optional:
    • subcategories: Array of specific subcategories to evaluate within each dataset
    • extra_args: Additional parameters for cloud-based models (credentials, settings)

Architecture

The package consists of two main components:

  1. Model Server: Serves your custom models, making them accessible to the VirtueRed system
  2. CLI Tool: Provides command-line interface for managing VirtueRed operations

The typical setup involves:

  1. Running the Model Server to serve your custom models
  2. Running the VirtueRed Docker container, configured to connect to your Model Server
  3. Using the CLI tool to manage and monitor operations

Notes

  • The Model Server must be running and accessible to the Docker container
  • For local setup, use http://127.0.0.1:4299 as client address
  • For remote setup, use http://<server-ip>:4299 as client address
  • Ensure your models directory is properly structured with required model files

Support

If you need further assistance, please contact our support team at contact@virtueai.com

Project details


Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

virtuered-2.0.0b1.tar.gz (16.7 kB view details)

Uploaded Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

virtuered-2.0.0b1-py3-none-any.whl (15.3 kB view details)

Uploaded Python 3

File details

Details for the file virtuered-2.0.0b1.tar.gz.

File metadata

  • Download URL: virtuered-2.0.0b1.tar.gz
  • Upload date:
  • Size: 16.7 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/5.1.1 CPython/3.12.2

File hashes

Hashes for virtuered-2.0.0b1.tar.gz
Algorithm Hash digest
SHA256 717f6b95ac61c373cfa379124146d7677984b111c33042587530d52ac277d044
MD5 c1d4bd0a5342665108f0773bc2cb4648
BLAKE2b-256 af71a03cf89fd4b3dfabffbe7fcd964c22c52696d4ffdec4654ad187336c08f0

See more details on using hashes here.

File details

Details for the file virtuered-2.0.0b1-py3-none-any.whl.

File metadata

  • Download URL: virtuered-2.0.0b1-py3-none-any.whl
  • Upload date:
  • Size: 15.3 kB
  • Tags: Python 3
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/5.1.1 CPython/3.12.2

File hashes

Hashes for virtuered-2.0.0b1-py3-none-any.whl
Algorithm Hash digest
SHA256 0e96dd8996e6784460c41f9c6e50d2f9f8892562563288d198758a30ef5d2e16
MD5 08570cdb9ba344b9a44572d4d4bae26b
BLAKE2b-256 a143397ce73e708319cec6719fb2b0569b55228afc6b7076b9c4290f6e94f1ed

See more details on using hashes here.

Supported by

AWS Cloud computing and Security Sponsor Datadog Monitoring Depot Continuous Integration Fastly CDN Google Download Analytics Pingdom Monitoring Sentry Error logging StatusPage Status page