VertexAI Memory integration for Autogen agents
Project description
autogen-vertexai-memory
VertexAI Memory integration for Autogen agents. Store and retrieve agent memories using Google Cloud's VertexAI Memory service with semantic search capabilities and intelligent caching.
Features
- Persistent Memory Storage - Store agent memories in Google Cloud VertexAI
- Semantic Search - Find relevant memories using natural language queries
- Intelligent Caching - Reduce API calls with configurable cache TTL (default 5 minutes)
- Automatic Cache Invalidation - Cache updates automatically on write operations
- Automatic Context Updates - Seamlessly inject memories into chat contexts
- Async/Await Support - Full async API compatible with Autogen's runtime
- User-Scoped Isolation - Multi-tenant memory management
- Tool Integration - Ready-to-use tools for agent workflows
Installation
pip install autogen-vertexai-memory
Prerequisites
- Google Cloud Project with VertexAI API enabled
- Authentication configured (Application Default Credentials)
- VertexAI Memory Resource created in your project
# Set up authentication
gcloud auth application-default login
# Enable VertexAI API
gcloud services enable aiplatform.googleapis.com
Quick Start
Basic Memory Usage
from autogen_vertexai_memory import VertexaiMemory, VertexaiMemoryConfig
from autogen_core.memory import MemoryContent, MemoryMimeType
# Configure memory with caching enabled (default)
config = VertexaiMemoryConfig(
api_resource_name="projects/my-project/locations/us-central1/......./",
project_id="my-project",
location="us-central1",
user_id="user123",
cache_enabled=True, # Enable caching (default)
cache_ttl_seconds=300 # Cache for 5 minutes (default)
)
memory = VertexaiMemory(config=config)
# Store a memory (invalidates cache)
await memory.add(
content=MemoryContent(
content="User prefers concise responses and uses Python",
mime_type=MemoryMimeType.TEXT
)
)
# Semantic search for relevant memories
results = await memory.query(query="programming preferences")
for mem in results.results:
print(mem.content)
# Output: User prefers concise responses and uses Python
# Retrieve all memories
all_memories = await memory.query(query="")
Using with Autogen Agents
from autogen_core.model_context import ChatCompletionContext
from autogen_core.models import UserMessage
# Create chat context
context = ChatCompletionContext()
# Add user message
await context.add_message(
UserMessage(content="What programming language should I use?")
)
# Inject relevant memories into context (uses caching)
# First call: Fetches from VertexAI and caches
# Subsequent calls: Returns cached results if still valid
result = await memory.update_context(context)
print(f"Added {len(result.memories.results)} memories to context")
# Now the agent has access to stored preferences
Environment Variables
You can also configure using environment variables:
export VERTEX_PROJECT_ID="my-project"
export VERTEX_LOCATION="us-central1"
export VERTEX_USER_ID="user123"
export VERTEX_API_RESOURCE_NAME="projects/my-project/locations/us-central1/memories/agent-memory"
# Auto-loads from environment
config = VertexaiMemoryConfig()
memory = VertexaiMemory(config=config)
Memory Tools for Agents
Integrate memory capabilities directly into your Autogen agents:
from autogen_agentchat.agents import AssistantAgent
from autogen_ext.models.openai import OpenAIChatCompletionClient
from autogen_vertexai_memory.tools import (
SearchVertexaiMemoryTool,
UpdateVertexaiMemoryTool,
VertexaiMemoryToolConfig
)
# Configure memory tools
memory_config = VertexaiMemoryToolConfig(
project_id="my-project",
location="us-central1",
user_id="user123",
api_resource_name="projects/my-project/locations/us-central1/memories/agent-memory"
)
# Create memory tools
search_tool = SearchVertexaiMemoryTool(config=memory_config)
update_tool = UpdateVertexaiMemoryTool(config=memory_config)
# Create agent with memory tools
agent = AssistantAgent(
name="memory_assistant",
model_client=OpenAIChatCompletionClient(model="gpt-4"),
tools=[search_tool, update_tool],
system_message="""You are a helpful assistant with memory capabilities.
Use search_vertexai_memory_tool to retrieve relevant information about the user.
Use update_vertexai_memory_tool to store important facts you learn during conversations.
"""
)
# Now the agent can search and store memories automatically!
# Example conversation:
# User: "I prefer Python for data analysis"
# Agent uses update_vertexai_memory_tool to store this preference
#
# Later...
# User: "What language should I use for my data project?"
# Agent uses search_vertexai_memory_tool, retrieves the preference, and responds accordingly
API Reference
VertexaiMemoryConfig
Configuration model for VertexAI Memory with caching support.
VertexaiMemoryConfig(
api_resource_name: str, # Full resource name: "projects/{project}/locations/{location}/"
project_id: str, # Google Cloud project ID
location: str, # GCP region (e.g., "us-central1", "europe-west1")
user_id: str, # Unique user identifier for memory isolation
cache_ttl_seconds: int = 300, # Cache time-to-live in seconds (0 to disable)
cache_enabled: bool = True # Whether to enable caching
)
Caching Behavior:
- Cache is used by
update_context()method to reduce repeated API calls - Cache is automatically invalidated on
add()andclear()operations - Set
cache_ttl_seconds=0orcache_enabled=Falseto disable caching query()method does NOT use caching as queries may vary
Environment Variables:
VERTEX_API_RESOURCE_NAMEVERTEX_PROJECT_IDVERTEX_LOCATIONVERTEX_USER_ID
VertexaiMemory
Main memory interface implementing Autogen's Memory protocol with intelligent caching.
VertexaiMemory(
config: Optional[VertexaiMemoryConfig] = None,
client: Optional[Client] = None
)
Methods:
add(content, cancellation_token=None)
Store a new memory and invalidate the cache.
await memory.add(
content=MemoryContent(
content="Important fact to remember",
mime_type=MemoryMimeType.TEXT
)
)
query(query="", cancellation_token=None, **kwargs)
Search memories or retrieve all. Does NOT use caching.
# Semantic search (top 3 results)
results = await memory.query(query="user preferences")
# Get all memories
all_results = await memory.query(query="")
Returns: MemoryQueryResult with list of MemoryContent objects
update_context(model_context)
Inject memories into chat context as system message. Uses caching to reduce API calls.
context = ChatCompletionContext()
result = await memory.update_context(context)
# Context now includes relevant memories
Caching Details:
- First call: Fetches from VertexAI and caches results
- Subsequent calls: Returns cached results if still valid
- After cache expiry: Fetches fresh data and updates cache
Returns: UpdateContextResult with retrieved memories
clear()
Permanently delete all memories and invalidate cache (irreversible).
await memory.clear() # Use with caution!
close()
Release resources and clear cache.
await memory.close()
Memory Tools
VertexaiMemoryToolConfig
Shared configuration for memory tools.
VertexaiMemoryToolConfig(
project_id: str,
location: str,
user_id: str,
api_resource_name: str
)
Environment Variables:
VERTEX_PROJECT_IDVERTEX_LOCATIONVERTEX_USER_IDVERTEX_API_RESOURCE_NAME
SearchVertexaiMemoryTool
Tool for semantic memory search. Automatically used by agents to retrieve relevant memories.
SearchVertexaiMemoryTool(config: Optional[VertexaiMemoryToolConfig] = None, **kwargs)
Tool Name: search_vertexai_memory_tool
Description: Perform a search with given parameters using vertexai memory bank
Parameters:
query(str): Semantic search query to retrieve information about usertop_k(int, default=5): Maximum number of relevant memories to retrieve
Returns: SearchQueryReturn with list of matching memory strings
UpdateVertexaiMemoryTool
Tool for storing new memories. Automatically used by agents to save important information.
UpdateVertexaiMemoryTool(config: Optional[VertexaiMemoryToolConfig] = None, **kwargs)
Tool Name: update_vertexai_memory_tool
Description: Store a new memory fact in the VertexAI memory bank for the user
Parameters:
content(str): The memory content to store as a fact in the memory bank
Returns: UpdateMemoryReturn with success status and message
Advanced Examples
Configuring Cache Behavior
# Disable caching completely
config = VertexaiMemoryConfig(
api_resource_name="projects/my-project/locations/us-central1/...",
project_id="my-project",
location="us-central1",
user_id="user123",
cache_enabled=False
)
# Short cache TTL (30 seconds)
config = VertexaiMemoryConfig(
api_resource_name="projects/my-project/locations/us-central1/...",
project_id="my-project",
location="us-central1",
user_id="user123",
cache_ttl_seconds=30
)
# Long cache TTL (1 hour)
config = VertexaiMemoryConfig(
api_resource_name="projects/my-project/locations/us-central1/...",
project_id="my-project",
location="us-central1",
user_id="user123",
cache_ttl_seconds=3600
)
Custom Client Configuration
from vertexai import Client
# Create custom client with specific settings
client = Client(
project="my-project",
location="us-central1"
)
memory = VertexaiMemory(config=config, client=client)
Multi-User Isolation
# User 1's memories
user1_config = VertexaiMemoryConfig(
api_resource_name="projects/my-project/locations/us-central1/...",
project_id="my-project",
location="us-central1",
user_id="user1"
)
user1_memory = VertexaiMemory(config=user1_config)
# User 2's memories (isolated from User 1)
user2_config = VertexaiMemoryConfig(
api_resource_name="projects/my-project/locations/us-central1/...",
project_id="my-project",
location="us-central1",
user_id="user2"
)
user2_memory = VertexaiMemory(config=user2_config)
Sharing Config Across Tools
from autogen_agentchat.agents import AssistantAgent
from autogen_ext.models.openai import OpenAIChatCompletionClient
from autogen_vertexai_memory.tools import (
SearchVertexaiMemoryTool,
UpdateVertexaiMemoryTool,
VertexaiMemoryToolConfig
)
# Create config once
config = VertexaiMemoryToolConfig(
project_id="my-project",
location="us-central1",
user_id="user123",
api_resource_name="projects/my-project/locations/us-central1/..."
)
# Share across multiple tools
search_tool = SearchVertexaiMemoryTool(config=config)
update_tool = UpdateVertexaiMemoryTool(config=config)
# Use in multiple agents
agent1 = AssistantAgent(
name="agent1",
model_client=OpenAIChatCompletionClient(model="gpt-4"),
tools=[search_tool, update_tool]
)
agent2 = AssistantAgent(
name="agent2",
model_client=OpenAIChatCompletionClient(model="gpt-4"),
tools=[search_tool] # This agent can only search, not update
)
# Both agents use the same VertexAI client and configuration
Development
Setup
# Clone repository
git clone https://github.com/thelaycon/autogen-vertexai-memory.git
cd autogen-vertexai-memory
# Install dependencies with Poetry
poetry install
# Run tests
poetry run pytest
# Run tests with coverage
poetry run pytest --cov=autogen_vertexai_memory --cov-report=html
# Type checking
poetry run mypy src/autogen_vertexai_memory
# Linting
poetry run ruff check src/
Project Structure
autogen-vertexai-memory/
├── src/
│ └── autogen_vertexai_memory/
│ ├── __init__.py
│ ├── memory/
│ │ ├── __init__.py
│ │ └── _vertexai_memory.py # Main memory implementation with caching
│ └── tools/
│ ├── __init__.py
│ └── _vertexai_memory_tools.py # Tool implementations
├── tests/
│ ├── conftest.py
│ └── test_vertexai_memory.py
├── pyproject.toml
└── README.md
Running Tests
The test suite uses mocking to avoid real VertexAI API calls:
# Run all tests
poetry run pytest
# Run with verbose output
poetry run pytest -v
# Run specific test class
poetry run pytest tests/test_vertexai_memory.py::TestVertexaiMemoryConfig
# Run with coverage report
poetry run pytest --cov=autogen_vertexai_memory --cov-report=term-missing
Troubleshooting
Authentication Issues
# Verify authentication
gcloud auth application-default print-access-token
# Set explicit credentials
export GOOGLE_APPLICATION_CREDENTIALS="/path/to/service-account-key.json"
Empty Query Results
# Check if memories exist
all_memories = await memory.query(query="")
print(f"Total memories: {len(all_memories.results)}")
# Verify user_id matches
print(f"Using user_id: {memory.user_id}")
Cache Not Working
# Check cache configuration
print(f"Cache enabled: {memory._cache_enabled}")
print(f"Cache TTL: {memory._cache_ttl_seconds}")
# Manually invalidate cache if needed
memory._invalidate_cache()
Contributing
Contributions are welcome! Please follow these steps:
- Fork the repository
- Create a feature branch (
git checkout -b feature/amazing-feature) - Make your changes with tests
- Run tests (
poetry run pytest) - Commit your changes (
git commit -m 'Add amazing feature') - Push to the branch (
git push origin feature/amazing-feature) - Open a Pull Request
Development Guidelines
- Write tests for new features
- Follow existing code style
- Update documentation for API changes
- Ensure all tests pass before submitting PR
License
MIT License - see LICENSE file for details.
Support
- GitHub Issues - Bug reports and feature requests
- GitHub Discussions - Questions and community support
- VertexAI Documentation - Official VertexAI docs
- Autogen Documentation - Autogen framework docs
Project details
Release history Release notifications | RSS feed
Download files
Download the file for your platform. If you're not sure which to choose, learn more about installing packages.
Source Distribution
Built Distribution
Filter files by name, interpreter, ABI, and platform.
If you're not sure about the file name format, learn more about wheel file names.
Copy a direct link to the current filters
File details
Details for the file autogen_vertexai_memory-0.1.17.tar.gz.
File metadata
- Download URL: autogen_vertexai_memory-0.1.17.tar.gz
- Upload date:
- Size: 14.0 kB
- Tags: Source
- Uploaded using Trusted Publishing? No
- Uploaded via: poetry/2.2.1 CPython/3.13.5 Linux/6.6.87.2-microsoft-standard-WSL2
File hashes
| Algorithm | Hash digest | |
|---|---|---|
| SHA256 |
bf7583ad1d309dfa53c4245ad61997bdd2ff4f8cd127bf2f606c29579eed7683
|
|
| MD5 |
0bbb7c1bff4ecfbec3656df1754182f7
|
|
| BLAKE2b-256 |
7d7448e94eba164be35b3c1478e6b59f1807d20878d598e6a772c5935b62a467
|
File details
Details for the file autogen_vertexai_memory-0.1.17-py3-none-any.whl.
File metadata
- Download URL: autogen_vertexai_memory-0.1.17-py3-none-any.whl
- Upload date:
- Size: 13.7 kB
- Tags: Python 3
- Uploaded using Trusted Publishing? No
- Uploaded via: poetry/2.2.1 CPython/3.13.5 Linux/6.6.87.2-microsoft-standard-WSL2
File hashes
| Algorithm | Hash digest | |
|---|---|---|
| SHA256 |
44f49fcd7228fa5962662fd80329dc54795813e203ad27758aa8680dbb2ebb5a
|
|
| MD5 |
15e60685f095d51a179eee80b8146ae1
|
|
| BLAKE2b-256 |
f339f2da382083dd45700b1528158464f4313efa698cc30a006440977d0d1112
|