Compressor Reflex MCP
Compressor Reflex MCP is a Model Context Protocol (MCP) server and transparent proxy that provides high-fidelity tool-output compression for Cursor, Antigravity IDE, Claude Desktop, and other MCP-compliant developer environments.
Powered by aialchemist-dev/compressor-reflex (fine-tuned ModernBERT-151M), the model reduces tool-output token consumption by up to 90% while guaranteeing 100% retention of compiler errors, test failures, target locations, and decisive anchors.
Key Capabilities
- 89.7% Measured Tool Compression: Condenses extensive test suites, git diffs, directory listings, and file dumps into their essential operational lines.
- 100.0% Critical Anchor Retention: Calibrated at decision threshold tau* = 0.50 across held-out evaluations. Error tracebacks, fail signatures, and query targets remain intact.
- Fail-Open Bypass Policy: Outputs containing <= 5 physical lines or <= 64 tokens automatically bypass compression and pass through verbatim, avoiding latency overhead on short outputs.
- Dual Deployment Modes:
- Direct MCP Tools: Exposes standard callable tools (
compress_tool_output,compress_file) for explicit agent invocation. - Transparent Proxy: Wraps any standard MCP server (such as filesystem, git, or terminal) to automatically compress output streams before delivery to the context window.
- Direct MCP Tools: Exposes standard callable tools (
- Automated Weight Management: Model weights (INT8 ONNX) and tokenizers are automatically retrieved from Hugging Face Hub on initial startup and cached locally.
Installation
From Source or Git
Install directly via pip:
pip install git+https://github.com/ericmaddox/compressor-reflex-mcp.git
Or clone the repository and install in editable mode:
git clone https://github.com/ericmaddox/compressor-reflex-mcp.git
cd compressor-reflex-mcp
pip install -e .
Optional Model Pre-Caching
To download the model weights ahead of time:
compressor-reflex-mcp download
Model files are cached in the standard user cache directory (~/.cache/compressor-reflex/ or %LOCALAPPDATA%/compressor-reflex/). The cache path can be overridden using the COMPRESSOR_MODEL_DIR environment variable.
IDE Configuration
Cursor
Add the server definition to your workspace .cursor/mcp.json or global Cursor settings:
{
"mcpServers": {
"compressor-reflex": {
"command": "python",
"args": ["-m", "compressor_reflex_mcp.server"]
}
}
}
If utilizing uvx:
{
"mcpServers": {
"compressor-reflex": {
"command": "uvx",
"args": ["--from", "git+https://github.com/ericmaddox/compressor-reflex-mcp", "compressor-reflex-mcp"]
}
}
}
Antigravity IDE
Add to ~/.gemini/config/mcp_config.json or your project .gemini/mcp_config.json:
{
"mcpServers": {
"compressor-reflex": {
"command": "python",
"args": ["-m", "compressor_reflex_mcp.server"]
}
}
}
Transparent Proxy Mode (Antigravity and Cursor)
Wrap existing tools to automatically compress outputs from heavy providers (for example, filesystem or git inspection):
{
"mcpServers": {
"filesystem-compressed": {
"command": "compressor-reflex-mcp",
"args": [
"proxy",
"--",
"npx",
"-y",
"@modelcontextprotocol/server-filesystem",
"."
]
}
}
}
Claude Desktop
Update claude_desktop_config.json:
{
"mcpServers": {
"compressor-reflex": {
"command": "python",
"args": ["-m", "compressor_reflex_mcp.server"]
}
}
}
Tool Reference
| Tool | Parameters | Description |
|---|---|---|
compress_tool_output |
text (required), intent (optional), threshold (optional, default: 0.50) |
Extracts relevant lines from raw terminal stdout, test logs, or diffs with optional intent routing. |
compress_file |
file_path (required), intent (optional), threshold (optional, default: 0.50) |
Reads a file from disk and extracts decisive lines based on the provided intent. |
get_model_status |
None | Returns metadata on local model cache status, Hugging Face Hub link, and threshold settings. |
Command-Line Interface
# Launch the stdio MCP server
compressor-reflex-mcp serve
# Run as transparent proxy wrapping another command
compressor-reflex-mcp proxy -- npx -y @modelcontextprotocol/server-filesystem /path/to/project
# Compress a file or standard input directly
compressor-reflex-mcp compress tests/results.log --intent "find failures"
# Verify model cache and runtime configuration
compressor-reflex-mcp info
# Pre-fetch weights from Hugging Face
compressor-reflex-mcp download
Empirical Performance
Metrics collected across real multi-turn developer sessions in IDE environments:
| Metric | Measured Value | Methodology / Context |
|---|---|---|
| Tool Compression Ratio | 89.69% | Measured over 181 tool outputs (129,405 raw to 13,339 kept tokens) |
| Critical Anchor Retention | 100.0% | 611/611 anchor lines preserved at calibrated threshold tau* = 0.50 |
| Fail-Open Bypass Rate | 9.39% | Automatically bypassed on outputs with <= 5 physical lines or <= 64 tokens |
| Maximum Single-Session Savings | 60.54% | Measured in deep multi-module code exploration session |
| Model Size | 143 MB | INT8 quantized ONNX (ModernBERT-151M) |
Agent Instructions
For system prompt guidelines and autonomous agent behavior rules, refer to AGENTS.md.
License
This project is licensed under the MIT License. See LICENSE for details.
Model architecture and pre-trained weights are hosted at Hugging Face.
Metadata
Release files for compressor-reflex-mcp 0.2.1
For a detailed explanation of source distributions (sdists) and built distributions (wheels), please see the package formats documentation.
Source distribution (sdist)
| File | Size | Uploaded | |
|---|---|---|---|
| compressor_reflex_mcp-0.2.1.tar.gz | 21.5 kB | Details |
Built distribution (wheel)
| File | Interpreter | ABI | Platform | Reset |
|---|---|---|---|---|
| compressor_reflex_mcp-0.2.1-py3-none-any.whl | Python 3 | none | any | Details |
Total release size: 41.5 kB
Release files / compressor_reflex_mcp-0.2.1.tar.gz
| Download URL | compressor_reflex_mcp-0.2.1.tar.gz |
|---|---|
| Size | 21.5 kB |
| Tags | Source |
|
SHA-256 checksum How to use checksums |
68da322170f2fea1ccde183a6671d80a16be8615010cb33d301162e517de1401
|
|
BLAKE2b-256 checksum How to use checksums |
86fc6a9ecffe0fe81cdc00b412e71b4d9338225978ee393a1ca605540908be3d
|
| Upload date | |
|
Uploaded using Trusted Publishing? What is trusted publishing? |
Yes |
| Uploaded via |
twine/7.0.0 CPython/3.13.14
|
Provenance
Provenance describes where a file came from. On PyPI, provenance is shared via attestations, which provide a verifiable record of the build or publishing details. View details, limitations and caveats.
PyPI Publish Attestation
PyPI verified that this artifact, at this checksum, originated from the publisher listed below.
Signed by GitHub Actions, verified by PyPI on Oct 2, 2026.
Transparency logRelease files / compressor_reflex_mcp-0.2.1-py3-none-any.whl
| Download URL | compressor_reflex_mcp-0.2.1-py3-none-any.whl |
|---|---|
| Size | 20.0 kB |
| Tags | Python 3 |
|
SHA-256 checksum How to use checksums |
9a39b16f6718cb2be4a6bd7916e90ad32994300550881580f96e8be96787f194
|
|
BLAKE2b-256 checksum How to use checksums |
ffa2f147e8e7796a251bade5359b5cfd75aba7eb4638f6897a80f4cf478b2f54
|
| Upload date | |
|
Uploaded using Trusted Publishing? What is trusted publishing? |
Yes |
| Uploaded via |
twine/7.0.0 CPython/3.13.14
|
Provenance
Provenance describes where a file came from. On PyPI, provenance is shared via attestations, which provide a verifiable record of the build or publishing details. View details, limitations and caveats.
PyPI Publish Attestation
PyPI verified that this artifact, at this checksum, originated from the publisher listed below.
Signed by GitHub Actions, verified by PyPI on Oct 2, 2026.
Transparency log