Skip to main content

An MCP server that reviews your code the way a kernel maintainer would: bluntly, technically, without diplomacy.

Project description

Code Review Agent MCP

AI code review agent that doesn't sugarcoat. If your code is bad, it tells you. Bluntly.

CI Python 3.10+ License: MIT Code style: ruff

An AI-powered MCP (Model Context Protocol) server that reviews your code the way experienced maintainers review patches on serious projects: technically, directly, without diplomacy.

What it does

  • AI code review — reviews code snippets, files, git diffs, and commits
  • Returns findings with severity labels: CRITICAL, MAJOR, MINOR, NIT, CLEAN
  • Cites specific line numbers for every finding
  • Explains why something is bad, then how to fix it
  • Ends with a clear verdict: "Ship it." or "Do not merge until..."
  • No false praise. No hedging. No apologies. No filler.

Why?

Most AI code reviewers are trained to be polite. They preface criticism with "great work but...", they hedge with "you might want to consider...", they soften findings with "this could potentially be an issue".

That's not how serious code review works. On real projects (Linux kernel, PostgreSQL, Redis, SQLite), maintainers review patches directly. They find problems and call them out clearly. They don't exist to make the author feel good — they exist to make the code better.

This MCP encodes that tradition. The personality IS the product.

Anti-RLHF engineering

LLMs are RLHF-trained to be polite. This AI agent fights that with three layers:

  1. System prompt — 12 core principles (code-not-coder, verdict-first, severity labels, line citations, no false praise, no hedging, no apologies, no filler, second person, no demographic references, no refusal, no hallucinated bugs)
  2. Post-processor — strips 40+ banned phrases (hedging, softeners, apologies, filler) from LLM output
  3. Validator — checks output has severity labels, line citations, and verdict; flags vague problem statements as hallucination signals

Tools

Tool Description
review_code Review a code snippet
review_file Review a file from disk (sandboxed)
review_diff Review a git diff
review_commit Review a git commit (git show <ref>)
list_severities List severity labels with definitions

Harshness levels

Level Behavior
gentle Softens language slightly. Still reports all findings.
standard Default blunt review. Direct, technical, no diplomacy.
brutal No softening. "This is wrong." not "This should be changed."
kernel-maintainer Maximum bluntness. Short sentences. Imperative voice.

Installation

pip install code-review-agent-mcp

Or with uv:

uv pip install code-review-agent-mcp

Configuration

Claude Desktop

Add to claude_desktop_config.json:

{
  "mcpServers": {
    "code-review-agent": {
      "command": "python",
      "args": ["-m", "blunt_codereview.server"]
    }
  }
}

Or if installed via pip:

{
  "mcpServers": {
    "code-review-agent": {
      "command": "blunt-codereview-mcp"
    }
  }
}

Cursor

Add to .cursor/mcp.json:

{
  "mcpServers": {
    "code-review-agent": {
      "command": "python",
      "args": ["-m", "blunt_codereview.server"]
    }
  }
}

Usage examples

Review a code snippet

User: Review this code for me

def get_user(username):
    import sqlite3
    conn = sqlite3.connect("users.db")
    cursor = conn.cursor()
    query = f"SELECT * FROM users WHERE username = '{username}'"
    cursor.execute(query)
    return cursor.fetchone()

MCP response:

## Code Review: snippet

### Findings

**CRITICAL** `snippet:7` — SQL injection
The query uses an f-string with user input, allowing SQL injection. Use parameterized queries: `cursor.execute("SELECT * FROM users WHERE username = ?", (username,))`.

### Verdict

Do not merge until CRITICAL is fixed.

Review a file

User: Review src/auth.py

MCP calls review_file with file_path="src/auth.py"
Returns blunt review with line citations.

Review a commit

User: Review the last commit

MCP calls review_commit with commit_ref="HEAD"
Returns blunt review of the diff.

Severity labels

Label When to use
CRITICAL Security vulnerability, data loss, deadlock, RCE, anything that ships broken
MAJOR Logic error, race condition, resource leak, broken edge case, wrong abstraction
MINOR Style, naming, missing test, redundant code, brittle assumption
NIT Cosmetic, formatting, comment wording
CLEAN Explicitly state when a section is fine. Prevents invented-bug bias.

Security

This MCP server implements security sandboxing:

  • File access is sandboxed to the current working directory by default
  • Sensitive paths (.ssh, .aws, .env, /etc/passwd, etc.) are refused
  • Git refs are validated against a strict character whitelist to prevent option injection
  • Subprocess calls use shell=False and disable global git config

See SECURITY.md for the full threat model.

Development

# Install in development mode
pip install -e ".[dev]"

# Run tests
pytest

# Run tests with coverage
pytest --cov=blunt_codereview

Benchmark snippets

The benchmarks/ directory contains 5 regression snippets that verify the reviewer:

  1. SQL injection — expects CRITICAL, line citation, "Do not merge"
  2. Mutable default — expects MAJOR
  3. Clean code (binary search) — expects CLEAN, "Ship it" (anti-hallucination test)
  4. Swallowed exception — expects MAJOR
  5. Off-by-one — expects MAJOR

The clean code benchmark is the most important — it catches hallucination. If the reviewer invents bugs in correct code, the anti-RLHF system is broken.

License

MIT

Acknowledgments

This project encodes the kernel maintainer tradition of code review — a methodology practiced by many senior engineers across many projects (Linux kernel, PostgreSQL, Redis, SQLite, and others). We cite the tradition, not any single practitioner.

Project details


Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

code_review_agent_mcp-0.1.0.tar.gz (34.0 kB view details)

Uploaded Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

code_review_agent_mcp-0.1.0-py3-none-any.whl (25.1 kB view details)

Uploaded Python 3

File details

Details for the file code_review_agent_mcp-0.1.0.tar.gz.

File metadata

  • Download URL: code_review_agent_mcp-0.1.0.tar.gz
  • Upload date:
  • Size: 34.0 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/6.2.0 CPython/3.13.12

File hashes

Hashes for code_review_agent_mcp-0.1.0.tar.gz
Algorithm Hash digest
SHA256 e128c9ab00524ab59bc557d8558dc6a25e3fc97f43a5b8d44e91ebb10f0b7eb4
MD5 c257e7e35635e353e5f4c3734e96105e
BLAKE2b-256 92e856a48da6977fc0f2983237492f4d414763583be993ca0f4a135778aaf558

See more details on using hashes here.

File details

Details for the file code_review_agent_mcp-0.1.0-py3-none-any.whl.

File metadata

File hashes

Hashes for code_review_agent_mcp-0.1.0-py3-none-any.whl
Algorithm Hash digest
SHA256 4f43645ef4d3c86fe3cd1d70bf7204f55be5d2596d07c79d60002fe09d22965e
MD5 bb89635daac21d869d80c9c909e73105
BLAKE2b-256 fe6477a7d8080efc3ae369e8fc034dbe6d8350b5e8c76085ed91921244556531

See more details on using hashes here.

Supported by

AWS Cloud computing and Security Sponsor Datadog Monitoring Depot Continuous Integration Fastly CDN Google Download Analytics Pingdom Monitoring Sentry Error logging StatusPage Status page