Skip to main content

MCP server for orchestrating multi-subagent code runs with minimized context and patch-only outputs.

Project description

agent-code-squad

Python MCP server that queues code-writing subagents with minimal context packs, isolated git worktrees, and patch-first collection. Heavy work (git worktree add + codex exec start) runs in a background worker to avoid tool-call timeouts.

Quickstart

cd mcp-tools/agent-code-squad
poetry env use python3
poetry install
poetry run agent-code-squad

MCP client config (example)

[mcp_servers.agent-code-squad]
command = "poetry"
args = ["run", "agent-code-squad"]
cwd = "/Volumes/workspace/dzrlab/k3s-test/mcp-tools/agent-code-squad"

Project config (recommended)

Create <repo_root>/code-squad/config.json to set defaults per project.

Priority: code-squad/config.json → CLI flags → built-in defaults.

Example:

{
  "dispatch_dir": "~/.codex/code-squad",
  "defaults": {
    "model": "gpt-5.2-codex",
    "sandbox": "workspace-write",
    "approval_policy": "never",
    "worker_concurrency": 2
  },
  "energy_saver": {
    "default_reasoning_effort": "medium",
    "max_reasoning_effort": "medium"
  }
}

Recommended flow

  1. code_squad_execute: start tasks immediately and return quickly (default non-blocking).
  2. Optional manual flow:
    • code_squad_run: queue tasks and enqueue workers right away.
    • code_squad_tick: optional bounded status refresh (enqueues already-queued tasks if needed).
    • code_squad_status / code_squad_events: poll light status or stream JSONL events.
    • code_squad_collect: extract patches from job stdout and persist under run artifacts (works even after worktree cleanup).
    • code_squad_verify: optional; runs compileall/pytest or custom options.commands arrays; results saved under artifacts.
    • code_squad_report: emit run/task summary as report.json and report.md.
    • code_squad_prune: clean up old runs/worktrees by retention policy.
    • code_squad_cancel (optional): cancel a job or all tasks in a run.
    • code_squad_cleanup: drop worktrees and optionally delete run artifacts.

Modify Multiple Modules

When the user request is “modify several modules”, prefer one task per module and enforce boundaries so parallel work stays precise.

  • Put each module under a distinct directory (example: modules/<name>/).
  • Use allow_globs per task to restrict which files a task is allowed to touch.
  • Use deny_globs to protect shared areas (configs, app entrypoints, shared libs) from module tasks.
  • allow_globs / deny_globs are also injected into the subagent prompt as hard rules to reduce scope drift.
  • If allow_globs is set, the worktree attempts a best-effort git sparse-checkout to physically hide non-allowed paths (further reducing context + accidental edits). The mapping prefers the tightest directory prefix (e.g. modules/user/**modules/user).
  • If scope is violated, collection marks scope.ok=false and writes a scope artifact under runs/<run_id>/artifacts/scope/<slug>.json.

Example code_squad_execute input:

{
  "cwd": "/path/to/repo",
  "tasks": [
    {
      "name": "user module",
      "scope_name": "modules/user",
      "allow_globs": ["modules/user/**"],
      "deny_globs": ["shared/**", "config/**", "app/**"],
      "prompt": "Fix bug in user module: ... (patch only)"
    },
    {
      "name": "billing module",
      "scope_name": "modules/billing",
      "allow_globs": ["modules/billing/**"],
      "deny_globs": ["shared/**", "config/**", "app/**"],
      "prompt": "Fix bug in billing module: ... (patch only)"
    }
  ],
  "options": {
    "poll_interval": 1.0,
    "wait_seconds": 0,
    "cleanup": true,
    "keep_failed": true
  }
}

Auto Context Packs (Optional)

If a task does not provide context_pack, you can enable automatic context packing:

  • options.auto_context_pack: dict with keys glob, max_files, max_snippets, max_total_chars, and optional hints.
  • Query selection per task: task.context_querytask.scope_nametask.name.
  • If a task specifies allow_globs / deny_globs, the context pack generation applies the same filters to reduce cross-scope leakage.
  • Context packs are cached under <dispatch_base>/context_packs/ to avoid repeated rg scans.

Change Size Limits (Optional)

To prevent large/low-signal edits:

  • Per-task: task.max_touched_files, task.max_patch_bytes
  • Or defaults: options.max_touched_files, options.max_patch_bytes

If limits are exceeded, scope is marked ok=false and code_squad_execute fails the task.

Cleanup and Retention

Worktree cleanup is controlled by options.cleanup / options.keep_failed (used by the background finalizer).

  • options.cleanup (default true): remove worktrees after execution
  • options.keep_failed (default true): keep failed and timeout task worktrees for debugging
  • options.cleanup_delete_run_artifacts (default false): also delete run artifacts (use with care)

For periodic cleanup of old runs/worktrees, use code_squad_prune:

  • Defaults: keep last 5 runs, keep successful runs for 3 days, keep failed/timeout runs for 1 day
  • dry_run=true by default; set dry_run=false to actually delete

Logs and Artifacts

Each run writes an index file to make review/debug easier:

  • runs/<run_id>/artifacts/index.json: run/task summary plus expected artifact paths
  • code_squad_execute returns quickly with index_path and next_poll_ms; heavy work is done asynchronously.
  • runs/<run_id>/artifacts/timeline.jsonl: structured timeline events (used by recent_events in status/execute/collect).
  • code_squad_status returns progress, next_action, next_poll_ms, and a per-task message.
  • code_squad_events(include_noise=false) filters out thread/turn/heartbeat noise and command/tool-call items by default

Reasoning effort cap (eco)

This server enforces model_reasoning_effort<=max_reasoning_effort for all tasks (default max is medium).

  • If you want lower effort, set options.model_reasoning_effort="low" (or options.reasoning_effort="low").
  • Requests above the configured max are clamped down to the max.

Tools

  • code_squad_capabilities_get: defaults and dispatch paths.
  • code_squad_context_pack: ripgrep-based context pack with snippet/char caps.
  • code_squad_run: queue tasks with per-task model/sandbox/approval/extra_config (auto-enqueues workers).
  • code_squad_execute: start run and return quickly (optional short wait via options.wait_seconds).
  • code_squad_tick: optional bounded worker pump (refresh a few running).
  • code_squad_status: summarize task states with compact last messages.
  • code_squad_events: stream compacted JSONL events per job using cursors (supports noise filtering).
  • code_squad_collect: extract patch + touched files from job stdout and persist under artifacts.
  • code_squad_verify: run compileall/pytest or custom commands using sys.executable, persisting results.
  • code_squad_report: write report.json/report.md under run artifacts.
  • code_squad_prune: clean up old runs/worktrees by retention policy.
  • code_squad_cancel: cancel a job or whole run (sets state to cancelled).
  • code_squad_cleanup: remove worktrees and (optional) run artifacts under .codex/code-squad/.

Debug output

All tools accept options.debug=true to return a debug block (paths, timings, raw stdout/stderr/events). Default responses stay minimal and omit worktree/run/dispatch paths, trace IDs, PIDs, and full prompts.

For robustness, options is treated as best-effort: non-dict inputs are coerced to {} rather than hard-failing validation.

Paths and defaults

  • Run metadata: <dispatch_base>/runs/<run_id>/run.json.
  • Job artifacts: <dispatch_base>/runs/<run_id>/jobs/<job_id>/{meta.json,stdout.jsonl,stderr.log,last_message.txt}.
  • Worktrees: <repo>/.codex/code-squad/worktrees/<run_id>/<task_slug>.
  • Dispatch base (default): <repo>/.codex/code-squad.
  • Defaults: model gpt-5.1-codex-max, sandbox workspace-write, approval policy never, worker concurrency 2.
  • Override via code-squad/config.json or CLI flags (config takes priority over CLI).

Project details


Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

agent_code_squad-0.1.6.tar.gz (30.2 kB view details)

Uploaded Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

agent_code_squad-0.1.6-py3-none-any.whl (32.4 kB view details)

Uploaded Python 3

File details

Details for the file agent_code_squad-0.1.6.tar.gz.

File metadata

  • Download URL: agent_code_squad-0.1.6.tar.gz
  • Upload date:
  • Size: 30.2 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? No
  • Uploaded via: poetry/2.2.1 CPython/3.14.2 Darwin/24.6.0

File hashes

Hashes for agent_code_squad-0.1.6.tar.gz
Algorithm Hash digest
SHA256 2d0676fff21499af0cd70b2625c2964082baea8ad478251609e1b958e94c389c
MD5 fe9422479b741096ed46d16bb7e3d593
BLAKE2b-256 88f23597c33d6412097854b3b8ca99636cb10adae02becbd72b62ed445a79104

See more details on using hashes here.

File details

Details for the file agent_code_squad-0.1.6-py3-none-any.whl.

File metadata

  • Download URL: agent_code_squad-0.1.6-py3-none-any.whl
  • Upload date:
  • Size: 32.4 kB
  • Tags: Python 3
  • Uploaded using Trusted Publishing? No
  • Uploaded via: poetry/2.2.1 CPython/3.14.2 Darwin/24.6.0

File hashes

Hashes for agent_code_squad-0.1.6-py3-none-any.whl
Algorithm Hash digest
SHA256 039b9c3b423fbe94a9124f6ebfceef52fcf730470906ee779073e1aba7c2335c
MD5 1e8bc0bcc2e322439f2f8c81fed77a11
BLAKE2b-256 dc1fac76115c7af53571839827cdbc0bd0fd390b98659688f993a636b6702505

See more details on using hashes here.

Supported by

AWS Cloud computing and Security Sponsor Datadog Monitoring Depot Continuous Integration Fastly CDN Google Download Analytics Pingdom Monitoring Sentry Error logging StatusPage Status page