agentless-mcp
agentless-mcp provides tree-sitter-based code navigation, repository
structure analysis, and patch validation for local repositories. It exposes
the same functionality through a command-line interface and a read-only
MCP server.
Install
uv tool install "agentless-mcp[mcp]"
That installs two console scripts. agentless-mcp is the CLI, for a human or
for an agent driving it over a shell. agentless-mcp-server is the MCP
server, which an MCP client launches for you rather than something you run in
a terminal. The server needs the mcp extra, which is why the install above
leads with it; install without the extra when you want the CLI alone:
uv tool install agentless-mcp
Without the extra, agentless-mcp-server --help and --version still
answer, and any real invocation exits with the install command for the
extra.
Installing from a local checkout takes the same command with the path to the
checkout in place of the package name. Add --reinstall --refresh whenever
you rebuild: uv tool install --force reuses a cached wheel while the version
string is unchanged, so a changed tree does not reach the installed tool.
The history view needs git 2.25 or newer, the first release whose git log -L documents --no-patch as suppressing the patch. On an older git the patch
text reaches the log parser, which refuses the answer rather than misread it.
Every other view works without git.
Both entry points warm cold grammars in the background at startup (one
digest-verified bundle fetch at most; --no-auto-warm or
AGENTLESS_MCP_NO_AUTO_WARM opts out, AGENTLESS_MCP_NO_DOWNLOAD forbids
all fetching). To warm explicitly and fail loudly instead:
agentless-mcp warmup
agentless-mcp guide prints the full agent usage guide, which ships with the
package; agentless-mcp guide --section NAME prints one section, and an
unknown name lists them all.
CLI
Most commands analyze the repository containing the current directory. Use
--repo PATH to select another repository. Add --json where supported for
machine-readable output.
Navigate code
agentless-mcp map --focus src/app.py
agentless-mcp tree --depth 3
agentless-mcp skeleton src/app.py
agentless-mcp expand py:src/app.py::App.run
agentless-mcp slice src/app.py --lines 40:80
agentless-mcp find-symbol App
agentless-mcp refs App.run
agentless-mcp explain App.run
agentless-mcp history py:src/app.py::App.run
These commands provide repository maps, directory trees, symbol overviews,
full symbol bodies, source slices, symbol lookup, references, symbol context,
and the commits that touched a symbol's lines. Symbol IDs are printed by map
and skeleton and can be passed to expand, refs, explain, and related
commands.
history takes a symbol id and prints the commits that changed that symbol's
lines, newest first, with their message bodies. --limit caps the commits
(default 10) and --budget caps the tokens spent on the bodies (default
12000); at most 120 commits render, however large the limit is. A span no
commit touched, a file that is not in HEAD, and a directory without git are
each a refusal rather than an empty answer.
Analyze structure
agentless-mcp path App.run Database.connect
agentless-mcp cycles
agentless-mcp communities
agentless-mcp health
agentless-mcp diagram > modules.mmd
agentless-mcp html > modules.html
These commands find relationships between symbols or files, report import cycles, group related files, list orphan candidates, unused exports and hubs, and export Mermaid or interactive HTML graphs.
Validate patches
The CLI also supports deterministic patch workflows:
agentless-mcp patch parse --file change.patch
agentless-mcp patch check --file change.patch --repo /path/to/repo
agentless-mcp patch apply --file change.patch --repo /path/to/repo
agentless-mcp lint --candidates ./candidates --repo /path/to/repo
agentless-mcp lint --diff change.patch --repo /path/to/base-checkout
agentless-mcp validate --candidates ./candidates --repo /path/to/repo \
--test-cmd 'pytest -q'
agentless-mcp vote --verdicts verdicts.jsonl
Patch candidates can use SEARCH/REPLACE text or the package's edits.json
format. One request carries at most 500 edits and 2,000,000 bytes of search
and replace text; a larger one is refused when it is read, before any file is
opened. validate normalizes them against HEAD, runs byte-identical resulting
file states once while preserving every candidate's vote, and skips a
reproduction command when regression has already failed. vote ranks the
candidates that pass.
lint --diff runs the same checks over a branch's or a pull request's unified
diff, so a change that already exists does not have to be hand-converted first.
The checks compare the diff against --repo as it stands, which means --repo
must be a checkout of the diff's base, not a tree with the diff already
applied — otherwise every symbol the diff adds is already in the file and the
report would describe the change against itself. The usual shape is a second
worktree at the merge-base:
git diff main...HEAD > change.patch
git worktree add /tmp/base $(git merge-base main HEAD)
agentless-mcp lint --diff change.patch --repo /tmp/base
Pointing --repo at the branch instead is not silently wrong: each affected
file is reported as a not_checked coverage gap naming the remedy. Binary files
and mode-only changes are reported the same way, and a construct one edit cannot
express — a rename, a -U0 diff with no context — is refused with the reason.
Cache and capabilities
Parsing happens on demand. Build an optional repository cache to improve repeated queries:
agentless-mcp index --repo /path/to/repo
agentless-mcp capabilities --repo /path/to/repo
The MCP server builds and refreshes this cache itself, in the background,
the first time it serves a repository whose index is absent or stale
(--no-auto-index or AGENTLESS_MCP_NO_AUTO_INDEX opts out); the CLI
indexes only through the explicit command above. Use --no-cache on
repository-scoped commands to bypass the cache.
Every index run then releases cold repository caches until
$XDG_CACHE_HOME/agentless-mcp/ fits under 5 GiB, least-recently-used first.
Because the server indexes on its own, nothing else bounds that directory.
Caches used in the last 24 hours are never released, so a working set larger
than the ceiling exceeds it instead of thrashing. Set
AGENTLESS_MCP_MAX_CACHE_BYTES to another ceiling, or to 0 to keep every
cache forever. An evicted repository loses no answers, only the speed: the
next call parses on demand and the one after that re-indexes.
MCP server
The server exposes read-only repository tools. It talks over stdio by default, which is what a client that launches the server as a child expects. For a single-user machine, register it once and let the client's advertised workspace authorize repositories: whatever repository you open a session in is served on the first tool call, with nothing to enable per repo.
claude mcp add --scope user agentless -- agentless-mcp-server --allow-client-roots
For a locked-down server, omit --allow-client-roots and pass an explicit
allowlist instead; then only the listed repositories are servable, and a
client-advertised root can only select among them, never add one:
agentless-mcp-server --root /path/to/repo --root /path/to/other
Over HTTP the server also watches its own install: when the package is
upgraded or reinstalled, it finishes in-flight requests and replaces itself
with the new code (--no-auto-restart or AGENTLESS_MCP_NO_AUTO_RESTART
opts out). A long-running process otherwise serves the code it loaded at
startup forever -- reconnecting clients refreshes the connection, never the
process. On Windows the server exits cleanly instead and a supervisor's
Restart= completes the loop; docs/deploy/mcp-agentless.service is a
ready example unit.
--roots-from FILE reads that same list from a file, one path per line.
The file is re-read whenever it changes on disk, so appending a line enrolls
a repository on the next call without a restart, and the refusal an agent
sees for an unlisted repository names the file to append to. Blank lines and
whole-line # comments are skipped, and the flag is repeatable and combines
with --root:
cat > ~/.config/agentless-mcp/roots <<'EOF'
# one repository path per line
/path/to/repo
/path/to/other
EOF
claude mcp add --scope user agentless -- \
agentless-mcp-server --roots-from ~/.config/agentless-mcp/roots
Serving over HTTP
A client that cannot spawn a child process gets the same tools over FastMCP's
streamable-http transport, from one long-lived server that several clients
share. The endpoint is http://HOST:PORT/mcp:
agentless-mcp-server --transport http --port 8766 \
--roots-from ~/.config/agentless-mcp/roots
The bind address is loopback-only and is checked, not merely defaulted: this
server authenticates nobody, so the --root allowlist decides which
repositories are readable and says nothing about who may read them. On a
routable address that is unauthenticated read access to every enrolled
repository, so a non-loopback --host is refused before the socket opens.
Put an authenticating proxy in front if you need it off-host.
--host and --port apply to the HTTP transport only; passing either under
stdio is refused rather than ignored, because there is no socket to bind.
Several clients sharing one server also share its worker threads, so
--max-concurrency N bounds how many tool calls run at once; it accepts 1 to
64 and defaults to 8. Every admitted call may hold a tag-cache connection and
a git child, and the ungated ceiling is whatever the interpreter's default
thread pool sized itself to -- 32 on a large host, 5 on a small one. Raise it
for a server with many clients, lower it to keep the machine's git and disk
load predictable. It applies to stdio too, where the default is already the
right answer for a single client.
Every tool takes repo_root first. It may be omitted only when the server
holds one repository, or when the client advertises a root that selects
exactly one; otherwise the refusal lists the roots to choose from.
The MCP tools are six intent-shaped surfaces; three of them fold their
questions behind an operation parameter:
| Tool | Operations | Purpose |
|---|---|---|
orient |
map, communities, cycles, diagram, path, health |
Where does this live, how is the repository put together |
symbols |
find, overview, expand, explain, locate |
Look up, skeleton, expand, or explain symbols; resolve locations |
find_referencing_symbols |
Find references and callers (blast radius) | |
history |
Why does this code exist: the commits that touched a symbol's lines, bodies included | |
read |
slice, dir |
Read selected source lines; list the repository tree |
capabilities |
Report loaded grammars and cache state |
One worked call per surface:
orient(operation="map", focus=["src/app.py", "quote"])
symbols(operation="expand", stable_ids=["py:src/app.py::App.run"])
find_referencing_symbols(target="App.run")
history(target="py:src/app.py::App.run")
read(operation="slice", path="src/app.py", lines=[[40, 80]])
capabilities()
A wrong operation is answered with the valid list, and a parameter foreign
to the selected operation is refused with a message naming what that
operation accepts and requires.
This v2 surface is the default. For the transition, --surface v1 publishes
the previous twelve per-question tools (repo_map, expand_symbols, and the
rest) and --surface both publishes the union; v1 remains for one release.
find_referencing_symbols, capabilities and history stand alone on both
surfaces, so the union is fifteen names rather than eighteen. The mapping
between the surfaces is in
agentless-mcp guide --section the-two-surfaces.
What the server runs
The MCP server does not apply patches or execute repository commands. Git is
the only program the read surface spawns, and every git call the package makes
goes through one bounded runner: a fixed configuration prefix on the argv, an
environment with git's redirection variables stripped, LC_ALL=C, a deadline,
and a cap on the output read from the pipe.
The prefix holds the settings a repository's own configuration could otherwise
use to name a program. core.fsmonitor, diff.external and core.pager are
fixed on the argv, and log.showSignature=false keeps a repository-local
gpg.program from running on every git log, which is what the receipt's
churn count and history read. The CLI's patch sandbox passes --no-textconv
on its diff, the one control git offers over a repository-named textconv
driver.
Two GIT_ variables survive the environment scrub: GIT_CONFIG_GLOBAL and
GIT_CONFIG_NOSYSTEM. They move where git reads configuration from rather
than which repository it reads, so a caller that pins git's configuration
reaches these calls too. Every other GIT_ variable is removed, because that
family can redirect git away from the repository the call named.
Each tool answers on a worker thread rather than on the event loop. One slow
call -- a cold map, a 30-second git log -L -- therefore does not stall the
answers to the others, and calls on one connection can overlap.
Keeping the tools enabled in Claude Code
Claude Code asks for approval the first time a session calls each MCP tool.
Every prompt is a chance to fall back to Grep, and the approval does not
carry to the next session. Pre-approve the server once instead, in
permissions.allow in ~/.claude/settings.json for every project, or in
the repository's .claude/settings.json for one:
{
"permissions": {
"allow": [
"mcp__agentless__orient",
"mcp__agentless__symbols",
"mcp__agentless__find_referencing_symbols",
"mcp__agentless__read",
"mcp__agentless__capabilities",
"mcp__agentless__history"
]
}
}
A rule of the form mcp__agentless approves every tool the server
publishes, which survives a surface change and a new operation. List the
tools one by one, as above, when you want each addition to ask once before
it runs unattended. Claude Code matches these rules literally: a wildcard
such as mcp__agentless__* matches nothing.
The prefix carries the server name you registered. These entries assume the
claude mcp add ... agentless ... line at the top of this section. Under
--surface v1 or --surface both, add the per-question tool names as
well: repo_map, list_dir, get_symbols_overview, expand_symbols,
read_slice, find_symbol, explain_symbol, analyze_structure, and
resolve_locations.
Eager tool schemas
Claude Code can defer an MCP server's tools: they arrive as bare names, and
the schema is fetched before the tool can be called. Five of the six tools
publish an alwaysLoad hint that asks a deferring client to hold their
schemas from the first turn, so the gate below redirects an agent that already
knows what each tool answers. Measured, the hint costs nothing: the deferred
and eager arms were indistinguishable on every localization metric, and
deferral spent a round trip fetching the schemas anyway. The gate is the
load-bearing part either way. history publishes no hint by design: it
answers why rather than where, so a deferring client fetches its schema only
when an agent asks.
Structural-first gate (recommended hooks)
Prose asks for the structural pass. A hook enforces it. Two scripts in
contrib/hooks/ deny repository-wide or directory-wide Grep, every Glob,
and any Bash command that parses as a tree search, until the session has
made a localizing Agentless call.
orient(map|path), symbols(find|overview|expand|explain), read(slice) and
find_referencing_symbols unlock the session; diagnostics, read(dir),
symbols(locate), history, and the shape listings orient(communities|cycles|diagram|health)
do not. read(slice) unlocks on the same rule that allows an exact-file Grep:
naming a file and a line range is the localization. A directory listing is how
you look for a file, not evidence that you found one. The equivalent v1 tools unlock servers running the temporary
compatibility surface. Grep scoped to one existing file remains available
before unlock because the caller has already localized that search. The mark
hook reads the call and not its result, so an Agentless call that errored
still unlocks.
The constraint is an order, not a ban: once the gate opens, broad search keeps
the one job a symbol map cannot do, which is string literals, error messages,
config keys, and fixtures.
Search routed through the shell is covered too. The check hook reads
Bash commands rather than trusting the tool name, and denies only what
parses cleanly as a tree search: rg with no path operand or a directory
operand, grep with a recursive flag, and find with -name/-path.
Everything else passes, including a pipe filter over another command's output
(git log | grep fix) and a search scoped to one existing file. Shell text
cannot be parsed in general, so this half fails open: git grep, xargs,
subshells, and variable command names all pass unexamined, and this
heuristic has not been benchmarked the way the tool-name half has.
This is the recommended install rather than an optional extra, and the reason
is measured: on SWE-Explore-Bench (n=60 issue-localization tasks, Sonnet,
agentless-mcp 0.6.1), an arm restricted to the agentless tools plus Read
beat a free-choice arm on all six localization metrics, every 95% confidence
interval excluding 0. Read that as evidence for the ordering, not as a
prediction for your repository: the measured arm removed the native search
tools, and this gate only defers them.
Publishing history costs nothing detectable on the same instrument, measured
twice. Two paired 60-instance runs put the hooked arm on 0.7.3 against the
hooked arm on 0.8.0, differing only in the server build. No acceptance metric
moved significantly against the release in either run, and nDCG@100 favoured
it in both. A same-configuration replicate measured the run-to-run noise floor
those deltas are read against. The full comparison,
its guardrails, and the schema-policy arms are in
docs/analysis/benchmark-methodology.md.
Claude Code
Install the gate by copying the two scripts and adding one hooks block.
- Copy
contrib/hooks/agentless_gate_check.pyandcontrib/hooks/agentless_gate_mark.pyto a stable path, for example~/.claude/hooks/. - Merge the block in
contrib/hooks/settings-example.jsoninto~/.claude/settings.jsonfor every project, or into the repository's.claude/settings.jsonfor one. JSON allows no comments, so the file carries placeholder paths and this section carries the explanation. - Replace each
/ABSOLUTE/PATH/TO/...placeholder with the absolute path of the copied script. A relative path does not resolve. - Keep the
/usr/bin/env python3prefix. Claude Code runs thecommandstring as a shell command, not as an argv list, so the interpreter has to be named. - Start a new session. Claude Code reads
settings.jsonat session start.
{
"hooks": {
"PreToolUse": [
{
"matcher": "Grep|Glob|Bash",
"hooks": [
{
"type": "command",
"command": "/usr/bin/env python3 /home/you/.claude/hooks/agentless_gate_check.py"
}
]
}
],
"PostToolUse": [
{
"matcher": "mcp__agentless__.*",
"hooks": [
{
"type": "command",
"command": "/usr/bin/env python3 /home/you/.claude/hooks/agentless_gate_mark.py"
}
]
}
]
}
}
Verify the scripts outside a session first. Feed each one a fake hook payload on stdin and read the exit code:
rm -rf /tmp/agentless_gate
# No structural call yet: exit 2, and the denial text goes to stderr.
echo '{"session_id":"probe","tool_name":"Grep","tool_input":{"pattern":"needle"}}' \
| python3 ~/.claude/hooks/agentless_gate_check.py; echo "exit=$?"
# One localizing structural call marks the session.
echo '{"session_id":"probe","tool_name":"mcp__agentless__orient","tool_input":{"operation":"map"}}' \
| python3 ~/.claude/hooks/agentless_gate_mark.py
# Now the same Grep is allowed: exit 0, no output.
echo '{"session_id":"probe","tool_name":"Grep","tool_input":{"pattern":"needle"}}' \
| python3 ~/.claude/hooks/agentless_gate_check.py; echo "exit=$?"
rm -rf /tmp/agentless_gate
The first call prints the denial text and reports exit=2. The third call
prints nothing and reports exit=0. Inside a session, a denied Grep shows
as a blocked tool call, and the model receives the same denial text as an
instruction to call orient first. The unlock marker lives under
/tmp/agentless_gate/.
Both hooks fail open. Malformed stdin, a missing session id, an
unwritable /tmp, or any internal error exits 0 and allows the call: the
failure mode to expect is a Grep that runs early, never a session that
cannot search. Set AGENTLESS_GATE_LOG to a file path to append one JSONL
line per decision and confirm the gate fired. Remove the gate by deleting
the two hook blocks from settings.json, the copied scripts, and
/tmp/agentless_gate; nothing else on disk changes.
Codex CLI
The two scripts are unchanged between clients. Codex uses the same
PreToolUse and PostToolUse events, the same session_id, tool_name and
tool_input payload fields, and the same exit-2-blocks-with-stderr contract.
Three things differ, and all three are configuration.
-
Name the MCP server
agentless. The server name is the tool-name prefix: an entry namedagentless-mcpproducesmcp__agentless-mcp__orient, which the mark hook'smcp__agentless__.*matcher never sees, and the unlock half goes dormant.[mcp_servers.agentless] command = "/ABSOLUTE/PATH/TO/agentless-mcp-server" args = ["--allow-client-roots", "--root", "/ABSOLUTE/PATH/TO/repo"]
Pass every repository you work in as a
--root. A server given one root is blind in every other repository, and answers there look like an empty map rather than a misconfiguration. -
Write
~/.codex/hooks.json(or the same block inline under[hooks]in~/.codex/config.toml; a repository-local.codex/hooks.jsonalso works). ThePreToolUsematcher gainsexec_command, Codex's unified-exec tool name.{ "hooks": { "PreToolUse": [ { "matcher": "Grep|Glob|Bash|exec_command", "hooks": [{ "type": "command", "command": "/usr/bin/env python3 /ABSOLUTE/PATH/TO/agentless_gate_check.py" }] } ], "PostToolUse": [ { "matcher": "mcp__agentless__.*", "hooks": [{ "type": "command", "command": "/usr/bin/env python3 /ABSOLUTE/PATH/TO/agentless_gate_mark.py" }] } ] } }
-
Trust the hooks once. Codex does not run a non-managed hook until it has been reviewed. Open the TUI and run
/hooksto inspect and trust both. Until then they are skipped silently -- no error, no log line, and the gate simply does not fire.
Codex searches through its shell tool rather than a Grep tool, so there the
shell half of the gate does the work. The anthropic/alwaysLoad hint does not
port -- Codex always defers MCP schemas -- so Codex runs deferred schemas plus
the gate, which measured indistinguishable from eager on every metric.
Navigation craft for the agent
The gate decides when the structural pass happens; the routing knowledge
behind it ships with the package. agentless-mcp guide prints the full agent
usage guide -- the canonical recipe, the escalation funnel, and per-tool
usage -- and --section NAME prints one section for pasting into CLAUDE.md
or a dispatch prompt. The rules that matter most:
- Seed the map with what the task already names: paths, module stems,
symbol names, or a decorated spelling like a traceback frame. One exact
path outweighs a name that matched twenty files. A seed that matches
nothing comes back in
unresolved_seeds; reseed rather than trusting an unfocused ranking. - Escalate one rung at a time and stop at the rung that answers:
orient(map)to locate,symbols(overview)for what a file declares,symbols(expand)for the few bodies that matter,symbols(explain)for one suspect symbol's wiring,find_referencing_symbolsonly when callers decide the answer (it is the expensive one),read(slice)last. - Weigh reference tiers:
same-fileandresolved-via-importare bindings;uniqueandname-only-ambiguousare candidates to inspect. - Run
Grepafter the structural pass for what a symbol map cannot hold: string literals, error messages, config keys, fixtures. - Read an empty or thin map as a report, not an absence, and call
capabilitiesto separate a repository that never parsed from one that declares nothing there.
Supported languages
The bundled grammars support Bash, C, C++, C#, Go, HCL, Java, JavaScript, JSON,
Kotlin, Lua, PHP, Python, Ruby, Rust, Scala, SQL, Swift, TOML, TSX, TypeScript,
and YAML. Run capabilities to see the grammars available in the current
installation.
License
Download files
Download the file for your platform. If you're not sure which to choose, learn more about installing packages.
Source Distribution
Built Distribution
Filter files by name, interpreter, ABI, and platform.
If you're not sure about the file name format, learn more about wheel file names.
Copy a direct link to the current filters
File details
Details for the file agentless_mcp-0.8.1.tar.gz.
File metadata
- Download URL: agentless_mcp-0.8.1.tar.gz
- Upload date:
- Size: 486.2 kB
- Tags: Source
- Uploaded using Trusted Publishing? No
- Uploaded via:
uv/0.9.28 {"installer":{"name":"uv","version":"0.9.28","subcommand":["publish"]},"python":null,"implementation":{"name":null,"version":null},"distro":{"name":"Fedora Linux","version":"43","id":"","libc":null},"system":{"name":null,"release":null},"cpu":null,"openssl_version":null,"setuptools_version":null,"rustc_version":null,"ci":null}
File hashes
| Algorithm | Hash digest | |
|---|---|---|
| SHA256 |
caedc1e1c111bfb706b79866f32f2d9cc97a16b75790aa21b98fa9fe6e629f94
|
|
| MD5 |
dcfc00543d9714bcd1332de0cd22b0bc
|
|
| BLAKE2b-256 |
ba52b2f803f0f3d387aa55d06caad28164a30969300beeacb4a998f6af417b05
|
File details
Details for the file agentless_mcp-0.8.1-py3-none-any.whl.
File metadata
- Download URL: agentless_mcp-0.8.1-py3-none-any.whl
- Upload date:
- Size: 525.6 kB
- Tags: Python 3
- Uploaded using Trusted Publishing? No
- Uploaded via:
uv/0.9.28 {"installer":{"name":"uv","version":"0.9.28","subcommand":["publish"]},"python":null,"implementation":{"name":null,"version":null},"distro":{"name":"Fedora Linux","version":"43","id":"","libc":null},"system":{"name":null,"release":null},"cpu":null,"openssl_version":null,"setuptools_version":null,"rustc_version":null,"ci":null}
File hashes
| Algorithm | Hash digest | |
|---|---|---|
| SHA256 |
815a7725d00935f43b98722fcf91a760b71e3d869cfbbb39009d8c10c45de99c
|
|
| MD5 |
9cd22216ed2f907e98e26f20e17e1006
|
|
| BLAKE2b-256 |
01772fc070dc6bb1b231e833db79775306bf445779ad53f409127db583e0c31a
|