Skip to main content

retrieval-mcp

An MCP server for the Retrieval academic-paper API - semantic paper search, document matching, ACE journal memory, and index inventory. Self-contained: it talks to the backend over HTTP only (just mcp + httpx), so it installs anywhere with uvx / pip - no repo checkout, no GPU, no models.

By default it targets the compute box on the lab LAN (http://10.100.100.111:8000), which trusts LAN callers so no key is needed. Off-LAN, point RETRIEVAL_API_URL at the public gateway (https://retrieval.rnarket.com) and set RETRIEVAL_API_KEY (sk-...).

Tools

Every tool's full docstring (purpose + each argument with its default + an example) is what your LLM sees - call them by name. Summary:

Paper retrieval

Tool What it does
search_papers Semantic hybrid search over 95k+ top-venue CS papers (filters: venue, year, title_only)
search_within_paper Every matching passage inside one paper
match_document / match_paper Content-nearest papers to a passage / to a paper
list_conferences / corpus_stats Venue registry / corpus size

Journal work-memory (scoped to the current project by default)

Tool What it does
journal_record Record a work note (memory) or a file's current content (doc, latest-wins)
journal_search Search memory - keyword (FTS5, no embedding) or hybrid/dense/sparse
journal_recent List recent entries
journal_index_dir Batch-index a local dir's files into the journal (latest-wins per file)

Code KB (source stays local - only chunks are uploaded)

Tool What it does
index_code AST-chunk a repo locally (40+ languages) and index it, scoped to you
search_code Semantic code search with path:line citations
index_inventory Your indexed-file tree: user -> host -> project -> dir -> file

ACE playbook (accumulated, curated lessons per project)

Tool What it does
ace_context_aware / ace_playbook Retrieve relevant / list all curated bullets
ace_enhance_prompt / ace_smart_generate Attach playbook lessons to a prompt (no LLM call)
ace_smart_reflect Curate a transferable lesson into the playbook (grow-and-refine dedup)

Code KB language coverage

index_code chunks 40+ languages structurally via tree-sitter (chonkie CodeChunker): Python, TypeScript/TSX/JS/JSX (React), Java, Kotlin (incl. Jetpack Compose .kt/.kts), Swift, Go, Rust, C/C++, C#, Ruby, PHP, Lua, Scala, Dart, R, Julia, Elixir, Erlang, Haskell, OCaml, SQL, GraphQL, Protobuf, HTML, CSS/SCSS (Tailwind = CSS classes), Vue, Svelte, shell, PowerShell, Dockerfile, Terraform/HCL, CMake, YAML/JSON/TOML/XML, and more. Grammarless config/text files fall back to line-window chunks; docs (.md) and binaries are skipped (docs belong in the journal via journal_index_dir).

Install

Claude Code

# LAN (no key):
claude mcp add retrieval -- uvx --from retrieval-mcp==0.2.37 retrieval-mcp
# Off-LAN (public gateway + key):
claude mcp add retrieval \
  --env RETRIEVAL_API_URL=https://retrieval.rnarket.com \
  --env RETRIEVAL_API_KEY=sk-... \
  -- uvx --from retrieval-mcp==0.2.37 retrieval-mcp

Claude Desktop / any MCP client

claude_desktop_config.json (macOS: ~/Library/Application Support/Claude/, Windows: %APPDATA%\Claude\):

{
  "mcpServers": {
    "retrieval": {
      "command": "uvx",
      "args": ["--from", "retrieval-mcp==0.2.37", "retrieval-mcp"],
      "env": {
        "RETRIEVAL_API_URL": "https://retrieval.rnarket.com",
        "RETRIEVAL_API_KEY": "sk-..."
      }
    }
  }
}

No uv? pip install retrieval-mcp then use "command": "retrieval-mcp".

Automatic code-index refresh

index_code(...) returns a local job ID before repository walking, hashing, AST chunking, upload, GPU embedding, or Qdrant upsert completes. Poll that same ID with index_code_status() through the preparing, backend queue, and terminal phases. A second index request for the same scope reuses the active job instead of starting another scan.

With auto_refresh=True, the client then keeps one filesystem watcher for that absolute repository path. Ordinary file events hash and chunk only the touched paths; ignore-rule changes trigger a full reconcile. search_code() never waits for the watcher or indexing: it returns the last completed snapshot and reports freshness separately.

Git repositories continue to honor Git's ignore rules by default. Any directory, including non-Git projects, can add scope-relative patterns to .retrievalignore; callers can add temporary patterns with exclude_globs=["generated/**", "private.py"]. Explicit delete_code() cancels/fences stale refresh work and removes vectors, inventory, and the saved refresh policy.

Config (env)

Var Default Notes
RETRIEVAL_API_URL http://10.100.100.111:8000 LAN compute box (no key). Off-LAN, set to https://retrieval.rnarket.com.
RETRIEVAL_API_KEY - sk-... key for the gateway (create under /auth/keys). Required off-LAN.
Journal/code scope current absolute directory path The client derives scope from the directory you run or pass to each tool, so projects do not leak into each other.

License

MIT

Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

retrieval_mcp-0.2.37.tar.gz (40.6 kB view details)

Uploaded Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

retrieval_mcp-0.2.37-py3-none-any.whl (41.3 kB view details)

Uploaded Python 3

File details

Details for the file retrieval_mcp-0.2.37.tar.gz.

File metadata

  • Download URL: retrieval_mcp-0.2.37.tar.gz
  • Upload date:
  • Size: 40.6 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/6.2.0 CPython/3.13.7

File hashes

Hashes for retrieval_mcp-0.2.37.tar.gz
Algorithm Hash digest
SHA256 b8c58875bc35e76abe0cf39cc63de7ecb3c0e391ae3dde4cef5c56b2defc1618
MD5 1b18977b61ad1b7953aeff3a7d2d8349
BLAKE2b-256 472ee00a06346d2339f58ebdd6fad7492e4d3321d96819a26b43287e914401bf

See more details on using hashes here.

File details

Details for the file retrieval_mcp-0.2.37-py3-none-any.whl.

File metadata

  • Download URL: retrieval_mcp-0.2.37-py3-none-any.whl
  • Upload date:
  • Size: 41.3 kB
  • Tags: Python 3
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/6.2.0 CPython/3.13.7

File hashes

Hashes for retrieval_mcp-0.2.37-py3-none-any.whl
Algorithm Hash digest
SHA256 8f3f6711e549d60e1a269b8eb86863dbc5ccc5b130d3eec3884e5c7f3a28385f
MD5 2c907714b452c1d02713dc7a5a6210e9
BLAKE2b-256 499e1341241ed483a3a5437e05c5afa265419fc6257e185930e1a7c2bc33601e

See more details on using hashes here.

Release history Release notifications | RSS feed

0.4.11

2 files

0.4.10

2 files

0.4.9

2 files

0.4.8

2 files

0.4.7

2 files

0.4.6

2 files

0.4.5

2 files

0.4.4

2 files

0.4.3

2 files

0.4.2

2 files

0.4.1

2 files

0.4.0

2 files

0.3.2

1 file

0.3.1

2 files

0.3.0

2 files

This release

0.2.37 This release

2 files

0.2.36

2 files

0.2.35

2 files

0.2.34

2 files

0.2.33

2 files

0.2.32

2 files

0.2.31

2 files

0.2.30

2 files

0.2.29

2 files

0.2.28

2 files

0.2.27

2 files

0.2.26

2 files

0.2.25

2 files

0.2.24

2 files

0.2.23

2 files

0.2.22

2 files

0.2.21

2 files

0.2.20

2 files

0.2.19

2 files

0.2.18

2 files

0.2.17

2 files

0.2.16

2 files

0.2.15

2 files

0.2.14

2 files

0.2.13

2 files

0.2.12

2 files

0.2.11

2 files

0.2.10

2 files

0.2.9

2 files

0.2.8

2 files

0.2.7

2 files

0.2.6

2 files

0.2.5

2 files

0.2.4

2 files

0.2.3

2 files

0.2.2

2 files

0.2.1

2 files

0.2.0

2 files

0.1.2

2 files

0.1.1

2 files

0.1.0

2 files

Anthropic, PBC Visionary sponsor Bloomberg Visionary sponsor Hudson River Trading Visionary sponsor Meta Visionary sponsor NVIDIA Visionary sponsor Microsoft Sustainability sponsor Depot Continuous Integration AWS Cloud computing and Security Sponsor Datadog Monitoring Fastly CDN Google Download Analytics Sentry Error logging StatusPage Status page