Skip to main content

Cortex - RAG MCP for a knowledge base

Francais | English

Cortex is an MCP (Model Context Protocol) server that exposes semantic search over a local knowledge base. It lets Claude, Codex and Gemini find the right passage in your documents without wasting their context window. Search is semantic (by meaning, not keyword), in French and in English. Cortex processes and indexes the knowledge base locally without sending its content; the MCP client may still pass requested chunks to its model under its own policy. The optional Confluence writer only downloads explicitly allowlisted spaces; the generated Markdown, vector index, and lexical index remain local.

Starting with 2026.0906.01, Companion offers a guided home screen, operation history and Confluence setup from one page or space link. Connection and measured scope confirmation stay in the flow; successful collection is followed by indexing. Use the combined installer to keep Cortex and Companion compatible.

Source management in Companion (2026.0906.02)

The workflow includes search, remote page trees, impact review, evidence-based readiness, save-and-update, last-removal undo and recovery actions. See the contracts.

Paired version 2026.0906.02 provides visible Mes sources cards, a prefilled selection editor, confirmed page/space removal and browser links to originals. Removal only changes Cortex tracking; search reflects it after successful collection and indexing.

Removing the final source explicitly writes spaces = [] in a schema v2 or v3 configuration. This intentional empty allowlist permits an empty publication with tombstones for prior documents. An absent spaces key remains incomplete configuration and collection is refused. Connection and credential validation and publication safeguards still apply. Both updated components are required; this capability is absent from the installed 2026.0906.01 release.

Installation

Windows, no Python (recommended)

The simplest path: one installer for Cortex, Cortex Companion, the windowless Confluence converter, and the offline models. No separate Python or .NET runtime is required.

  1. Download Cortex-Setup.exe and SHA256SUMS from the latest release.
  2. Before running the unsigned installer, calculate its digest with Get-FileHash .\Cortex-Setup.exe -Algorithm SHA256 in PowerShell and verify that it exactly matches the Cortex-Setup.exe line in SHA256SUMS.
  3. Double-click only after that check. If SmartScreen still warns, select More info, then Run anyway.
  4. Choose the folder that holds your documents, keep Index everything in this folder, and finish. Cortex Companion opens when installation completes.
  5. In Companion, open Réglages (Settings) and verify the knowledge-base folder. The Cortex executable installed with Companion is detected automatically.
  6. Drop your documents in that folder, open Base locale (Local knowledge base), then select Synchroniser les documents locaux (Synchronize local documents).
  7. Restart your AI application: Cortex shows up there as an MCP server.

Companion then lets you synchronize, schedule, diagnose, and configure Cortex without a terminal. Details, silent mode and reinstall: Windows install.

Standalone archives (Windows x64, macOS Apple Silicon, Linux x64)

Every release also ships one ZIP archive per platform. It contains the single cortex or cortex.exe binary (MCP server + CLI, no Python) and the licenses for every embedded dependency. See Standalone distribution.

From PyPI (Python, advanced)

py -m pip install --upgrade cortex-local-rag
cortex setup

This path installs the CLI and MCP server, but not Cortex Companion. The model is downloaded on first use if its cache is empty.

From source (Python, advanced)

:: From the folder where you cloned Cortex
install.bat

install.bat initializes the configuration, installs the dependencies, offers to register Cortex in the detected MCP clients, and validates the installation. Details: Setup.

How it works

Documents folder (.md, .pdf)       Optional Confluence writer (REST)
      |                                      |
      |                              current Markdown generation
      +------------------+-------------------+
                         |
                         v
  cortex sync           <- Split, hash, vectorize, update FTS5
                         |
                         v
  %LOCALAPPDATA%\Cortex\  <- ChromaDB + lexical.db
      |
      v
  cortex serve          <- MCP server (FastMCP)
      |
      v
  MCP clients           <- Claude / Codex / Gemini / Antigravity / LM Studio / Cursor / Windsurf / VS Code

The embedding model is the multilingual ONNX paraphrase-multilingual-MiniLM-L12-v2. The Windows installer bundles it; a source installation or standalone binary downloads it when the local cache is empty.

Two indexing modes

  • Whole folder (default): anything you place in the chosen folder, at the root or in any subfolder, becomes searchable. Nothing to configure.
  • Sections (advanced): limits indexing to named subfolders you can search separately (defaults knowledge, projects, notes).

These modes govern the user-selected document folder. Generated ingestion documents are indexed separately from the current published generation with source_kind=doc and section sources.

Details: Configuration.

The cortex command

The installed package exposes a single command:

Subcommand Purpose
cortex setup Config + index + client registration in one go (--kb-path, --yes, --no-index, --reset).
cortex serve Runs the MCP server (used by clients).
cortex sync Incremental index synchronization.
cortex search Searches the index from the console (debugging aid).
cortex ingestion Shows source health and whether catch-up is due.
cortex confluence Stores the PAT interactively or runs the allowlisted writer.
cortex config Reads or changes configuration through an atomic JSON contract, notably for Companion.
cortex bundle Describes or verifies an encrypted portable archive.
cortex doctor Installation diagnostics (read-only).
cortex register / cortex unregister Adds or removes Cortex from MCP clients.
cortex init Creates the single per-user configuration.
cortex check Verifies the installation.

cortex --help describes every subcommand and cortex <command> --help describes its options. A prompt-free installation scripts as:

cortex setup --yes --kb-path "D:\Documents\Knowledge"

Exposed MCP tools

Tool Description
cortex_search Hybrid search. Parameters: query, section, top_k (1-10), source/author filters, and occurred/updated date ranges.
cortex_sync Triggers an incremental sync of the selected folder and, on a full sync, the current published document generation.
cortex_list_sections Lists included sections and "out of policy" folders.
cortex_freshness Read-only vault and ingestion freshness summary. Parameters: section (optional), include_entries (false by default).

Documentation

Prerequisites

Path Requirements
Windows installer No separate Python or .NET runtime. At least ~500 MB of space (applications, model + index).
Standalone archive No Python. ~500 MB of space (model + index).
From source Python 3.10+. ~500 MB of space.
Client Claude Desktop/Code, Codex, Gemini, Antigravity, LM Studio, Cursor, Windsurf or VS Code (MCP support).

License

Apache 2.0. See LICENSE.

Retrieval validation

See the validation guide for the desktop JSON search contract, the isolated FR/EN relevance corpus, performance measurements and exact paired-commit checks.

Release files for cortex-local-rag 2026.906.2

For a detailed explanation of source distributions (sdists) and built distributions (wheels), please see the package formats documentation.

Built distribution (wheel)

Table of built distributions (wheels) for cortex-local-rag 2026.906.2
File Interpreter ABI Platform
cortex_local_rag-2026.906.2-py3-none-any.whl Python 3 none any Details

Release files / cortex_local_rag-2026.906.2-py3-none-any.whl

Download URL cortex_local_rag-2026.906.2-py3-none-any.whl
Size 193.5 kB
Tags Python 3
SHA-256 checksum
How to use checksums
743e08fe5f82b59bb447d10e2f686d83b8758806a659c2a2fbc2bb8969608c3f
BLAKE2b-256 checksum
How to use checksums
8791b43d920c479116727dada63ca6c91679ec6760ff4affd82b21e1354c1b60
Upload date
Uploaded using Trusted Publishing?
What is trusted publishing?
Yes
Uploaded via twine/6.1.0 CPython/3.13.13

Provenance

Provenance describes where a file came from. On PyPI, provenance is shared via attestations, which provide a verifiable record of the build or publishing details. View details, limitations and caveats.

PyPI Publish Attestation

PyPI verified that this artifact, at this checksum, originated from the publisher listed below.

Signed by GitHub Actions, verified by PyPI on Sep 6, 2026.

Transparency log
Anthropic, PBC Visionary sponsor Bloomberg Visionary sponsor Hudson River Trading Visionary sponsor Meta Visionary sponsor NVIDIA Visionary sponsor Microsoft Sustainability sponsor Depot Continuous Integration AWS Cloud computing and Security Sponsor Datadog Monitoring Fastly CDN Google Download Analytics Sentry Error logging StatusPage Status page