Cortex - RAG MCP for a knowledge base
Francais | English
Cortex is an MCP (Model Context Protocol) server that exposes semantic search over a local knowledge base. It lets Claude, Codex and Gemini find the right passage in your documents without wasting their context window. Search is semantic (by meaning, not keyword), in French and in English. Cortex processes and indexes the knowledge base locally without sending its content; the MCP client may still pass requested chunks to its model under its own policy. The optional Confluence writer only downloads explicitly allowlisted spaces; the generated Markdown, vector index, and lexical index remain local.
Installation
Windows, no Python (recommended)
The simplest path: one installer for Cortex, Cortex Companion, and the offline models. No separate Python or .NET runtime is required.
- Download
Cortex-Setup.exeandSHA256SUMSfrom the latest release. - Before running the unsigned installer, calculate its digest with
Get-FileHash .\Cortex-Setup.exe -Algorithm SHA256in PowerShell and verify that it exactly matches theCortex-Setup.exeline inSHA256SUMS. - Double-click only after that check. If SmartScreen still warns, select
More info, thenRun anyway. - Choose the folder that holds your documents, keep
Index everything in this folder, and finish. Cortex Companion opens when installation completes. - In Companion, open
Réglages(Settings) and verify the knowledge-base folder. The Cortex executable installed with Companion is detected automatically. - Drop your documents in that folder, open
Base locale(Local knowledge base), then selectSynchroniser les documents locaux(Synchronize local documents). - Restart your AI application: Cortex shows up there as an MCP server.
Companion then lets you synchronize, schedule, diagnose, and configure Cortex without a terminal. Details, silent mode and reinstall: Windows install.
Standalone archives (Windows x64, macOS Apple Silicon, Linux x64)
Every release also ships one ZIP archive per platform. It contains the single
cortex or cortex.exe binary (MCP server + CLI, no Python) and the licenses
for every embedded dependency. See
Standalone distribution.
From PyPI (Python, advanced)
py -m pip install --upgrade cortex-local-rag
cortex setup
This path installs the CLI and MCP server, but not Cortex Companion. The model is downloaded on first use if its cache is empty.
From source (Python, advanced)
:: From the folder where you cloned Cortex
install.bat
install.bat initializes the configuration, installs the dependencies, offers
to register Cortex in the detected MCP clients, and validates the installation.
Details: Setup.
How it works
Documents folder (.md, .pdf) Optional Confluence writer (REST)
| |
| current Markdown generation
+------------------+-------------------+
|
v
cortex sync <- Split, hash, vectorize, update FTS5
|
v
%LOCALAPPDATA%\Cortex\ <- ChromaDB + lexical.db
|
v
cortex serve <- MCP server (FastMCP)
|
v
MCP clients <- Claude / Codex / Gemini / Antigravity / LM Studio / Cursor / Windsurf / VS Code
The embedding model is the multilingual ONNX
paraphrase-multilingual-MiniLM-L12-v2. The Windows installer bundles it; a
source installation or standalone binary downloads it when the local cache is
empty.
Two indexing modes
- Whole folder (default): anything you place in the chosen folder, at the root or in any subfolder, becomes searchable. Nothing to configure.
- Sections (advanced): limits indexing to named subfolders you can search
separately (defaults
knowledge,projects,notes).
These modes govern the user-selected document folder. Generated ingestion
documents are indexed separately from the current published generation with
source_kind=doc and section sources.
Details: Configuration.
The cortex command
The installed package exposes a single command:
| Subcommand | Purpose |
|---|---|
cortex setup |
Config + index + client registration in one go (--yes, --no-index, --reset). |
cortex serve |
Runs the MCP server (used by clients). |
cortex sync |
Incremental index synchronization. |
cortex ingestion |
Shows source health and whether catch-up is due. |
cortex confluence |
Stores the PAT interactively or runs the allowlisted writer. |
cortex config |
Reads or changes configuration through an atomic JSON contract, notably for Companion. |
cortex bundle |
Describes or verifies an encrypted portable archive. |
cortex doctor |
Installation diagnostics (read-only). |
cortex register / cortex unregister |
Adds or removes Cortex from MCP clients. |
cortex check |
Verifies the installation. |
Exposed MCP tools
| Tool | Description |
|---|---|
cortex_search |
Hybrid search. Parameters: query, section, top_k (1-10), source/author filters, and occurred/updated date ranges. |
cortex_sync |
Triggers an incremental sync of the selected folder and, on a full sync, the current published document generation. |
cortex_list_sections |
Lists included sections and "out of policy" folders. |
cortex_freshness |
Read-only vault and ingestion freshness summary. Parameters: section (optional), include_entries (false by default). |
Documentation
- Table of contents
- Windows install: unified Cortex + Companion + models installer, corpus choice, silent mode, reinstall.
- Standalone distribution: per-platform archives and reproducible builds.
- Install from source: prerequisites, MCP clients.
- User guide: sync, search, tools, doctor, logs.
- FAQ: installation, local data, sync, and diagnostics.
- Release notes: user-visible changes by version and the published-history notice.
- Technical changelog: complete changes by version.
- Configuration:
config.toml, indexing modes, sections, data home, migration. - Ingestion scheduling: source health, catch-up, retries, and Task Scheduler.
- Metadata v2 migration: structured search metadata, backup, migration, and restore.
- Confluence writer: allowlisted REST ingestion, Windows Credential Manager, conversion, and atomic generations.
- Reproducible install:
requirements.lock,--require-hashes, regenerating the lock. - Public specification: MCP surface, index contracts, data, distribution, and limits.
- Architecture: end-to-end and technical choices.
- Security: local runtime, telemetry off, single-writer.
Prerequisites
| Path | Requirements |
|---|---|
| Windows installer | No separate Python or .NET runtime. At least ~500 MB of space (applications, model + index). |
| Standalone archive | No Python. ~500 MB of space (model + index). |
| From source | Python 3.10+. ~500 MB of space. |
| Client | Claude Desktop/Code, Codex, Gemini, Antigravity, LM Studio, Cursor, Windsurf or VS Code (MCP support). |
License
Apache 2.0. See LICENSE.
Download files
Download the file for your platform. If you're not sure which to choose, learn more about installing packages.
Source Distributions
Built Distribution
Filter files by name, interpreter, ABI, and platform.
If you're not sure about the file name format, learn more about wheel file names.
Copy a direct link to the current filters
File details
Details for the file cortex_local_rag-2026.808.0-py3-none-any.whl.
File metadata
- Download URL: cortex_local_rag-2026.808.0-py3-none-any.whl
- Upload date:
- Size: 177.6 kB
- Tags: Python 3
- Uploaded using Trusted Publishing? Yes
- Uploaded via: twine/6.1.0 CPython/3.13.13
File hashes
| Algorithm | Hash digest | |
|---|---|---|
| SHA256 |
25c1d9685cef9fa7acae2b89d1fc6c63157d3d2474081d27595c967190b82fb5
|
|
| MD5 |
72def9069f41686ea913076d3c68b748
|
|
| BLAKE2b-256 |
84e900a587be0875be9d5f1ffe93f11edd4cd0ff7989aa3c99920d95af107ed6
|
Provenance
The following attestation bundles were made for cortex_local_rag-2026.808.0-py3-none-any.whl:
Publisher:
release.yml on VBlackJack/Cortex
-
Statement:
-
Statement type:
https://in-toto.io/Statement/v1 -
Predicate type:
https://docs.pypi.org/attestations/publish/v1 -
Subject name:
cortex_local_rag-2026.808.0-py3-none-any.whl -
Subject digest:
25c1d9685cef9fa7acae2b89d1fc6c63157d3d2474081d27595c967190b82fb5 - Sigstore transparency entry: 2386519371
- Sigstore integration time:
-
Permalink:
VBlackJack/Cortex@3710260ffd628b51214735afcf7b9559434b5ac2 -
Branch / Tag:
refs/tags/v2026.0808.00 - Owner: https://github.com/VBlackJack
-
Access:
public
-
Token Issuer:
https://token.actions.githubusercontent.com -
Runner Environment:
github-hosted -
Publication workflow:
release.yml@3710260ffd628b51214735afcf7b9559434b5ac2 -
Trigger Event:
push
-
Statement type: