holarch
One console for your AI work. It uses the cheapest path that can do the job, and it keeps working when the internet or your cloud credits don't.
Install
pipx install holarch # or: pip install holarch ; add [intent] for the free intent router
holarch doctor # what is set up
holarch index ~/notes # your notes become memory (public/ personal/ private/ folders = rings)
holarch "what did I decide about the launch?"
claude mcp add holarch -- holarch mcp # use it from Claude Code (holarch ide prints Cursor / VS Code config)
Local model: install Ollama and ollama pull qwen3:1.7b (and nomic-embed-text for better note search).
Cloud: set any of DEEPSEEK_API_KEY, OPENAI_API_KEY, OPENROUTER_API_KEY, GEMINI_API_KEY, or a holarch AI Credits key.
Keys stay in your environment; holarch never stores or prints them.
How it answers
- Your tools first. Ask for something by name and holarch runs it directly. No model call.
- Your notes next. It searches your own files and notes before it asks any model.
- A small model on your machine. It writes the answer from what it found. No internet needed.
- A long-context brain for planning. Big "what's the plan / summarize everything" questions go to your NotebookLM notebooks. The answer comes back with citations and is saved into your notes, so next time it's local.
- Cloud models last. Only when they're up and worth it. holarch notices a dead or out-of-credit provider and stops calling it.
Use it from anywhere
- Terminal:
holarch, plain language or commands. - MCP server: for Claude Code, Cursor, VS Code and ChatGPT. Ships with notes, receipts and health tools; add more as plugins (the
holarch.toolsentry point). - IDE: through the MCP config, plus the terminal.
Measured (not estimated)
| Offline, it picked the right tool | 4 of 4, and matched cloud answer quality (0.875 vs 0.875) |
| A dead provider costs | under 1 ms instead of 0.7–2.7 s per wasted call |
| Free intent routing | 81% accurate at $0, ~1 ms (a paid model: 88%) |
| Planning answers from your notebooks | median 40 s, cited, 20 of 20 answered |
Plans
- holarch Core: free. Your keys, your machine, the offline stack, the MCP server.
- holarch Pro: the notebook brain set up for you, sync, priority support.
- holarch AI Credits: prepaid packs to use cloud models through holarch.
- Holon setup: we set holarch up for you or your team on your own notes and workflows. Starts with an onboarding call.
Every plan starts the same way: pick it, pay (or not, for Core), and book your onboarding.
Metadata
Release files for holarch 0.2.0
For a detailed explanation of source distributions (sdists) and built distributions (wheels), please see the package formats documentation.
Source distribution (sdist)
| File | Size | Uploaded | |
|---|---|---|---|
| holarch-0.2.0.tar.gz | 45.8 kB | Details |
Built distribution (wheel)
| File | Interpreter | ABI | Platform | Reset |
|---|---|---|---|---|
| holarch-0.2.0-py3-none-any.whl | Python 3 | none | any | Details |
Total release size: 92.6 kB
Release files / holarch-0.2.0.tar.gz
| Download URL | holarch-0.2.0.tar.gz |
|---|---|
| Size | 45.8 kB |
| Tags | Source |
|
SHA-256 checksum How to use checksums |
fbd075510663d46b1d3a4b5afd69085e6ab5cdb98d2f00d4259088668f8d35f1
|
|
BLAKE2b-256 checksum How to use checksums |
adc38fb8075b5ef7d81e48cfd7ddb9ac104b8c301f3e537b49b46fbd29d001c8
|
| Upload date | |
|
Uploaded using Trusted Publishing? What is trusted publishing? |
No |
| Uploaded via |
twine/7.0.0 CPython/3.14.7
|
Release files / holarch-0.2.0-py3-none-any.whl
| Download URL | holarch-0.2.0-py3-none-any.whl |
|---|---|
| Size | 46.8 kB |
| Tags | Python 3 |
|
SHA-256 checksum How to use checksums |
400289fdfb4d9bfc6fb7974c8660414eb8b1b0882fddd71734fe3fda66e47f7a
|
|
BLAKE2b-256 checksum How to use checksums |
452c7246631272ba126b1d6028fcf2e888930cc78f49744c28a4a20e539fc9ba
|
| Upload date | |
|
Uploaded using Trusted Publishing? What is trusted publishing? |
No |
| Uploaded via |
twine/7.0.0 CPython/3.14.7
|