CPersona
MCP Memory Server
Give Claude persistent memory across sessions. Single SQLite file. 30 tools. Zero LLM dependency.
Documentation · Getting Started · Architecture · Tools · PyPI · Zenn Book (JP)
Standalone repository — This is the standalone version for use with Claude Desktop, Claude Code, and any MCP client. If you are a ClotoCore user, install CPersona from the in-app marketplace (ClotoHub) instead — it distributes this same repository.
Project status — 2.4.x is Stable; 2.5.x is Current, an internal stabilization line where all fixes land, pending production-soak certification. The DB schema is preserved across the line. Additive, rollback-safe features may land here as well (lifecycle standard §2.6); a change that cannot be rolled back waits for 2.6. Which version to run, and how long each line keeps receiving fixes: SUPPORT.md.
Upgrading from 2.5.2 or earlier? Two things need a decision from you. v2.5.3 will not start the HTTP transport without
CPERSONA_AUTH_TOKEN, wherever it binds — set one, or opt out withCPERSONA_ALLOW_UNAUTHENTICATED_HTTP=true(why; stdio is unaffected). v2.5.2 changed tool response shapes — branch onok is false, and treat any response carryingerroras a failure whether or notokis present (contract §10).
The Problem
Claude forgets everything between sessions. Every conversation starts from zero — no context about your project, your preferences, or what you discussed yesterday.
cpersona fixes this. It's an MCP server that stores memories in a local SQLite file and retrieves them through hybrid search. Claude remembers you. It runs against any MCP-compatible host — Claude Desktop, Claude Code, ClotoCore (the AI agent platform where cpersona originated, and whose memory layer it is), or a client of your own.
Quick Start
Claude Code? Let the agent do the setup. The wheel ships an Agent Skill that installs everything and teaches Claude when to store, recall and archive. Copy it in, then say "Set up CPersona."
python -c "import cpersona,pathlib,shutil; s=pathlib.Path(cpersona.__file__).parent/'skills'/'cpersona-memory'; shutil.copytree(s, pathlib.Path.home()/'.claude/skills/cpersona-memory', dirs_exist_ok=True)"
1. Install — Python 3.11+, and uv for the one-command path.
uvx cpersona # run directly, no install step
pip install cpersona # or install it
2. Run an embedding server (recommended — it powers the vector layer)
uvx --from "cembedding[onnx]" cembedding-download-model --model jina-v5-nano
EMBEDDING_PROVIDER=onnx_jina_v5_nano uvx --from "cembedding[onnx]" cembedding # serves http://127.0.0.1:8401/embed
Any endpoint implementing the embedding contract works. Without one, cpersona runs on FTS5 + keyword search and tells you it is degraded.
3. Register it with your MCP client
claude mcp add-json cpersona '{"type":"stdio","command":"uvx","args":["cpersona"],"env":{"CPERSONA_DB_PATH":"/home/you/.claude/cpersona.db","EMBEDDING_MODE":"http","EMBEDDING_HTTP_URL":"http://127.0.0.1:8401/embed"}}' -s user
That's it. Ask Claude to store something and recall it in a later session.
Claude Desktop config, Windows paths, installing from source and the full walkthrough: Getting Started.
What You Get
- Hybrid search — vector, FTS5 (trigram, so it works on Japanese and other space-less scripts) and keyword, fused by rank or relative score. The FTS and keyword layers rescue what vectors miss: identifiers, error strings, exact names.
- Three memory types — facts, session summaries and an accumulated profile.
- Zero LLM dependency — cpersona never calls a generative model; your agent summarizes and hands over the result. Recall is deterministic given a calibrated gate, but the gate is sampled, so two installs on identical data can settle differently.
- Single-file SQLite — no external database;
sqlite3 .backupcopies the corpus (the calibration sidecar beside it needs copying too). - Operable — auto-calibrated thresholds, a health check with auto-repair, an advisory when the embedding layer dies, JSONL export/import, agent-to-agent merge.
- Isolation —
agent_id,project_idandchannellet several agents and projects share one database without bleeding into each other.
How it fits together: Architecture · what the tools do: Tools · what you may rely on: Behavior Contracts.
Benchmarks
Measured on LMEB (Long-horizon Memory Embedding Benchmark, arXiv:2603.12572) — 22 datasets subsuming LoCoMo and LongMemEval, measured here as 22 retrieval tasks. The metric is Mean NDCG@10 across all 22 tasks. Track A is the raw embedding model alone; Track B routes the same embeddings through cpersona's real store/recall code paths (SQLite + FTS5 + RRF fusion + per-agent auto-calibration).
| Embedding Model | Params | Dim | Track A (raw) | Track B (cpersona) | Δ |
|---|---|---|---|---|---|
| all-MiniLM-L6-v2 | 22M | 384 | 43.67 | 50.10 | +6.43 |
| bge-m3 | 568M | 1024 | 56.83 | 57.66 | +0.83 |
Track B lands at or above Track A on both models: the fusion layers add signal rather than merely persisting vectors, and a weaker embedding gains more because the FTS5/keyword layers rescue what its vectors miss. How to read the deltas, the noise envelope, the measurement harness and the reproduction regime: benchmarks/.
Documentation
cloto-dev.github.io/CPersona is canonical — when this README disagrees with it, the site wins.
| Getting Started | Install, embedding server, client registration, verification |
| Behavior Contracts | What you may rely on: recall ordering, dedup, scan window, response shapes |
| Tools | All 30 tools, grouped by what you reach for them for |
| Architecture | Storage, the retrieval pipeline, isolation axes |
| Operations Runbook | Backup, degradation detection, tuning, CJK guidance, corpus sync |
| Configuration | Every environment variable and its default |
| Quality Assurance | How a release is gated: audits, the bug ledger, structural and mutation gates |
| FAQ | Short answers to the questions operators actually ask |
Japanese translations are in the language selector (English is canonical) and
agents can read llms.txt.
Longer reads in Japanese: a book
on the design and setup, and an article
on the token economics of session-end → /clear → recall.
Quality Assurance
Every release is gated by a machine-verifiable process: multi-agent audit rounds with adversarial verification, a bug ledger that fails CI if a fix marker disappears or a removed defect returns, structural gates for invariants a plain test cannot express, a mutation proof that those gates go red when the invariant is broken, and gates holding the documented counts, defaults and version claims to the source that defines them.
Behind it: ~1,130 test functions across ~95 test modules (~1,440 cases parametrised, more test code than server code), on Schema v13 — how a release is gated.
Support
Three tiers — Stable (production-certified, critical fixes only), Current (newest line, all fixes land here) and Experimental (opt-in pre-releases). A superseded line keeps critical-fix support for 30 more days. Read SUPPORT.md § Known issues before pinning a version — some of them change what you should run.
Found a bug, or something the docs do not explain? Open a bug report or feature request, even when you are not certain — a configuration problem mistaken for a bug means the documentation was unclear, which is a defect of its own. Report security vulnerabilities privately via SECURITY.md.
License
MIT — free to use from any MCP host without restriction.
Release files for cpersona 2.5.7
For a detailed explanation of source distributions (sdists) and built distributions (wheels), please see the package formats documentation.
Source distribution (sdist)
| File | Size | Uploaded | |
|---|---|---|---|
| cpersona-2.5.7.tar.gz | 286.5 kB | Details |
Built distribution (wheel)
| File | Interpreter | ABI | Platform | Reset |
|---|---|---|---|---|
| cpersona-2.5.7-py3-none-any.whl | Python 3 | none | any | Details |
Total release size: 564.6 kB
Release files / cpersona-2.5.7.tar.gz
| Download URL | cpersona-2.5.7.tar.gz |
|---|---|
| Size | 286.5 kB |
| Tags | Source |
|
SHA-256 checksum How to use checksums |
a20d3294cf2918dbbd9b27466c83266815384d245a3018cd60718fa986181a33
|
|
BLAKE2b-256 checksum How to use checksums |
02ab2db3e825968c7a9b5b074928eca5e4079b630c79cce068a3271f02357e08
|
| Upload date | |
|
Uploaded using Trusted Publishing? What is trusted publishing? |
Yes |
| Uploaded via |
twine/7.0.0 CPython/3.13.14
|
Provenance
Provenance describes where a file came from. On PyPI, provenance is shared via attestations, which provide a verifiable record of the build or publishing details. View details, limitations and caveats.
PyPI Publish Attestation
PyPI verified that this artifact, at this checksum, originated from the publisher listed below.
Signed by GitHub Actions, verified by PyPI on Aug 31, 2026.
Transparency logRelease files / cpersona-2.5.7-py3-none-any.whl
| Download URL | cpersona-2.5.7-py3-none-any.whl |
|---|---|
| Size | 278.1 kB |
| Tags | Python 3 |
|
SHA-256 checksum How to use checksums |
b96463d0c93f8d9045829047137bd967d4464b8272fcbdea383f3f64117934db
|
|
BLAKE2b-256 checksum How to use checksums |
7df502c0243652240feb2ce655ad682fc95c15190a6d7b414ed1df042ca34aa9
|
| Upload date | |
|
Uploaded using Trusted Publishing? What is trusted publishing? |
Yes |
| Uploaded via |
twine/7.0.0 CPython/3.13.14
|
Provenance
Provenance describes where a file came from. On PyPI, provenance is shared via attestations, which provide a verifiable record of the build or publishing details. View details, limitations and caveats.
PyPI Publish Attestation
PyPI verified that this artifact, at this checksum, originated from the publisher listed below.
Signed by GitHub Actions, verified by PyPI on Aug 31, 2026.
Transparency log