WanYi Memory Core 万忆中枢
永不遗忘的全量记忆系统 — Event-sourced long-term memory for AI agents: process memory, mistake books, experience crystallization, confidence-based decision blocking, counterfactual branches, cross-domain analogy, trajectory replay, proactive partner, semantic vector retrieval, reranker, memory graph, time decay and metacognitive knowledge-gaps. Ships as a local-first MCP server with 23 tools. Your data never leaves your machine.
🎬 Live Demo: 交互式演示页 · 源码
docs/demo.html
Why this is different
Most memory systems store your data in the cloud, need a heavy dependency stack, or only do keyword search. wanyimem is local-first, single-file SQLite, and installs with one dependency (numpy).
| wanyimem | typical memory server | |
|---|---|---|
| Data | never leaves your machine (zero telemetry) | cloud / SaaS |
| Infra | SQLite single file, no separate vector DB | Qdrant / Neo4j / Postgres |
| Defaults | shadows the agent, blocks high-risk actions, opens counterfactual branches | stores & retrieves |
| Resources | runs on 2-core / 2GB | heavier |
| Evolves | zero-participation (learns from your mistakes automatically) | manual "remember this" |
23 MCP tools, open source (MIT), Python 3.10+, pip install wanyimem.
Why
LLM agents forget. Every chat window is amnesia: preferences, lessons, and hard-won experience evaporate when the session ends.
WanYi Memory Core is a local-first, full-quantity, self-evolving memory system:
- Event sourcing — an append-only WAL is the single source of truth. Nothing is ever deleted; decay only affects retrieval ranking.
- Semantic recall — hybrid retrieval: BM25 keywords + local Chinese embedding (BAAI/bge-small-zh-v1.5) + reranker (BAAI/bge-reranker-base) + knowledge-graph expansion + explicit time decay. Vector search is hybrid itself: exact cosine below
ANN_MIN_COUNT, and asqlite-vecANN pre-filter + exact re-rank above it — so it stays fast at scale without sacrificing recall quality. - Metacognition — when recall is weak, the system admits it and records a knowledge-gap instead of hallucinating an answer.
- Decision guardrails — high-risk actions (all-in, revenge-trading, force-push, rm -rf) trigger confidence-based blocking with counterfactual branches: you see what would have happened if you had listened.
- Zero-participation evolution — no need to say "remember this"; the system decides what to store, consolidates overnight, and surfaces weekly trajectory reviews.
Install
pip install wanyimem # core
pip install "wanyimem[all]" # + vector/reranker models + sqlite-vec ANN (use [ann] for ANN only)
Requires Python 3.10+. Models (embedding ~95MB, reranker ~1.1GB) are downloaded on first use from HuggingFace; set HF_ENDPOINT=https://hf-mirror.com if you are in mainland China.
Before the PyPI release lands, you can also install directly from GitHub (identical code):
pip install "git+https://github.com/17861102832/wanyimem.git"
CLI & Automation
Beyond the MCP server, wanyimem ships two console commands after pip install wanyimem:
wanyi-export --db memory.db --out memory.md # readable, diff-able Markdown mirror of all memory (read-only)
wanyi-auto --db memory.db # one AutoMoat pass: honest counterfactual auto-settlement + consolidation + analog patrol
wanyi-auto --db memory.db --loop 3600 # periodic scheduler (daemon background; off by default)
wanyi-exportrenders the event-sourced store into a human-readable, version-controllable Markdown mirror (grouped by 道/法/术, plus mistakes / experiences / knowledge-gaps / counterfactual branches / cross-domain patterns).wanyi-autoautomates the guardrails: it honestly settles overdue counterfactual branches (marking themexpiredrather than fabricating a winner), runs sleep + deep consolidation, and surfaces the cross-domain analog patterns most worth recalling.
Quick Start (MCP)
Add to your mcp.json (Claude Desktop, Cursor, Trae, etc.):
{
"mcpServers": {
"wanyi": {
"command": "python",
"args": ["-m", "wanyi.memory_core"],
"env": {
"WANYI_STORE_DIR": "C:/path/to/your/memory"
}
}
}
}
Env keys are "Chinese-first, ASCII-fallback": the new
WANYI_STORE_DIR(recommended, more portable) and the legacy万忆中枢_STORE_DIRboth work. Barepythondepends on PATH and may fail; prefer an absolute interpreter path, orpip install wanyimemthen use"command": "wanyi".
Then any agent can call the 23 tools, e.g.:
万忆记录见闻 → "2026年5月基金大跌时我死扛不止损,亏了18%才割肉。"
万忆召回记忆 → query "认赔离场到底对不对" # semantic match even with zero shared keywords
万忆置信度决策检查 → "我要全仓梭哈" # BLOCK if confidence is low, with historical mistakes
Quick Start (Library)
from wanyi import WanYiCore
engine = WanYiCore()
engine.tool_record_memory(
content="止损纪律:亏损超过8%必须无条件卖出",
layer="法", mem_type="principle",
)
resp = engine.tool_recall_memory("认赔离场到底对不对", limit=5)
for m in resp["memories"]:
print(m["content"], m.get("_rerank_score"))
Features
| Area | Capability |
|---|---|
| Storage | SQLite + append-only event WAL; 道/法/术 three-layer half-lives |
| Retrieval | Keyword BM25 + vector (bge-small-zh) + reranker (bge-reranker-base) + graph expansion + time-decay fields |
| Metacognition | knowledge-gap auto-record, stats self-check, honest "I don't know" |
| Guardrails | confidence-based decision blocking, counterfactual branches with auto-settlement, cross-domain analogy bridging |
| Proactivity | daily brief on LOAD, due-branch reminders, weekly trajectory replay, risk-keyword alert |
| Growth | mistake book, experience crystallization, overnight consolidation, evolution queries |
| Privacy | fully local, zero telemetry, no cloud dependency |
Public benchmark (LongMemEval) — session-level retrieval, full results in benchmark/RESULTS.md. Core (BM25 + graph, no models) reaches Recall@5 = 0.960 / MRR = 0.907 on s_cleaned (with ~40 distractor sessions). Reproduce via python benchmark/longmemeval_run.py.
Benchmark — reproducible mini LongMemEval (14 keyword-mismatched cross-session fact queries, run via python benchmark/recall_benchmark.py):
| Version | Recall@5 | MRR |
|---|---|---|
| Core (keyword BM25 + knowledge-graph, no models) | 1.000 (14/14) | 0.857 |
| Full (bge-small-zh vector + bge-reranker-base rerank) | 1.000 (14/14) | 0.857 |
Every query is intentionally phrased with different keywords than its answer (e.g. 本地数据库怎么提高并发写 → WAL模式, 记忆系统最怕什么 → 事件溯源), so 14/14 reflects genuine semantic recall, not string matching. The knowledge-graph channel (active in the core, model-free) already lifts BM25 to parity here; the vector + reranker path shows its edge on larger-scale semantic expansion ("pip install wanyimem[all]" downloads the models).
Docs
Contributing
See CONTRIBUTING.md. Report vulnerabilities privately via SECURITY.md.
License
MIT © 2026 Zhao Xikun
Download files
Download the file for your platform. If you're not sure which to choose, learn more about installing packages.
Source Distribution
Built Distribution
Filter files by name, interpreter, ABI, and platform.
If you're not sure about the file name format, learn more about wheel file names.
Copy a direct link to the current filters
File details
Details for the file wanyimem-1.0.6.tar.gz.
File metadata
- Download URL: wanyimem-1.0.6.tar.gz
- Upload date:
- Size: 123.8 kB
- Tags: Source
- Uploaded using Trusted Publishing? Yes
- Uploaded via:
twine/7.0.0 CPython/3.13.14
File hashes
| Algorithm | Hash digest | |
|---|---|---|
| SHA256 |
cf1b3b396db5c88dd0d00f96ac5da410a2870903a75cb896674f9b49dd9d0af0
|
|
| MD5 |
c5bc430dde7e1898f3b3abe3f595c434
|
|
| BLAKE2b-256 |
571fef3109c18e51f53dbb86aa4276331ae7e5549c9459b1d405cb0a82c1ff5b
|
Provenance
The following attestation bundles were made for wanyimem-1.0.6.tar.gz:
Publisher:
publish.yml on 17861102832/wanyimem
-
Statement:
-
Statement type:
https://in-toto.io/Statement/v1 -
Predicate type:
https://docs.pypi.org/attestations/publish/v1 -
Subject name:
wanyimem-1.0.6.tar.gz -
Subject digest:
cf1b3b396db5c88dd0d00f96ac5da410a2870903a75cb896674f9b49dd9d0af0 - Sigstore transparency entry: 2578646803
- Sigstore integration time:
-
Permalink:
17861102832/wanyimem@5313c6b4ed66a6308bae6517d6e80a07e7c2b37a -
Branch / Tag:
refs/tags/v1.0.6 - Owner: https://github.com/17861102832
-
Access:
public
-
Token Issuer:
https://token.actions.githubusercontent.com -
Runner Environment:
github-hosted -
Publication workflow:
publish.yml@5313c6b4ed66a6308bae6517d6e80a07e7c2b37a -
Trigger Event:
push
-
Statement type:
File details
Details for the file wanyimem-1.0.6-py3-none-any.whl.
File metadata
- Download URL: wanyimem-1.0.6-py3-none-any.whl
- Upload date:
- Size: 98.7 kB
- Tags: Python 3
- Uploaded using Trusted Publishing? Yes
- Uploaded via:
twine/7.0.0 CPython/3.13.14
File hashes
| Algorithm | Hash digest | |
|---|---|---|
| SHA256 |
d6a5edc30b41ee1f19995a7f0bc4ba4b22c62d74aa599206cd6afc5f59ecb355
|
|
| MD5 |
d81d49babf904fdde2ae8b6cad234656
|
|
| BLAKE2b-256 |
b60a6887677ea1c3e0a8c883d0c8932b63dc55bdef9313e3d8c055c9b1ef3479
|
Provenance
The following attestation bundles were made for wanyimem-1.0.6-py3-none-any.whl:
Publisher:
publish.yml on 17861102832/wanyimem
-
Statement:
-
Statement type:
https://in-toto.io/Statement/v1 -
Predicate type:
https://docs.pypi.org/attestations/publish/v1 -
Subject name:
wanyimem-1.0.6-py3-none-any.whl -
Subject digest:
d6a5edc30b41ee1f19995a7f0bc4ba4b22c62d74aa599206cd6afc5f59ecb355 - Sigstore transparency entry: 2578647109
- Sigstore integration time:
-
Permalink:
17861102832/wanyimem@5313c6b4ed66a6308bae6517d6e80a07e7c2b37a -
Branch / Tag:
refs/tags/v1.0.6 - Owner: https://github.com/17861102832
-
Access:
public
-
Token Issuer:
https://token.actions.githubusercontent.com -
Runner Environment:
github-hosted -
Publication workflow:
publish.yml@5313c6b4ed66a6308bae6517d6e80a07e7c2b37a -
Trigger Event:
push
-
Statement type: