Skip to main content

Hybrid BM25 + vector search for Obsidian vaults with frontmatter awareness

Project description

qkb — Query Knowledge Base

An on-device hybrid search engine for Obsidian vaults that understands YAML frontmatter metadata. Combines BM25 keyword search (SQLite FTS5) and vector semantic search (sqlite-vec) with metadata filtering, sibling-document surfacing, and two first-class interfaces: a CLI for humans and an MCP server for LLM agents.

Status: Phase 1 (ingest, search tiers 1–3, CLI, MCP stdio).

Quickstart

1. Install (isolated, like pipx / npm -g):

uv tool install qkb-search

2. Point qkb at your vault — create ~/.config/qkb/config.toml:

[vault]
path = "~/Documents/MyVault"   # your Obsidian vault (read-only to qkb)
name = "MyVault"               # used to build obsidian:// links

3. Opt notes in. Only notes whose frontmatter has a context and/or source property are indexed — and an opted-in note also needs an id and a parseable date (created or date):

---
id: f47ac10b-58cc-4372-a567-0e02b2c3d401
context: homelab
created: 2026-03-15
---

4. Check, index, search:

qkb status                       # verify config, vault, and model resolve
qkb ingest                       # build the index (downloads the ~310 MB model once)
qkb status                       # see documents/chunks/vectors counts
qkb query "certificate renewal"  # hybrid search
qkb mcp                          # stdio MCP server for Claude Code / Desktop

No separate service, no compile: embeddings run in-process via ONNX Runtime, whose prebuilt wheels install with the package. The default model is embeddinggemma-300M (multilingual — the same embedding model QMD uses), cached after the first download. Re-running qkb ingest is incremental: unchanged notes are skipped, so it's cheap to re-index after editing notes.

Claude Code MCP registration:

claude mcp add qkb -- qkb mcp

Documents

The Short Version

Notes opt in to indexing via frontmatter (context and/or source properties). An ingestion pipeline walks the vault, chunks markdown with structure-aware break-point scoring, embeds in-process (fastembed/ONNX by default; Ollama or GGUF optional), and stores everything in a single SQLite file. A search engine layers BM25 (document-level, weighted columns), vector similarity (chunk-level), and Reciprocal Rank Fusion on top — exposed as qkb search / vsearch / query, qkb get <UUID>, and qkb mcp.

Inspired by QMD's search architecture, adapted for structured knowledge systems with frontmatter metadata.

Installation

qkb is a command-line tool, so install it into an isolated environment — the same idea as pipx or npm i -g:

# Recommended (uv):
uv tool install qkb-search

# Run without installing:
uvx --from qkb-search qkb query "certificate renewal"

# Alternatives (pipx isolates like uv; plain pip uses the current env):
pipx install qkb-search
pip install qkb-search

That's the whole setup — no service, no compile. The default embedding provider runs in-process via fastembed / ONNX Runtime, whose prebuilt wheels ship with the package (the C/C++ work is done upfront by the wheel builders, the way QMD relies on node-llama-cpp's prebuilt native binaries). Requires Python ≥3.11.

The default model is embeddinggemma-300M — the same embedding model QMD uses. GGUF (QMD) and ONNX (qkb) are just different packagings of the same weights for different runtimes; search quality comes from the model, not the file format. The ~310 MB quantized ONNX downloads once on first qkb ingest and is cached.

Embedding providers

Three interchangeable providers, set via [embedding].provider:

provider how it runs when to use
local (default) in-process fastembed / ONNX (prebuilt wheels) just works — no service, no compile
ollama the Ollama HTTP API you already run Ollama (e.g. a Linux box)
gguf in-process llama-cpp-python (the [gguf] extra) you want a specific GGUF; compiles on install

Switching provider or model changes the vectors, so run qkb ingest --full afterward to re-embed. Any model in fastembed's catalog also works — e.g. a smaller/faster one:

# ~/.config/qkb/config.toml
[embedding]
provider = "local"
model = "sentence-transformers/paraphrase-multilingual-MiniLM-L12-v2"  # 384-dim, ~220 MB
dimension = 384

License

MIT

Project details


Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

qkb_search-0.2.3.tar.gz (129.7 kB view details)

Uploaded Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

qkb_search-0.2.3-py3-none-any.whl (48.6 kB view details)

Uploaded Python 3

File details

Details for the file qkb_search-0.2.3.tar.gz.

File metadata

  • Download URL: qkb_search-0.2.3.tar.gz
  • Upload date:
  • Size: 129.7 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? Yes
  • Uploaded via: twine/6.1.0 CPython/3.13.12

File hashes

Hashes for qkb_search-0.2.3.tar.gz
Algorithm Hash digest
SHA256 60fe4811e12b2e30ba783443a6cf005e8d098a6d0e3be09927be7dc8a7a90e3b
MD5 1b6c43e1a616f943ff169c7be9fe474e
BLAKE2b-256 ba5e3c0e76e16880b7742f0f91d60113540b1393a2dd7d0b388da982ce47724c

See more details on using hashes here.

Provenance

The following attestation bundles were made for qkb_search-0.2.3.tar.gz:

Publisher: release.yml on miguelarios/qkb

Attestations: Values shown here reflect the state when the release was signed and may no longer be current.

File details

Details for the file qkb_search-0.2.3-py3-none-any.whl.

File metadata

  • Download URL: qkb_search-0.2.3-py3-none-any.whl
  • Upload date:
  • Size: 48.6 kB
  • Tags: Python 3
  • Uploaded using Trusted Publishing? Yes
  • Uploaded via: twine/6.1.0 CPython/3.13.12

File hashes

Hashes for qkb_search-0.2.3-py3-none-any.whl
Algorithm Hash digest
SHA256 68b2634be74443f4251c680e16f752b67c20f95dc5cdb209a11dba556dbacd78
MD5 104497359dd42fd126fe2c3c1e8137e6
BLAKE2b-256 1e2931820c368e56a6e5b57f9789163eadc11856fce387b490af9910c64a8187

See more details on using hashes here.

Provenance

The following attestation bundles were made for qkb_search-0.2.3-py3-none-any.whl:

Publisher: release.yml on miguelarios/qkb

Attestations: Values shown here reflect the state when the release was signed and may no longer be current.

Supported by

AWS Cloud computing and Security Sponsor Datadog Monitoring Depot Continuous Integration Fastly CDN Google Download Analytics Pingdom Monitoring Sentry Error logging StatusPage Status page