semble

Fast and Accurate Code Search for Agents

These details have been verified by PyPI

Project links

Owner

Minish

GitHub Statistics

These details have not been verified by PyPI

Project description

Fast and Accurate Code Search for Agents
_{Uses ~98% fewer tokens than grep+read}

Quickstart • CLI • MCP Server • Installation • Benchmarks

Semble is a code search library built for agents. It returns the exact code snippets they need instantly, using ~98% fewer tokens than grep+read. Indexing and searching a full codebase end-to-end takes under a second, with ~200x faster indexing and ~10x faster queries than a code-specialized transformer, at 99% of its retrieval quality (see benchmarks). Everything runs on CPU with no API keys, GPU, or external services. Run it as an MCP server or call it from the shell via AGENTS.md and any agent (Claude Code, Cursor, Codex, OpenCode, etc.) gets instant access to any repo.

Quickstart

Your agent queries Semble in natural language (e.g. "How is authentication handled?") and gets back only the relevant code snippets, without grepping or reading full files.

The fastest way to get started is the interactive installer. Install uv, then run:

uv tool install semble
semble install

semble install detects installed coding agents such as Claude Code, Codex, and OpenCode, and then lets you choose which integrations to enable:

MCP server: lets the agent call Semble directly as a tool.
Instructions: adds CLI usage guidance to AGENTS.md / CLAUDE.md.
Sub-agent: installs a dedicated semble-search sub-agent.

To undo the setup, run semble uninstall.

For manual setup instructions (MCP config per agent, AGENTS.md snippet, sub-agent files), see the installation docs.

Updating Semble

uv tool upgrade semble   # upgrade
uv cache clean semble    # for MCP users (restart your MCP client after)

Main Features

Fast: indexes an average repo in ~250 ms and answers queries in ~1.5 ms, all on CPU.
Accurate: NDCG@10 of 0.854 on our benchmarks, on par with code-specialized transformer models, at a fraction of the size and cost.
Token-efficient: returns only the relevant chunks, using ~98% fewer tokens than grep+read.
Zero setup: runs on CPU with no API keys, GPU, or external services required.
MCP server: works with Claude Code, Cursor, Codex, OpenCode, VS Code, and any other MCP-compatible agent.
Local and remote: pass a local path or a git URL.

CLI

Semble also ships as a standalone CLI. This is useful in scripts or anywhere you want search results without an MCP session. Indexes are built and cached on first run, and invalidated automatically when files change.

# Search a local repo (index is built and cached automatically)
semble search "authentication flow" ./my-project

# Search a remote repo (cloned on demand)
semble search "save model to disk" https://github.com/MinishLab/model2vec

# Limit results
semble search "save model to disk" ./my-project --top-k 10

# Search docs/config/everything instead of just code
semble search "deployment guide" ./my-project --content docs   # or: config, all

# Find code similar to a known location
semble find-related src/auth.py 42 ./my-project

--content accepts code (default), docs, config, or all. path defaults to the current directory when omitted; git URLs are accepted. If semble is not on $PATH, use uvx --from "semble[mcp]" semble in its place.

Controlling which files are indexed

Semble reads .gitignore and .sembleignore files to determine which files to index. Both files use standard gitignore syntax and their patterns are merged. .sembleignore lets you add semble-specific rules without touching .gitignore. Rules are applied recursively, so a .sembleignore in a subdirectory applies to that subtree.

Excluding files: add patterns the same way you would in .gitignore:

# .sembleignore
generated/     # exclude generated dir
*.pb.go.       # exclude Go protobuf files

Including non-default extensions: prefix the extension pattern with ! to force-include files that semble wouldn't index by default:

# .sembleignore
!*.proto       # include Protobuf files
!*.cob         # include COBOL files

Semble also always skips a set of well-known non-source directories regardless of ignore files (e.g. node_modules/, .venv/, dist/, build/, __pycache__/, and similar).

Savings

semble savings shows how many tokens semble has saved across all your searches:

semble savings           # summary by period
semble savings --verbose # also show breakdown by call type

  Semble Token Savings
  ════════════════════════════════════════════════════════════════
  Period        Calls   Savings
  ────────────────────────────────────────────────────────────────
  Today         42      [███████████████░]  ~58.4k tokens (95%)
  Last 7 days   287     [██████████████░░]  ~312.4k tokens (90%)
  All time      1.4k    [██████████████░░]  ~1.2M tokens (89%)

Savings are calculated as follows: for each call, semble records the total character count of the unique files containing returned chunks and the character count of the snippets returned. Estimated tokens saved is (file chars − snippet chars) / 4 (4 chars per token). This is a conservative estimate: the baseline is reading matched files in full, which is how coding agents often explore unfamiliar code.

By default, stats are stored in the OS cache folder (~/Library/Caches/semble/ on macOS, ~/.cache/semble/ on Linux, %LOCALAPPDATA%\semble\Cache\ on Windows). To override this location you can supply an environment variable SEMBLE_CACHE_LOCATION which should be the full path to the target cache location e.g. 'd:\caches\storemysemblecachehere'.

Library usage

Semble can also be used as a Python library for programmatic access, useful when building custom tooling or integrating search directly into your own code.

from semble import ContentType, SembleIndex

# Index a local directory (code only, the default)
index = SembleIndex.from_path("./my-project")

# Index docs and prose (markdown, rst, etc.)
index = SembleIndex.from_path("./my-project", content=ContentType.DOCS)

# Index everything (code, docs, and config)
index = SembleIndex.from_path("./my-project", content=[ContentType.CODE, ContentType.DOCS, ContentType.CONFIG])

# Index code and docs together
index = SembleIndex.from_path("./my-project", content=[ContentType.CODE, ContentType.DOCS])

# Index a remote git repository
index = SembleIndex.from_git("https://github.com/MinishLab/model2vec")

# Search the index with a natural-language or code query
results = index.search("save model to disk", top_k=3)

# Find code similar to a specific result
related = index.find_related(results[0], top_k=3)

# Each result exposes the matched chunk
result = results[0]
result.chunk.file_path   # "model2vec/model.py"
result.chunk.start_line  # 127
result.chunk.end_line    # 150
result.chunk.content     # "def save_pretrained(self, path: PathLike, ..."

MCP Server

Semble runs as an MCP server so agents can search any codebase directly as a native tool call. Repos are indexed on demand and cached; local paths are re-indexed automatically on file changes.

Tool	Description
`search`	Search a codebase with a natural-language or code query. Pass `repo` as a local path or an https:// git URL.
`find_related`	Given a file path and line number, return chunks semantically similar to the code at that location.

For per-agent setup instructions, see the installation docs.

Benchmarks

We benchmark quality and speed across ~1,250 queries over 63 repositories in 19 languages (left), and token efficiency against grep+read at equivalent recall levels (right).

Token efficiency: recall vs. retrieved tokens

The quality benchmark (left) scores retrieval quality (NDCG@10) against total latency; semble achieves 99% of the quality of the 137M-parameter CodeRankEmbed Hybrid while indexing 218x faster. The token efficiency benchmark (right) measures how many tokens each method needs to reach a given recall level; semble uses 98% fewer tokens on average and hits 94% recall at only 2k tokens, while grep+read needs a full 100k context window to reach 85%. See benchmarks for per-language results, ablations, and full methodology.

How it works

Semble splits each file into code-aware chunks using tree-sitter, then scores every query against the chunks with two complementary retrievers: static Model2Vec embeddings using the code-specialized potion-code-16M model for semantic similarity, and BM25 for lexical matches on identifiers and API names. The two score lists are fused with Reciprocal Rank Fusion (RRF).

After fusing, results are reranked with a set of code-aware signals:

Ranking signals

Adaptive weighting. Symbol-like queries (Foo::bar, _private, getUserById) get more lexical weight, while natural-language queries stay balanced between semantic and lexical retrievers.
Definition boosts. A chunk that defines the queried symbol (a class, def, func, etc.) is ranked above chunks that merely reference it.
Identifier stems. Query tokens are stemmed and matched against identifier stems in a chunk, giving an additional weight to chunks that contain them. For example, querying parse config boosts chunks containing parseConfig, ConfigParser, or config_parser.
File coherence. When multiple chunks from the same file match the query, the file is boosted so the top result reflects broad file-level relevance rather than a single out-of-context chunk.
Noise penalties. Test files, compat//legacy/ shims, example code, and .d.ts declaration stubs are down-ranked so canonical implementations surface first.

Because the embedding model is static with no transformer forward pass at query time, all of this runs in milliseconds on CPU.

Indexes are cached to disk automatically on the first search. On subsequent runs, Semble walks the file tree and compares modification times; if any file was added, removed, or changed, the index is fully rebuilt. In MCP mode, a file watcher detects changes and triggers a rebuild automatically so the index is always current within the same session.

Acknowledgements

Thanks to Greptile for providing free access to their AI code review platform.

License

MIT

Citing

If you use Semble in your research, please cite the following:

@software{minishlab2026semble,
  author       = {{van Dongen}, Thomas and Stephan Tulkens},
  title        = {Semble: Fast and Accurate Code Search for Agents},
  year         = {2026},
  publisher    = {Zenodo},
  doi          = {10.5281/zenodo.19785932},
  url          = {https://github.com/MinishLab/semble},
  license      = {MIT}
}

Project details

These details have been verified by PyPI

Project links

Owner

Minish

GitHub Statistics

These details have not been verified by PyPI

Release history Release notifications | RSS feed

0.3.3

Jun 5, 2026

This version

0.3.2

Jun 3, 2026

0.3.1

May 29, 2026

0.3.0

May 27, 2026

0.2.0

May 21, 2026

0.1.10

May 19, 2026

0.1.9

May 19, 2026

0.1.8

May 18, 2026

0.1.7

May 12, 2026

0.1.6

May 11, 2026

0.1.5

May 11, 2026

0.1.4

May 11, 2026

0.1.3

May 5, 2026

0.1.2

May 4, 2026

0.1.1

Apr 30, 2026

0.1.0

Apr 26, 2026

0.0.1

Apr 7, 2026

Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

semble-0.3.2.tar.gz (86.8 kB view details)

Uploaded Jun 3, 2026 Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

The dropdown lists show the available interpreters, ABIs, and platforms. Enable javascript to be able to filter the list of wheel files.

semble-0.3.2-py3-none-any.whl (61.1 kB view details)

Uploaded Jun 3, 2026 Python 3

File details

Details for the file semble-0.3.2.tar.gz.

File metadata

Download URL: semble-0.3.2.tar.gz
Upload date: Jun 3, 2026
Size: 86.8 kB
Tags: Source
Uploaded using Trusted Publishing? Yes
Uploaded via: twine/6.1.0 CPython/3.13.12

File hashes

Hashes for semble-0.3.2.tar.gz
Algorithm	Hash digest
SHA256	`cd3333f00ab47b54c4e08dd299389792c410bdbec8b8f9e05ce7cbfeaa0170e0`
MD5	`9a6c0b8206397e81c3530f17bb6df32e`
BLAKE2b-256	`8c6f71f950222b38264db5516b15767949abe04fc7898cb95a2252285010f8a1`

See more details on using hashes here.

Provenance

The following attestation bundles were made for semble-0.3.2.tar.gz:

Publisher: release.yaml on MinishLab/semble

Attestations: Values shown here reflect the state when the release was signed and may no longer be current.

Statement:
- Statement type: https://in-toto.io/Statement/v1
- Predicate type: https://docs.pypi.org/attestations/publish/v1
- Subject name: semble-0.3.2.tar.gz
- Subject digest: cd3333f00ab47b54c4e08dd299389792c410bdbec8b8f9e05ce7cbfeaa0170e0
- Sigstore transparency entry: 1710878911
- Sigstore integration time: Jun 3, 2026
Source repository:
- Permalink: MinishLab/semble@c7e52fa5665e065d2bf4ab283a57deac394de2cb
- Branch / Tag: refs/tags/v0.3.2
- Owner: https://github.com/MinishLab
- Access: public
Publication detail:
- Token Issuer: https://token.actions.githubusercontent.com
- Runner Environment: github-hosted
- Publication workflow: release.yaml@c7e52fa5665e065d2bf4ab283a57deac394de2cb
- Trigger Event: workflow_dispatch

File details

Details for the file semble-0.3.2-py3-none-any.whl.

File metadata

Download URL: semble-0.3.2-py3-none-any.whl
Upload date: Jun 3, 2026
Size: 61.1 kB
Tags: Python 3
Uploaded using Trusted Publishing? Yes
Uploaded via: twine/6.1.0 CPython/3.13.12

File hashes

Hashes for semble-0.3.2-py3-none-any.whl
Algorithm	Hash digest
SHA256	`2e0ce9e9eb53b4d119611442890a00a739bc27256e69b07c15a9d9c045f7ba21`
MD5	`f486d818340751e32729a5b703efb69f`
BLAKE2b-256	`ab3ec4e79972151fee12c69107e7e1418c71239f391bcf443f8f6bd1abba9bcb`

See more details on using hashes here.

Provenance

The following attestation bundles were made for semble-0.3.2-py3-none-any.whl:

Publisher: release.yaml on MinishLab/semble

Attestations: Values shown here reflect the state when the release was signed and may no longer be current.

Statement:
- Statement type: https://in-toto.io/Statement/v1
- Predicate type: https://docs.pypi.org/attestations/publish/v1
- Subject name: semble-0.3.2-py3-none-any.whl
- Subject digest: 2e0ce9e9eb53b4d119611442890a00a739bc27256e69b07c15a9d9c045f7ba21
- Sigstore transparency entry: 1710878946
- Sigstore integration time: Jun 3, 2026
Source repository:
- Permalink: MinishLab/semble@c7e52fa5665e065d2bf4ab283a57deac394de2cb
- Branch / Tag: refs/tags/v0.3.2
- Owner: https://github.com/MinishLab
- Access: public
Publication detail:
- Token Issuer: https://token.actions.githubusercontent.com
- Runner Environment: github-hosted
- Publication workflow: release.yaml@c7e52fa5665e065d2bf4ab283a57deac394de2cb
- Trigger Event: workflow_dispatch

semble 0.3.2

Navigation

Verified details

Project links

Owner

GitHub Statistics

Unverified details

Meta

Classifiers

Project description

Fast and Accurate Code Search for Agents
_{Uses ~98% fewer tokens than grep+read}

Quickstart

Main Features

CLI

MCP Server

Benchmarks

How it works

Acknowledgements

License

Citing

Project details

Verified details

Project links

Owner

GitHub Statistics

Unverified details

Meta

Classifiers

Release history Release notifications | RSS feed

Download files

Source Distribution

Built Distribution

File details

File metadata

File hashes

Provenance

File details

File metadata

File hashes

Provenance

semble 0.3.2

Navigation

Verified details

Project links

Owner

GitHub Statistics

Unverified details

Meta

Classifiers

Project description

Fast and Accurate Code Search for Agents Uses ~98% fewer tokens than grep+read

Quickstart

Main Features

CLI

MCP Server

Benchmarks

How it works

Acknowledgements

License

Citing

Project details

Verified details

Project links

Owner

GitHub Statistics

Unverified details

Meta

Classifiers

Release history Release notifications | RSS feed

Download files

Source Distribution

Built Distribution

File details

File metadata

File hashes

Provenance

File details

File metadata

File hashes

Provenance

Fast and Accurate Code Search for Agents
_{Uses ~98% fewer tokens than grep+read}