bastionskill
Static scanner for skill-poisoning. Point it at an agent skill (a SKILL.md
plus its bundled scripts) and it inspects the bundled executable code for
malicious behavior — then reports the shadow: what the code does that the
skill's description never declared.
Agent skills bundle scripts that run when the skill is invoked, and can install hooks that run afterward. That is an arbitrary-code-execution surface. bastionskill is the code-layer leg of the bastion suite; the prompt-layer (malicious SKILL.md text) is bastionsupply's job.
Install
pip install bastionskill
# optional: full prompt-layer scanning via bastionsupply
pip install "bastionskill[prompt]"
Zero required dependencies. Python 3.10+.
Use
bastionskill scan ./some-skill # scan a local skill dir
bastionskill scan ~/.claude/skills # batch-scan every skill under a dir
bastionskill scan owner/repo # pre-flight a REMOTE skill (shallow clone, no exec)
bastionskill scan https://github.com/o/r # ... by full URL
bastionskill scan ./skill --prompt # + hidden-unicode / prompt-layer
bastionskill scan ./skill --json # machine-readable
bastionskill scan ./skill --report out.json # signable manifest (per-file hashes, verdict)
bastionskill scan ./skill --record # append result to the local ledger
bastionskill scan ./skill --fail-on block # CI gate: block|review|none (default: review)
bastionskill harden ./skill -o skill-policy.yaml # v2 skill verdict for loaders/CI
bastionskill ledger # list previously scanned skills + dates
Verdict, not a wall of severities
The scan ends in one of three verdicts, because capability is not malice — a legit power-tool exercises network, secrets, and hooks too:
- allow — clean, or capability the skill legitimately has (even a lot of it).
- review — a poisoning signal a human should eyeball: a shadow (the code
exercises a capability
SKILL.mdnever declared), obfuscation, an opaque binary, or the exfil pattern (reads secrets and has egress). - block — hard malice with no honest use: a staged-exec (decode piped to a shell).
Capability findings are reported as informational context, not as blockers.
--fail-on (block|review|none, default review) is the CI gate. See
docs/github-action.md.
--sarif emits SARIF 2.1.0 (file + line per finding) for GitHub code scanning:
- run: bastionskill scan ./skill --sarif > bastionskill.sarif
continue-on-error: true
- uses: github/codeql-action/upload-sarif@v3
with:
sarif_file: bastionskill.sarif
Remote pre-flight shallow-clones the repo to a temp dir, scans statically, and deletes it. The skill's own code is never executed.
Ledger & rug-pull. --record writes each scan to ~/.bastionskill/ledger.jsonl
(source, content hash, date, verdict). Re-scan the same source after it changes and
you get a ! DRIFT warning — the poisoned-update vector.
What it catches (code-layer)
| Detector | Example |
|---|---|
| hook-install (lead) | a script that writes a PostToolUse hook into settings.json = persistence |
| network egress | socket.connect, requests.post, curl/wget, fetch() |
| secret read | ~/.aws/credentials, id_rsa, .env |
| obfuscation | `base64 -d |
| dynamic exec | exec(), eval(), getattr(m,n)() (Python AST tier) |
| destructive | rm -rf, Remove-Item -Recurse |
| lateral-tamper | writes to CLAUDE.md, MCP config, or other skills |
| opaque-binary | bundles a compiled/loadable file it can't inspect (incl. renamed binaries, magic-byte sniffed) |
| shadow | code exercises a capability SKILL.md never declared |
Python files get a real ast pass (stdlib) on top of regex, so dynamic exec /
import / attribute-built calls survive reflow. Bash and JS use regex heuristics.
Findings are reported regardless of dead-code or if False: / env-flag guards —
the scanner reads source, it never runs it, and malware hides behind guards too.
How it fits the suite
- Prompt-layer → bastionsupply (dependency, optional extra)
- Runtime gating → bastiongate
hardenemits apolicy_version: 2skill verdict (allow/deny + the checks and capabilities that tripped it) under theskill:block, for whatever installs or loads skills. agentbastion and bastiongate don't run skills: they load the file and ignore the block, and it sets no tool policy. The enforcement today isscan --fail-onin CI or before install.
Test fixture
The inert, defanged demo skill this scanner is built against lives at
Rinkia/poisoned-skill-demo — a
"markdown formatter" that actually exfiltrates and installs a hook. See its
EXPECTED.md for the findings oracle.
bastionskill scan Rinkia/poisoned-skill-demo # scan the demo straight off GitHub
License
MIT © 2026 Stefano Rizzello
Metadata
Release files for bastionskill 0.4.0
For a detailed explanation of source distributions (sdists) and built distributions (wheels), please see the package formats documentation.
Source distribution (sdist)
| File | Size | Uploaded | |
|---|---|---|---|
| bastionskill-0.4.0.tar.gz | 29.2 kB | Details |
Built distribution (wheel)
| File | Interpreter | ABI | Platform | Reset |
|---|---|---|---|---|
| bastionskill-0.4.0-py3-none-any.whl | Python 3 | none | any | Details |
Total release size: 55.7 kB
Release files / bastionskill-0.4.0.tar.gz
| Download URL | bastionskill-0.4.0.tar.gz |
|---|---|
| Size | 29.2 kB |
| Tags | Source |
|
SHA-256 checksum How to use checksums |
619b23ce21f36ce8788739f66db2b5cf4667755fd1610fd089451698cfda63d3
|
|
BLAKE2b-256 checksum How to use checksums |
0cafbb7c1a5192482242ddab78f85eb7b3e6ae1ad5259c4b0ac0dec9a6cc4ba0
|
| Upload date | |
|
Uploaded using Trusted Publishing? What is trusted publishing? |
Yes |
| Uploaded via |
twine/7.0.0 CPython/3.13.14
|
Provenance
Provenance describes where a file came from. On PyPI, provenance is shared via attestations, which provide a verifiable record of the build or publishing details. View details, limitations and caveats.
PyPI Publish Attestation
PyPI verified that this artifact, at this checksum, originated from the publisher listed below.
Signed by GitHub Actions, verified by PyPI on Sep 29, 2026.
Transparency logRelease files / bastionskill-0.4.0-py3-none-any.whl
| Download URL | bastionskill-0.4.0-py3-none-any.whl |
|---|---|
| Size | 26.5 kB |
| Tags | Python 3 |
|
SHA-256 checksum How to use checksums |
904850cabf188cf4255b658b5008afd9b4bed72df6767c2d8659389fd6a2a458
|
|
BLAKE2b-256 checksum How to use checksums |
f95c36d3a4f9f9365f41b6b48a5dbb138d9dc13bccf64ed6d6e8f5630d99c78f
|
| Upload date | |
|
Uploaded using Trusted Publishing? What is trusted publishing? |
Yes |
| Uploaded via |
twine/7.0.0 CPython/3.13.14
|
Provenance
Provenance describes where a file came from. On PyPI, provenance is shared via attestations, which provide a verifiable record of the build or publishing details. View details, limitations and caveats.
PyPI Publish Attestation
PyPI verified that this artifact, at this checksum, originated from the publisher listed below.
Signed by GitHub Actions, verified by PyPI on Sep 29, 2026.
Transparency log