Skip to main content

bastionskill

Static scanner for skill-poisoning. Point it at an agent skill (a SKILL.md plus its bundled scripts) and it inspects the bundled executable code for malicious behavior — then reports the shadow: what the code does that the skill's description never declared.

Agent skills bundle scripts that run when the skill is invoked, and can install hooks that run afterward. That is an arbitrary-code-execution surface. bastionskill is the code-layer leg of the bastion suite; the prompt-layer (malicious SKILL.md text) is bastionsupply's job.

Install

pip install bastionskill
# optional: full prompt-layer scanning via bastionsupply
pip install "bastionskill[prompt]"

Zero required dependencies. Python 3.10+.

Use

bastionskill scan ./some-skill              # scan a local skill dir
bastionskill scan ~/.claude/skills          # batch-scan every skill under a dir
bastionskill scan owner/repo                # pre-flight a REMOTE skill (shallow clone, no exec)
bastionskill scan https://github.com/o/r    #   ... by full URL
bastionskill scan ./skill --prompt          # + hidden-unicode / prompt-layer
bastionskill scan ./skill --json            # machine-readable
bastionskill scan ./skill --report out.json # signable manifest (per-file hashes, verdict)
bastionskill scan ./skill --record          # append result to the local ledger
bastionskill scan ./skill --fail-on block   # CI gate: block|review|none (default: review)
bastionskill harden ./skill -o skill-policy.yaml   # v2 skill verdict for loaders/CI
bastionskill ledger                         # list previously scanned skills + dates

Verdict, not a wall of severities

The scan ends in one of three verdicts, because capability is not malice — a legit power-tool exercises network, secrets, and hooks too:

  • allow — clean, or capability the skill legitimately has (even a lot of it).
  • review — a poisoning signal a human should eyeball: a shadow (the code exercises a capability SKILL.md never declared), obfuscation, an opaque binary, or the exfil pattern (reads secrets and has egress).
  • block — hard malice with no honest use: a staged-exec (decode piped to a shell).

Capability findings are reported as informational context, not as blockers. --fail-on (block|review|none, default review) is the CI gate. See docs/github-action.md.

--sarif emits SARIF 2.1.0 (file + line per finding) for GitHub code scanning:

- run: bastionskill scan ./skill --sarif > bastionskill.sarif
  continue-on-error: true
- uses: github/codeql-action/upload-sarif@v3
  with:
    sarif_file: bastionskill.sarif

Remote pre-flight shallow-clones the repo to a temp dir, scans statically, and deletes it. The skill's own code is never executed.

Ledger & rug-pull. --record writes each scan to ~/.bastionskill/ledger.jsonl (source, content hash, date, verdict). Re-scan the same source after it changes and you get a ! DRIFT warning — the poisoned-update vector.

What it catches (code-layer)

Detector Example
hook-install (lead) a script that writes a PostToolUse hook into settings.json = persistence
network egress socket.connect, requests.post, curl/wget, fetch()
secret read ~/.aws/credentials, id_rsa, .env
obfuscation `base64 -d
dynamic exec exec(), eval(), getattr(m,n)() (Python AST tier)
destructive rm -rf, Remove-Item -Recurse
lateral-tamper writes to CLAUDE.md, MCP config, or other skills
opaque-binary bundles a compiled/loadable file it can't inspect (incl. renamed binaries, magic-byte sniffed)
shadow code exercises a capability SKILL.md never declared

Python files get a real ast pass (stdlib) on top of regex, so dynamic exec / import / attribute-built calls survive reflow. Bash and JS use regex heuristics.

Findings are reported regardless of dead-code or if False: / env-flag guards — the scanner reads source, it never runs it, and malware hides behind guards too.

How it fits the suite

  • Prompt-layer → bastionsupply (dependency, optional extra)
  • Runtime gating → bastiongate
  • harden emits a policy_version: 2 skill verdict (allow/deny + the checks and capabilities that tripped it) under the skill: block, for whatever installs or loads skills. agentbastion and bastiongate don't run skills: they load the file and ignore the block, and it sets no tool policy. The enforcement today is scan --fail-on in CI or before install.

Test fixture

The inert, defanged demo skill this scanner is built against lives at Rinkia/poisoned-skill-demo — a "markdown formatter" that actually exfiltrates and installs a hook. See its EXPECTED.md for the findings oracle.

bastionskill scan Rinkia/poisoned-skill-demo   # scan the demo straight off GitHub

License

MIT © 2026 Stefano Rizzello

Metadata

Release files for bastionskill 0.4.0

For a detailed explanation of source distributions (sdists) and built distributions (wheels), please see the package formats documentation.

Source distribution (sdist)

Source distribution for bastionskill 0.4.0
File Size Uploaded
bastionskill-0.4.0.tar.gz 29.2 kB Details

Built distribution (wheel)

Table of built distributions (wheels) for bastionskill 0.4.0
File Interpreter ABI Platform
bastionskill-0.4.0-py3-none-any.whl Python 3 none any Details

Total release size: 55.7 kB

Release files / bastionskill-0.4.0.tar.gz

Download URL bastionskill-0.4.0.tar.gz
Size 29.2 kB
Tags Source
SHA-256 checksum
How to use checksums
619b23ce21f36ce8788739f66db2b5cf4667755fd1610fd089451698cfda63d3
BLAKE2b-256 checksum
How to use checksums
0cafbb7c1a5192482242ddab78f85eb7b3e6ae1ad5259c4b0ac0dec9a6cc4ba0
Upload date
Uploaded using Trusted Publishing?
What is trusted publishing?
Yes
Uploaded via twine/7.0.0 CPython/3.13.14

Provenance

Provenance describes where a file came from. On PyPI, provenance is shared via attestations, which provide a verifiable record of the build or publishing details. View details, limitations and caveats.

PyPI Publish Attestation

PyPI verified that this artifact, at this checksum, originated from the publisher listed below.

Signed by GitHub Actions, verified by PyPI on Sep 29, 2026.

Transparency log

Release files / bastionskill-0.4.0-py3-none-any.whl

Download URL bastionskill-0.4.0-py3-none-any.whl
Size 26.5 kB
Tags Python 3
SHA-256 checksum
How to use checksums
904850cabf188cf4255b658b5008afd9b4bed72df6767c2d8659389fd6a2a458
BLAKE2b-256 checksum
How to use checksums
f95c36d3a4f9f9365f41b6b48a5dbb138d9dc13bccf64ed6d6e8f5630d99c78f
Upload date
Uploaded using Trusted Publishing?
What is trusted publishing?
Yes
Uploaded via twine/7.0.0 CPython/3.13.14

Provenance

Provenance describes where a file came from. On PyPI, provenance is shared via attestations, which provide a verifiable record of the build or publishing details. View details, limitations and caveats.

PyPI Publish Attestation

PyPI verified that this artifact, at this checksum, originated from the publisher listed below.

Signed by GitHub Actions, verified by PyPI on Sep 29, 2026.

Transparency log

Release history Release notifications | RSS feed

0.9.0

2 release files

0.8.0

2 release files

0.7.0

2 release files

0.6.0

2 release files

0.5.0

2 release files

This release

0.4.0 This release

2 release files

0.3.0

2 release files

0.2.0

2 release files

0.1.0

2 release files

Anthropic, PBC Visionary sponsor Bloomberg Visionary sponsor Hudson River Trading Visionary sponsor Meta Visionary sponsor NVIDIA Visionary sponsor Microsoft Sustainability sponsor Depot Continuous Integration AWS Cloud computing and Security Sponsor Datadog Monitoring Fastly CDN Google Download Analytics Sentry Error logging StatusPage Status page