Skip to main content

pystreamliner

PyPI Python License

Automatically clean up messy Python files — without breaking anything.

pystreamliner uses Python's AST (abstract syntax tree) to safely detect and fix common code issues. It operates on two tiers: things it can fix automatically with zero risk, and things it flags for you to review manually.

Supports single files, multiple files, and recursive directory cleaning with a tight summary mode for large runs. Emits JSON and SARIF 2.1.0 for CI. Optional mtime cache for repeated local runs.

Discord: https://discord.gg/Z6cXxhSKS


What it does

Auto-fixes (Tier 1 — applied immediately):

  • Removes unused imports, or trims partially unused from x import y statements
  • Removes consecutive duplicate lines
  • Caps excessive blank lines

Warnings (Tier 2 — reported, never auto-changed):

  • Unused variables
  • Unused top-level functions
  • Unused classes
  • Vague variable names (x, tmp, foo, bar, etc.)
  • Shadowed built-ins
  • Dangerous calls (eval, exec, pickle, os.system, subprocess(..., shell=True), unsafe yaml.load)
  • Possible hardcoded secrets
  • Assert statements
  • Broad except: / except Exception

pystreamliner never touches code it isn't certain about. If there's any doubt, it warns you instead.


Install

pip install pystreamliner

No dependencies. Runs on Python 3.13+.


Usage

Single file:

pystreamliner your_file.py

Multiple files:

pystreamliner file1.py file2.py utils/*.py

Entire project (recursive):

pystreamliner .
# or
pystreamliner src/ tests/

Directories are walked recursively. Common junk directories (.git, __pycache__, venv, node_modules, etc.) are automatically skipped when they appear as sub-directories.

Preview without modifying:

pystreamliner --dry-run .

CI mode (exit non-zero on issues):

pystreamliner --check --quiet .

SARIF for Code Scanning / security dashboards:

pystreamliner --sarif --dry-run . > results.sarif

JSON for scripts:

pystreamliner --json --dry-run .

Faster repeated local runs (mtime cache):

pystreamliner --cache .
# optional custom cache path
pystreamliner --cache --cache-file /tmp/ps-cache.json .

Parallelism:

# default is sequential (-j 1) — safest for small trees
pystreamliner .

# auto (capped workers)
pystreamliner -j 0 .

# explicit workers; prefer threads on many small files
pystreamliner -j 4 --threads .

Cross-file unused functions / classes (opt-in):

pystreamliner --project --dry-run src/

Name-based: if another file in the same run imports or references the name, the unused_function / unused_class warning is dropped. Default remains per-file. Zero extra dependencies.


Big runs / Summary mode

When you process 5 or more files (configurable with --summary-threshold), pystreamliner switches to a compact summary instead of dumping a full report for every file.


CLI reference (high-signal flags)

Flag Purpose
-d, --dry-run Analyze / report only; do not write
-c, --check Exit 1 if changes or warnings (CI)
-q, --quiet Suppress human report
--json Machine-readable JSON (includes import_details)
--sarif SARIF 2.1.0 report on stdout
--cache Skip unchanged files (mtime + size)
--cache-file PATH Cache location (default .pystreamliner_cache.json)
-j, --jobs N Workers; default 1; 0 = auto (capped)
--threads Use threads instead of processes when jobs > 1
--project Suppress unused function/class warnings if the name is referenced in another file in this run
-w, --warn-only Report only; never rewrite
--fix-only Tier-1 fixes only; suppress Tier-2 warnings
--select / --ignore Filter warning categories
--exclude-path Glob path excludes (repeatable)
--aggressive Stricter blank-line collapsing

Config file support (zero deps): .pystreamliner.toml or [tool.pystreamliner] in pyproject.toml. CLI always wins.


Limitations / By design

These behaviours are intentional. They keep the tool zero-dependency, fast, and conservative.

Unused function / class detection is per-file by default

By default pystreamliner analyses each file independently using only that file's AST.
It does not follow imports across modules unless you pass --project.

Consequence without --project: a function or class that is defined in one file and imported + used in another file will be reported as unused when you run the tool on the definition file alone.

--project (also project = true in config) builds a cheap name index over every file in the current run and suppresses unused_function / unused_class when the name is referenced elsewhere (imports, attribute access, identifier strings, __all__). Still zero third-party deps. It is not a full import resolver: it does not understand types, import *, or dynamic getattr beyond literal strings.

This stays opt-in so the default remains conservative and single-pass cheap.

Work-arounds without --project:

  • Put public API names in __all__ — they are automatically treated as used.
  • Use --ignore unused_function,unused_class (or the config equivalent).
  • Run with --project on the whole package so definitions and call sites share one index.

Directory name collisions with the ignore list

The built-in ignore list contains common junk directories (__pycache__, .git, venv, coverage, htmlcov, etc.).
These are only skipped when they appear as sub-directories of a path you gave the tool.

If you explicitly pass a directory that happens to be named one of those (e.g. pystreamliner coverage/), its contents are processed. (This was fixed in 1.19.1.)

Nested junk directories inside that tree are still skipped as expected.

Summary mode vs detailed reports

When ≥ 5 files are processed (configurable), output switches to a compact summary that shows counts only.
Detailed per-file reports (with every warning message) appear only for smaller runs. This is intentional so large projects stay readable.

Parallelism defaults

Default is sequential (-j 1). Process pools have non-trivial spawn cost; for many small files prefer --threads or leave the default alone. Use -j 0 only when you know you want capped multi-core.


Contributing

See CONTRIBUTING.md.

Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

pystreamliner-1.21.0.tar.gz (27.6 kB view details)

Uploaded Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

pystreamliner-1.21.0-py3-none-any.whl (28.8 kB view details)

Uploaded Python 3

File details

Details for the file pystreamliner-1.21.0.tar.gz.

File metadata

  • Download URL: pystreamliner-1.21.0.tar.gz
  • Upload date:
  • Size: 27.6 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? Yes
  • Uploaded via: twine/7.0.0 CPython/3.13.14

File hashes

Hashes for pystreamliner-1.21.0.tar.gz
Algorithm Hash digest
SHA256 e29e8adca8b59999d5a6dfc4d4bbc89814af92e997988bdc27d21990cbaec8eb
MD5 ac0b6f83bb29513356e863dc189f31cd
BLAKE2b-256 e7703e5cb26be9844e78625b570ecd33e93eb2dc0afc08c80fd7789d22272947

See more details on using hashes here.

Provenance

The following attestation bundles were made for pystreamliner-1.21.0.tar.gz:

Publisher: main.yml on Supe232323/pystreamliner

Attestations: Values shown here reflect the state when the release was signed and may no longer be current.

File details

Details for the file pystreamliner-1.21.0-py3-none-any.whl.

File metadata

  • Download URL: pystreamliner-1.21.0-py3-none-any.whl
  • Upload date:
  • Size: 28.8 kB
  • Tags: Python 3
  • Uploaded using Trusted Publishing? Yes
  • Uploaded via: twine/7.0.0 CPython/3.13.14

File hashes

Hashes for pystreamliner-1.21.0-py3-none-any.whl
Algorithm Hash digest
SHA256 c6a17f08b6fb64ed5e0e62912273659cecc0cbb4859c3873bb02015b83094d98
MD5 18ab25dd555db5740dca0148e9b20cd6
BLAKE2b-256 93a706c8797f774c31c8c30e8ad8bf75647941ec33fbfdf31d2155665ac62f35

See more details on using hashes here.

Provenance

The following attestation bundles were made for pystreamliner-1.21.0-py3-none-any.whl:

Publisher: main.yml on Supe232323/pystreamliner

Attestations: Values shown here reflect the state when the release was signed and may no longer be current.

Release history Release notifications | RSS feed

1.21.1

2 files

This release

1.21.0 This release

2 files

1.20.3

2 files

1.20.2

2 files

1.20.0

2 files

1.19.1

2 files

1.19.0

2 files

1.18.0

2 files

1.17.0

2 files

1.16.0

2 files

1.15.0

2 files

1.14.0

2 files

1.13.1

2 files

1.12.1

2 files

1.12.0

2 files

1.11.0

2 files

1.3

2 files

Anthropic, PBC Visionary sponsor Bloomberg Visionary sponsor Hudson River Trading Visionary sponsor Meta Visionary sponsor NVIDIA Visionary sponsor Microsoft Sustainability sponsor Depot Continuous Integration AWS Cloud computing and Security Sponsor Datadog Monitoring Fastly CDN Google Download Analytics Sentry Error logging StatusPage Status page