Skip to main content

The AI Bench — CLI

Find the best local-AI model and setup for your hardware, from your terminal.

Tell it your platform, GPU/VRAM (or Mac unified memory), and what you want to do, and it returns a dated, opinionated recommendation: which models to run, the runner to use, quantization advice, and an expected tokens/sec band — the same logic that powers theaibench.ai.

$ uvx --from git+https://github.com/theaibench/theaibench-cli theaibench --platform windows --vram 24 --use-case coding

  THE AI BENCH  ·  Comfortable  (tier 5.25 · high band)
  Comfortable for midsize local models

  Strong for daily local use, coding, and experimentation.

  TOP PICKS
  1. Qwen3-Coder-30B-A3B (MoE, fits 24GB)
     3B-active MoE — benchmark champion for local coding at this tier.
  2. Qwen 3.5 35B-A3B (generalist MoE)
     Often wins real mixed-codebase work over the Coder variant; Apache 2.0.
  3. gpt-oss-20b
     OpenAI Apache 2.0; 21B MoE with 3.6B active; near o4-mini reasoning; fits 16GB.

  RUNNER Ollama or LM Studio — LM Studio for UI, Ollama for CLI + API.
  QUANT  Q4_K_M is the sweet spot at 14B. MoE 30B-A3B runs at 3B-dense speed.
  SPEED  50–70 tok/s on 8B, 30–50 on 14B Q4, 20–35 on 30B-A3B MoE.

  Full planner → https://theaibench.ai

What this is (and isn't): a recommender, not a benchmark. It calls The AI Bench's free planner API and prints the recommendation — it does not detect your GPU, run models, or measure tok/s on your machine. The speed figures are calibrated bands from the site's published methodology, not a live test of your hardware. Use it as a fast, opinionated starting point, then verify against real benchmarks for your exact GPU + model + quant before a big purchase.

Install / run

No install needed with uv — run straight from this repo:

uvx --from git+https://github.com/theaibench/theaibench-cli theaibench                # interactive
uvx --from git+https://github.com/theaibench/theaibench-cli theaibench --platform mac --memory 64 --use-case docs

Or install it:

pipx install git+https://github.com/theaibench/theaibench-cli    # isolated
pip install git+https://github.com/theaibench/theaibench-cli     # into the current environment

PyPI release is coming — once it lands, the short forms (uvx theaibench, pip install theaibench) will work too.

Usage

theaibench [options]

  --platform     windows | windows-laptop | mac | linux
  --vram         GPU VRAM in GB (PC):  none 8 12 16 20 24 32 48 64 96 128
  --memory       Mac unified memory in GB:  16 24 32 36 48 64 96 128
  --ram          system RAM in GB (PC):  16 32 48 64 96 128 192 256
  --gpu-family   nvidia | amd | cpu        (cpu = no dedicated GPU)
  --use-case     coding | chat | docs | image | agents | voice
  --priority     privacy | speed | cost
  --context      4096 | 16384 | 65536 | 200000   (typical prompt length)
  --mode         current | new            (new = planning a build, uses --budget)
  --budget       under1500 | 1500to3000 | 3000to6000 | 6000plus
  -i, --interactive   answer questions one at a time
  --json              print the raw API response
  --no-color          disable colored output

Run with no flags in a terminal and it walks you through the questions. Any flags you omit fall back to sensible defaults (Windows · 16 GB VRAM · coding · speed · NVIDIA).

Examples

# Mac Studio for long-document work
theaibench --platform mac --memory 96 --use-case docs --context 65536

# Budget NVIDIA box for image generation
theaibench --platform windows --vram 16 --use-case image

# CPU-only, privacy-first
theaibench --platform linux --gpu-family cpu --ram 64 --priority privacy

# Pipe the raw data somewhere
theaibench --platform mac --memory 64 --json | jq '.result.picks'

How it works

This is a thin, zero-dependency (standard-library-only) client over the free public planner API:

GET https://theaibench.ai/api/v1/plan

Because it reads the live API, recommendations stay as current as the website — model and hardware data are refreshed there on an ongoing basis. The full API is documented at theaibench.ai/api/ (OpenAPI schema included), and the agent/automation integration guide is at theaibench.ai/for-agents/.

Notes

  • Free to use and cite, with attribution.
  • No tracking, no accounts, no ads. The CLI sends only the hardware/use-case parameters you provide.
  • Not affiliated with NVIDIA, Apple, Ollama, LM Studio, or any model vendor.
  • Requires Python 3.9+.

License

MIT.

Metadata

Release files for theaibench 0.1.0

For a detailed explanation of source distributions (sdists) and built distributions (wheels), please see the package formats documentation.

Source distribution (sdist)

Source distribution for theaibench 0.1.0
File Size Uploaded
theaibench-0.1.0.tar.gz 9.1 kB Details

Built distribution (wheel)

Table of built distributions (wheels) for theaibench 0.1.0
File Interpreter ABI Platform
theaibench-0.1.0-py3-none-any.whl Python 3 none any Details

Total release size: 18.4 kB

Release files / theaibench-0.1.0.tar.gz

Download URL theaibench-0.1.0.tar.gz
Size 9.1 kB
Tags Source
SHA-256 checksum
How to use checksums
ddc00e9f5fe1e164671d62247798491d172eee4a5ddcdd0505dc421f083affd6
BLAKE2b-256 checksum
How to use checksums
13e1853ce75e44c0adb4308e4e93d891c46868ab3b071d332fb6fcce0272d32f
Upload date
Uploaded using Trusted Publishing?
What is trusted publishing?
No
Uploaded via uv/0.11.19 {"installer":{"name":"uv","version":"0.11.19","subcommand":["publish"]},"python":null,"implementation":{"name":null,"version":null},"distro":{"name":"macOS","version":null,"id":null,"libc":null},"system":{"name":null,"release":null},"cpu":null,"openssl_version":null,"setuptools_version":null,"rustc_version":null,"ci":null}

Release files / theaibench-0.1.0-py3-none-any.whl

Download URL theaibench-0.1.0-py3-none-any.whl
Size 9.3 kB
Tags Python 3
SHA-256 checksum
How to use checksums
18ce3623a4b6ead2b29cd42d9befa60decfbd667e84c96d70d2e1e9108f2a850
BLAKE2b-256 checksum
How to use checksums
3b1e8308fde105789c91fdad09fccb90b1d4e6f73a0a9216caf99109178b903d
Upload date
Uploaded using Trusted Publishing?
What is trusted publishing?
No
Uploaded via uv/0.11.19 {"installer":{"name":"uv","version":"0.11.19","subcommand":["publish"]},"python":null,"implementation":{"name":null,"version":null},"distro":{"name":"macOS","version":null,"id":null,"libc":null},"system":{"name":null,"release":null},"cpu":null,"openssl_version":null,"setuptools_version":null,"rustc_version":null,"ci":null}

Release history Release notifications | RSS feed

This release

0.1.0 This release

2 release files

Anthropic, PBC Visionary sponsor Bloomberg Visionary sponsor Hudson River Trading Visionary sponsor Meta Visionary sponsor NVIDIA Visionary sponsor Microsoft Sustainability sponsor Depot Continuous Integration AWS Cloud computing and Security Sponsor Datadog Monitoring Fastly CDN Google Download Analytics Sentry Error logging StatusPage Status page