Last released Aug 22, 2026
LLM deployment optimizer (performance, cost, and safety) — backed by ~204,000 real measurements on consumer GPUs
Quantize an LLM and check it still refuses what it should — a GPU-aware quantization CLI that measures the safety drift of the quantization it just performed.
Supported by