Last released Oct 7, 2026
LLM router that picks the model and reasoning effort per query from accuracy and real billed cost. Includes a correct per-request LLM cost engine.
Last released Oct 5, 2026
LLM cost routing with correct per-request cost: output-length forecasting, cache-aware tiered pricing, and $/quality ranking.