Skip to main content

Oriora c2 — a local proxy that plugs Oriora's model-routing decision into any OpenAI-style agent. Your vendor key and prompts stay on your machine; only the routing decision crosses.

Project description

oriora-c2

Use Oriora's model-routing decision in any OpenAI-style agent — without your key or prompts ever leaving your machine.

oriora-c2 runs a tiny local proxy on 127.0.0.1. Your agent points at it; before each call it asks Oriora's /api/select "which model is best for this task?", then dispatches the call directly to the vendor on your own key. Only the routing decision (task type + your candidate models) crosses to Oriora — never the key, the prompt, or the response.

This is for off-the-shelf agents (Cursor, Aider, Continue, the raw openai SDK, LangChain ChatOpenAI, …) that can't easily insert a "ask Oriora first" step. Writing your own code? You don't need this — call /api/select directly (or pip install oriora and use model_select()).

Install

pip install oriora-c2     # Python 3.10–3.13

Python 3.14 isn't supported yet — a dependency (litellm's pinned orjson) has no 3.14 build. On 3.14 the install succeeds but the tool exits with this exact guidance: create the venv with python3.13 -m venv and install there.

Run

oriora-c2 init                       # scaffolds config.yaml + .env.oriora-c2.example
# set your keys:
export ORIORA_API_KEY=sk_oriora_...  # the decision call only
export DEEPSEEK_API_KEY=...          # your own vendor keys (the actual call runs on these, locally)
export MINIMAX_API_KEY=...
oriora-c2 serve                      # local proxy on http://127.0.0.1:4000

Point any OpenAI client at it:

from openai import OpenAI
c = OpenAI(base_url="http://127.0.0.1:4000/v1", api_key="anything")
c.chat.completions.create(model="oriora-auto", messages=[{"role":"user","content":"…"}])
# task type is auto-detected locally (free); or force it: model="oriora-auto:coding" (or :coding_hard)

Reasoning models behave vendor-native through the proxy. The proxy never touches the response, so models like deepseek-v4-pro return their reasoning_content raw — and a tiny max_tokens budget can be consumed by reasoning before any answer text appears. Budget max_tokens generously (or pick non-reasoning candidates) for short-answer use.

How it works

  1. Agent → http://127.0.0.1:4000 (model="oriora-auto"). Prompt never leaves your box.
  2. The pre-call hook classifies the task locally (free regex rules + overshoot-biased difficulty escalation to *_hard, no LLM) → calls POST /api/select {task_type, models} — the one Oriora touch ($0.001/decision).
  3. It rewrites data["model"] to the recommended model.
  4. LiteLLM dispatches direct to the vendor on your local key; the stream flows vendor → you.

Privacy / c2 invariant: the proxy is customer-hosted (127.0.0.1). Only {task_type, model candidates} reach Oriora. If a third party ever hosted this, it would no longer be c2.

Fail-open: if /api/select is slow (>ORIORA_SELECT_TIMEOUT_S, default 2.5s) or down, the hook falls back to ORIORA_FALLBACK_MODEL so your agent is never blocked.

Configuration (env)

Var Purpose
ORIORA_API_KEY Oriora key for the decision call (required)
DEEPSEEK_API_KEY, MINIMAX_API_KEY, … your own vendor keys (the call runs on these)
ORIORA_CANDIDATES comma-sep catalog ids you hold keys for (sent to /api/select)
ORIORA_FALLBACK_MODEL model used if the decision call fails (default deepseek-v4-flash)
ORIORA_SELECT_TIMEOUT_S decision-call budget before fail-open (default 2.5)

Add a vendor = add its key + a model_list entry in config.yaml + its catalog id to ORIORA_CANDIDATES. v1 ships configured for DeepSeek + MiniMax (OpenAI-format).

MIT © Orioralabs OÜ · https://orioralabs.com

Project details


Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

oriora_c2-0.1.3.tar.gz (10.3 kB view details)

Uploaded Source

Built Distribution

If you're not sure about the file name format, learn more about wheel file names.

oriora_c2-0.1.3-py3-none-any.whl (12.7 kB view details)

Uploaded Python 3

File details

Details for the file oriora_c2-0.1.3.tar.gz.

File metadata

  • Download URL: oriora_c2-0.1.3.tar.gz
  • Upload date:
  • Size: 10.3 kB
  • Tags: Source
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/6.2.0 CPython/3.13.11

File hashes

Hashes for oriora_c2-0.1.3.tar.gz
Algorithm Hash digest
SHA256 d48befb977db6d1e3f953fa5b9a52766972b4cce3b78a628f27f6e038e56fb66
MD5 543dc8360a8baf96d9b97838d2d9854e
BLAKE2b-256 588911ee330d39cc03b2220b033dbceee51a555dcfc1a1a04bfa431045be72ba

See more details on using hashes here.

File details

Details for the file oriora_c2-0.1.3-py3-none-any.whl.

File metadata

  • Download URL: oriora_c2-0.1.3-py3-none-any.whl
  • Upload date:
  • Size: 12.7 kB
  • Tags: Python 3
  • Uploaded using Trusted Publishing? No
  • Uploaded via: twine/6.2.0 CPython/3.13.11

File hashes

Hashes for oriora_c2-0.1.3-py3-none-any.whl
Algorithm Hash digest
SHA256 1740f53676d9a7c08d7a5fc1716e7a196745e36bf5de39d55b705ca5222ca773
MD5 bb4933882bff04a9d3d04cdcdf106438
BLAKE2b-256 d8b9f9eadfed5f9f5c1bf43d6a9fda7ec045385d7bec482c0f1d2e59a2e726a6

See more details on using hashes here.

Supported by

AWS Cloud computing and Security Sponsor Datadog Monitoring Depot Continuous Integration Fastly CDN Google Download Analytics Pingdom Monitoring Sentry Error logging StatusPage Status page