4 projects
openjoule
Plant-level inference and control engine for AI infrastructure. Estimates plant state under partial observation, forecasts outcomes under specified actions, imposes interventions, and measures results.
smol-vllm
From-scratch paged-attention inference engine: paged KV cache, continuous batching, preemption
abidex
Zero-code OpenTelemetry tracing for AI agents
relayserve
Relay: minimal LLM inference server for heterogeneous devices