Last released Sep 21, 2026
The LLM stack you can read and understand: Gemma 4 and Llama on CUDA, built from explicit C++23 components.