6 projects
evoke-vllm
Relevance-driven CPU KV-cache offload policy for stock vLLM
incr-concurrent
Incremental computation engine with thread-safe runtime for Python
incr-compute
The fastest incremental computation engine for Python
cognitive-cache
Task-aware context selection for LLMs
skillprobe
Automated end-to-end skill testing for LLM coding tools
fuse-grader
Assignment grader.