4 projects
conditioning
A Conditioning Analyzer that tiers its claims: it detects cue, behavior, consequence, reinforcement loops from event sequences and never asserts a pattern from two ordinary points. Ships episode-dedupe.
cognitive-governance
Traits, not prompts: persistent dispositions above your model — anti-manipulation governance, investigative drive, moral drift detection. Deterministic, disclosed, zero dependencies.
capability-honesty
Make an AI report the lowest proven state: advertise abilities only after a probe passes, and call work done only when a verifiable artifact backs it. Fail-closed.
adversarial-honesty-tests
Test an AI for honesty by trying to make it lie: probes that hand the system a dishonesty shape and pass only when it declines and discloses. A silent success on a rigged input is the bug.