Proof index
Everything on this site is backed by something you can open. Here it all is, in one place.
#Public code
Clone it and run it yourself — these prove the method is real, not a slide. llm-eval-gate is a working CI eval-gate you can drop into a repo; playwright-sdet-regression-suite ships its evidence folder (traces, screenshots) alongside a 37/37 zero-flake run.
#Verbatim runs & case studies
These prove the work survives contact with reality. The nexural-qa-os pair is the honest one to open first: the gate went red on a real CVE finding, blocked the release, and shows the green rerun after the fix — no quiet override.
#Live & interactive
Proof you can operate yourself, right now. The live eval grades an AI in your browser — and its “bring your own AI” mode scores a real answer of yours against the same rubric I build for clients.