AI architect. I build agent tooling and measure it: benchmarks with published losses, silent-failure hunts on real repos.
Latest: an A/B retrieval benchmark of memory stacks — harness-comparison, cited in claude-mem #3693, write-up at ai-architect.tools/notes.
Tools: Cortex · Zetetic Agents · ai-architect.tools




