AI architect. I build agent tooling and measure it: benchmarks with published losses, silent-failure hunts on real repos.
Latest: an A/B retrieval benchmark of memory stacks — harness-comparison, cited in claude-mem #3693, write-up at ai-architect.tools/notes.
Open source: DeusData/codebase-memory-mcp #1832 — open upstream PR proposing Markdown-to-file REFERENCES_FILE graph edges, with focused tests, after a documentation fan-in blind spot surfaced in my A/B retrieval work.
Tools: Cortex · Zetetic Agents · ai-architect.tools




