I build the AI systems a team actually works with: automating the busywork, keeping people in the loop where it counts.
I spent 8 years running community operations at scale: 250K+ player community, 131K peak concurrent, 15 broadcast languages at Com2uS. Now I build the AI workflows, evals, and enablement systems that let small teams operate like bigger ones, and prove they actually get used.
- job-match-radar: self-hosted n8n + Supabase pipeline that scrapes listings, scores them against a personal rubric, and emails a weekly digest, with an eval + guardrail harness regression-testing the LLM scorer. Production automation, not a demo.
- llm-judge-evals: a dependency-free harness for keeping an LLM-as-judge honest: gate tests on documented mis-scores, a human-in-the-loop guardrail, and version-keyed drift detection.
- claude-code-skills: Claude Code skills I built for my own stack: capture, handoff, and knowledge-ops utilities across Obsidian, Notion, and NotebookLM.
- second-brain-system: a self-documenting knowledge + automation system: Obsidian as the canonical store, Claude Code + local LLMs as the activation layer, drift tripwires as guardrails.
Last updated: June 2026
