Independent agent-tooling builder. I ship deterministic, no-LLM software that agents and teams can trust: 4 projects live in the Official MCP Registry (skill security audit, evidence-gated claim reporting, benchmark-grading hygiene, secret scrubbing), all MIT + fully audited. I also documented and fixed a real benchmark-harness bug (VulcanBench #79) where a leaked pytest --cov option silently scored passing tasks as failures. I build custom MCP connectors, agent-verification layers, and eval/benchmark hygiene for teams that need auditable AI tooling — reproducible outputs, no LLM in the hot path.