Engineer focused on the reliability of long-running systems — incident response, SLOs, and safe change to systems that must keep running.
- pgincident — live terminal dashboard for the first 30 seconds of a PostgreSQL incident: connections, locks, long queries, idle transactions
- colref — find every live reference to a database column before a risky schema change (Rails / Django)
- gomarklint — catch broken docs and links before your readers do
- tinytap — eBPF-based traffic capture, making server-side request flows visible (work in progress)




