Summary
Dylan Bochman is a Sr. Site Reliability Engineer and Technical Incident Manager with eight years of experience building resilient systems and driving incident response at companies like NVIDIA, Groq, HashiCorp (now IBM), and Spotify. He blends SRE practice with product management sensibilities to improve service availability, on-call effectiveness, and post-incident learning across distributed teams. Dylan has led high-severity incident command, shaped service-level strategy, and scaled detection and response workflows, translating operational metrics into leadership-level decisions. Based in Boston, he brings a hands-on automation mindset and a track record of turning messy outages into repeatable processes that reduce toil and risk. An early IT support background and time as a product manager give him uncommon empathy for both operators and product teams when balancing reliability and speed.
7 years of coding experience
10 years of employment as a software developer
UMass Lowell
High School Diploma, Computer and Information Sciences, General, High School Diploma, Computer and Information Sciences, General at Bedford High School
Spanish