Summary
Andrew Timmes is a Senior Site Reliability Engineer with 13 years of experience bridging HPC systems administration and modern SRE practices, currently supporting NVIDIA's DGX Cloud after platform leadership roles at SeatGeek. He’s a proven outage engineer who has kept large-scale services resilient at Twitter and Bloomberg, and who brings deep expertise in distributed caching, data-center ops, and cloud platform engineering. Andrew began in academic HPC at Princeton, which informs his pragmatic approach to performance and research-compute workloads at scale. Colleagues rely on him for steady incident leadership and for turning messy, legacy infrastructure into observable, reliable platforms.
13 years of coding experience
6 years of employment as a software developer
Bachelor of Science - BS Computer Science, Bachelor of Science - BS Computer Science at The College of New Jersey