Summary
Tai Mai is a Senior Site Reliability Engineer with 11+ years of experience building and operating large-scale, production systems across cloud and on-prem environments. Currently at NVIDIA, he brings deep operational expertise from past roles at Ooyala, Yahoo!, EMC and Amazon, including running high-throughput mail systems, building Docker-based PaaS, and standing up Kubernetes/EKS clusters across regions. Comfortable in languages from Go and Python to C/C++, he pairs hands-on coding with infrastructure automation (Chef, Jenkins, Datadog) to streamline deployments and improve system reliability. Colleagues know him for thorough testing, collaborative problem solving, and pragmatically improving customer experience under tight deadlines. An early-career firmware background and work on routing automation (route53 CNAME automation for k8s) hint at a blend of low-level systems thinking and modern cloud-native tooling. Based in California, he focuses on shipping maintainable, well-tested server-side software that scales.
11 years of coding experience
13 years of employment as a software developer
BS, Computer Engineering, BS, Computer Engineering at University of California, Berkeley
Parelta Colleges