Summary
Ryan Wright is a Senior Site Reliability Engineer with 10 years of experience building and operating large-scale cloud-native systems across AWS and GCP, specializing in Kubernetes, Terraform, and observability. He’s a pragmatic SRE who has repeatedly been promoted into high-visibility roles—owning fleets of thousands of nodes and hundreds of clusters supporting petabyte-scale data pipelines and 1,300+ customers. Ryan blends hands-on automation (Python, Bash, Go) with strong documentation and runbook discipline, having written onboarding guides, RFCs, and the majority of team docs wherever he’s worked. He’s shipped production Argo/Argo Workflows and Prometheus Blackbox monitoring for customer-facing login reliability and designed serverless RAG AI services and custom LLM inference infrastructure. Notably, he survived and thrived through a NOC downsizing by coding the tooling and processes that kept operations running and later scaled to FedRAMP-compliant Data Pipelines. Based in Denver, he pairs deep technical ownership with a knack for turning chaotic incidents into repeatable, documented solutions.
10 years of coding experience
4 years of employment as a software developer
A&P Airframe / Planet License, A&P Airframe / Planet License at Enterprise - Ozark Avaition School