Summary
Mohammad Amirkhani is a Senior Service Reliability Engineer with 9 years of experience building and operating cloud-native infrastructure for high-throughput platforms. He has led teams to deliver systems that handle billions of daily requests with >99.9% uptime, cut resource usage by 30%+, and maintain sub-50ms ad response times through reactive architectures and efficient orchestration. Mohammad combines deep SRE and DevOps expertise—Kubernetes, ArgoCD/Workflows, Prometheus/Grafana, Terraform—with hands-on backend work in Kotlin, Java, and Go to automate complex ML and data workflows at scale. He has designed storage- and CDN-scale systems (100TB+) and engineered MLASS improvements using Apache Druid, Delta Lake, and Spark to reclaim idle cluster capacity. Based in London, he pairs research-oriented training in AI from Sharif University with pragmatic platform engineering, and is notable for building a Kubernetes Operator to automate backups and helm manifest generation for 1000+ deployments.
9 years of coding experience
10 years of employment as a software developer
Master's Degree Artificial Intelligence, Master's Degree Artificial Intelligence at Sharif University of Technology