Summary
Anton Kirillov is a Principal Engineer specializing in distributed systems, ML infrastructure, and cloud-native compute with over a decade of hands-on experience building production platforms at startups, scaleups, and NVIDIA and Airbnb. He has repeatedly led greenfield designs and complex migrations—shipping MLOps and Spark operators, multi-tenant on‑prem and air‑gapped solutions, and GPU-enabled integration testing—while balancing deep implementation work with cross-functional product delivery. Expert in Kubernetes, Kubeflow, Spark, and Cassandra, Anton drives platforms that make ML workloads reliable and accessible to data scientists via developer-friendly SDKs and operators. His career blends research depth (PhD-level mathematical modelling) with practical performance engineering, from low-latency RTB systems to geo-distributed data pipelines. Based in Boulder, he thrives on tackling dynamic infrastructure challenges that span orchestration, security, and scalability. An uncommon through-line is his knack for shipping both operator-driven automation (KUDO/Kaptain) and developer ergonomics that materially accelerate ML adoption.
13 years of coding experience
17 years of employment as a software developer
Engineer, Information Technology, Engineer, Information Technology at Russian State Technological University named after K.E. Tsiolkovsky (MATI)
Doctor of Philosophy (Ph.D.), Mathematical Modelling, Doctor of Philosophy (Ph.D.), Mathematical Modelling at State University — Higher School of Economics
Russian, English