Anoop Johnson is a Principal Software Engineer with 13 years of experience building high-performance distributed systems for search, analytics and cloud data platforms. He has led low-latency services at Amazon/A9 powering autocomplete and spelling corrections at massive scale, improved query engines and access controls on EMR and Athena, and helped bring BigQuery capabilities to data lakes at Google. Currently at Databricks, he blends deep backend systems expertise with practical cloud integrations—his open-source contributions to Trino include spatial joins, ST_WITHIN, and enhanced S3/AWS credential handling that improved spatial query performance and observability. Known for squeezing latency out of complex systems, he has a track record of designing separation-of-compute-and-storage architectures and building production-grade query engines and storage integrations. He pairs hands-on coding with cross-team technical leadership and a penchant for measurable performance wins.
13 years of coding experience
13 years of employment as a software developer
Graduate-level coursework (non-degree) in Computer Science, Distributed Systems, Database System Principles, Graduate-level coursework (non-degree) in Computer Science, Distributed Systems, Database System Principles at Stanford University
Bachelor of Science (B.S.), Computer Science, Bachelor of Science (B.S.), Computer Science at University of Kerala
Official repository of Trino, the distributed SQL query engine for big data, formerly known as PrestoSQL (https://trino.io)
Role in this project:
Back-end Developer
Contributions:23 reviews, 3 PRs, 82 comments in 5 years 7 months
Contributions summary:Anoop contributed to the Trino SQL query engine by implementing spatial join functionality, improving the performance of spatial queries. They added the `ST_WITHIN` spatial function and optimized the execution plan. Further contributions involved adding spilled data metrics to the query statistics, web UI, and CLI, providing better query monitoring capabilities. Additionally, the user made changes to the S3 file system and AWS credential handling to improve integration with AWS services like S3 and Glue.
An open-source storage framework that enables building a Lakehouse architecture with compute engines including Spark, PrestoDB, Flink, Trino, and Hive and APIs
Contributions:45 pushes, 9 branches in 1 year 5 months
Find and Hire Top DevelopersWe’ve analyzed the programming source code of over 60 million software developers on GitHub and scored them by 50,000 skills. Sign-up on Prog,AI to search for software developers.