Anoop Johnson

Principal Software Engineer at Databricks

San Francisco Bay Area United States
email-iconphone-icongithub-logolinkedin-logotwitter-logostackoverflow-logofacebook-logo
Join Prog.AI to see contacts
email-iconphone-icongithub-logolinkedin-logotwitter-logostackoverflow-logofacebook-logo
Join Prog.AI to see contacts

Summary

🤩
Rockstar
🎓
Top School
Anoop Johnson is a Principal Software Engineer with 13 years of experience building high-performance distributed systems for search, analytics and cloud data platforms. He has led low-latency services at Amazon/A9 powering autocomplete and spelling corrections at massive scale, improved query engines and access controls on EMR and Athena, and helped bring BigQuery capabilities to data lakes at Google. Currently at Databricks, he blends deep backend systems expertise with practical cloud integrations—his open-source contributions to Trino include spatial joins, ST_WITHIN, and enhanced S3/AWS credential handling that improved spatial query performance and observability. Known for squeezing latency out of complex systems, he has a track record of designing separation-of-compute-and-storage architectures and building production-grade query engines and storage integrations. He pairs hands-on coding with cross-team technical leadership and a penchant for measurable performance wins.
code13 years of coding experience
job13 years of employment as a software developer
bookGraduate-level coursework (non-degree) in Computer Science, Distributed Systems, Database System Principles, Graduate-level coursework (non-degree) in Computer Science, Distributed Systems, Database System Principles at Stanford University
bookBachelor of Science (B.S.), Computer Science, Bachelor of Science (B.S.), Computer Science at University of Kerala
stackoverflow-logo

Stackoverflow

Stats
46reputation
3kreached
2answers
0questions
github-logo-circle

Github Skills (13)

javas10
query-engine10
databases10
presto10
distributed-database10
sql10
java10
database10
aws9
big-data9
hive-metastore6
amazon-athena6
amazon-web-services6

Programming languages (5)

JavaRustScalaHTMLPython

Github contributions (5)

github-logo-circle
trinodb/trino

May 2019 - Dec 2024

Official repository of Trino, the distributed SQL query engine for big data, formerly known as PrestoSQL (https://trino.io)
Role in this project:
userBack-end Developer
Contributions:23 reviews, 3 PRs, 82 comments in 5 years 7 months
Contributions summary:Anoop contributed to the Trino SQL query engine by implementing spatial join functionality, improving the performance of spatial queries. They added the `ST_WITHIN` spatial function and optimized the execution plan. Further contributions involved adding spilled data metrics to the query statistics, web UI, and CLI, providing better query monitoring capabilities. Additionally, the user made changes to the S3 file system and AWS credential handling to improve integration with AWS services like S3 and Glue.
big-datasql-querytrinojavapresto
anoopj/delta

Feb 2025 - Jul 2026

An open-source storage framework that enables building a Lakehouse architecture with compute engines including Spark, PrestoDB, Flink, Trino, and Hive and APIs
Contributions:45 pushes, 9 branches in 1 year 5 months
Find and Hire Top DevelopersWe’ve analyzed the programming source code of over 60 million software developers on GitHub and scored them by 50,000 skills. Sign-up on Prog,AI to search for software developers.
Request Free Trial