Ahir Reddy is a software engineer based in Berkeley with 14 years of experience building scalable backend systems, currently at Databricks in the San Francisco Bay Area. He has made notable open-source contributions to Apache Spark, enhancing the PySpark API, SparkSQL integration, and robust job cancellation and worker monitoring features used by large-scale data processing workloads. A UC Berkeley graduate in EECS and Business Administration, he blends strong systems engineering skills with product-aware thinking. Early internships at Google reflect a background in high-performance, production-focused engineering. Colleagues rely on him for pragmatic solutions that bridge Python interfaces with distributed JVM systems.
14 years of coding experience
Bachelor of Science (B.S.) Electrical Engineering and Computer Science, Bachelor of Science (B.S.) Electrical Engineering and Computer Science at University of California, Berkeley
Apache Spark - A unified analytics engine for large-scale data processing
Role in this project:
Back-end Developer
Contributions:2 commits, 1 PR, 1 comment in 1 day
Contributions summary:Ahir primarily contributed to the PySpark API, focusing on enhancing its functionality. They developed features to integrate PySpark with SparkSQL, enabling SQL query execution on RDDs. The user also improved job cancelation capabilities within PySpark, including a monitor thread to manage and terminate Python workers. Additionally, they fixed issues related to character encoding and proper HiveContext setting.
Contributions:1 review, 226 pushes, 4 branches in 9 years 11 months
scalabazelrules
Find and Hire Top DevelopersWe’ve analyzed the programming source code of over 60 million software developers on GitHub and scored them by 50,000 skills. Sign-up on Prog,AI to search for software developers.