Software Engineer at The Apache Software Foundation
Indore, Madhya Pradesh, India
Join Prog.AI to see contacts
Join Prog.AI to see contacts
Summary
🤩
Rockstar
🎓
Top School
Parag Jain is a software engineer with 11 years of experience specializing in distributed systems and big data, currently building BI-as-code at Rill Data from Indore, India. He is a long-time contributor and committer to Apache Druid, where his work on indexing, real-time processing, failure handling, and handoff metrics has improved reliability for high-throughput analytics. Previously he helped productionize Druid and its Kafka indexing service at Yahoo and served on Lyft’s data infrastructure team, giving him deep operational experience with streaming ingestion and monitoring stacks. At Rill he extended SQL tooling—working on Calcite parsing and protobuf-serialized ASTs—bridging low-level systems work with higher-level analytics UX. He holds a Master’s in Computer Science from UIUC and brings a pragmatic balance of production hardening, open-source stewardship, and tooling that accelerates data teams. An observer of system behavior as much as a builder, he often focuses on small reliability improvements that yield outsized production stability.
11 years of coding experience
7 years of employment as a software developer
Bachelor of Engineering (BE) Information Technology, Bachelor of Engineering (BE) Information Technology at SGSITS, Indore
Master's degree Computer Science, Master's degree Computer Science at University of Illinois Urbana-Champaign
Rill is a tool for effortlessly transforming data sets into powerful, opinionated dashboards using SQL. BI-as-code.
Role in this project:
Back-end Developer
Contributions:631 reviews, 22 commits, 264 PRs in 5 months
Contributions summary:Parag primarily contributed to the development of a SQL library within the Rill project. Their work included extending the Calcite parser to support custom syntax for model creation and query expansion. They also focused on generating protobuf serialized parsed and validated ASTs with optional type information. Furthermore, the user was responsible for integrating and testing the SQL library, as evidenced by the build, testing, and dependency management for the SQL library using GraalVM and Protobuf.
Apache Druid: a high performance real-time analytics database.
Role in this project:
Back-end Developer
Contributions:72 reviews, 87 commits, 255 PRs in 6 years 7 months
Contributions summary:Parag contributed significantly to the Apache Druid codebase by focusing on improvements to the indexing and real-time processing components. Their work included implementing failure handling within task lifecycles, ensuring robust task behavior in the face of exceptions. Furthermore, the user added metrics for monitoring handoff counts. These contributions enhance Druid's reliability and performance, particularly within the real-time data ingestion pipeline.
real-timebig-datadruiddatabasehadoop
Find and Hire Top DevelopersWe’ve analyzed the programming source code of over 60 million software developers on GitHub and scored them by 50,000 skills. Sign-up on Prog,AI to search for software developers.
Request Free Trial
Parag Jain - Software Engineer at The Apache Software Foundation