Rahil Chertara

Apache Hudi Contributor at The Apache Software Foundation

Seattle, Washington, United States
email-iconphone-icongithub-logolinkedin-logotwitter-logostackoverflow-logofacebook-logo
Join Prog.AI to see contacts
email-iconphone-icongithub-logolinkedin-logotwitter-logostackoverflow-logofacebook-logo
Join Prog.AI to see contacts

Summary

🤩
Rockstar
🎓
Top School
Rahil Chertara is a software engineer with 8 years’ experience specializing in data and AI infrastructure, currently helping lead Apache Hudi work at Onehouse after a multi-year tenure at AWS where he contributed to SageMaker Lakehouse and S3 Tables. He is an active Apache contributor across Hudi, Iceberg, XTable, and Polaris, with deep hands-on expertise in query engines (Spark, Trino), data catalogs, open table and file formats (Parquet, Avro, Arrow, Lance), and distributed systems. Rahil’s contributions to Apache Hudi include bootstrap and partition handling, Spark optimizations, and AWS DMS integration—work that improves reliability and incremental processing at scale. Based in Seattle and grounded in practical production experience from EMR to startup, he blends backend engineering rigor with open-source stewardship to enable new AI/ML use cases on unstructured data.
code8 years of coding experience
job5 years of employment as a software developer
bookBachelor of Science - BS, Computer Science, Bachelor of Science - BS, Computer Science at Rutgers University–New Brunswick
languagesEnglish, Hindi
github-logo-circle

Github Skills (10)

javas10
apache-spark10
apache-hudi10
java10
data-engineering10
big-data9
avro9
data-integration8
apache-flink6
stream-processing5

Programming languages (7)

JavaRustCScalaJavaScriptThriftPython

Github contributions (5)

github-logo-circle
apache/hudi

Jul 2022 - Jan 2023

Upserts, Deletes And Incremental Processing on Big Data.
Role in this project:
userBack-end Developer / Data Engineer
Contributions:337 reviews, 19 commits, 95 PRs in 6 months
Contributions summary:Rahil contributed to the Apache Hudi project by addressing various issues and implementing improvements related to data processing and table management. Their work includes adding logging functionality, fixing Avro-related issues, and optimizing Spark-based operations like table invalidation. The user also focused on bootstrap and partition handling within Hudi, modifying the file index to accurately manage bootstrapped tables. Further contributions involved AWS DMS integration and enhancements.
big-dataincremental-processinghudiapachehudidatalake
rahil-c/hudi

Nov 2021 - Jul 2026

Upserts, Deletes And Incremental Processing on Big Data.
Contributions:16 reviews, 13 PRs, 501 pushes in 4 years 9 months
big-dataincremental-processing
Find and Hire Top DevelopersWe’ve analyzed the programming source code of over 60 million software developers on GitHub and scored them by 50,000 skills. Sign-up on Prog,AI to search for software developers.
Request Free Trial