Sahil Takiar

San Diego, California, United States
email-iconphone-icongithub-logolinkedin-logotwitter-logostackoverflow-logofacebook-logo
Join Prog.AI to see contacts
email-iconphone-icongithub-logolinkedin-logotwitter-logostackoverflow-logofacebook-logo
Join Prog.AI to see contacts

Summary

🤩
Rockstar
🎓
Top School
Sahil Takiar is an engineering leader and backend systems expert with 13 years building large-scale data infrastructure, most recently managing Platform and Data Engineering teams at Flock Freight. He brings deep hands-on experience from Cloudera and LinkedIn working on core Apache projects—Impala, Hive and Gobblin—where he focused on performance, stability and storage/read-path improvements. Sahil’s contributions span low-level HDFS enhancements, decryption/readFully APIs, and query-level optimizations like Impala’s topn_bytes_limit, reflecting a strong blend of performance engineering and pragmatic design. As an Apache committer and project chair across Hive, Impala and Gobblin, he pairs open-source stewardship with production delivery. Based in San Diego and UC Berkeley–trained, he’s equally comfortable leading teams as he is digging into byte-buffer APIs or refactoring complex configuration code. Colleagues cite him for improving reliability in long-running distributed systems and shipping subtle but high-impact fixes that prevent crashes and boost throughput.
code13 years of coding experience
job10 years of employment as a software developer
bookBachelor of Science (B.S.), Electrical Engineering and Computer Science, Bachelor of Science (B.S.), Electrical Engineering and Computer Science at University of California, Berkeley
bookThe Harker School
languagesSpanish
stackoverflow-logo

Stackoverflow

Stats
1reputation
0reached
0answers
0questions
github-logo-circle

Github Skills (24)

c-language10
back-end-development10
configuration-management10
query-optimization10
hdfs10
hadoop10
java10
bytebuffer10
javas10
sql10
apache-hive10
performance-tuning10
imp10
cprogramming-language10
impala10

Programming languages (6)

JavaC++ShellScalaRAMLPython

Github contributions (5)

github-logo-circle
apache/gobblin

Feb 2015 - Jun 2017

A distributed data integration framework that simplifies common aspects of big data integration such as data ingestion, replication, organization and lifecycle management for both streaming and batch data ecosystems.
Role in this project:
userBack-end Developer
Contributions:1 release, 189 commits, 288 PRs in 2 years 4 months
Contributions summary:Sahil primarily contributed to the Gobblin framework by refactoring and enhancing existing code. They focused on cleaning up and improving the configuration usage across the codebase, specifically addressing the usage of "fork config". They also added comprehensive JavaDocs to clarify the functionality of the methods and classes, and updated configuration key values. Their work touched upon multiple core components of the framework.
datadcosdata-streambig-data-integrationbatch-data
apache/impala

Sep 2018 - Oct 2020

Apache Impala
Role in this project:
userBack-end Developer & Performance Engineer
Contributions:89 commits in 2 years 1 month
Contributions summary:Sahil primarily addressed performance and stability issues within the Apache Impala project. Their contributions focused on fixing arithmetic overflows, which could lead to crashes, and optimizing performance of existing features. Furthermore, the user enhanced the codebase by introducing query options, such as 'topn_bytes_limit,' to improve performance for large queries. In addition to improving performance of core features, the user also added tests to existing test suites.
apache-impalaparquetolapsqlapache
Find and Hire Top DevelopersWe’ve analyzed the programming source code of over 60 million software developers on GitHub and scored them by 50,000 skills. Sign-up on Prog,AI to search for software developers.
Request Free Trial