Saptarshee Panda

Lead Data Engineer

Atlanta, Georgia, United States
email-iconphone-icongithub-logolinkedin-logotwitter-logostackoverflow-logofacebook-logo
Join Prog.AI to see contacts
email-iconphone-icongithub-logolinkedin-logotwitter-logostackoverflow-logofacebook-logo
Join Prog.AI to see contacts

Summary

🤩
Rockstar
🎓
Top School
Saptarshee Panda is a Lead Data Engineer based in Atlanta with over a decade of software experience and six years focused in data engineering roles across enterprises like Land O'Lakes, Fractal, and PepsiCo. He designs and operationalizes ETL pipelines and data warehouses using Azure Databricks, ADF, PySpark, and traditional platforms such as Teradata and IBM Netezza, with hands-on production support and L3 incident resolution. His open-source contributions to the Nebula graph database demonstrate deeper systems-level expertise—implementing TTL, index-scan enhancements, and admin tooling that improve query performance and storage behavior. Comfortable across cloud-native and legacy data stacks, he combines practical delivery in large organizations with a curiosity-driven approach to tackling database internals. Colleagues describe him as someone who seeks continuous learning and applies low-level optimizations to make high-level analytics more reliable.
code6 years of coding experience
job10 years of employment as a software developer
bookBachelor of Technology (B.Tech.) Mechanical Engineering, Bachelor of Technology (B.Tech.) Mechanical Engineering at Biju Patnaik University of Technology, Odisha
languagesEnglish, Hindi, Odia
stackoverflow-logo

Stackoverflow

Stats
1reputation
0reached
0answers
0questions
github-logo-circle

Github Skills (17)

c-language10
databases10
query-optimization10
graph-database10
optimisation10
pact10
cpp10
pac10
cprogramming-language10
distributed-database10
database-optimization10
optimization10
database10
testing9
sql8

Programming languages (4)

DockerfileC++ShellHTML

Github contributions (5)

github-logo-circle
vesoft-inc/nebula

Feb 2020 - Dec 2022

A distributed, fast open-source graph database featuring horizontal scalability and high availability
Role in this project:
userBack-end Developer & Database Engineer
Contributions:430 reviews, 131 commits, 67 PRs in 2 years 10 months
Contributions summary:Saptarshee made significant contributions to the storage capabilities of the Nebula graph database. They implemented time-to-live (TTL) functionality, enabling data expiration in storage, and also added test cases for the implemented TTL. The user was involved in adding and refactoring several tests to improve performance. Further, they were responsible for enhancing boundary condition checks.
nebulagraphnebula-graphnebulabig-datadistributed-systems
vesoft-inc/nebula-graph

Nov 2020 - May 2021

A distributed, fast open-source graph database featuring horizontal scalability and high availability. This is an archived repo for v2.5 only, from 2.6.0 +, NebulaGraph switched back to https://github.com/vesoft-inc/nebula
Role in this project:
userBack-end Developer & Database Engineer
Contributions:27 reviews, 9 commits, 10 PRs in 6 months
Contributions summary:Saptarshee's contributions center around enhancing the index scanning functionality within the NebulaGraph database. They modified the `IndexScanRule` and `IndexScanValidator` classes to support lookup operations on tag and edge names, and create indexes without fields. They also introduced a "SHOW STATUS" command to display database statistics and fixed error information. These changes indicate a focus on database query optimization and admin functionalities.
nebulagraphnebula-graphnebuladatabasehigh-availability
Find and Hire Top DevelopersWe’ve analyzed the programming source code of over 60 million software developers on GitHub and scored them by 50,000 skills. Sign-up on Prog,AI to search for software developers.
Request Free Trial
Saptarshee Panda - Lead Data Engineer