Ole Sasse

Senior Software Engineer at Databricks

Berlin, Germany
email-iconphone-icongithub-logolinkedin-logotwitter-logostackoverflow-logofacebook-logo
Join Prog.AI to see contacts
email-iconphone-icongithub-logolinkedin-logotwitter-logostackoverflow-logofacebook-logo
Join Prog.AI to see contacts

Summary

🤩
Rockstar
🎓
Top School
Ole Sasse is a Senior Software Engineer based in Berlin with six years of professional experience building high-performance SQL and logistics systems. Currently at Databricks, he focuses on SQL query execution performance and has contributed notable back-end and optimizer fixes to Apache Spark, including timestamp handling in subqueries and metadata-aware filtering. His prior work spans logistics algorithms at Zalando and long-term routing and navigation engineering at TomTom, reflecting deep expertise in algorithms and distributed data processing. An algorithm engineer who also shapes team productivity, he has a practical knack for improving robustness in open-source projects such as Delta Lake and Spark—work that often surfaces as subtle correctness and metrics fixes rather than flashy features.
code6 years of coding experience
job14 years of employment as a software developer
bookDiplom, Informatik, Diplom, Informatik at Universität Karlsruhe (TH)
languagesGerman, English
github-logo-circle

Github Skills (21)

apache-spark10
delta-lake10
operation10
testing10
big-data10
mergetool10
scala10
merge10
mergesort10
sql10
error-handling10
data-lake10
optimization10
data-engineering9
spark9

Programming languages (2)

CScala

Github contributions (5)

github-logo-circle
delta-io/delta

Apr 2022 - Jan 2023

An open-source storage framework that enables building a Lakehouse architecture with compute engines including Spark, PrestoDB, Flink, Trino, and Hive and APIs
Role in this project:
userBack-end Developer
Contributions:35 reviews, 9 commits, 15 PRs in 8 months
Contributions summary:Ole primarily contributed to improving the error handling and messaging within the Delta Lake framework, specifically when dealing with table paths and metrics. They fixed issues related to incorrect operational metrics for DELETE commands and refactored code related to MergeIntoSQLSuite and OptimisticTransaction. The user also implemented features around consistent timestamps in the MergeIntoCommand and handling concurrent operations involving row tracking. Their work focused on enhancing the robustness and accuracy of Delta Lake operations.
prestodbsparktrinobig-dataanalytics
apache/spark

Jun 2022 - Jan 2023

Apache Spark - A unified analytics engine for large-scale data processing
Role in this project:
userBack-end Developer & Data Engineer
Contributions:49 reviews, 5 commits, 15 PRs in 7 months
Contributions summary:Ole primarily contributed to the Apache Spark codebase, focusing on the SQL and data processing aspects. Their commits involve modifications to the optimizer, specifically addressing timestamp evaluation in subqueries and adding new metadata column types for file source. Furthermore, they addressed issues related to filtering by row index, ensuring correct results and added tests for filtering based on metadata. They also improved code by re-using literal objects in ComputeCurrentTime rule and added support for custom metrics for V1Fallback writers.
apache-sparkpythonscalarjava
Find and Hire Top DevelopersWe’ve analyzed the programming source code of over 60 million software developers on GitHub and scored them by 50,000 skills. Sign-up on Prog,AI to search for software developers.
Request Free Trial