Tim Harsch

Engineering Manager at Unite Genomics, Inc.

San Francisco Bay Area United States
email-iconphone-icongithub-logolinkedin-logotwitter-logostackoverflow-logofacebook-logo
Join Prog.AI to see contacts
email-iconphone-icongithub-logolinkedin-logotwitter-logostackoverflow-logofacebook-logo
Join Prog.AI to see contacts

Summary

🤩
Rockstar
🎓
Top School
Tim Harsch is an engineering manager and seasoned software engineer with 21+ years of product development experience, currently leading engineering at Unite Genomics in the San Francisco Bay Area. He blends deep hands-on expertise in big data, Spark, distributed systems and Java/Scala toolchains with strong team-building and Scrum leadership, having architected automation and validation frameworks at Cray and led data engineering efforts at Teradata/Think Big. Tim has a track record of integrating complex platforms—evidenced by his Kylo contributions (improving Spark stability and adding EMR, VCS and JupyterHub integrations)—that bridge data pipelines, analytics, and reproducible data science. Comfortable from low-level OS and cluster orchestration to cloud-native stacks (AWS, Docker, Helm), he frequently pairs practical systems thinking with a long-view product mindset. Colleagues see him as an insightful mentor who prefers solving hard reliability and performance problems while enabling teams to deliver scalable, production-grade data platforms.
code11 years of coding experience
job12 years of employment as a software developer
bookBachelor of Science (BS), Computer Science, Bachelor of Science (BS), Computer Science at University of California, Davis
bookAS, Computer Science, AS, Computer Science at Mendocino College
stackoverflow-logo

Stackoverflow

Stats
31reputation
41reached
2answers
1question
github-logo-circle

Github Skills (22)

spark10
data-pipelines10
data-engineering10
java10
javas10
data-pipeline10
hadoop9
rest-api8
sql8
api-rest8
restful-api8
api-design8
thrift7
maven6
react-router6

Programming languages (5)

JavaShellScalaHTMLPython

Github contributions (5)

github-logo-circle
Teradata/kylo

Jan 2017 - Dec 2018

Kylo is a data lake management software platform and framework for enabling scalable enterprise-class data lakes on big data technologies such as Teradata, Apache Spark and/or Hadoop. Kylo is licensed under Apache 2.0. Contributed by Teradata Inc.
Role in this project:
userBack-end Developer & Data Engineer
Contributions:248 commits, 4 PRs, 158 pushes in 1 year 11 months
Contributions summary:Tim primarily worked on improving the functionality and stability of the Kylo data lake management platform. Their commits focused on fixing issues with long-running Spark scripts and addressing failures within the Spark file schema parser service. Furthermore, they made changes to the template import and table creation processes, suggesting involvement in the data pipeline aspects of the platform. The user demonstrated a strong understanding of Spark and its associated processes, as well as a focus on addressing errors and improving performance.
data-lakelicensedbig-dataetlapache-spark
harschware/kylo

May 2018 - Jan 2019

Contributions:49 pushes, 18 branches in 8 months
thinkdata-lakelicensedbig-datakylo
Find and Hire Top DevelopersWe’ve analyzed the programming source code of over 60 million software developers on GitHub and scored them by 50,000 skills. Sign-up on Prog,AI to search for software developers.
Request Free Trial