Timothy Carbone is Head of Data at Unsplash with 11 years of experience building large-scale data platforms and pipelines. He designed Unsplash’s data architecture from scratch—collecting 600+ GB/day and ~1,500 behavioural events/sec—and led efforts across batch and real-time systems on AWS and GCP using tools like Airflow, dbt, Redshift, BigQuery and Snowplow. A hands-on engineer who moved from sole-data-team implementer to leader, he published the Unsplash Dataset, one of the largest open image datasets for ML research. His background blends computer engineering and business strategy, and he’s particularly strong at bridging data engineering, product analytics and operational monitoring. Based in Montreal, he combines deep systems know-how (Python, Node.js, PostgreSQL) with a knack for turning high-throughput logs into product features and actionable business insight.
11 years of coding experience
8 years of employment as a software developer
Master of Science (M.Sc.) Business Management & Strategy, Master of Science (M.Sc.) Business Management & Strategy at KEDGE Business School
High School Diploma Engineering Science, High School Diploma Engineering Science at Lycée Marseilleveyre
DUT Network & Telecommunications Computer Systems Networking and Telecommunications, DUT Network & Telecommunications Computer Systems Networking and Telecommunications at IUT d'Aix Marseille
Master of Engineering (M.Eng.) Computer Engineering, Master of Engineering (M.Eng.) Computer Engineering at Polytech Marseille (ESIL)
🎁 4,800,000+ Unsplash images made available for research and machine learning
Role in this project:
Data Engineer & Database Engineer
Contributions:5 releases, 46 commits, 22 PRs in 1 year 1 month
Contributions summary:Timothy primarily focused on improving the data loading and storage capabilities for the Unsplash datasets. They added instructions and sample code for loading data into PostgreSQL and Pandas, demonstrating expertise in data ingestion and analysis workflows. Furthermore, they designed and implemented the database schema for various datasets, including photos, keywords, collections, colors, and conversions, ensuring data integrity and efficient storage. They also updated the schema to include additional data fields for enhanced data analysis.
Contributions:11 commits, 11 pushes in 2 years 2 months
herokulogsdatadog
Find and Hire Top DevelopersWe’ve analyzed the programming source code of over 60 million software developers on GitHub and scored them by 50,000 skills. Sign-up on Prog,AI to search for software developers.