Max Jakob is a seasoned tech lead and principal engineer with 15 years of experience building scalable search, NLP, and data extraction systems from research prototypes to production. Based in Berlin, he has led engineering and data science teams at Elastic, Ecosia, and SoundCloud, blending hands-on backend development with people leadership and product-facing delivery. His academic background is in NLP with dual master's degrees and a bachelor’s in the field, which underpins deep expertise in language understanding and information extraction. A long-time contributor to DBpedia projects, he has improved core annotation, disambiguation, and Wikipedia extraction pipelines—work that sits at the intersection of knowledge graphs and practical NLP. Colleagues rely on him for clarifying complex data-quality issues and shipping robust integration fixes across large codebases. He combines research-grade thinking with pragmatic engineering to drive measurable improvements in search relevance and data reliability.
15 years of coding experience
12 years of employment as a software developer
Bachelor’s Degree Natural Language Processing, Bachelor’s Degree Natural Language Processing at Heidelberg University
Master’s Degree (double degree) Natural Language Processing, Master’s Degree (double degree) Natural Language Processing at Universität des Saarlandes
Master’s Degree (double degree) Natural Language Processing, Master’s Degree (double degree) Natural Language Processing at Charles University
DBpedia Spotlight is a tool for automatically annotating mentions of DBpedia resources in text.
Role in this project:
Back-end Developer
Contributions:123 commits in 2 years 8 months
Contributions summary:Max primarily contributed to back-end code, focusing on improving and adding to the functionality of the DBpedia Spotlight tool. The commits show changes in Java and Scala files, indicating modifications to core data loading, context extraction, and disambiguation logic. Several commits added libraries and updated the Project Object Model (POM), reflecting dependency management work in the project.
The software used to extract structured data from Wikipedia
Role in this project:
Back-end Developer
Contributions:253 commits, 1 issue in 1 year
Contributions summary:Max made several code changes focused on resolving namespace conflicts, adding comments to Freebase scripts, and deprecating/replacing methods within the codebase. These modifications indicate a focus on improving code clarity and resolving issues related to data extraction and integration. The contributions also included bug fixes related to redirect functionality, suggesting involvement in data quality improvements.
pythonstructured-datastructuredwikipedia
Find and Hire Top DevelopersWe’ve analyzed the programming source code of over 60 million software developers on GitHub and scored them by 50,000 skills. Sign-up on Prog,AI to search for software developers.