Hady Elsahar

Staff Research Scientist at Meta

Berlin, Germany
email-iconphone-icongithub-logolinkedin-logotwitter-logostackoverflow-logofacebook-logo
Join Prog.AI to see contacts
email-iconphone-icongithub-logolinkedin-logotwitter-logostackoverflow-logofacebook-logo
Join Prog.AI to see contacts

Summary

👤
Senior
🎓
Top School
Hady Elsahar is a Staff Research Scientist at Meta AI (FAIR) based in Berlin with 13 years of experience in NLP and ML, focused on generative AI, controlled natural language generation, and content provenance at scale. He leads research and engineering efforts on watermarking and provenance systems that are integrated into production pipelines serving billions of media assets daily, and his work regularly appears in top venues. His background spans academic rigor—a PhD in NLP and ML with influential publications and projects like Scribe for underserved languages—to practical systems work, including refactoring core parsing components for the widely used DBpedia extraction framework. Comfortable bridging research and production, he has steered teams at NAVER LABS Europe and Meta while contributing hands-on to large open-source and industrial codebases. An underappreciated thread through his career is applying generation and summarization techniques to reduce language gaps and tailor content for diverse, low-resource communities.
code13 years of coding experience
job7 years of employment as a software developer
bookDoctor of Philosophy - PhD Natural Language Processing - Machine Learning, Doctor of Philosophy - PhD Natural Language Processing - Machine Learning at Université de Lyon
bookMaster's degree Informatics, Master's degree Informatics at Nile University - NU
bookBachelor's degree Computer Engineering, Bachelor's degree Computer Engineering at Ain sham university
languagesEnglish, Arabic, French
stackoverflow-logo

Stackoverflow

Stats
2,131reputation
243kreached
22answers
46questions
github-logo-circle

Github Skills (16)

data-extraction10
json-parser10
json10
jsonp10
wikidata10
parsing10
scala10
semantic-web6
printing6
html6
css6
nlp6
matlab6
sparql6
signal-processing6

Programming languages (8)

CScalaTeXJavaScriptLuaHTMLJupyter NotebookPython

Github contributions (5)

github-logo-circle
dbpedia/extraction-framework

Dec 2013 - Mar 2014

The software used to extract structured data from Wikipedia
Role in this project:
userBack-end Developer
Contributions:50 commits in 2 months
Contributions summary:Hady primarily focused on refactoring the core components of the extraction framework, specifically targeting the Wikidata JSON parsing functionality. Their work involved modifying the JSON parsing logic to align with changes in the Wikidata data format. They also addressed type conflicts and modified the extraction of language links and labels within the JSON data, demonstrating a deep understanding of the underlying data structures and parsing processes.
pythonstructured-datastructuredwikipedia
hadyelsahar/RE-NLG-Dataset

Jan 2017 - Jan 2019

T-Rex : A Large Scale Alignment of Natural Language with Knowledge Base Triples
Contributions:224 commits, 2 PRs, 97 pushes in 2 years
corpus-linguisticsknowledge-basetriplesdbpedianatural-language-processing
Find and Hire Top DevelopersWe’ve analyzed the programming source code of over 60 million software developers on GitHub and scored them by 50,000 skills. Sign-up on Prog,AI to search for software developers.
Request Free Trial