Marco Lui

Head Of Data Science (post-acquisition By Omio Group)

Greater Melbourne Area Australia
email-iconphone-icongithub-logolinkedin-logotwitter-logostackoverflow-logofacebook-logo
Join Prog.AI to see contacts
email-iconphone-icongithub-logolinkedin-logotwitter-logostackoverflow-logofacebook-logo
Join Prog.AI to see contacts

Summary

🤩
Rockstar
🎓
Top School
Marco Lui is a Head of Data Science with 16 years’ experience leading analytics and ML teams in travel tech, currently shaping data strategy at Rome2rio post-acquisition by Omio. He combines deep NLP research (PhD-level work in language identification) with hands-on engineering, having built production systems like geocoders and transit search optimizations. Marco hires and coaches cross-functional teams of data scientists, analysts and engineers to embed data-driven decision-making across product, acquisition and revenue functions. He contributed to an open-source Python language-identification project implementing multiprocessing tokenizers and experimental skew-based classification, reflecting a focus on pragmatic, performant solutions. Based in Melbourne, he pairs academia-grade modelling expertise with a track record of shipping scalable web services that directly impact user growth and monetization.
code17 years of coding experience
job10 years of employment as a software developer
bookThe University of Melbourne
languagesEnglish, Chinese, Chinese, Italian
stackoverflow-logo

Stackoverflow

Stats
1reputation
0reached
0answers
0questions
github-logo-circle

Github Skills (9)

python10
natural-language-processing9
nlp9
python-multiprocessing9
multiprocessing9
multi-process9
machine-learning8
data-analysis8
webservice8

Programming languages (4)

C#JavaC++Python

Github contributions (5)

github-logo-circle
saffsd/langid.py

Feb 2011 - Jul 2017

Stand-alone language identification system
Role in this project:
userBack-end Developer & Data Scientist
Contributions:241 commits, 7 PRs, 13 pushes in 6 years 6 months
Contributions summary:Marco contributed to the development of a standalone language identification system written in Python. Their work involved the implementation of core components like a tokenizer utilizing multiprocessing for performance and a naive Bayes classifier. They introduced an experimental approach to language identification, including a skew-based method. They also added web service capabilities to the system.
nlpbootstrappingstand-alonestandphonetics
Contributions:5 commits, 1 push in 3 years 5 months
Find and Hire Top DevelopersWe’ve analyzed the programming source code of over 60 million software developers on GitHub and scored them by 50,000 skills. Sign-up on Prog,AI to search for software developers.
Request Free Trial