Martin Olveyra

Software Development And Data Science

Uruguay
email-iconphone-icongithub-logolinkedin-logotwitter-logostackoverflow-logofacebook-logo
Join Prog.AI to see contacts
email-iconphone-icongithub-logolinkedin-logotwitter-logostackoverflow-logofacebook-logo
Join Prog.AI to see contacts

Summary

🤩
Rockstar
🎓
Top School
Martin Olveyra is a software developer and data scientist with 15 years of experience specializing in web scraping, distributed data mining, and machine learning at Zyte (formerly Scrapinghub). He has made notable open-source contributions to Scrapy and related projects—enhancing extractors, autothrottle, and scraping schemas—which reflect deep expertise in backend systems for large-scale web crawling. Trained as a telecommunications engineer, Martin combines systems-level thinking with cloud and parallel computing practices to build robust, production-grade data pipelines. Beyond professional work, he pursues a rigorous self-directed study in cell biology, genomics, and aging research with the aim of transitioning into bioinformatics and doctoral research. Colleagues describe him as a pragmatic problem-solver who bridges engineering discipline with curiosity-driven scientific learning. Based in Uruguay, he brings a rare blend of long-term scraping platform experience and emerging expertise in deep learning applied to scientific domains.
code16 years of coding experience
job6 years of employment as a software developer
bookTelecommunications Engineering, Telecommunications Engineering at Universidad ORT Uruguay
languagesSpanish, English, French, Portuguese, Japanese
github-logo-circle

Github Skills (19)

application-framework10
lib10
python10
app-framework10
webscraping10
regular-expression10
scrapy10
html-parsing10
web-framework10
http9
xml9
automations8
automation8
json8
unit-testing8

Programming languages (4)

JavaScriptHTMLVim scriptPython

Github contributions (5)

github-logo-circle
scrapy/scrapely

Jul 2011 - Aug 2014

A pure-python HTML screen-scraping library
Role in this project:
userBack-end Developer
Contributions:21 commits, 3 PRs, 3 pushes in 3 years 2 months
Contributions summary:Martin primarily contributed to the `scrapely` library by adding new extractors, refining existing ones, and improving template handling. They implemented a price extractor and enhanced the extraction process by allowing per-template item descriptors. The commits also include bug fixes related to variant detection and the application of extra required attributes, as well as improvements to text extraction logic. Overall, the contributions focused on enhancing the functionality and robustness of the HTML screen-scraping library.
pythonscreen-scraping
scrapy/scrapy

Sep 2011 - Nov 2014

Scrapy, a fast high-level web crawling & scraping framework for Python.
Role in this project:
userBack-end Developer
Contributions:15 commits, 3 comments, 3 issues in 3 years 2 months
Contributions summary:Martin primarily contributed to the Scrapy framework, focusing on improving its functionality and addressing existing issues. Their work involved enhancing the autothrottle extension by introducing features like configurable maximum concurrency and minimum download delay. They also made changes to improve the flexibility of the start requests and sitemap spiders. Furthermore, they implemented settings support for handling HTTP error defaults and added an FTP handler.
crawlingpythonscrapyscrapingframework
Find and Hire Top DevelopersWe’ve analyzed the programming source code of over 60 million software developers on GitHub and scored them by 50,000 skills. Sign-up on Prog,AI to search for software developers.
Request Free Trial