Martin Olveyra is a software developer and data scientist with 15 years of experience specializing in web scraping, distributed data mining, and machine learning at Zyte (formerly Scrapinghub). He has made notable open-source contributions to Scrapy and related projects—enhancing extractors, autothrottle, and scraping schemas—which reflect deep expertise in backend systems for large-scale web crawling. Trained as a telecommunications engineer, Martin combines systems-level thinking with cloud and parallel computing practices to build robust, production-grade data pipelines. Beyond professional work, he pursues a rigorous self-directed study in cell biology, genomics, and aging research with the aim of transitioning into bioinformatics and doctoral research. Colleagues describe him as a pragmatic problem-solver who bridges engineering discipline with curiosity-driven scientific learning. Based in Uruguay, he brings a rare blend of long-term scraping platform experience and emerging expertise in deep learning applied to scientific domains.
16 years of coding experience
6 years of employment as a software developer
Telecommunications Engineering, Telecommunications Engineering at Universidad ORT Uruguay
Contributions:21 commits, 3 PRs, 3 pushes in 3 years 2 months
Contributions summary:Martin primarily contributed to the `scrapely` library by adding new extractors, refining existing ones, and improving template handling. They implemented a price extractor and enhanced the extraction process by allowing per-template item descriptors. The commits also include bug fixes related to variant detection and the application of extra required attributes, as well as improvements to text extraction logic. Overall, the contributions focused on enhancing the functionality and robustness of the HTML screen-scraping library.
Scrapy, a fast high-level web crawling & scraping framework for Python.
Role in this project:
Back-end Developer
Contributions:15 commits, 3 comments, 3 issues in 3 years 2 months
Contributions summary:Martin primarily contributed to the Scrapy framework, focusing on improving its functionality and addressing existing issues. Their work involved enhancing the autothrottle extension by introducing features like configurable maximum concurrency and minimum download delay. They also made changes to improve the flexibility of the start requests and sitemap spiders. Furthermore, they implemented settings support for handling HTTP error defaults and added an FTP handler.
crawlingpythonscrapyscrapingframework
Find and Hire Top DevelopersWe’ve analyzed the programming source code of over 60 million software developers on GitHub and scored them by 50,000 skills. Sign-up on Prog,AI to search for software developers.