Summary
Michel De Arruda is a data scientist and machine learning engineer with 11 years of experience who specializes in building end-to-end AI systems and MLOps pipelines, currently leading development of an LLM-based Retrieval-Augmented Generation platform for Pernambuco’s IT agency. He pairs deep expertise in NLP, LangChain, vector databases (Milvus), and production tooling (FastAPI, Docker, AWS/Azure) with practical wins in industry—such as a lead-classification model that boosted conversions by 50% and saved $12M annually. Comfortable across data engineering, feature engineering, and model deployment at scale, he has delivered ML platforms and observability tooling at OLX and architected data lakes and pipelines using Spark, Airflow and AWS services. Earlier in his career he led tech teams for public-sector products that reached millions of users, demonstrating an ability to ship high-impact solutions under resource constraints. Notably, he blends research-grade knowledge from a master’s in systems and ongoing AI specialization with a hands-on track record of turning LLM research into production assistants for government transparency and efficiency.
11 years of coding experience
7 years of employment as a software developer
Technician in Industrial Informatics, Technician in Industrial Informatics at Centro Federal de Educação Tecnológica Celso Suckow da Fonseca
Postgraduate Certificate, Artificial Intelligence and Machine Learning, Artificial Intelligence and Machine Learning, Postgraduate Certificate, Artificial Intelligence and Machine Learning, Artificial Intelligence and Machine Learning at PUC Minas
Master's in Systems and Computer Engineering, Data Science and Machine Learning, Master's in Systems and Computer Engineering, Data Science and Machine Learning at Universidade Federal do Rio de Janeiro
Bachelor's Degree in Computer Science, Data Mining and Data Analisys, Bachelor's Degree in Computer Science, Data Mining and Data Analisys at Universidade Federal Rural do Rio de Janeiro
Portuguese, English