Mykola Melnyk is a Principal AI Engineer with 13+ years of hands-on experience and a 15+-year background in ML/data engineering, specializing in large-scale Spark pipelines, Visual NLP/OCR, and privacy-preserving data anonymization. He led development of Spark OCR and contributed an open-source spark-pdf datasource, bridging Spark internals with real-world document processing across PDFs, DOCX and DICOM. Comfortable from low-level C/C++ systems to modern LLM stacks (LLama, Gemini, Hugging Face, LangChain), he builds end-to-end solutions that scale and meet regulatory requirements like GDPR and HIPAA. Mykola has deep domain expertise in healthcare, biotech and medtech, and is the co-founder of an AI-powered PDF redaction product delivered as online, API and on-premise options. He combines research-grade ML (PyTorch, CV, NLP) with production engineering (Scala, PySpark, Databricks) and a track record of timely delivery and transparent communication. Based in Warsaw, he prefers long-term collaborations and often converts complex compliance challenges into automated, auditable pipelines.
13 years of coding experience
19 years of employment as a software developer
National Technical University "Kharkiv Polytechnic Institute"
Contributions:4 PRs, 54 pushes, 63 branches in 6 months
declarativeclassifiermachine-learning
Find and Hire Top DevelopersWe’ve analyzed the programming source code of over 60 million software developers on GitHub and scored them by 50,000 skills. Sign-up on Prog,AI to search for software developers.