Stijn Van Dongen

Principal Data Operations Engineer

Cambridge, England, United Kingdom
email-iconphone-icongithub-logolinkedin-logotwitter-logostackoverflow-logofacebook-logo
Join Prog.AI to see contacts
email-iconphone-icongithub-logolinkedin-logotwitter-logostackoverflow-logofacebook-logo
Join Prog.AI to see contacts

Summary

👤
Senior
🎓
Top School
Stijn Van Dongen is a Principal Data Operations Engineer based in Cambridge with over eight years of experience building high-performance bioinformatics software and data infrastructure. He specializes in algorithm and workflow development for large-scale genomic data, leveraging C, R, Nextflow, and Linux batch systems to deliver robust, distributed pipelines. His work includes authoring widely used scientific tools—MCL, Sylamer and contributing to Kraken—demonstrating a rare blend of discrete mathematics, network analysis and practical engineering. Comfortable across low-level HPC and scripting ecosystems (Perl, Python, Bash), he also emphasizes clear communication through tutoring, documentation and cross-disciplinary collaboration. A PhD in Computational and Applied Mathematics informs his rigorous approach to algorithm design and statistical computing. Colleagues rely on him to tame messy, massive datasets and translate complex methods into repeatable production workflows.
code8 years of coding experience
job21 years of employment as a software developer
bookMaster's degree Mathematics, Master's degree Mathematics at Eindhoven University of Technology
bookDoctor of Philosophy - PhD Computational and Applied Mathematics, Doctor of Philosophy - PhD Computational and Applied Mathematics at Utrecht University
github-logo-circle

Github Skills (69)

slurm10
dataflow10
seq10
weighted10
pipeline-framework10
genomics10
clustering10
reproducible-science10
next-generation-sequencing10
singularity10
hierarchical-clustering10
nextflow10
htslib10
pipeline10
rna-seq10

Programming languages (12)

JavaRC++ShellCOCamlTeXNextflow

Github contributions (5)

github-logo-circle
micans/mcl

Apr 2021 - Dec 2022

MCL, the Markov Cluster algorithm, also known as Markov Clustering, is a method and program for clustering weighted or simple networks, a.k.a. graphs.
Contributions:2 reviews, 249 commits, 2 PRs in 1 year 8 months
methodgraph-clusteringmarkov-clusteringclusteringnearest-neighbors
micans/cimfomfa

Apr 2021 - Sep 2022

Contributions:25 commits, 16 pushes, 1 branch in 1 year 6 months
Find and Hire Top DevelopersWe’ve analyzed the programming source code of over 60 million software developers on GitHub and scored them by 50,000 skills. Sign-up on Prog,AI to search for software developers.
Request Free Trial
Stijn Van Dongen - Principal Data Operations Engineer