Karan Singh is a pragmatic software engineer with 10 years of experience building scalable full-stack systems across cloud-native, enterprise, and research-oriented environments. Currently at Apple after recent senior roles that drove API modernizations and migrations, he combines Go, Python, TypeScript, and Java expertise with production experience in Kubernetes, Docker, AWS, and GCP. He has shipped high-throughput data pipelines and multi-tenant SaaS platforms with strong security and compliance focus (GDPR/HIPAA), and led codebase migrations that cut latency and costs in half. An active open-source committer and PMC contributor to Apache Nutch and Sparkler, he brings full-stack fixes from Selenium integration to crawler de-duplication, reflecting deep familiarity with crawling and large-scale data ingestion. Comfortable mentoring teams, he pairs hands-on engineering with measurable quality improvements—such as boosting test coverage and dramatically shrinking container build sizes—while staying curious about new tooling and architectures.
10 years of coding experience
4 years of employment as a software developer
Bachelor of Technology - B.Tech Computer Science, Bachelor of Technology - B.Tech Computer Science at SRM Institute of Science and Technology (SRMIST)
Spark-Crawler: Apache Nutch-like crawler that runs on Apache Spark.
Role in this project:
Back-end Developer
Contributions:91 commits, 20 PRs, 46 pushes in 4 years 4 months
Contributions summary:Karan contributed to the core functionality of the Sparkler crawler, focusing on configuration, data processing, and integration with Solr. They made changes to the crawler pipeline, including adding the text extraction and title to CrawlDB. The user refactored code for YAML configuration. They worked on integrating the FetcherJBrowser and improving the De-duplication logic.
Apache Nutch is an extensible and scalable web crawler
Role in this project:
Full-stack Developer
Contributions:7 commits, 4 PRs, 14 comments in 1 month
Contributions summary:Karan's contributions focused on fixing bugs and implementing improvements across the Nutch codebase. They addressed issues related to Selenium integration, specifically focusing on handling timeouts and exceptions in the Selenium web driver. The user also updated the HTMLUnit integration, including fixes for redirects, screenshots, and general build issues. Their work spans both protocol plugins and related libraries, highlighting a full-stack involvement within the Nutch project.
scalableweb-crawlernutchosgiapache
Find and Hire Top DevelopersWe’ve analyzed the programming source code of over 60 million software developers on GitHub and scored them by 50,000 skills. Sign-up on Prog,AI to search for software developers.