Ankush Khanna

Member Of Technical Staff at Cohere

Berlin, Germany
email-iconphone-icongithub-logolinkedin-logotwitter-logostackoverflow-logofacebook-logo
Join Prog.AI to see contacts
email-iconphone-icongithub-logolinkedin-logotwitter-logostackoverflow-logofacebook-logo
Join Prog.AI to see contacts

Summary

🤩
Rockstar
🎓
Top School
Ankush Khanna is a Berlin-based Member of Technical Staff with 11 years of experience building scalable data and streaming systems across startups and large enterprises. He has led data engineering teams and shipped production pipelines using Kafka, Spark, Flink, Kafka Connect, and cloud data lakes while mentoring engineers on best practices. His recent roles at Cohere, Billie, Shopify and Wayfair reflect a blend of hands-on implementation and technical leadership across real-time and batch architectures. Ankush is an active contributor and co-instructor to the popular DataTalksClub Data Engineering Zoomcamp, where he added BigQuery ML examples and end-to-end Kafka+Avro pipelines. Comfortable in Scala and Python, he brings a pragmatic focus on partitioning, performance and operationalizing ML workflows—often tackling tricky geodata and streaming challenges. Curious by nature, he pairs deep systems knowledge with a habit of teaching others, which surfaces in both open-source work and internal guild leadership.
code11 years of coding experience
job11 years of employment as a software developer
bookBachelor of Engineering (BE), Computer Science, Bachelor of Engineering (BE), Computer Science at Maharshi Dayanand University
bookHigh School, 10+2, High School, 10+2 at Chinmaya Vidyalaya, New Delhi
bookMaster's Degree, Computer Science, Master's Degree, Computer Science at The University of Bonn
languagesHindi, English
github-logo-circle

Github Skills (14)

kafka10
bigquery10
data-pipeline10
data-pipelines10
sql10
data-warehouse10
datamart10
data-engineering10
dockers9
docker9
avro9
dbt8
python8
spark4

Programming languages (5)

JavaCSSScalaGoJupyter Notebook

Github contributions (5)

github-logo-circle
Data Engineering Zoomcamp is a free 9-week course on building production-ready data pipelines. Join the course here 👇🏼
Role in this project:
userData Engineer
Contributions:3 reviews, 48 commits, 37 PRs in 1 year 3 months
Contributions summary:Ankush primarily contributed to data engineering tasks within the repository. Their work included writing SQL queries for data warehousing in BigQuery, creating external and partitioned tables, and demonstrating the impact of partitioning on query performance. They also added examples of machine learning models in BigQuery and implemented and tested a data pipeline using Kafka and Avro. Furthermore, the user worked on integrating various stream processing examples using Faust.
data-engineeringdata-pipelinedata-pipelineskafkaspark
Contributions:12 PRs, 37 pushes, 12 branches in 3 years 1 month
Find and Hire Top DevelopersWe’ve analyzed the programming source code of over 60 million software developers on GitHub and scored them by 50,000 skills. Sign-up on Prog,AI to search for software developers.
Request Free Trial