Summary
Srishti Saha is a Senior Data Scientist and NLP engineer with 7 years of experience building production-grade ML and metadata-driven data pipelines for reputation intelligence, survey analytics, and media monitoring. At RepTrak she designs scalable Airflow/CI-CD workflows on AWS, maintains production LLM and RAG systems, and translates legacy analytics into reproducible Python pipelines that serve 100+ clients. Her background spans decision science at Mu Sigma, Bayesian modeling in academic research at Duke, and a Wells Fargo internship studying linguistic drift—demonstrating strong grounding in both rigorous statistics and applied NLP. She also volunteers as Chief Data Officer for a gaming nonprofit, where she builds retrieval systems and causal models linking game features to human behavior. Comfortable bridging product, engineering, and client teams, she combines technical depth with an appetite for new languages, blockchain concepts, and cross-domain problem solving.
7 years of coding experience
7 years of employment as a software developer
Master of Science - MS, Interdisciplinary Data Science, Master of Science - MS, Interdisciplinary Data Science at Duke University
Bachelor’s Degree, Electrical, Electronics and Communications Engineering, Bachelor’s Degree, Electrical, Electronics and Communications Engineering at Manipal Institute of Technology
English, Hindi, Bengali, Spanish, Japanese