Fahim Khan is a Software Engineer based in California with seven years of industry experience and a current focus as an LLM Performance Engineer at NVIDIA. He brings practical expertise in optimizing large language model performance, bridging low-level systems work with applied ML deployment. Known for tackling performance bottlenecks and scaling inference workloads, he combines engineering rigor with a product-minded approach. Colleagues can expect a pragmatic problem-solver who marries profiling-driven optimization with production reliability.
Contributions:1 release, 9 commits, 3 pushes in 1 year 6 months
Find and Hire Top DevelopersWe’ve analyzed the programming source code of over 60 million software developers on GitHub and scored them by 50,000 skills. Sign-up on Prog,AI to search for software developers.