Rohan Badlani is a research scientist at NVIDIA AI Research with 11 years of experience building and deploying deep learning systems for real-time, multi-modal applications. A Stanford MS in Computer Science with dual specializations in AI and Information Management, he has bridged industry and academia—contributing to audio event detection at CMU, ML-for-systems at TUM, and semantic knowledge graphs at TIFR before product work at Microsoft. At NVIDIA he helped develop the Audio Foundation Model Fugatto, invented a model-composition inference technique for privacy-preserving expert models, and co-created scalable unsupervised TTS alignment and multilingual voice-preserving TTS systems. Comfortable taking ideas from research to production, he blends graph and transformer approaches (applied previously to source code at X) and has a track record of shipping high-impact, real-time ML products. Based in Palo Alto, he pairs systems-level engineering pragmatism with deep research rigor, often tackling problems that require novel algorithmic and deployment trade-offs.
12 years of coding experience
4 years of employment as a software developer
BITS Pilani, Birla Institute of Technology and Science
Master's degree Computer Science, Master's degree Computer Science at Stanford University
Stanford Ignite Business Administration and Management General, Stanford Ignite Business Administration and Management General at Stanford University Graduate School of Business
This is a project involving learning sounds of different objects and evolve relationships between objects.
Contributions:44 commits, 15 PRs, 21 pushes in 1 year 11 months
deep-learningsoundsmachine-learningevolve
Find and Hire Top DevelopersWe’ve analyzed the programming source code of over 60 million software developers on GitHub and scored them by 50,000 skills. Sign-up on Prog,AI to search for software developers.