Balarama Buddharaju

Senior Software Engineer at NVIDIA

Pittsburgh, Pennsylvania, United States
email-iconphone-icongithub-logolinkedin-logotwitter-logostackoverflow-logofacebook-logo
Join Prog.AI to see contacts
email-iconphone-icongithub-logolinkedin-logotwitter-logostackoverflow-logofacebook-logo
Join Prog.AI to see contacts

Summary

🤩
Rockstar
🎓
Top School
Balarama Buddharaju is a Senior Software Engineer in Pittsburgh focused on applying machine learning to computer vision and robotics, with two years of professional experience currently working on TensorRT for deep learning inference at NVIDIA. He previously built learning-based motion forecasting and prediction systems for autonomous vehicles at Motional and developed lidar perception and multi-sensor fusion pipelines at Cyngn. His background blends academic research—teaching deep reinforcement learning at CMU and working on vision-in-the-loop bipedal robot control—with hands-on production engineering for safety-critical systems. Trained in computational design and manufacturing (CMU) and mechanical engineering (BITS Pilani), he brings a rare mix of control, perception, and systems expertise that helps bridge research and real-world autonomous platforms.
code2 years of coding experience
job9 years of employment as a software developer
bookBITS Pilani, Birla Institute of Technology and Science
bookMaster's degree, Computational Design & Manufacturing, Master's degree, Computational Design & Manufacturing at Carnegie Mellon University
github-logo-circle

Github Skills (29)

pytorch10
runtimes10
python10
tensorrt10
inference10
large-language-models10
llm10
nvidia10
deep-learning10
cuda10
cpp10
gpu-acceleration10
audio9
backend9
nlp8

Programming languages (3)

C++RustPython

Github contributions (5)

github-logo-circle
brb-nv/TensorRT-LLM

Mar 2025 - Jul 2026

TensorRT-LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and build TensorRT engines that contain state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT-LLM also contains components to create Python and C++ runtimes that execute those TensorRT engines.
Contributions:1322 pushes, 253 branches in 1 year 4 months
cppinferencelarge-language-modelsllmnvidia
NVIDIA/TensorRT-LLM

Mar 2025 - Jul 2026

TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. TensorRT LLM also contains components to create Python and C++ runtimes that orchestrate the inference execution in a performant way.
Contributions:314 reviews, 109 PRs, 67 pushes in 1 year 3 months
cppinferencelarge-language-modelsllmnvidia
Find and Hire Top DevelopersWe’ve analyzed the programming source code of over 60 million software developers on GitHub and scored them by 50,000 skills. Sign-up on Prog,AI to search for software developers.
Request Free Trial