Ben Caine

Principal Engineer at Google DeepMind

Cambridge, Massachusetts, United States
email-iconphone-icongithub-logolinkedin-logotwitter-logostackoverflow-logofacebook-logo
Join Prog.AI to see contacts
email-iconphone-icongithub-logolinkedin-logotwitter-logostackoverflow-logofacebook-logo
Join Prog.AI to see contacts

Summary

👤
Senior
🎓
Top School
Ben Caine is a Senior Staff Research Engineer at Google DeepMind with 12 years of experience building deep learning systems for computer vision, robotics, and autonomous vehicles. He leads Gemini Vision post-training efforts that enable image and video inputs to multimodal models, after a multi-year stint applying research to Waymo and prototyping an L3 driving system at a startup. His background spans applied ML at scale—from contributing to tensorflow/lingvo integrations with the Waymo Open Dataset and 3D detection pipelines to production-focused research roles across Google Brain and DeepMind. Based in Cambridge, MA, he combines research rigor with hands-on engineering, shipping evaluation metrics and refactors that improve real-world perception stacks. Notably, his career threads academic time-series work and early product-facing roles, showing an ability to move ideas from lab prototypes into operational systems.
code13 years of coding experience
job11 years of employment as a software developer
bookBachelor of Science (B.S.), Computer Engineering, Bachelor of Science (B.S.), Computer Engineering at Northeastern University
github-logo-circle

Github Skills (16)

computer-vision10
machine-learning10
3d-object-detection10
tensorflow10
python10
speech-recognition8
natural-language-processing7
data-engineering7
nlp7
gtts6
freetts6
machine-translation6
speech-to-text6
distribute6
speech6

Programming languages (4)

C++ShellJupyter NotebookPython

Github contributions (5)

github-logo-circle
tensorflow/lingvo

Oct 2019 - Dec 2021

Lingvo
Role in this project:
userML Engineer
Contributions:28 commits, 1 comment in 2 years 2 months
Contributions summary:Ben made significant contributions to the `tensorflow/lingvo` repository, a project focused on machine learning, particularly in the context of computer vision and speech recognition. Their work involved modifying existing models and metrics, including adding error metrics and enabling velocity breakdowns for improved analysis. The user also updated the codebase to integrate with the Waymo Open Dataset, demonstrating an understanding of 3D object detection and related evaluation pipelines. Further contributions included refactoring code and making improvements to existing functionality.
lingvospeech-recognitiontranslationspeech-to-textmachine-translation
bcaine/VLC-Transceiver

Sep 2015 - Feb 2016

Contributions:158 commits, 132 pushes, 3 branches in 4 months
Find and Hire Top DevelopersWe’ve analyzed the programming source code of over 60 million software developers on GitHub and scored them by 50,000 skills. Sign-up on Prog,AI to search for software developers.
Request Free Trial