Summary
Chao Jia is a Senior Staff Researcher in GenAI based in the San Francisco Bay Area with six years of industry experience building and productionizing large multimodal models. He has driven multimodal vision research and deployment across top-tier labs and products at Google DeepMind, Google, Apple, and Waymo, including core contributions to Gemini and Google's ALIGN and MUM foundation models. Chao combines deep academic training (PhD, UT Austin) with hands-on systems engineering, shipping multimodal perception and reasoning features in consumer APIs and large-scale products like multisearch and Gemini API. He has repeatedly led cross-functional teams to convert research prototypes into robust post-training and API-grade capabilities, spanning image, video, and audio modalities. A less obvious strength is his track record of originating foundation-model efforts within product organizations (e.g., Waymo’s vision-language FM) and scaling them into multi-team priorities. He excels at bridging cutting-edge research with product-quality engineering to deliver multimodal AI that is both innovative and deployable.
6 years of coding experience
15 years of employment as a software developer
Doctor of Philosophy (Ph.D.) Electrical and Computer Engineering, Doctor of Philosophy (Ph.D.) Electrical and Computer Engineering at The University of Texas at Austin
B.S. Electrical and Computer Engineering, B.S. Electrical and Computer Engineering at Tsinghua University