Summary
Yinfei Yang is a machine learning and multimodal systems leader with a decade of experience building production-grade vision and language models across industry leaders and startups. As Co-Founder and Chief Multimodal Architect at Elorian AI, he is focused on next-generation intelligence that bridges computer vision and NLP, drawing on research leadership roles at Apple and Google. His work spans high-impact projects like ALIGN, MURAL, multimodal LLM capabilities (Ferret), and AFM multimodel research, demonstrating a rare blend of foundational research and product delivery. Prior roles at Redfin and Amazon grounded his expertise in applied computer vision for real-world products, and he holds advanced graduate training in computer vision, robotics, and NLP from the University of Pennsylvania. Based in Mountain View, he combines deep research authorship with hands-on engineering and entrepreneurship, often translating academic advances into scalable systems. A not-obvious strength is his consistent track record of moving from core-model innovations to usable multimodal products across both hyperscalers and startups.
10 years of coding experience
12 years of employment as a software developer
Master's degree (Ph.D quite), Computer Vision, Robotics, NLP, Master's degree (Ph.D quite), Computer Vision, Robotics, NLP at University of Pennsylvania
Chinese, English