Summary
Haoxuan You is an AI research scientist with eight years of experience building cutting-edge multimodal and vision-language systems, currently driving multimodal post-training at Meta after leading vision-language continual pretraining for an Apple foundation model. A Columbia Ph.D. candidate in computer science, he has held research roles across Apple, Google, and Microsoft, blending deep academic rigor with product-scale model engineering. He specializes in continual pretraining and multimodal alignment, with hands-on experience taking research prototypes toward foundation-model deployments. Based in Menlo Park, he combines strong systems intuition from industry internships and academic collaborations with a knack for translating complex research into robust, scalable models.
8 years of coding experience
2 years of employment as a software developer
Bachelor of Engineering - BE, Electronic Information Engineering, Bachelor of Engineering - BE, Electronic Information Engineering at Xidian University
Doctor of Philosophy - PhD, Computer Science, Doctor of Philosophy - PhD, Computer Science at Columbia University