Afshin Dehghan is an AI/ML lead with a decade of experience translating cutting-edge research into production systems, currently heading Apple’s Multimodal INTelligence (MINT) group in Cupertino. He has driven multimodal foundation models, agentic AI systems, and large-scale perception tech that power features across Apple products—from FaceID and real-time camera semantic reasoning to 3D perception for Vision Pro. Trained as a PhD computer scientist, Afshin blends deep academic credentials and award-winning research with hands-on startup experience building mobile-friendly vision models and large-scale image and video systems. He excels at bridging bold research and product engineering, often leading cross-disciplinary teams to deliver deployable, privacy-conscious perception and reasoning capabilities. A less obvious strength is his track record of shepherding long-term research (e.g., 3D scene understanding and agentic AI) from early incubation to core platform technologies used company-wide.
10 years of coding experience
6 years of employment as a software developer
Electrical Electronics and Communications Engineering, Electrical Electronics and Communications Engineering at University of Tehran
Doctor of Philosophy (PhD) Computer Science, Doctor of Philosophy (PhD) Computer Science at University of Central Florida
This repo accompanies the research paper, ARKitScenes - A Diverse Real-World Dataset for 3D Indoor Scene Understanding Using Mobile RGB-D Data and contains the data, scripts to visualize and process assets, and training code described in our paper.
Contributions:3 commits, 2 PRs, 8 pushes in 2 months
Find and Hire Top DevelopersWe’ve analyzed the programming source code of over 60 million software developers on GitHub and scored them by 50,000 skills. Sign-up on Prog,AI to search for software developers.