Summary
Weixing Zhang is a seasoned software engineer with 11+ years building high-performance ML frameworks, distributed training/inference, and GPU-accelerated systems; he currently serves as Member of Technical Staff at OpenAI after leading AI Framework and ONNX Runtime efforts at Microsoft. He brings deep C/C++ expertise, kernel-level Linux knowledge, and hands-on experience with CUDA/OpenCL, GPU architecture, graphics drivers and OpenGL—skills honed through roles at AMD and AWS. Weixing has a track record of performance tuning and productionizing large-scale LLM inference, and he pairs technical delivery with cross-functional alignment and OKR-driven execution. Comfortable bridging low-level systems and ML stacks, he’s as likely to debug kernel memory issues as to optimize model throughput on specialized hardware.
11 years of coding experience
19 years of employment as a software developer
Bachelor Automation, Bachelor Automation at Wuhan University
Master Automation, Master Automation at Shanghai Jiao Tong University
English, Chinese