Youjin Chung is a Machine Learning Engineer with a decade of experience building and deploying LLM and generative AI solutions across startups and regulated enterprise environments. He has shipped production RAG systems, low-latency speech-to-text pipelines (sub-300ms) and optimized models for real-time edge and cloud inference using quantization, pruning, ONNX Runtime and TensorRT. Comfortable from research to production, he architects scalable ML pipelines on AWS, FastAPI and Docker while exploring MLOps tooling like MLflow to close the loop on continuous optimization. His work has improved response accuracy and throughput for customer support and financial document workflows, and he has hands-on experience benchmarking top-tier models (GPT-4, Claude) for domain-specific tasks. Based in New York and with interdisciplinary training spanning cognitive science and interactive telecommunications, he brings a research-minded approach to practical product impact. An unusual strength is his mix of GAN compression and synthetic data generation experience alongside large-scale LLM deployment, enabling end-to-end solutions from dataset creation to low-latency serving.
10 years of coding experience
6 years of employment as a software developer
Master of Science - MS Computer Science, Master of Science - MS Computer Science at Georgia Institute of Technology
Masters Data Science, Masters Data Science at The Graduate Center, City University of New York
Master of Professional Studies Interactive Telecommunications, Master of Professional Studies Interactive Telecommunications at New York University
Bachelor of Engineering (B.E.) Electrical Electronics and Communications Engineering, Bachelor of Engineering (B.E.) Electrical Electronics and Communications Engineering at Korea University
Find and Hire Top DevelopersWe’ve analyzed the programming source code of over 60 million software developers on GitHub and scored them by 50,000 skills. Sign-up on Prog,AI to search for software developers.