Summary
Logan Robbins is a Principal AI Architect in the San Francisco Bay Area with 11 years of experience building production-grade AI systems and enterprise platforms across Disney, Apple, Intel, and IBM. He combines deep research in transformer architectures and agentic systems—authoring the Parallel Decoder Transformer paper and releasing models and code on Hugging Face—with hands-on platform engineering for secure, scalable GenAI deployments. Logan has led cross-functional teams to operationalize RAG, MLOps, and observability at scale, translating responsible-AI requirements into actionable governance and deployment patterns. His background spans high-throughput e-commerce and real-time personalization to multi-cloud ML pipelines, emphasizing performance, cost efficiency, and reliability. Notably, he engineered custom CUDA attention kernels and speculative decoding techniques (Speculative Note Conditioning) to tackle coherence drift in parallel decoding. He’s equally fluent at shaping technical strategy for enterprise adoption and shipping the low-level optimizations that make cutting-edge models production-ready.
11 years of coding experience
14 years of employment as a software developer
Calabasas High School
Computer Science, Computer Science at Charles University (Univerzita Karlova)