Sam Mcveety is a Distinguished Engineer at Google with 13 years of experience shaping cloud data analytics and leading technical strategy for Internet-scale, data-parallel processing tools. Based in Seattle, he mentors technical leads across thousands of engineers to balance tactical delivery with long-term investments that democratize data access and insights. His hands-on contributions to high-profile open-source projects like Apache Beam and Google Cloud Dataflow reflect deep backend expertise—improving SDKs, I/O robustness, and runtime configuration for large-scale streaming and batch pipelines. Sam pairs rigorous technical leadership with a commitment to equity and inclusion, running training and team initiatives to build more diverse engineering cultures. With advanced degrees from MIT and a Master of Jurisprudence from the University of Washington, he blends technical depth, legal perspective, and systems-level thinking to guide complex, cross-functional programs.
13 years of coding experience
12 years of employment as a software developer
Master of Jurisprudence, Master of Jurisprudence at University of Washington
Master of Engineering - MEng, Computer Science, Master of Engineering - MEng, Computer Science at Massachusetts Institute of Technology
Google Cloud Dataflow provides a simple, powerful model for building both batch and streaming parallel data processing pipelines.
Role in this project:
Back-end Developer
Contributions:53 commits, 20 PRs, 32 comments in 2 years
Contributions summary:Sam's contributions primarily involved fixing issues related to Google Cloud Storage (GCS) staging and implementing AvroIO validation within the Dataflow Java SDK. Their work included refactoring the retry mechanism for uploading classpath elements to GCS and adding validation checks for input and output paths associated with AvroIO operations. Additionally, the user was involved in a rollback of a previous change related to the First.of transform. The commits show a focus on improving the reliability and functionality of core data processing components.
Apache Beam is a unified programming model for Batch and Streaming data processing.
Role in this project:
Back-end Developer
Contributions:43 commits, 68 PRs, 151 comments in 1 year 1 month
Contributions summary:Sam's contributions primarily revolve around enhancing Apache Beam's core Java SDK. Their work includes adding a property name to the `RuntimeValueProvider` class and implementing a utility for handling JSON option manipulation. The user also added a new experimental ServiceAccount option. Furthermore, they introduced a NestedValueProvider and enhanced TextIO.Read to support ValueProvider.
golangpythonstreaming-databeambatch
Find and Hire Top DevelopersWe’ve analyzed the programming source code of over 60 million software developers on GitHub and scored them by 50,000 skills. Sign-up on Prog,AI to search for software developers.