Job Closed
This listing is no longer active.
Bring everyone the inspiration to create a life they love.
Principal Machine Learning Engineer
Location
United States
Posted
190 days ago
Salary
0
Seniority
Lead
Job Description
Principal Machine Learning Engineer
This description is a summary of our understanding of the job description. Click on 'Apply' button to find out more. Role Description We are looking for a Principal Machine Learning Engineer, a senior technical visionary, to be the Principal Technical Lead for the Growth Engineering team, responsible for setting up overall technical strategy, unified technical architecture and defining a roadmap for industry leading methodology for user and engagement growth. Strong hands-on machine learning background including deep learning architectures, generative AI, and large scale deployment and measurement of ML systems is required. As the Principal Machine Learning Engineer you'll be responsible for the technical direction, strategy and health of our Growth Engineering org. You'll ensure that our technology can deliver on the business/product requirements necessary to keep Pinterest as an engaging platform for everyone. This means working with other leads to set and execute a long-term strategy for the overall engagement growth of Pinterest, aligning the strategy with other internal partners where it makes sense and communicating to leadership our current status and path to having world-class Growth capabilities. You'll also foster a healthy community where all ML engineers can learn best practices, collaborate effectively and understand our technical direction. What you’ll do: - Develop strong partnerships with product teams to understand and proactively address future technology needs and current developer pain points. - Champion and drive large-scale, cross-functional initiatives that grow user visitation and engagement depth of our platform. - Act as the ultimate "advocate" for engineers on Growth including representing needs to leadership and prioritizing projects on the platform teams that ensure high quality capabilities and a world-class Pinner experience. - Scale your leadership through both direct mentorship and via best practices, processes, training and tools. - Ensure solid technical plans are in place for projects within Growth via direct review or delegation. - Be the technical point of contact for decisions that impact the whole Pinterest platform via the Growth initiatives and for cross-functional partners for an 125+ member org. Qualifications - Deep expertise building large scale ML systems at scale with modern frameworks. - Knowledge of (and a passion for) building responsible and quality first discovery surfaces to drive user visitations. - Track record of innovating and delivering large, cross-functional projects across multiple organizations. - Strong written and verbal communication skills and proven ability to collaborate cross-functionally. - Degree in Computer Science, Machine Learning, Statistics or related field. - 10+ years of professional experience as a hands-on engineer and technical leader leading multiple projects. Requirements - This role will need to be in the office for in-person collaboration 1-2 times every 6-months and therefore can be situated anywhere in the country. - This position is not eligible for relocation assistance. Benefits - Visit our PinFlex page to learn more about our working model.
Job Requirements
- Deep expertise building large scale ML systems at scale with modern frameworks.
- Knowledge of (and a passion for) building responsible and quality first discovery surfaces to drive user visitations.
- Track record of innovating and delivering large, cross-functional projects across multiple organizations.
- Strong written and verbal communication skills and proven ability to collaborate cross-functionally.
- Degree in Computer Science, Machine Learning, Statistics or related field.
- 10+ years of professional experience as a hands-on engineer and technical leader leading multiple projects.
- This role will need to be in the office for in-person collaboration 1-2 times every 6-months and therefore can be situated anywhere in the country.
- This position is not eligible for relocation assistance.
Benefits
- Visit our PinFlex page to learn more about our working model.
Related Guides
Related Job Pages
More Machine Learning Engineer Jobs
Machine Learning Engineer – Platform
Artera.netArtera is a Swiss ISP that produces premium hosting and cloud services.
• Work on the AI Platform team focusing on scalable and efficient pipelines for model training, evaluation, and data processing • Build and evolve core libraries used by AI scientists to develop, launch, and monitor AI products • Optimize GPU and CPU efficiency and data throughput of large-scale foundation models • Ensure Artera’s observability infrastructure provides a clear picture of model performance optimization
Machine Learning Engineer
CalendlyThe scheduling automation platform for eliminating the back-and-forth emails to find the perfect time — and so much more
• Own ML powered features from design through deployment • Understand and share domain knowledge • Prioritize your work independently • Proactively seek and offer support to teammates • Understand and troubleshoot our deployment pipelines • Use our monitoring and observability tools • Serve as a subject matter expert for the features and services
ML Engineer, Foundation Model Infrastructure
WaymoWaymo is an autonomous driving technology company creating a new way forward in mobility.
This description is a summary of our understanding of the job description. Click on 'Apply' button to find out more. Role Description The mission of the Waymo AI Foundations team is to develop machine learning solutions addressing open problems in autonomous driving, towards the goal of safely operating Waymo vehicles in dozens of cities and under all driving conditions. This role follows a hybrid work schedule and you will report to a Senior Research Scientist. - Build and operate the petabyte-scale data systems and ML pipelines at the heart of Waymo's foundation model development - Shepherd cutting-edge foundation models from research prototypes to robust components within the Waymo Driver - Create the automated infrastructure for rigorously benchmarking, continuously monitoring, and safely releasing models - Wield large-scale compute and frameworks like Flume and JAX to process massive datasets and train/deploy complex models - Drive significant leaps in the speed, reliability, and efficiency of the end-to-end ML development lifecycle - Partner with AI Foundations, ML, and Platform experts to transform model innovations into tangible on-road improvements Qualifications - Masters degree in Computer Science, Machine Learning, Robotics, similar technical field of study, or equivalent practical experience - Proficiency in Python - Proficiency in C++ - Familiarity with one of the modern deep learning frameworks (e.g. Pytorch, JAX, Tensorflow) - Experience building or maintaining large-scale data pipelines or ML infrastructure (e.g., Flume, Spark, Borg, Kubeflow) Requirements - Strong hands-on SWE skills, able to drive development of large, complex shared codebases - Experience in AV planning and related research - Experience designing and building distributed systems or MLOps platforms (e.g., model versioning, experiment tracking, CI/CD for ML) - Prior work in an industrial or research setting developing methodologies for the evaluation of ML models Benefits - Eligible to participate in Waymo’s discretionary annual bonus program - Equity incentive plan - Generous Company benefits program, subject to eligibility requirements Salary Range $204,000 — $259,000 USD
Member of Technical Staff, Inference
RunwayBusiness financials got stuck in the 15th century so we're showing them today’s computers 🖥
This description is a summary of our understanding of the job description. Click on 'Apply' button to find out more. Role Description We're looking for an ML infrastructure engineer to bridge the gap between research and production at Runway. You'll work directly with our research teams to productionize cutting-edge generative models—taking checkpoints from training to staging to production, ensuring reliability at scale, and building the infrastructure that enables fast iteration. You'll be embedded within research teams, providing platform support throughout the entire model development lifecycle. Your work will directly impact how quickly we can ship new models and features to millions of users. A peek at our technical stack - API endpoints for real-time collaboration and media asset management written in TypeScript, running in ECS containers on AWS Fargate. - Leverage multiple AWS-native components, such as S3, CloudFront, Lambda, Kinesis, and SQS. - Inference backend written in Python (PyTorch, TorchScript), deployed across multiple clusters/cloud providers. - Use Kubernetes for container orchestration, with k8s-native components such as Flyte, Kueue, and Kyverno for efficient job orchestration. - Invest in Prometheus and Grafana for monitoring, and Terraform to manage infrastructure. Qualifications - 4+ years of experience running ML model inference at scale in production environments. - Strong experience with PyTorch and multi-GPU inference for large models. - Experience with Kubernetes for ML workloads—deploying, scaling, and debugging GPU-based services. - Comfortable working across multiple cloud providers and managing GPU driver compatibility. - Experience with monitoring and observability for ML systems (errors, throughput, GPU utilization). - Self-starter who can work embedded with research teams and move fast. - Strong systems thinking and pragmatic approach to production reliability. - Humility and open-mindedness; at Runway we love to learn from one another. Requirements - Experience building custom inference frameworks or serving systems (Nice to Have). - Deep understanding of distributed training and inference patterns (FSDP, data parallelism, tensor parallelism) (Nice to Have). - Ability to debug low-level issues: NCCL networking problems, CUDA errors, memory leaks, performance bottlenecks (Nice to Have). - Experience with diffusion models or video generation systems (Nice to Have). - Knowledge of real-time or latency-sensitive ML applications (Nice to Have). Benefits - Salary range: $240,000 - $290,000. - Commitment to creating a space where employees can bring their full selves to work and have equal opportunity to succeed. Company Description Runway strives to recruit and retain exceptional talent from diverse backgrounds while ensuring pay equity for our team. Our salary ranges are based on competitive market rates for our size, stage, and industry, and salary is just one part of the overall compensation package we provide. There are many factors that go into salary determinations, including relevant experience, skill level and qualifications assessed during the interview process, and maintaining internal equity with peers on the team. The range shared below is a general expectation for the function as posted, but we are also open to considering candidates who may be more or less experienced than outlined in the job description. In this case, we will communicate any updates in the expected salary range. Lastly, the provided range is the expected salary for candidates in the U.S. Outside of those regions, there may be a change in the range, which again, will be communicated to candidates. We're excited to be recognized as a best place to work by Crain's, InHerSight, BuiltIn NYC, and INC.




