Senior Machine Learning Engineer, Training

Posted 4 Days Ago
Be an Early Applicant
Mountain View, CA
192K-243K Annually
Mid level
Automotive
The Role
As a Senior Machine Learning Engineer, you will develop infrastructure components for distributed training, implement automation solutions for monitoring and scaling, diagnose system health issues, identify performance bottlenecks, and enhance developer experience within Waymo's scalable ML framework.
Summary Generated by Built In

Waymo is an autonomous driving technology company with the mission to be the most trusted driver. Since its start as the Google Self-Driving Car Project in 2009, Waymo has focused on building the Waymo Driver—The World's Most Experienced Driver™—to improve access to mobility while saving thousands of lives now lost to traffic crashes. The Waymo Driver powers Waymo One, a fully autonomous ride-hailing service, and can also be applied to a range of vehicle platforms and product use cases. The Waymo Driver has provided over one million rider-only trips, enabled by its experience autonomously driving tens of millions of miles on public roads and tens of billions in simulation across 13+ U.S. states.

The Waymo ML Infrastructure team works with Research and Production teams to develop models in Perception and Planning that are core to our autonomous driving software. We ensure our partners by offering the best solutions for the entire model development lifecycle. These solutions are developed in close collaboration with teams at Google. They are geared towards both scaling models and solving problems unique to ML for autonomous driving.

We develop a set of libraries and tools that enhance TensorFlow and JAX, and address scalability, reliability, and performance challenges faced by Waymo's ML practitioners: increasing ML accelerator efficiency, fine-tuning multimodal LLMs for autonomous driving tasks, and validating newly trained DNNs when deployed into the full onboard software stack.

In this Hybrid role, you will report to our TLM of Machine Learning Training.

You will:

  • Develop the infrastructure components necessary for distributed training
  • Implement automation solutions for provisioning, deployment, monitoring, and scaling of distributed training infrastructure
  • Monitor system health, diagnose and perform routine maintenance tasks to ensure the reliability of the distributed training infrastructure.
  • Identify performance bottlenecks and optimization opportunities
  • Improve the developer experience and performance of our scalable ML framework

You have:

  • Bachelor's degree in Computer Science, Engineering, or related field, or 4+ years equivalent experience
  • Experience building distributed systems for production environments.
  • Solid Python or C++ skills
  • Prior experience with Machine Learning frameworks (e.g., TensorFlow, PyTorch) and distributed training algorithms

We prefer:

  • Practical familiarity using ML accelerator profiling tools to uncover performance bottlenecks
  • Experience deploying and managing distributed systems in cloud environments
  • Knowledge of optimization and deep learning algorithms

#LI-Hybrid

The expected base salary range for this full-time position across US locations is listed below. Actual starting pay will be based on job-related factors, including exact work location, experience, relevant training and education, and skill level. Your recruiter can share more about the specific salary range for the role location or, if the role can be performed remote, the specific salary range for your preferred location, during the hiring process. 

Waymo employees are also eligible to participate in Waymo’s discretionary annual bonus program, equity incentive plan, and generous Company benefits program, subject to eligibility requirements. 

Salary Range

$192,000$243,000 USD

Top Skills

C++
Python
The Company
Mountain View, CA
2,359 Employees
On-site Workplace
Year Founded: 2009

What We Do

Waymo is an autonomous driving technology company with a mission to make it safe and easy for people and things to move around. With the Waymo Driver, we can improve the world’s mobility while saving thousands of lives.
Waymo reaches out to candidates from official channels only (e.g. directly from @waymo.com email addresses, or through our recruiters or sourcers who are noted as such on LinkedIn). We do not contact candidates about career opportunities through instant messaging apps like Telegram, email addresses from domains other than waymo.com (such as Gmail addresses), direct messages on Twitter, Facebook, and Instagram, or text messages. Visit waymo.com to check out our official job listings.

Similar Jobs

Remote
3 Locations
431 Employees
174K-261K Annually

Atlassian Logo Atlassian

Principal Engineer, Distribution at Loom

Cloud • Information Technology • Productivity • Security • Software • App development • Automation
Remote
San Francisco, CA, USA
11000 Employees
171K-274K Annually

Crunchyroll Logo Crunchyroll

Staff Software Engineer, Content Delivery

Digital Media • eCommerce • Gaming • Mobile • News + Entertainment
Hybrid
San Francisco, CA, USA
1200 Employees
190K-239K Annually

Block Logo Block

Software Engineer (Backend), Buyer Foundations

Blockchain • eCommerce • Fintech • Payments • Software • Financial Services • Cryptocurrency
Remote
Hybrid
7 Locations
12000 Employees
139K-245K Annually

Similar Companies Hiring

Chamberlain Group Thumbnail
Software • PropTech • Mobile • Internet of Things • Hardware • Automotive • App development
Oak Brook, IL
5637 Employees
Cox Enterprises Thumbnail
Software • Other • Information Technology • Greentech • Cybersecurity • Cloud • Automotive
Atlanta, GA
50000 Employees
UL Solutions Thumbnail
Software • Renewable Energy • Professional Services • Energy • Consulting • Chemical • Automotive
Chicago, IL
15000 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account