Waymo logo
WaymoPosted 2 months ago

Staff Software Engineer, Compute Reliability

$251,000–$310,000 year

HybridMountain View Santa Clara County, California, United States

Full TimeSenior LevelDoctorate Or Professional DegreeLargeAutonomous Driving

Job Summary

Staff Software Engineer, Compute Reliability at Waymo. In this hybrid role, you will drive the design and development of instrumentation and onboard architecture to proactively prevent, detect, and debug reliability issues across Waymo’s compute platforms. You will design and implement long-term strategies to mitigate compute hardware reliability and numerical stability across heterogeneous platforms (CPU, TPU, GPU), creating scalable tools to automate triage, debugging, and resolution. You will lead cross-functional initiatives across Waymo, Alphabet, and external partners to architect proactive solutions that prevent issues originating from compilers (LLVM, XLA, JAX), optimization (AutoFDO), or model changes. You will architect frameworks and diagnostics to proactively identify and eliminate complex software and firmware faults, including deadlocks, memory corruption, memory leaks, and tail latency spikes. You will influence the design and tooling of next-generation hardware platforms, serving as a technical authority to ensure a coherent, forward-looking reliability architecture. You have: BS/MS in Computer Science, Electrical Engineering, Robotics, Physics, Math, or related field (or equivalent experience); Experience in C++; Experience with CUDA/GPU/TPU acceleration and model reliability or optimization; Experience with writing GPU kernels and evaluation/debugging of GPU workloads. Preferred: MS or PhD in Computer Science, Robotics, similar field, or equivalent practical experience; Experience with autonomous vehicles (L4) or ADAS systems; Experience in software reliability space. The salary range is provided for US locations and can be discussed during hiring, with eligibility for Waymo’s bonus and equity programs.

Required Qualifications

  • BS/MS in Computer Science, Electrical Engineering, Robotics, Physics, Math, or related field (or equivalent experience)
  • Experience in C++
  • Experience with CUDA/GPU/TPU acceleration and model reliability or optimization
  • Experience with writing GPU kernels and evaluation/debugging of GPU workloads

Hiring someone like this?

Get your role in front of qualified candidates on Sorce.

Get started

Apply to this job in one click with Sorce

Apply on Sorce