Before you apply
Listed location: Mountain View, CA, USA; San Francisco, CA, USA
Work arrangement: onsite. A remote label does not confirm worldwide eligibility or visa sponsorship.
Read the employer’s description for qualifications, compensation and work eligibility. Confirm the position is still open on the application page.
Job description supplied by Waymo; category and skill labels may be inferred. How our listings work · Report a problem
Job Description
Waymo is an autonomous driving technology company with the mission to be the world's most trusted driver. Since its start as the Google Self-Driving Car Project in 2009, Waymo has focused on building the Waymo Driver—The World's Most Experienced Driver™—to improve access to mobility while saving thousands of lives now lost to traffic crashes. The Waymo Driver powers Waymo’s fully autonomous ride-hail service and can also be applied to a range of vehicle platforms and product use cases. The Waymo Driver has provided over ten million rider-only trips, enabled by its experience autonomously driving over 100 million miles on public roads and tens of billions in simulation across 15+ U.S. states.
Waymo’s Software Reliability Engineers (SREs) are responsible for the stable operation of Waymo’s fully autonomous systems and supporting infrastructure. As an SRE, you combine software and systems engineering techniques to build and run large-scale, fault-tolerant, reliable systems. You focus on optimizing existing systems, building new infrastructure, eliminating manual, error-prone or time-consuming work through automation, and ensuring products that are fast, efficient, and effective.
This role follows a hybrid work schedule and reports to the Tech Lead Manager.
You will:
- Become a Waymo production expert, and collaborate with other engineers to build reliable systems for autonomous vehicle operations, including depot logistics, automation flow, and critical vehicle state infrastructure
- Manage end-to-end availability and performance for core fleet services, ensuring we have enough usable vehicle supply available to meet targeted user demand, and developing observability and automation to support this goal
- Involvement in the whole lifecycle of services - from inception and design, through deployment, operation and refinement
- Write designs and implement software to improve system architecture, telemetry or deployment for fleet-specific mission-critical services, preventing outages that could hinder vehicle launch or maintenance.
- Write designs and code software/automation for global infrastructure
- Serve as the first responder for fleet and supply infrastructure by leading incident response efforts. You’ll participate in a sustainable on-call rotation, while championing a culture of blameless retrospectives to drive continuous improvement
- Be a technical leader -- work with SREs, SWE partners and PMs to develop and set the technical direction and architectural guidelines for Waymo's software development
You have:
- 6+ years of experience architecting and maintaining mission-critical systems in C++, Java, or Python
- Proven ability to conduct deep-dive performance profiling and lead large-scale refactoring efforts to improve system maintainability and latency
- Demonstrated success managing massive distributed systems and a drive to solve the unique production engineering challenges found at the intersection of software and physical vehicle fleets
- A proven track record defining SLIs/SLOs/SLA frameworks. Experience designing and deploying observability systems to enhance system visibility and reliability through sophisticated monitoring aligned to critical service health and the CUJ needs of users, devs and SREs
- Demonstrated experience leading cross-functional initiatives between Engineering and Dev
- A history of mentoring junior and mid-level engineers, fostering a culture of operational excellence, and driving technical roadmap decisions for a high-growth department
- A Bachelor's degree in a relevant field or 8+ years similar experience in a high growth environment with leadership experience
We prefer:
- Demonstrated engineering leadership ability of a mission critical system at scale
- A demonstrated track record of translating reliability needs into technical roadmaps and rigorous, data-driven SLO frameworks
- Proven ability to lead deep-dive architectural investigations and resolve complex, high-impact system failures across the stack
- A Bachelors of Computer Science (or similar)
The expected base salary range for this full-time position across US locations is listed below. Actual starting pay will be based on job-related factors, including exact work location, experience, relevant training and education, and skill level. Your recruiter can share more about the specific salary range for the role location or, if the role can be performed remote, the specific salary range for your preferred location, during the hiring process.
Waymo employees are also eligible to participate in Waymo’s discretionary annual bonus program, equity incentive plan, and generous Company benefits program, subject to eligibility requirements.
Skills mentioned
Categories
Frequently asked questions
Is the Senior Site Reliability Engineer, Waymo Fleet position at Waymo remote?
The Senior Site Reliability Engineer, Waymo Fleet role at Waymo does not have a confirmed remote arrangement in our data. Check the employer description for its work location.
What type of employment is the Senior Site Reliability Engineer, Waymo Fleet role?
Waymo is hiring for a full-time Senior Site Reliability Engineer, Waymo Fleet position.
Which skills are mentioned for the Senior Site Reliability Engineer, Waymo Fleet job at Waymo?
Detected skill labels include Python, C++, Java, Distributed Systems. Check the employer description to distinguish required skills from preferred experience.
How do I apply for the Senior Site Reliability Engineer, Waymo Fleet position at Waymo?
You can apply for the Senior Site Reliability Engineer, Waymo Fleet role directly through Waymo's official application link provided on this page.
Similar AI jobs
Sr. Specialist Solutions Architect -AI&ML Engineer
Databricks · fulltime
Detection & Response Lead
Nebius · fulltime
Software Engineer (Compute Efficiency), London
Isomorphic Labs · fulltime
Senior Automation Engineer, Compute
Crusoe · fulltime
Senior Staff Optics Engineer, Network Qualification & Strategy
Crusoe · fulltime
Mechanical Engineer Internship, Robotics Hardware
Field AI · internship