Role at a glance
- Salary
- Not Disclosed
- Location
- Warsaw, Masovian Voivodeship, Poland
- Work arrangement
- Hybrid
- Employment
- Full-time
- Experience
- 6+ years of experience architecting and maintaining mission-critical systems in C++, Java, or Python
- Education
- A Bachelors degree in a relevant field or 8+ years similar experience in a high growth environment with leadership experience
Spotted an issue?
We’ll check it against the original posting.
Role Summary
This software reliability engineering role supports Waymo’s fully autonomous systems and infrastructure, with a focus on reliable fleet operations, depot logistics, automation flow, and critical vehicle state infrastructure. The role manages availability and performance for core fleet services and contributes to the technical direction of software development.
What You'll Do
- Build reliable systems for autonomous vehicle operations, including depot logistics, automation flow, and critical vehicle state...
- Manage end-to-end availability and performance for core fleet services
- Develop observability and automation to support usable vehicle supply and targeted user demand
- Design and implement software to improve system architecture, telemetry, and deployment for mission-critical services
- Write designs and code software and automation for global infrastructure
- Lead incident response efforts and participate in an on-call rotation
Generated from the employer's posting. Verify important details before applying.
View full postingQualifications
6+ years of experience architecting and maintaining mission-critical systems in C++, Java, or Python; experience with performance profiling, large-scale refactoring, massive distributed systems, SLI/SLO/SLA frameworks, observability systems, cross-functional initiatives, mentoring engineers, and technical roadmap decisions.
Required
- 6+ years of experience architecting and maintaining mission-critical systems in C++, Java, or Python
- Deep-dive performance profiling
- Large-scale refactoring efforts
- Managing massive distributed systems
- Defining SLI/SLO/SLA frameworks
- Designing and deploying observability systems
- Leading cross-functional initiatives between Engineering and Dev
- Mentoring junior and mid-level engineers
Preferred
- Demonstrated engineering leadership ability of a mission critical system at scale
- Translating reliability needs into technical roadmaps and rigorous, data-driven SLO frameworks
- Leading deep-dive architectural investigations and resolving complex, high-impact system failures across the stack
- A Bachelors of Computer Science (or similar)
Original job description
Content provided by the employer
Original job description
Content provided by the employer
Waymo is an autonomous driving technology company with the mission to be the world's most trusted driver. Since its start as the Google Self-Driving Car Project in 2009, Waymo has focused on building the Waymo Driver—The World's Most Experienced Driver™—to improve access to mobility while saving thousands of lives now lost to traffic crashes. The Waymo Driver powers Waymo’s fully autonomous ride-hail service and can also be applied to a range of vehicle platforms and product use cases. The Waymo Driver has provided over ten million rider-only trips, enabled by its experience autonomously driving over 100 million miles on public roads and tens of billions in simulation across 15+ U.S. states.
Waymo’s software reliability engineers (SRE) are responsible for the stable operation of Waymo’s fully autonomous systems and supporting infrastructure. As an SRE, you combine software and systems engineering techniques to build and run large-scale, fault-tolerant, reliable systems. You focus on optimizing existing systems, building new infrastructure, eliminating manual, error-prone or time-consuming work through automation, and ensuring products that are fast, efficient, and effective.
This role follows a hybrid work schedule and reports to the Tech Lead Manager.
You will:
- Become a Waymo production expert, and collaborate with other engineers to build reliable systems for autonomous vehicle operations, including depot logistics, automation flow, and critical vehicle state infrastructure
- Manage end-to-end availability and performance for core fleet services, ensuring we have enough usable vehicle supply available to meet targeted user demand, and developing observability and automation to support this goal
- Involvement in the whole lifecycle of services - from inception and design, through deployment, operation and refinement
- Write designs and implement software to improve system architecture, telemetry or deployment for fleet-specific mission-critical services, preventing outages that could hinder vehicle launch or maintenance
- Write designs and code software/automation for global infrastructure
- Serve as the first responder for fleet and supply infrastructure by leading incident response efforts. You’ll participate in a sustainable on-call rotation, while championing a culture of blameless retrospectives to drive continuous improvement
- Be a technical leader -- work with SREs, SWE partners and PMs to develop and set the technical direction and architectural guidelines for Waymo's software development
You have:
- 6+ years of experience architecting and maintaining mission-critical systems in C++, Java, or Python
- Proven ability to conduct deep-dive performance profiling and lead large-scale refactoring efforts to improve system maintainability and latency
- Demonstrated success managing massive distributed systems and a drive to solve the unique production engineering challenges found at the intersection of software and physical vehicle fleets
- A proven track record defining SLIs/SLOs/SLA frameworks. Experience designing and deploying observability systems to enhancing system visibility and reliability through sophisticated monitoring aligned to critical service health and the CUJ needs of users, devs and SREs
- Demonstrated experience leading cross-functional initiatives between Engineering and Dev
- A history of mentoring junior and mid-level engineers, fostering a culture of operational excellence, and driving technical roadmap decisions for a high-growth department
- A Bachelors degree in a relevant field or 8+ years similar experience in a high growth environment with leadership experience
We prefer:
- Demonstrated engineering leadership ability of a mission critical system at scale
- A demonstrated track record of translating reliability needs into technical roadmaps and rigorous, data-driven SLO frameworks.
- Proven ability to lead deep-dive architectural investigations and resolve complex, high-impact system failures across the stack
- A Bachelors of Computer Science (or similar)
About the company
Waymo
Large Enterprise
Waymo is a leading autonomous driving technology company, originally a part of Google's parent company Alphabet Inc. Established in 2009, Waymo focuses on developing self-driving cars and innovative transportation solutions designed to enhance mobility and safety on the roads. By employing advanced artificial intelligence and machine learning algorithms, the company aims to revolutionize personal and shared transportation, making it more accessible and efficient. Waymo's efforts contribute significantly to the future of autonomous vehicles and the ongoing shift towards smart transportation systems.
Waymo is a leading autonomous driving technology company, originally a part of Google's parent company Alphabet Inc. Established in 2009, Waymo focuses on developing self-driving cars and innovative transportation solutions designed to enhance mobility and safety on the roads. By employing advanced artificial intelligence and machine learning algorithms, the company aims to revolutionize personal and shared transportation, making it more accessible and efficient. Waymo's efforts contribute significantly to the future of autonomous vehicles and the ongoing shift towards smart transportation systems.