
Site Reliability Engineer (Vehicle Software)
Wayve
Job description
-
As a Site Reliability Engineer at Wayve, you will work across the full reliability stack for our vehicle software: observability, incident management, operator tooling, and the automation that lets engineering and operations teams detect, triage, and resolve issues faster
-
You will work closely with software engineers, field engineers, and operations, debugging real systems, improving real processes, and having direct impact on how our vehicles perform in production
-
Wayve is scaling its autonomous vehicle programs with partners, and you will help shape the SRE approach for vehicle-centric systems as that happens
-
This is not a role where reliability sits on top of the real work, it is central to it
-
Build and improve tooling, automation, observability, and incident-management processes for vehicle software reliability
-
Work closely with field engineers, software teams, and operations to diagnose reliability issues and improve system performance
-
Own hands-on debugging across Linux, low-level systems, logs, metrics, traces, and vehicle and production data
-
Develop reliability improvements that reduce manual intervention and make issue detection, triage, and recovery faster
-
Support vehicle-centric operations, including safety-operator tools and the reliability needs of our customer and partner programs
Benefits
-
Private healthcare: Choose our optional health insurance for comprehensive coverage for you and your family.
-
Paid time off: Paid vacation plus public holidays and additional leave programs, ensuring you have time to unwind.
-
Mental health resources: Through Spill, you can access therapy and mental health support.
-
Community and socials: Join clubs or attend team socials to connect over hobbies, sports, or just for fun.
-
Competitive compensation: Our compensation package includes cash and equity, making you a true partner in our success.
-
Learning and development: Budgets for books, courses, and company-wide training to support your continuous growth.- Based in Sunnyvale and able to be onsite at least three days a week to work alongside the field teams you will support
-
You will write production-quality code every day in Python, C++, or Rust. This is a software engineering role as much as a reliability role, and coding is central to the work
-
Experience with CI/CD, containerization, networking, distributed systems, databases, and observability and incident-management tooling, including DataDog, Prometheus, Grafana, OpenTelemetry, Splunk, or Humio
-
The communication skills to work effectively across engineering, operations, and field teams, and the judgment to know when to bring others in
-
Comfortable debugging at the Linux and systems level, reading logs, tracing failures, and finding root cause in complex environments
-
Experience diagnosing complex production or operational systems, not just escalating, but seeing problems through to resolution
-
We understand that everyone has a unique set of skills and experiences and that not everyone will meet all of the requirements listed above. If youโre passionate about self-driving cars and think you have what it takes to make a positive impact on the world, we encourage you to apply
-
Experience with autonomous vehicles, robotics, embedded systems, or safety-critical environments. We know this is rare, and it is not a barrier to applying
-
Experience with large-scale telemetry or high-volume data pipelines
-
Familiarity with vehicle-centric or on-device software systems