Site Reliability Engineer, Vehicle Software
Wayve
- Location
- Sunnyvale
- Work model
- On-Site
- Level
- Mid
- Salary
- $209.7k – $266.8k/yr
- Posted
- 2h ago
Skills
About this role
About us
Founded in 2017, Wayve is the leading developer of Embodied AI technology. Our advanced AI software and foundation models enable vehicles to perceive, understand, and navigate any complex environment, enhancing the usability and safety of automated driving systems.
Our vision is to create autonomy that propels the world forward. Our intelligent, mapless, and hardware-agnostic AI products are designed for automakers, accelerating the transition from assisted to automated driving.
In our fast-paced environment big problems ignite us—we embrace uncertainty, leaning into complex challenges to unlock groundbreaking solutions. We aim high and stay humble in our pursuit of excellence, constantly learning and evolving as we pave the way for a smarter, safer future.
At Wayve, your contributions matter. We value diversity, embrace new perspectives, and foster an inclusive work environment; we back each other to deliver impact.
Make Wayve the experience that defines your career!
Site Reliability Engineer - Vehicle Software
The role
As a Site Reliability Engineer at Wayve, you will work across the full reliability stack for our vehicle software: observability, incident management, operator tooling, and the automation that lets engineering and operations teams detect, triage, and resolve issues faster. You will work closely with software engineers, field engineers, and operations, debugging real systems, improving real processes, and having direct impact on how our vehicles perform in production.
Wayve is scaling its autonomous vehicle programs with partners, and you will help shape the SRE approach for vehicle-centric systems as that happens. This is not a role where reliability sits on top of the real work, it is central to it.
Key responsibilities
• Build and improve tooling, automation, observability, and incident-management processes for vehicle software reliability.
• Work closely with field engineers, software teams, and operations to diagnose reliability issues and improve system performance.
• Own hands-on debugging across Linux, low-level systems, logs, metrics, traces, and vehicle and production data.
• Develop reliability improvements that reduce manual intervention and make issue detection, triage, and recovery faster.
• Support vehicle-centric operations, including safety-operator tools and the reliability needs of our customer and partner programs.
About you
Essential skills
• You will write production-quality code every day in Python, C++, or Rust. This is a software engineering role as much as a reliability role, and coding is central to the work.
• Comfortable debugging at the Linux and systems level, reading logs, tracing failures, and finding root cause in complex environments.
• Experience with CI/CD, containerization, networking, distributed systems, databases, and observability and incident-management tooling, including DataDog, Prometheus, Grafana, OpenTelemetry, Splunk, or Humio.
• Experience diagnosing complex production or operational systems, not just escalating, but seeing problems through to resolution.
• The communication skills to work effectively across engineering, operations, and field teams, and the judgment to know when to bring others in.
• Based in Sunnyvale and able to be onsite at least three days a week to work alongside the field teams you will support.
Desirable skills
• Experience with autonomous vehicles, robotics, embedded systems, or safety-critical environments. We know this is rare, and it is not a barrier to applying.
• Experience with