Site Reliability Engineer, Vehicle Software

WayveSunnyvale, CA
Hybrid

About The Position

As a Site Reliability Engineer at Wayve, you will work across the full reliability stack for our vehicle software: observability, incident management, operator tooling, and the automation that lets engineering and operations teams detect, triage, and resolve issues faster. You will work closely with software engineers, field engineers, and operations, debugging real systems, improving real processes, and having direct impact on how our vehicles perform in production. Wayve is scaling its autonomous vehicle programs with partners, and you will help shape the SRE approach for vehicle-centric systems as that happens. This is not a role where reliability sits on top of the real work, it is central to it.

Requirements

  • Write production-quality code every day in Python, C++, or Rust.
  • Comfortable debugging at the Linux and systems level, reading logs, tracing failures, and finding root cause in complex environments.
  • Experience with CI/CD, containerization, networking, distributed systems, databases, and observability and incident-management tooling, including DataDog, Prometheus, Grafana, OpenTelemetry, Splunk, or Humio.
  • Experience diagnosing complex production or operational systems, not just escalating, but seeing problems through to resolution.
  • The communication skills to work effectively across engineering, operations, and field teams, and the judgment to know when to bring others in.
  • Based in Sunnyvale and able to be onsite at least three days a week to work alongside the field teams you will support.

Nice To Haves

  • Experience with autonomous vehicles, robotics, embedded systems, or safety-critical environments.
  • Experience with large-scale telemetry or high-volume data pipelines.
  • Familiarity with vehicle-centric or on-device software systems.

Responsibilities

  • Build and improve tooling, automation, observability, and incident-management processes for vehicle software reliability.
  • Work closely with field engineers, software teams, and operations to diagnose reliability issues and improve system performance.
  • Own hands-on debugging across Linux, low-level systems, logs, metrics, traces, and vehicle and production data.
  • Develop reliability improvements that reduce manual intervention and make issue detection, triage, and recovery faster.
  • Support vehicle-centric operations, including safety-operator tools and the reliability needs of our customer and partner programs.

Benefits

  • Competitive equity package
  • Inclusive interview experience
  • Accommodations or adjustments to participate fully in our interview process
  • Commitment to creating a diverse, fair and respectful culture that is inclusive of everyone
© 2026 Teal Labs, Inc
Privacy PolicyTerms of Service