Wayve
Site Reliability Engineer, Vehicle Software
Sunnyvale$209.7k–$266.8kfull-timemidAdded today
About this role
Wayve seeks a Site Reliability Engineer to own the full reliability stack for autonomous vehicle software, including observability, incident management, and automation. You'll debug production systems, collaborate across engineering and operations teams, and drive improvements that reduce manual intervention and accelerate issue resolution.
What you'll do
- Build and improve observability, automation, and incident-management tooling for vehicle software systems
- Diagnose reliability issues and improve system performance working with field engineers and software teams
- Perform hands-on debugging across Linux, logs, metrics, traces, and production vehicle data
- Develop reliability improvements that reduce manual work and speed up detection, triage, and recovery
- Support vehicle operations including safety-operator tools and partner program reliability needs
- Write production-quality code in Python, C++, or Rust as part of daily responsibilities
What they're looking for
- Python, C++, or Rust programming
- Linux and systems-level debugging
- CI/CD and containerization
- Observability and incident-management tools (DataDog, Prometheus, Grafana, OpenTelemetry, Splunk, Humio)
- Distributed systems and networking
- Database and data pipeline experience
- Cross-functional communication and collaboration
- Production system diagnosis and root-cause analysis
Benefits
- Competitive equity package
- Hybrid working policy combining office and remote work
- Work on cutting-edge autonomous vehicle technology
- Direct impact on vehicle performance and safety
- Collaborative environment across engineering, operations, and field teams
- Inclusive and diverse workplace culture
Opens the official application on the employer’s site. No login required.
Wayve
Wayve develops autonomous driving AI technology and systems for vehicles. The company is hiring for validation engineers, systems engineers, ML engineers, integration specialists, and program managers to build and deploy autonomous driving platforms across testing, AI development, hardware integration, and production scaling.
- Website
- wayve.ai
Likely interview questions
- Walk us through a complex production incident you debugged—how did you identify root cause and what was your approach to resolution?
- Describe your experience building observability systems. What metrics, logs, or traces have you found most valuable for detecting issues in distributed systems?