Wayve
Site Reliability Engineer, Vehicle SW
Salary not disclosedHybrid
- Engineering
- Leonberg
- Full time
- Yesterday
About the role
The Fleet Reliability Engineering team ensures that systems supporting autonomous vehicle operations are reliable, observable, and scalable. You will monitor live system health, investigate incidents, and build automation to reduce manual operational work. This role involves collaborating across software, infrastructure, and fleet operations to improve system resilience and safety.
Responsibilities
- Monitor live system health using logs, metrics, traces, alerts, and dashboards.
- Investigate incidents end-to-end, including data analysis and root cause identification.
- Participate in on-call rotations and contribute to post-incident reviews.
- Develop service-level indicators and objectives to improve monitoring and alerting.
- Build tools and automation to reduce manual operational work and improve incident resolution.
- Influence system design to turn operational pain points into lasting improvements.
Required skills
- Site Reliability Engineering
- Python
- C++
- Rust
- Linux
- Cloud platforms
- Containers
- Kubernetes
- CI/CD
- Observability
- Datadog
- Prometheus
- Grafana
- OpenTelemetry
- Splunk
Benefits
- Health insurance
- Dental insurance
- Enhanced maternity and paternity leave
- Retirement or pension
- Learning and development budget
- Relocation support
- Visa sponsorship
- Equity
About the Company
Wayve is building the leading AI platform for autonomous driving, pioneering an end-to-end AI approach that enables vehicles to learn directly from real-world experience. Their mapless and hardware-agnostic AI platform integrates with global OEM partners to enable continuous software evolution and advanced levels of automation.