Doctolib
Visit websiteSenior Site Reliability Engineer - Observability
Salary not disclosed3+ yearsHybrid
- Engineering
- Berlin
- Yesterday
About the role
The Senior Site Reliability Engineer will join the Core Reliability & Observability team to shape the observability strategy and ensure the platform remains reliable, debuggable, and scalable. You will work in a feature team developing logging, metrics, tracing, and alerting capabilities to support a large-scale healthcare platform.
Responsibilities
- Lead the observability strategy across the platform, with an emphasis on building scalable, developer-friendly logging and tracing capabilities
- Identify and lead large-scale cross-cutting reliability initiatives, including improvements to our incident detection, response, and postmortem analysis capabilities
- Take part in the on-call rotation, and actively contribute to improving our on-call experience by refining alerting, reducing noise, and ensuring actionable telemetry
Required skills
- AWS
- Azure
- Google Cloud
- Docker
- Kubernetes
- Helm
- ArgoCD
- GitOps
- Observability
- Logging
- Tracing
- Metrics
- Infrastructure as Code
Nice to have
- Open-source contribution
- High-growth environment experience
- Developer experience
Benefits
- Deutschlandticket
- 28 vacation days
- Work from abroad policy
- Company health insurance
- Company pension scheme
- ParentCare Program
- DoctoGrowth employee value sharing plan
- Mental health and coaching services
- Subsidized sports membership
About the Company
Doctolib is a cloud-native platform that supports web and mobile app interfaces for healthcare professionals and patients. The company leverages AI ethically across its products to improve the daily lives of care teams and patients.