Back to results

Senior Site Reliability Engineer - Observability

Salary not disclosed3+ yearsHybrid

  • Engineering
  • Berlin
  • Yesterday
Newly posted

About the role

The Senior Site Reliability Engineer will join the Core Reliability & Observability team to shape the observability strategy and ensure the platform remains reliable, debuggable, and scalable. You will work in a feature team developing logging, metrics, tracing, and alerting capabilities to support a large-scale healthcare platform.

Responsibilities

  • Lead the observability strategy across the platform, with an emphasis on building scalable, developer-friendly logging and tracing capabilities
  • Identify and lead large-scale cross-cutting reliability initiatives, including improvements to our incident detection, response, and postmortem analysis capabilities
  • Take part in the on-call rotation, and actively contribute to improving our on-call experience by refining alerting, reducing noise, and ensuring actionable telemetry

Required skills

  • AWS
  • Azure
  • Google Cloud
  • Docker
  • Kubernetes
  • Helm
  • ArgoCD
  • GitOps
  • Observability
  • Logging
  • Tracing
  • Metrics
  • Infrastructure as Code

Nice to have

  • Open-source contribution
  • High-growth environment experience
  • Developer experience

Benefits

  • Deutschlandticket
  • 28 vacation days
  • Work from abroad policy
  • Company health insurance
  • Company pension scheme
  • ParentCare Program
  • DoctoGrowth employee value sharing plan
  • Mental health and coaching services
  • Subsidized sports membership

About the Company

Doctolib is a cloud-native platform that supports web and mobile app interfaces for healthcare professionals and patients. The company leverages AI ethically across its products to improve the daily lives of care teams and patients.

Senior Site Reliability Engineer - Observability at Doctolib · Grasshire