Senior Site Reliability Engineer (x/f/m) at Doctolib | Germany | Rezi

Senior Site Reliability Engineer (x/f/m) at Doctolib

Senior Site Reliability Engineer (x/f/m)

Doctolib · Germany

1 weeks ago

Senior Site Reliability Engineer (x/f/m)

Doctolib · Germany

7 days ago
Resume preview

Impress employers and recruiters.
Choose from hundreds of resume examples.

Target Resume Now

About the Role

We are looking for a Senior Site Reliability Engineer to join our SRE team dedicated to platform reliability within Platform Engineering. Your mission will be to ensure Doctolib's platform remains reliable, scalable, and resilient at a European scale across infrastructure, observability, and cross-cutting reliability initiatives. You will work within a team driving reliability standards across 170+ applications, contributing directly to supporting 520,000 health professionals and 90 million patients in their daily healthcare journey. Working in the tech team at Doctolib means taking ownership of critical systems, driving reliability improvements end-to-end, and partnering closely with product and engineering teams to enable fast, safe delivery.

Responsibilities

  • Build and maintain infrastructure automation and infrastructure-as-code at scale, with a strong focus on consistency, reliability, and developer experience across 170+ applications
  • Identify and lead large-scale cross-cutting reliability initiatives, including improvements to incident detection, response, and postmortem analysis capabilities
  • Design, build, and improve infrastructure components that support reliability, scalability, and observability across the platform
  • Define and drive SLOs, error budgets, and alerting standards across multiple product teams
  • Take part in the on-call rotation, and actively contribute to improving our on-call experience by reducing noise and ensuring actionable telemetry
  • Partner with software engineering teams to embed reliability practices early in the development lifecycle

Requirements

  • Solid hands-on experience (5y+) in a Site Reliability Engineering role within a large-scale, multi-team production environment
  • Proven experience with cloud platforms such as AWS, GCP, or Azure
  • Strong experience with containerization and orchestration technologies, Kubernetes is a must, its deployment and scaling strategies ecosystem
  • Experience working closely with software engineering teams to co-design reliable systems
  • Hands-on experience with infrastructure as code, particularly Terraform, applied in real production contexts
  • Implemented and operated SLIs, SLOs, and error budgets in production
  • Experience managing on-call rotations and leading incident response in high-stakes environments
  • Experience with GitOps workflows (ArgoCD, Helm)
  • Comfortable with at least one scripting or programming language (Python, Go, Ruby...) for automation
  • Fluent in English
  • Appreciate working in regulated environments, healthcare, fintech, or similar
  • Care about reliability enablement, golden paths, runbooks, shared libraries

Skills

  • Kubernetes
  • Terraform
  • AWS
  • GCP
  • Azure
  • Python
  • Go
  • Ruby
  • ArgoCD
  • Helm
  • Prometheus
  • OpenTelemetry
  • Datadog

Location

  • Berlin, Germany

Work Type

  • Hybrid work setup (up to 2 remote days per week)
  • Permanent position
  • Full-time

Experience Level

  • Senior

Benefits

  • A Deutschlandticket (Germany-wide public transport pass) fully paid for by Doctolib
  • 28 vacation days + 1 additional day for each full calendar year of employment (up to a maximum of 30 days)
  • Work from abroad for up to 10 days per year thanks to our flexibility days policy
  • Company health insurance with great supplementary benefits through our partner Allianz
  • Company pension scheme (bAV) through Allianz with an employer subsidy of 40% (15% within the probationary period)
  • The Doctolib Parent Care program, which includes one month additional parental leave and much more
  • Enrollment in Doctolib's long-term employee value sharing plan called DoctoGrowth
  • Free mental health and coaching services through our partner Moka.care
  • Subsidized sports membership through our partner Urban Sports Club
  • A flexible workplace policy offering both hybrid and office-based mode
  • Alongside healthy snacks and our regular breakfast buffet, we provide a subsidized meal benefit
  • For caregivers and workers with disabilities, a package including an adaptation of the remote policy, extra days off for medical reasons, and psychological support
  • Relocation support in case of international mobility
  • Access to the best AI tools for coding, development and dedicated training

About the Company

  • Our solutions are built on a single fully cloud-native platform that supports web and mobile app interfaces, multiple languages, and is adapted to country and healthcare specialty requirements.
  • Our stack is composed of Rails, TypeScript, Java, Python, Kotlin, Swift, and React Native.
  • We leverage AI ethically across our products to empower patients and health professionals. Discover our AI vision here.
  • Want to learn more about our tech culture and environment? Visit the Doctolib Tech site.

Equal Opportunity

  • At Doctolib, we are committed to improving access to healthcare for everyone. This translates into our recruitment process. We evaluate candidates based solely on qualifications and motivation, without any form of discrimination.
  • The more diverse ideas are heard, the more our product will truly improve healthcare for all. You are welcome to apply to Doctolib, regardless of your gender, religion, age, sexual orientation, ethnicity, or disability.
  • To ensure equal opportunities, we invite you to exclude personal information (e.g., pictures, age) from your applications. If you require any accommodation, please let us know for support during the hiring process.