About Doctolib
At Doctolib, we’re reinventing the daily lives of healthcare professionals and empowering people to live healthier lives across Europe.
More than 3,000 Doctolibers across France, Germany, Italy, the UK, and the Netherlands are building the next generation of AI-powered healthcare technologies, trusted by over 520,000 healthcare professionals and 90 million patients.
Artificial Intelligence is at the heart of how we build our products and how we work. We foster a culture where AI fluency enables every team member to improve efficiency, quality, and impact at scale.
Your Impact
We are looking for a Senior Site Reliability Engineer (SRE) to join our Platform Engineering organization.
In this role, you will help ensure Doctolib’s platform remains reliable, scalable, and resilient across Europe by driving infrastructure, observability, and reliability initiatives that support more than 170 applications and millions of daily users.
You’ll take ownership of critical platform services, improve system reliability end-to-end, and work closely with engineering teams to enable fast and secure software delivery.
Your Responsibilities
- Infrastructure Automation
- Build and maintain infrastructure-as-code and automation at scale.
- Improve consistency, reliability, and developer experience across more than 170 applications.
- Reliability Engineering
- Lead cross-platform reliability initiatives.
- Improve incident detection, response, and postmortem processes.
- Platform Engineering
- Design and enhance infrastructure supporting platform scalability, observability, and resilience.
- Continuously improve platform reliability.
- Reliability Standards
- Define and implement SLIs, SLOs, error budgets, and alerting standards.
- Partner with multiple product teams to improve service reliability.
- Incident Management
- Participate in the on-call rotation.
- Improve the on-call experience through actionable monitoring and reduced alert fatigue.
- Engineering Collaboration
- Work closely with software engineers to embed reliability best practices early in the development lifecycle.
Your Profile
- Professional Experience
- 5+ years of Site Reliability Engineering experience in large-scale production environments.
- Experience working across multiple engineering teams.
- Cloud Platforms
- Hands-on experience with AWS, GCP, or Azure.
- Containerization & Orchestration
- Strong Kubernetes expertise, including deployment and scaling strategies.
- Infrastructure as Code
- Production experience with Terraform.
- Reliability Engineering
- Experience implementing SLIs, SLOs, and error budgets.
- Proven incident response and on-call management experience.
- GitOps & Automation
- Experience with ArgoCD, Helm, and GitOps workflows.
- Proficiency in at least one scripting or programming language such as Python, Go, or Ruby.
- Communication
- Fluent English language skills.
Nice to Have
- Experience working in regulated industries such as healthcare or fintech.
- Knowledge of reliability enablement, golden paths, runbooks, and shared engineering libraries.
- Passion for building scalable engineering standards and operational excellence.
Life at Doctolib Tech
Our products run on a fully cloud-native platform supporting web and mobile applications across multiple countries, languages, and healthcare specialties.
Technology Stack:
- Rails
- TypeScript
- Java
- Python
- Kotlin
- Swift
- React Native
- Kubernetes
- Terraform
- AWS / GCP
- Prometheus
- OpenTelemetry
- Datadog
- ArgoCD
We use AI responsibly to improve healthcare experiences for both patients and healthcare professionals.
What We Offer
- Deutschlandticket fully paid by Doctolib.
- 28 vacation days plus one additional day per completed calendar year (up to 30 days).
- Work abroad for up to 10 days per year.
- Comprehensive company health insurance through Allianz.
- Company pension scheme with up to 40% employer contribution.
- Enhanced parental leave through the Doctolib Parent Care program.
- Participation in the DoctoGrowth employee value-sharing plan.
- Mental health and coaching support via Moka.care.
- Subsidized Urban Sports Club membership.
- Flexible hybrid working model.
- Subsidized meals, healthy snacks, and breakfast buffet.
- Additional support for caregivers and employees with disabilities.
- Relocation assistance for international candidates.
- Access to leading AI development tools and dedicated training.
Interview Process
- Recruiter Interview
- Technical SRE Interview
- System Design Interview
- Behavioral Interview
- Reference Check
Job Details
- Position: Senior Site Reliability Engineer
- Employment Type: Permanent, Full-Time
- Location: Berlin, Germany
- Work Model: Hybrid (up to 2 remote days per week)
- Start Date: As soon as possible
Diversity & Inclusion
At Doctolib, we are committed to improving healthcare for everyone by fostering an inclusive workplace where all candidates are evaluated based on their qualifications and motivation.
We welcome applicants of every gender, religion, age, sexual orientation, ethnicity, and disability. Candidates are encouraged to remove personal information such as photographs and age from their application documents. If you require accommodations during the recruitment process, our team will be happy to support you.
Join us in building the healthcare we all dream of.
“`