Senior Site Reliability Engineer
- Salary
- €80K–€100K
- Hiring from
- Netherlands
- Work type
- Hybrid
- Posted
Show job descriptionHide job description
Job Summary
The Senior Site Reliability Engineer plays a vital role in ensuring the reliability, availability, and performance of DeepHealth software applications, which integrate AI algorithms to deliver clinically relevant information for enhanced decision support. This role takes ownership of the reliability of cloud components deployed at client sites, ensuring that the DeepHealth solution is scalable, resilient, and secure, providing support to the operational team, and providing technical leadership within the platform engineering practice.
Essential Duties and Responsibilities
Own the reliability, availability, and performance of cloud components deployed at client sites and of the DeepHealth solution.
Define and monitor service level objectives (SLOs), error budgets, and key reliability metrics.
Develop and implement automation tools and processes to eliminate toil and streamline deployment, monitoring, and incident response operations.
Design and maintain observability tooling (monitoring, logging, alerting, and tracing), and resolve issues before they impact clients.
Contribute to the writing of technical specifications and documentation, ensuring compliance with regulatory requirements and industry best practices.
Lead incident response, conduct blameless post-mortems, and drive the implementation of corrective and preventive actions.
Perform capacity planning and performance tuning to anticipate growth and ensure optimal resource utilization.
Support the deployment of software solutions at customer sites, ensuring smooth implementation and optimal performance.
Provide ongoing support and maintenance for deployed solutions and to the operational team, addressing any issues or challenges promptly to maintain high levels of customer satisfaction.
Mentor and onboard engineers, providing technical leadership in SRE and platform engineering best practices.
Implement and maintain secure infrastructure configurations per approved baselines.
Ensure CI/CD pipeline security, including integrity verification and access controls.
Perform day-to-day technical security controls including system hardening and log monitoring.
Document all infrastructure changes and maintain audit trails.
PLEASE NOTE: This is not an exhaustive list of all duties, responsibilities and requirements of the position described above. Other functions may be assigned and management retains the right to add or change duties at any time.
Minimum Qualifications, Education and Experience
Fluency in English, both written and spoken.
Significant hands-on experience (7+ years) in Site Reliability Engineering or DevOps.
In-depth knowledge of software development practices, including design, implementation, testing, and deployment.
Strong knowledge of networking and security best practices.
Experience with fleet management and GitOps (e.g., ArgoCD).
Proficiency in containerization technologies, such as Docker and Kubernetes and its ecosystem (e.g., Istio, KEDA), and with Linux systems.
Experience with virtualization technologies, automation tools (such as Ansible or Terraform) and with SQL database administration.
Proficiency with continuous integration and continuous deployment (CI/CD) pipelines.
Strong proficiency with cloud-based environments (GCP or AWS).
Strong experience with observability and monitoring tools (e.g., Prometheus, Grafana).
Excellent communication skills, both written and verbal, and strong problem-solving and analytical skills.
Strong technical leadership and mentoring abilities
Preferred:
Familiarity with the medical device industry and the specific requirements for software applications within this domain.
Understanding of AI and machine learning concepts, with the ability to integrate algorithms into software applications.
Experience with information security standards (e.g., ISO 27001, SOC 2).
Travel
Occasional travel may be required (typically less than 10%), primarily within Europe, for audits, customer meetings, partner / vendor visits, or company offsites.
Working Environment
France or The Netherlands – Remote-friendly.
The role can be based in France or The Netherlands with flexible remote working arrangements.
There are offices in Paris, Amsterdam and Rotterdam.
Periodic on-site presence may be required for team meetings, audits, or workshops.