hirq
← All jobs

DeepHealth

Senior Site Reliability Engineer

Amsterdam / Rotterdam, Noord-Holland, Netherlands · Remote · fulltime_permanent · R&D

Apply well, not just fast

Create a free account and upload your resume to get a match score, keyword gaps, a tailored resume, a cover letter and interview prep for this job.

About the role

SREIncident ResponseCI/CDObservability
Job Summary The Senior Site Reliability Engineer plays a vital role in ensuring the reliability, availability, and performance of DeepHealth software applications, which integrate AI algorithms to deliver clinically relevant information for enhanced decision support. This role takes ownership of the reliability of cloud components deployed at client sites, ensuring that the DeepHealth solution is scalable, resilient, and secure, providing support to the operational team, and providing technical leadership within the platform engineering practice. Essential Duties and Responsibilities - Own the reliability, availability, and performance of cloud components deployed at client sites and of the DeepHealth solution. - Define and monitor service level objectives (SLOs), error budgets, and key reliability metrics. - Develop and implement automation tools and processes to eliminate toil and streamline deployment, monitoring, and incident response operations. - Design and maintain observability tooling (monitoring, logging, alerting, and tracing), and resolve issues before they impact clients. - Contribute to the writing of technical specifications and documentation, ensuring compliance with regulatory requirements and industry best practices. - Lead incident response, conduct blameless post-mortems, and drive the implementation of corrective and preventive actions. - Perform capacity planning and performance tuning to anticipate growth and ensure optimal resource utilization. - Support the deployment of software solutions at customer sites, ensuring smooth implementation and optimal performance. - Provide ongoing support and maintenance for deployed solutions and to the operational team, addressing any issues or challenges promptly to maintain high levels of customer satisfaction. - Mentor and onboard engineers, providing technical leadership in SRE and platform engineering best practices. - Implement and maintain secure infrastructure configurations per approved baselines. - Ensure CI/CD pipeline security, including integrity verification and access controls. - Perform day-to-day technical security controls including system hardening and log monitoring. - Document all infrastructure changes and maintain audit trails. PLEASE NOTE: This is not an exhaustive list of all duties, responsibilities and requirements of the position described above. Other functions may be assigned and management retains the right to add or change duties at any time.