The posting
Senior Site Reliability Engineer IT Infrastructure Senior Site Reliability Engineer - schedule 32 - 40 uur - business_center Vast dienstverband - euro 102.000 - pin_drop Regio Den Haag - sell DevOps Are you an experienced Site Reliability Engineer who sees reliability as an engineering challenge rather than a series of incidents to resolve? Do you enjoy building self-healing infrastructure, eliminating repetitive operational work and improving large-scale cloud platforms? Then this could be an interesting opportunity for you! Solliciteren About the position As a Senior Site Reliability Engineer, you will be responsible for the reliability of a large international SaaS environment running primarily on Microsoft Azure. The platform operates across more than 10 global data centres, serves millions of end users and needs to deliver 24/7 availability backed by strict SLAs. This is not a traditional operations or ticket-driven infrastructure role. You will approach reliability from an engineering perspective. You define and manage SLOs and error budgets, improve observability, automate repetitive work and design systems that can recover automatically when something goes wrong. You will work closely with cloud engineers and software development teams. Instead of only becoming involved after an incident occurs, you will participate early in the development process and help engineering teams make architectural decisions that improve scalability, performance and reliability. What will you do? - Define and manage SLOs and error budgets for critical services across the Azure environment - Identify operational toil and replace repetitive manual work with automation and self-healing solutions - Improve and standardise observability, including metrics, alerting and tracing - Lead complex production incidents and facilitate blameless postmortems - Translate lessons from incidents into structural improvements and automation - Improve...



