Azure Site Reliability Engineer
- Hiring from
- Ireland
- Work type
- Hybrid
- Posted
Is this job info correct?
513,132 remote jobs, straight from company career pages
100% free · New jobs every hour
Show job descriptionHide job description
IT Contracting has gotten a boost with a great SRE opportunity with one of our clients in their Dublin offices!
About Your New Employer
- Join a global financial leader at the forefront of innovation within Market Risk Technology.
- Work in a collaborative, mission-critical environment driving modernisation and operational excellence for regulated financial platforms.
- Exposure to advanced AI engineering, large-scale platform upgrades, and complex remediation programs in a top-tier banking technology setting.
About Your New Job
- Embed Site Reliability Engineering principles into Azure-native platform solutions and product operating models.
- Support and improve enterprise Azure landing zones, ingress and egress DMZs, service integration layers, and shared platform components.
- Design and enhance platform observability through effective monitoring, logging, alerting, dashboards, and reliability metrics.
- Participate in incident response, troubleshooting, root-cause analysis, and post-incident reviews.
- Develop automation to improve operational efficiency, reduce manual intervention, and support consistent platform operations.
- Plan and execute resilience testing, recovery exercises, and operational readiness assessments.
- Develop, maintain, and test operational runbooks, playbooks, and recovery procedures.
- Support the onboarding and sustained operation of enterprise workloads across Azure.
- Work closely with platform engineering, networking, security, infrastructure, and AI teams to identify and resolve reliability risks.
- Help define and promote reliability, availability, scalability, and operational standards across horizontal engineering teams.
- Support secure AI infrastructure workloads, including platforms enabling Microsoft Foundry and private OpenAI access.
What Skills You Need In This Job
- Strong experience operating and supporting Azure platform services in complex enterprise environments.
- Proven hands-on experience with SRE practices, including observability, incident response, automation, resilience testing, and operational readiness.
- Experience supporting Azure landing zones, network perimeters, ingress and egress DMZs, service integration layers, and shared Azure services.
- Direct experience operating secure Azure environments and supporting AI infrastructure workloads at scale.
- Experience supporting platforms that enable Microsoft Foundry and private OpenAI access.
- Strong knowledge of highly available service integration layers and enterprise networking environments.
- Demonstrable experience developing and maintaining operational runbooks.
- Strong automation and scripting skills, using tools such as Azure CLI, PowerShell, Python, Terraform, or Bicep.
- Experience designing observability solutions for cloud platforms and shared services.
- The ability to define and apply reliability objectives, including availability, scalability, and operational performance targets.
- Strong communication and collaboration skills, with the ability to influence standards across platform engineering, networking, security, and infrastructure teams.
What’s on Offer In This Job
- Daily Rate of up to €550+
- Hybrid working environment that is 3 days a week in the office
- 12 Month Contract with a good likelihood of contract extensions
What’s Next
If you think this job is for you, apply now! Looking forward to talking to some great candidates!