Job Description This is a remote position. Location: Remote Engagement Type: Project-based Assignment; B2B Contract (Outside IR35) Duration: 6 months with auto renew Timezone : CET ± 2 hours Be part of our Global Engineering Network! One of our clients in the Telecommunications sector is looking for a Site Reliability Engineer (SRE). In this role, you will provide first-level support, monitor systems, respond to incidents, and perform operational maintenance to ensure our applications and infrastructure remain reliable and available. You will be requested to support on a shift schedule, including weekends and public holidays, monitoring critical systems, responding to alerts, resolving issues, and escalating complex incidents when needed. This role is ideal for someone with strong troubleshooting skills who wants to grow their expertise in cloud infrastructure, monitoring, and production operations. Key Deliverables Monitor production systems, applications, and infrastructure using monitoring and alerting tools. Respond to alerts, investigate incidents, and take corrective action following established procedures. Perform routine health checks and scheduled maintenance activities. Review application, system, and server logs to troubleshoot issues. Restart services and carry out operational tasks using documented runbooks. Manage and resolve support tickets within agreed service levels (SLAs). Escalate complex issues to second-level support or engineering teams as required. Track incidents through resolution and maintain accurate documentation. Conduct daily operational checks to ensure system stability and availability. Maintain and update operational documentation, runbooks, and troubleshooting guides. Participate in shift rotations, including weekends and public holidays. Support incident management activities and post-incident reviews. Identify recurring issues and recommend improvements to enhance reliability and efficiency. Requirements Ideal Profile 5+ years of experience in Technical Support, Operations, NOC, SOC, or Site Reliability Engineering roles. Strong troubleshooting, analytical, and problem-solving skills. Experience with monitoring and alerting tools, including: Zabbix Grafana Prometheus Experience reviewing and analyzing application, system, and server logs. Understanding of incident management and escalation processes. Experience using ticketing tools such as Jira. Experience with cloud platforms, particularly Google Cloud. Familiarity with Linux and common command-line tools. Ability to follow structured procedures and operational runbooks. Strong attention to detail and commitment to service reliability. Good written and verbal communication skills. Ability to work independently and effectively within a shift-based team. Why Partner with Us? Clear scope with no ambiguity over deliverables. Flexible working arrangements. Opportunity for repeat engagements based on performance. Selection Process Proposal Submission Submit your professional profile/CV by applying on the role. Business Alignment Call 30-min virtual discussion with Human Capital Consultant to review scope Verification Opportunity to complete Castille Vetting (background/compliance checks) Client Skills Review Direct interview with end client to discuss project specifics Project-specific technical assessment (if required) Ongoing Business Support Access to CX guidance and market insights through our professional network.
Technical Operations Specialist - Second Line Support
Netcraft
Operations & Sales Support EA/CS
Outsourcedin
Executive Assistant - Operations Support
Outsourcedin
Senior Marketing & Operations Support
Lioncrest People
Customer Operations Coordinator, Support
Canvas Medical
Operations Support Lead
AB InBev