What we do At Perlego, we are working hard to make education accessible to all. In this digital age, we believe that anyone should be able to learn anything at any time. Knowledge should be more accessible, not locked behind sky-high price tags. Over the past 9 years, our goal has been to support students across the UK & Europe to access quality books. Our ambition is to expand our support to students globally, specifically looking at the US, and build a product that goes beyond the book, a platform that helps students study smarter and more effectively. What we're looking for: We are looking for an experienced Cloud Infrastructure Engineer with a strong background in AWS services and monitoring tools. In this role, you will ensure the availability and reliability of our services. You will be integral to swiftly addressing issues, resolving incidents independently, and thriving in a fast-paced environment. What you’ll do: As a Cloud Infrastructure Engineer , your main focus will be to ensure our services remain highly available and performant. Key responsibilities include: Cloud Infrastructure Management: Manage and support AWS infrastructure , focusing on scalability, security, and reliability. Handle deployments, managing CI/CD pipelines for both containerised (Docker/ECS) and serverless (AWS Lambda) applications. Own infrastructure as code — provisioning resources declaratively so environments are reproducible, version-controlled, and safe to change. Ensure effective backup, recovery, and disaster recovery strategies to minimise downtime. Manage operational and analytical data stores (Aurora MySQL, DynamoDB) Drive cost optimization across the infrastructure — monitoring spend, eliminating waste, and rightsizing resources to balance performance and cost. Monitoring & Incident Management: Monitor and manage platform activity using tools like Prometheus , Grafana , or AWS CloudWatch Respond quickly to alerts and incidents, independently resolving issues and ensuring service uptime. Conduct post-incident reviews and help improve system resiliency through automation and monitoring enhancements. Review network activity with AWS Security Hub and Cloudflare Collaboration & Communication: Collaborate with cross-functional teams to implement platform improvements. Work independently and make swift decisions when managing service incidents outside core business hours. Assist in platform security, ensuring adherence to best practices for cloud security and compliance. Continuous Improvement: Automate manual processes to reduce human error and improve efficiency. Continuously enhance monitoring systems, ensuring robust early detection and resolution capabilities. Identify potential performance bottlenecks and contribute to overall platform optimisation.
Infrastructure & Operations - Platform Engineer - Support
Ibgllc
Senior IT Infrastructure / Systems Engineer
Flatrock
Platform Infrastructure Engineer (SRE Core)
Menlosecurity
Mainframe Hardware Infrastructure Engineer
Rbs
Infrastructure Engineer, Mainframe Capacity
Rbs
Security Infrastructure Engineer/ 6+ months/ Fully Remote in UK
Cloud Bridge Tech Recruitment