Senior HPC Engineer
Location: Bolney UK
Employment Type: Full-time
Working Pattern: Onsite
About the Role
We are seeking an experienced Senior HPC Engineer to join a global HPC, Linux and Digital Infrastructure team. This is a high-impact technical role combining advanced Linux system administration, HPC infrastructure, automation, DevOps and high-performance computing.
You will act as a senior technical escalation point, lead infrastructure projects, deploy new compute technologies and optimise Linux platforms for performance, scalability, availability and security. The role also offers an opportunity to act as Deputy to the HPC Manager, supporting team coordination, technical decision-making, stakeholder engagement and mentoring.
Key Responsibilities
- Act as the final escalation point for complex Linux and HPC infrastructure issues.
- Lead deployment, upgrades and optimisation of HPC clusters, Linux systems and compute platforms.
- Manage end-to-end technical projects including scoping, planning, implementation, documentation and handover.
- Evaluate and pilot emerging technologies through Proof of Concepts (PoCs).
- Develop and maintain IT standards, technical frameworks and infrastructure best practices.
- Automate infrastructure and configuration using Ansible, Puppet, GitLab and Terraform.
- Optimise Linux environments for performance, availability, scalability and compliance.
- Support Docker, Kubernetes, OpenStack, CI/CD and containerised environments.
- Collaborate with global teams on installations, upgrades and infrastructure enhancements.
- Use Jira Service Management and ITSM processes to manage incidents, changes and escalations.
- Participate in Change Control / CCRB processes and drive continual service improvement.
- Provide 24/7 on-call support on a rotational basis.
- Support in-house applications and ensure alignment with IT policies and standards.
- Mentor engineers and communicate technical information to technical and non-technical stakeholders.
- Act as deputy to the HPC Manager when required, supporting operational coordination and team performance.
Key Requirements
- Degree in Computer Science, Computer Engineering, Information Systems or related discipline.
- 5+ years of Linux system administration experience, ideally within an HPC environment.
- Strong experience troubleshooting Linux, HPC, compute and infrastructure platforms.
- Experience with automation and configuration management tools such as Ansible, Puppet, GitLab or Terraform.
- Experience with Docker and container orchestration.
- Strong scripting skills using Bash, Python and/or Perl.
- Experience delivering technical infrastructure projects.
- Strong ITSM, incident, change and problem management knowledge.
- Proven ability to mentor, guide and influence technical teams.
- Strong communication, stakeholder management and project/time management skills.
- Interest in developing towards technical leadership and people management.
Desirable Skills
- HPC clustering and high-performance computing
- Kubernetes / OpenStack / CI/CD / DevOps
- Cloud administration and virtualisation
- GPU computing, RAID and high-performance storage
- Linux certifications such as LPIC-2/3 or CompTIA Linux+
- ITIL Foundation or equivalent IT service management certification.
Benefits
- Competitive salary and annual bonus
- 22+ days annual leave with holiday purchase/sale options
- Generous employer pension contribution
- Life Assurance and Group Income Protection
- Employee Assistance and wellbeing programmes
- Private Medical & Dental care
- Flexible benefits and employee discounts
- Visa sponsorship and comprehensive relocation support
- Employee referral scheme
- Subsidised on-site canteen and wellness facilities
- Cycle purchase scheme and regular employee/social events