Senior HPC Engineer - Hybrid - Inside IR35
Location: Stevenage, UK
Contract: 12 months
Working Pattern: Hybrid - 3 days onsite
Rate: £340+ per day
The Role
We are looking for an experienced Senior HPC Engineer to support and maintain a scientific computing environment, with a strong focus on RHEL, Slurm and HPC infrastructure.
You will work closely with research scientists and technical teams to ensure HPC services are secure, reliable and high performing.
Key Responsibilities
- Administer, patch and maintain RHEL 7, 8 and 9 across HPC clusters and workstations.
- Deploy, configure and manage Slurm, including queues, partitions and scheduling.
- Monitor cluster health, performance, storage, networking and resource utilisation.
- Install and support scientific applications, compilers, libraries and MPI environments.
- Work with scientists to optimise workloads and resolve application issues.
- Troubleshoot hardware, operating system, scheduler and application problems.
- Manage incidents and service requests through ServiceNow.
- Collaborate with storage, networking, security and DevOps teams.
Essential Skills
- 10+ years' enterprise IT experience, including 3-5+ years in HPC or research computing.
- Strong hands-on experience with RHEL 7, 8 and 9.
- Proven experience managing HPC clusters and Slurm.
- Experience supporting scientific or research applications on Linux.
- Strong troubleshooting and root-cause analysis skills.
- Experience with ServiceNow or a similar ITSM platform.
- Strong communication and stakeholder-management skills.
- Able to work onsite in Stevenage 3 days per week.
Desirable Skills
- Docker or container technologies.
- Ansible or similar automation tools.
- GPU computing, CUDA and GPU-accelerated workloads.
- OpenMPI, MPICH or other MPI libraries.
- InfiniBand and high-speed networking.
- Web server and SSL certificate management.
- RHCSA, RHCE or equivalent Red Hat certification.