26 to 50 of 73 Slurm Workload Manager Jobs in England

AI Infrastructure Engineer

Hiring Organisation
LinuxRecruit
Location
London, UK
Employment Type
Full-time
will eliminate bottlenecks in the data path to ensure training is fast and as capital efficient as possible alongside managing cluster orchestration using slurm and Kubernetes while preparing to expand into specialised GPU providers. And finally you will master the stack from pytorch based learning libraries to complex data ...

Senior Software Engineer Scientific Data Oxford, England, United Kingdom

Location
Oxford, England, United Kingdom
Experience: Experience leveraging Kubernetes for application deployment, and familiarity with distributed computing frameworks (e.g., Ray, Spark), or specialised batch schedulers/resource managers (e.g., Slurm, Volcano, Kueue). Demonstratedexpertisewith specialised serving engines (e.g.,vLLM,SGLang) or techniques for deploying models in resource-constrained or high-throughput environments. Familiarity withGitOpsand ...

Bioinformatician Plant Biology Institute Oxford, England, United Kingdom

Location
Oxford, England, United Kingdom
bioinformatics tools for genome assembly, annotation and variant analysis. Familiarity operating within Linux environments, including operating within HPC and/or cloud environments (e.g. SLURM, OCI). Experience applying machine learning methods to genomic and transcriptomic data, with an understanding of different ML approaches and their suitability for different ...

Infrastructure Engineer

Location
Cambridge, England, United Kingdom
/reliable systems, Care about the experience of the people relying on your infrastructure. Nice to have/Beneficial HPC‐style job scheduling (Kueue, Slurm, MPI, or similar). Experience with large scientific or array datasets. Any background in photonics, semiconductors, EDA, or scientific computing. We are particularly interested ...

IT Lead Engineer London

Location
Greater London, England, United Kingdom
slow" needs a real root cause, not a restart. What you will do Own and evolve our HPC environment: cluster administration, job scheduling (e.g., Slurm/PBS/LSF), performance tuning, and capacity planning for compute-heavy engineering workloads. Experience with Entra ID Governance: Access Reviews, Identity Protection, Privileged ...

AI Inference Engineer

Hiring Organisation
Fuse Energy Supply
Location
London, UK
Employment Type
Full-time
scale inference; multi-tenant serving or SLA-driven infrastructure; background at a hyperscaler, frontier AI lab or large-scale distributed inference system; Kubernetes/Slurm; interest in energy markets, grid systems or sustainability-focused computeBenefitsCompetitive salary and eligibility for equityBiannual bonus schemeFully expensed tech to match your needsPrivate health ...

AI Inference Engineer

Location
Greater London, England, United Kingdom
scale inference; multi-tenant serving or SLA-driven infrastructure; background at a hyperscaler, frontier AI lab or large-scale distributed inference system; Kubernetes/Slurm; interest in energy markets, grid systems or sustainability-focused compute Competitive salary and eligibility for equity Biannual bonus scheme Fully expensed tech to match ...

AI Inference Engineer

Location
Greater London, England, United Kingdom
multi-tenant serving or SLA-driven infrastructure. Background at a hyperscaler, frontier AI lab, or large-scale distributed inference system. Familiarity with Kubernetes/Slurm for cluster orchestration. Interest or experience in energy markets, grid systems, or sustainability-focused compute. Benefits Competitive salary and an equity sign-on bonus. ...

Senior HPC Linux Engineer - Onsite, Slurm/PBS Platform Lead

Location
England, United Kingdom
Unknown is seeking an experienced HPC Engineer to join an elite engineering organisation in the East Midlands. This onsite role requires taking ownership of the HPC platform, improving services, and delivering best-in-class HPC ...

Senior HPC Engineer - Hybrid GPU Linux Clusters

Location
Stevenage, England, United Kingdom
researchers and technical teams. You will manage HPC clusters, ensure reliability and performance, and support scientific applications and complex workloads, including GPU computing and Slurm workload management. Hybrid onsite presence is required. #J-18808-Ljbffr ...

Senior Lead Software Engineer (C++, Python) – HPC & Cloud

Location
Greater London, England, United Kingdom
GenAI to improve monitoring, management, and efficiency of large compute platforms. You will collaborate with cross-functional teams, deploy on Kubernetes, and work with Slurm, LSF, Spark, Ray, and Symphony in cloud environments to deliver scalable, secure technologies #J-18808-Ljbffr ...

Linux HPC Specialist

Location
Stevenage, England, United Kingdom
fast paced development environment. Salary: Up to £75,000 depending on experience Dynamic (hybrid) working: 2-3 days per week on-site due to workload classification Security Clearance This role will require DV Clearance. Restrictions and/or limitations relating to nationality and/or rights to work … engineering and operations teams Performance & Scalability Ensure systems are designed for optimal throughput, latency, and scalability Lead performance benchmarking, capacity planning, and workload modellingIdentify and eliminate architectural bottlenecks Workload & Software Ecosystem Define strategies for workload orchestration (SLURM, Kubernetes for HPC, etc.) Guide software stack design ...

Senior Research HPC Engineer - Accelerate Discovery

Location
Greater London, England, United Kingdom
design, deploy, and maintain HPC services, containers, and workflows, aligning with strategic goals and research needs. The role requires extensive experience with Linux HPC, SLURM, and software packaging, plus strong collaboration with researchers and cross‐functional IT teams to deliver reliable, high‐performance solutions. #J-18808-Ljbffr ...

Senior Research HPC Engineer

Hiring Organisation
MRC Laboratory of Medical Sciences
Location
London, United Kingdom
Employment Type
Permanent
Salary
£65,000
operates its own dedicated HPC environment, which is extensively used by multidisciplinary research groups and Institute facilities. The LMS IT department currently manages a SLURM-based cluster comprising CPU, high-memory (HMEM), and GPU nodes, supporting a diverse range of computational research. About the role As the Institutes computational ...

Senior Research HPC Engineer

Location
Greater London, England, United Kingdom
administering Linux‐based systems in an HPC, research, academic, or production environment Experience in configuration and maintenance of multi‐queue job scheduling systems (e.g. SLURM) Experience in the use of scientific software compilation and deployment systems (e.g. Spack, EasyBuild, Lmod, conda) Virtualisation and containerisation deployment and management (e.g. Docker … verbal and written communication skills Able to self‐motivate when working independently, on projects, and collaboratively within a team Effectively plans, multitask, and prioritises workload to achieve results, adapting quickly while maintaining control in challenging situations Ensures appropriate engagement of colleagues with relevant stakeholders and issues Ability to develop ...

Senior HPC Engineer

Hiring Organisation
Hays Specialist Recruitment Limited
Location
London, South East England, United Kingdom
Employment Type
Full-Time
Salary
£60,000 - £70,000 per annum
Senior HPC Engineer Please read the advert below carefully, and if you are a good match I want to speak to you ASAP. Please call or email Lorenz Pasch at Hays Recruitment - my contact details ...

Senior HPC Engineer

Hiring Organisation
Hays
Location
London, United Kingdom
Reference: 4546575Job ID: 5416319Posted: 2026-09-23Closing date: 2026-12-21Location: LondonSalary: Up to 70k (per annum)Job type: PermanentWorking pattern: Full timeIndustry: Scientific and R&D/School TechniciansCompany: HaysConsultant: Lorenz PaschHays ...

Senior AI Platform Engineer

Location
Greater London, England, United Kingdom
engineering solutions. Partner with centralised infrastructure teams to design and deliver high-performance compute environments across AWS and on-premises platforms, including GPU infrastructure, Slurm clusters, and migration from ad hoc research workflows. Optimise LLM training and inference workloads, supporting research and product teams in maximising performance, scalability … Nsight, DCGM, and related ecosystem technologies. Strong background in AWS cloud services, high-performance computing, distributed systems, containerised environments, and infrastructure automation. Experience with workload orchestration technologies such as Slurm, Kubernetes, Ray, or equivalent distributed compute frameworks. Demonstrated success bridging research and production environments, enabling rapid experimentation while ...

Senior DevOps Engineer

Location
Greater London, England, United Kingdom
high-performance, high-IOPS cloud storage solutions, ensuring seamless data synchronisation, backup strategies, and optimal throughput. Automate the deployment, configuration, and auto-scaling of workload orchestrators, container platforms, and core service components to ensure high resource utilisation and cost efficiency. Enforce strict security protocols across all environments by implementing … benefit updates for leadership. Experience - Desirable Proven track record of architecting and deploying production HPC workloads on AWS using AWS ParallelCluster, SOCA or custom Slurm fleets. Ability to architect for maximum cost efficiency, implementing automated spot-instance utilization, auto-scaling strategies, and Savings Plans optimization. Proficiency in monitoring infrastructure ...

HPC Support Engineer

Location
Greater London, England, United Kingdom
challenges and continually improve the platform. You’ll work across Linux based compute environments, GPU infrastructure, high performance storage and networking, with technologies including Slurm, Python, Bash and Ansible. You’ll also have exposure to HPC software management through Spack or EasyBuild, containerisation with Docker or Singularity, and monitoring … from the platform and translating complex technical requirements into practical solutions. If you have solid HPC or Linux infrastructure experience and enjoy working with Slurm, GPUs, automation and large scale systems, this is a chance to take on a technically varied role within an environment where engineering ...

Linux HPC Specialist

Hiring Organisation
MBDA
Location
Stevenage, United Kingdom
dynamic, fast paced development environment.Salary: Up to 75,000 depending on experienceDynamic (hybrid) working: 2-3 days per week on-site due to workload classificationSecurity Clearance This role will require DV Clearance. Restrictions and/or limitations relating to nationality and/or rights to work may apply. … leadership and oversight to HPC engineering and operations teamsPerformance & ScalabilityEnsure systems are designed for optimal throughput, latency, and scalabilityLead performance benchmarking, capacity planning, and workload modellingIdentify and eliminate architectural bottlenecksWorkload & Software EcosystemDefine strategies for workload orchestration (SLURM, Kubernetes for HPC, etc.)Guide software stack design (MPI, CUDA ...

Linux HPC Specialist

Hiring Organisation
MBDA
Location
Bristol, United Kingdom
dynamic, fast paced development environment.Salary: Up to 75,000 depending on experienceDynamic (hybrid) working: 2-3 days per week on-site due to workload classificationSecurity Clearance This role will require DV Clearance. Restrictions and/or limitations relating to nationality and/or rights to work may apply. … leadership and oversight to HPC engineering and operations teamsPerformance & ScalabilityEnsure systems are designed for optimal throughput, latency, and scalabilityLead performance benchmarking, capacity planning, and workload modellingIdentify and eliminate architectural bottlenecksWorkload & Software EcosystemDefine strategies for workload orchestration (SLURM, Kubernetes for HPC, etc.)Guide software stack design (MPI, CUDA ...

Senior Cloud Engineer (K8S)

Location
Greater London, England, United Kingdom
platform s. Experience with solutions for monitoring and observability. e.g. Grafana, Prometheus, OpenSearch/ElasticSearch, Loki. Experience with High Performance Computing (HPC) environments using SLURM or similar batch workload solutions. Programming experience with Python3 utilising classes and inheritance. Benefits In addition to a competitive salary flexible working ...

Build Engineer

Location
Greater London, England, United Kingdom
hybrid working model, required onsite 3 days a week. Experience for the Build Engineer includes: Software development with Python Modern build systems, e.g. Bazel Workload management, e.g. Slurm, LSF or SGE Infrastructure as code or IAC, e.g. Ansible or Terraform Containerisation, e.g. Docker Desired: Experience supporting silicon ...

Platform Architect - Nvidia AI/GB300

Hiring Organisation
Oscar Associates (UK) Limited
Location
London, United Kingdom
Employment Type
Contract
Contract Rate
£700 - £765 per day
data-centre teams. Key Requirements Strong Platform/Infrastructure Architecture experience across compute, storage, networking and Linux. Expert-level Kubernetes architecture experience. Strong Slurm and Run experience - essential. Proven experience with GPU/HPC environments and large-scale AI platforms. Hands-on experience with NVIDIA HGX GB300/NVL72 … including NVLink, NVSwitch and Grace Blackwell architecture. Experience with NVIDIA RTX 6000 series GPU servers. Strong understanding of GPU workload scheduling, partitioning and sharing, including MIG, vGPU and time-slicing Strong understanding of InfiniBand, RoCE, Spectrum-X, GPUDirect RDMA/Storage and high-performance AI fabrics. Experience with Terraform ...