51 to 57 of 57 Slurm Workload Manager Jobs in the UK

Staff Software Engineer, Kubernetes Platform

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
controllers — so it stays responsive as object counts and node counts grow by orders of magnitude. And we build the core cluster services every workload depends on, like service discovery, so they hold up under the same pressure. We make sure the control plane is fast, correct, and always … accelerator fleets, including custom scheduling plugins and policies for gang scheduling, topology awareness, and preemption Scale the Kubernetes control plane (apiserver, etcd, controller‐manager) to support clusters far beyond typical limits, and find the next bottleneck before it finds us Design, build, and operate core cluster services such ...

Senior Staff+ Software Engineer, Kubernetes Platform

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
controllers — so it stays responsive as object counts and node counts grow by orders of magnitude. And we build the core cluster services every workload depends on, like service discovery, so they hold up under the same pressure. We make sure the control plane is fast, correct, and always … accelerator fleets, including custom scheduling plugins and policies for gang scheduling, topology awareness, and preemption Scale the Kubernetes control plane (apiserver, etcd, controller-manager) to support clusters far beyond typical limits, and find the next bottleneck before it finds us Design, build, and operate core cluster services such ...

Founding GPU Engineer

Hiring Organisation
Fuse Energy Supply
Location
London, United Kingdom
Salary
£ 80 K
center systems. You'll work on low-level performance engineering for large-scale compute clusters, helping Fuse build the software layer that ties GPU workload behaviour to energy availability and grid demand.The OpportunityDemand for high-performance compute capacity across the markets we operate in significantly outpaces what … multi-node scaling using NCCL, MPI, or similar communication libraries.Work with data center infrastructure teams on power capping, dynamic voltage/frequency scaling, and workload scheduling strategies that reduce energy cost and carbon intensity.Collaborate with ML/systems engineers to integrate custom kernels into training/inference pipelines.Benchmark against ...

AI Infrastructure Engineer — Scalable ML Clusters

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
technology firm in London is seeking an AI Infrastructure Engineer to join their team. The role involves designing and maintaining scalable Kubernetes clusters, optimizing Slurm-based HPC environments, and developing APIs for AI workloads. Candidates should have expertise in Kubernetes, experience with Slurm, and skills in Python ...

Research Software Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
surface live metrics. Write efficient, well-tested Python and systems code; enforce code review, CI, and observability. Design and optimise distributed services (Kubernetes/SLURM, thousands-of-GPU jobs). Prototype utilities (CLI, dashboards) and carry them through to stable, shared libraries. About the Research Engineering team Based … observability. Fluency in Python plus one systems language (C++, Rust, Go or Java). Hands-on with container orchestration and schedulers (Kubernetes/K8s, SLURM, or similar). Comfortable profiling performance, optimising I/O, and automating workflows. Self-starter, low-ego, collaborative, high-energy. Nice-to-haves Exposure ...

HPC Operations Lead

Hiring Organisation
LinuxRecruit
Location
London, United Kingdom
Salary
£ 70 K
operation of high performance storage services, supporting both internal workloads and external collaboration. The environment includes large scale HPC clusters, Linux based systems, workload schedulers such as Slurm, networking with Infiniband and parallel file systems such as GPFS. Experience with high performance storage at petabyte scale is particularly ...

Lead HPC Engineer

Hiring Organisation
LinuxRecruit
Location
London, United Kingdom
Salary
£ 80 K
forefront of designing, optimising, and managing advanced computational infrastructure. You’ll have a solid grasp of all things HPC, Linux, Slurm, and storage systems (bonus points if you’re familiar with GPFS). Your expertise will ensure the systems are reliable, scalable, and high-performing, ready to support researchers … keeping our infrastructure at the forefront of innovation. We’re looking for someone with deep expertise in HPC environments, including: Linux systems, workload management, parallel storage, and high-speed networking. You’ll also bring strong leadership skills, inspiring and managing teams, while rolling up your sleeves to tackle technical ...