51 to 73 of 73 Slurm Workload Manager Jobs in England

R&D Solution Architect

Location
Hook, England, United Kingdom
implementing solutions that adhere to FAIR data principles (Findable, Accessible, Interoperable, Reusable)." - Experience architecting for High-Performance Computing (HPC) environments, including knowledge of workload schedulers (e.g., SLURM) and applying cloud-native patterns to scientific, batch-processing workloads. - Familiarity with scientific workflow management tools (e.g., Nextflow, Snakemake ...

Solutions Engineer

Location
Greater London, England, United Kingdom
engineering, and operations teams to ensure proposed solutions are realistic, scalable, and aligned with platform standards Provide guidance on compute, networking, storage, orchestration, and workload optimisation for AI and machine learning use cases Help create repeatable demo environments, technical playbooks, reference architectures, and sales enablement materials … data centre, or platform environments Good understanding of cloud infrastructure, GPU compute, AI/ML workloads, or high-performance infrastructure Familiarity with containers, Kubernetes, Slurm, orchestration platforms, or workload deployment models Understanding of networking, storage, and distributed compute concepts in modern infrastructure environments Ability to quickly learn ...

High Performance Computing Architect (Linux)

Location
London, United Kingdom
parallel file systems, object storage, and tiered storage Provide technical leadership and oversight to engineering and operations teams Lead performance benchmarking, capacity planning, and workload modelling activities Identify and resolve system bottlenecks to ensure optimal throughput and scalability Establish standards, reference architectures, and best practices across HPC environments Collaborate … with internal stakeholders and external vendors to align infrastructure with evolving business needs Support the development of workload orchestration strategies using tools such as SLURM and Kubernetes The ideal candidate would have: Strong background in high-performance computing environments and infrastructure design Experience working with open-source technologies ...

High Performance Computing Architect (Linux)

Hiring Organisation
Eclectic Recruitment
Location
Stevenage, Hertfordshire, United Kingdom
Employment Type
Permanent
Salary
£65000 - £80000/annum
parallel file systems, object storage, and tiered storage Provide technical leadership and oversight to engineering and operations teams Lead performance benchmarking, capacity planning, and workload modelling activities Identify and resolve system bottlenecks to ensure optimal throughput and scalability Establish standards, reference architectures, and best practices across HPC environments Collaborate … with internal stakeholders and external vendors to align infrastructure with evolving business needs Support the development of workload orchestration strategies using tools such as SLURM and Kubernetes The ideal candidate would have: Strong background in high-performance computing environments and infrastructure design Experience working with open-source technologies ...

Python Software Engineer - Intraday Trading

Hiring Organisation
Millennium Management
Location
London, UK
Employment Type
Full-time
strong understanding of Linux operating systems Experience with grid scheduling and compute orchestration for real-time compute management at scale, including technologies such as SLURM Strong understanding of event-driven architecture and experience with messaging and caching technologies such as Kafka, Solace, Pulsar, Memcache, and Redis Experience building … scale real-time portfolio analytics tools, cloud platforms, containerization technologies such as Docker and Kubernetes, or multi-threaded C++ is a plusRecruiter: Ruby KazmiHiring Manager: Shashank GiriDepartment: Information Technology ...

Senior HPC Engineer

Location
West of England, England, United Kingdom
Linux system administration. Familiarity with at least one scripting language (e.g., Bash, Python). Interest in high-performance computing and willingness to learn Slurm, xCAT, and Ansible. Strong problem-solving skills and attention to detail. Good communication and collaboration skills. Ability to work independently with mentorship and as part … week. No hybrid/remote working option. Internship or academic experience in a research computing or HPC environment. Exposure to job schedulers (e.g., Slurm, LSF). Familiarity with version control systems (e.g., Git). Coursework or projects involving distributed systems, networking, or parallel computing. Understanding of basic cybersecurity concepts. ...

Staff Software Engineer, Kubernetes Platform

Hiring Organisation
Humanloop
Location
London, UK
Employment Type
Full-time
controllers — so it stays responsive as object counts and node counts grow by orders of magnitude. And we build the core cluster services every workload depends on, like service discovery, so they hold up under the same pressure. We make sure the control plane is fast, correct, and always … Anthropic's accelerator fleets, including custom scheduling plugins and policies for gang scheduling, topology awareness, and preemptionScale the Kubernetes control plane (apiserver, etcd, controller-manager) to support clusters far beyond typical limits, and find the next bottleneck before it finds usDesign, build, and operate core cluster services such ...

Senior Staff+ Software Engineer, Kubernetes Platform

Location
Greater London, England, United Kingdom
controllers — so it stays responsive as object counts and node counts grow by orders of magnitude. And we build the core cluster services every workload depends on, like service discovery, so they hold up under the same pressure. We make sure the control plane is fast, correct, and always … accelerator fleets, including custom scheduling plugins and policies for gang scheduling, topology awareness, and preemption Scale the Kubernetes control plane (apiserver, etcd, controller-manager) to support clusters far beyond typical limits, and find the next bottleneck before it finds us Design, build, and operate core cluster services such ...

Senior Specialist Field Engineer - HPC/AI/ML

Location
Greater London, England, United Kingdom
functionality, and performance, contributing regularly to discussions about product strategy and architecture. Conduct periodic technical reviews and assessments of customer workloads, pinpointing opportunities for workload optimization and suggesting suitable solutions. Stay informed of the latest developments and trends in Kubernetes, cloud computing and infrastructure, sharing your thought leadership with … related technical discipline, or equivalent experience 7+ years of proven experience as a Solutions Architect, Field Engineer, Engineer, Researcher, or Technical Account Manager in Cloud Infrastructure, focusing on building distributed systems or HPC/cloud services, with an expertise focused on AI/ML inference Fluency in cloud computing ...

Founding GPU Engineer

Hiring Organisation
Fuse Energy Supply
Location
London, UK
Employment Type
Full-time
Engineer to develop and optimise GPU-accelerated software for data centre systems: low-level performance engineering for large-scale compute clusters, tying GPU workload behaviour to energy availability and grid demand. This puts CUDA/GPU performance engineering at the centre of how Fuse scales its compute infrastructure. ResponsibilitiesDesign … multi-node scaling using NCCL, MPI, or similar communication librariesWork with data center infrastructure teams on power capping, dynamic voltage/frequency scaling, and workload scheduling strategies that reduce energy cost and carbon intensityCollaborate with ML/systems engineers to integrate custom kernels into training/inference pipelinesBenchmark against ...

Staff HPC Systems Software Engineer

Location
Greater London, England, United Kingdom
core HPC platform domain at Nscale. In this role, you will operate beyond a single team, shaping how multiple teams build, automate,and run Slurm-based capabilities within Nscale’s wider cloud-native platform. You’ll work acrossengineering boundaries to bring coherence to architecture, interfaces, lifecycle models, andoperational approaches … Domain Architecture & Technical Direction Own and evolve the technical direction for a defined HPC systems domain, such as Slurmplatform architecture, scheduler integrations, cluster lifecycle, workload environments orservice automation. Make architectural decisions that balance software quality, operational realities, customerneeds, and long-term maintainability. Define how proven Slurm implementations should ...

Founding GPU Engineer

Location
Greater London, England, United Kingdom
node scaling using NCCL, MPI, or similar communication libraries. Work with data center infrastructure teams on power capping, dynamic voltage/frequency scaling, and workload scheduling strategies that reduce energy cost and carbon intensity. Collaborate with ML/systems engineers to integrate custom kernels into training/inference pipelines. … data center power/thermal management or demand-response systems. Background in HPC, quantitative finance, or large-scale distributed systems. Familiarity with Kubernetes/Slurm for GPU cluster orchestration. Interest or experience in energy markets, grid systems, or sustainability-focused compute. Competitive salary and an equity sign-on bonus. ...

Sr Lead Software Engineer - C++, Python,

Location
Greater London, England, United Kingdom
solutions using Python, C++; and Go. Deploy and manage applications using Kubernetes (K8s) and HPC/Grid Computing/Big Data technologies such as Slurm, LSF, Spark, Ray and Symphony in cloud environments. Required Qualifications, Capabilities, And Skills Formal training or certification on Software Engineering concepts and 5+ years … applied experience. Expertise in C++, Python, Kubernetes (K8s), and HPC/Grid Computing technologies (Slurm, LSF, Symphony), GenAI, agent framework and tools, observability tooling such as OTel Proven experience with cloud technologies and environments Strong problem-solving skills and the ability to work collaboratively in a team setting. Experience ...

Principal - AI & HPC Data Centre Compute

Location
Greater London, England, United Kingdom
inference workloads (LLMs, multimodal and scientific AI) across distributed clusters for cost, throughput and latency targets Lead HPC cluster design and orchestration using Slurm, Kubernetes and parallel processing models (MPI) Enhance GPU utilisation by addressing bottlenecks across compute, memory and data pipelines; collaborate with energy teams on power-aware … infrastructure, accelerated computing or distributed systems architecture Deep knowledge of GPU architectures, AI workloads, networking and large-scale cluster operations Hands-on expertise with Slurm, Kubernetes and performance optimisation across multi-node environments Proven ability to advise senior stakeholders and influence technical strategy at C-level Demonstrated experience ...

Research Software Engineer

Location
Greater London, England, United Kingdom
surface live metrics. Write efficient, well-tested Python and systems code; enforce code review, CI, and observability. Design and optimise distributed services (Kubernetes/SLURM, thousands-of-GPU jobs). Prototype utilities (CLI, dashboards) and carry them through to stable, shared libraries. About the Research Engineering team Based … observability. Fluency in Python plus one systems language (C++, Rust, Go or Java). Hands-on with container orchestration and schedulers (Kubernetes/K8s, SLURM, or similar). Comfortable profiling performance, optimising I/O, and automating workflows. Self-starter, low-ego, collaborative, high-energy. Nice-to-haves Exposure ...

HPC Operations Lead

Hiring Organisation
LinuxRecruit
Location
London, United Kingdom
operation of high performance storage services, supporting both internal workloads and external collaboration. The environment includes large scale HPC clusters, Linux based systems, workload schedulers such as Slurm, networking with Infiniband and parallel file systems such as GPFS. Experience with high performance storage at petabyte scale is particularly ...

Senior Linux & HPC Systems Engineer

Location
Leatherhead, England, United Kingdom
Inc. seeks a Senior IT Systems Administrator (Linux & HPC) to own and operate Linux-based enterprise and HPC platforms, including SLURM scheduling, Nvidia Base Command Manager, and CycleCloud. You will work with engineering and scientific users to diagnose issues spanning compute, storage, and networking. The role emphasizes hands ...

Senior Manager – Counterparty Credit Risk & XVA

Location
Greater London, England, United Kingdom
Your role As a Senior Manager in Counterparty Credit Risk (CCR) and XVA at Zanders, you will join our global Financial Institutions team in London. Your remit is to lead quantitative traded risk engagements across CCR, XVA and the high-performance computing (HPC) that underpins them, working in multidisciplinary … sales in the CCR, XVA and HPC space. You can build on your existing UK network and expand it over time. As a Senior Manager you act as a career coach, responsible for the development of up to three direct reports, supporting them on project work and across their ...

HPC Architect: Storage & Infrastructure Lead (Hybrid)

Location
Stevenage, England, United Kingdom
will lead architecture, standards, and collaboration with engineering teams across national and international sites to deliver scalable, high-throughput infrastructure. You will drive workload strategies with SLURM, Kubernetes for HPC, MPI/CUDA stacks, and ensure reproducibility and portability of workloads while partnering with MBDA vendors #J ...

Senior Software Engineer - Research Technology

Location
Greater London, England, United Kingdom
fundamentals: data structures, algorithms, networking, OS, concurrency, and system design. Experience running compute at cluster scale: job scheduling, resource management, retries, and reliability. Slurm, Kubernetes, Ray, Spark, or custom internal schedulers all count. Proven data‐engineering experience: schema design, storage formats, compression, I/O trade-offs, and pipelines … ship production software safely and repeatedly, with an obsession for data driven quality. Desirable/nice-to-have Rust experience alongside C++ and Python. Slurm or other cluster scheduler expertise. Familiarity with ML/Deep Learning frameworks. Prior finance or market‐data experience, including low‐level market connectivity. ...

Software Engineer

Hiring Organisation
Microtech Global Ltd
Location
Cambridgeshire, East Anglia, United Kingdom
Employment Type
Contract
network storage principles, particularly NFS mount configurations and troubleshooting. Preferred/"Nice to Have" Experience Hands-on exposure to distributed systems, compute farms, or workload managers such as LSF (IBM Spectrum) , Grid Engine , or Slurm . ...

Lead HPC Engineer

Hiring Organisation
LinuxRecruit
Location
London, UK
Employment Type
Full-time
forefront of designing, optimising, and managing advanced computational infrastructure. You'll have a solid grasp of all things HPC, Linux, Slurm, and storage systems (bonus points if you're familiar with GPFS). Your expertise will ensure the systems are reliable, scalable, and high-performing, ready to support researchers … keeping our infrastructure at the forefront of innovation. We're looking for someone with deep expertise in HPC environments, including: Linux systems, workload management, parallel storage, and high-speed networking. You'll also bring strong leadership skills, inspiring and managing teams, while rolling up your sleeves to tackle technical ...

Staff HPC Systems Engineer - Slurm & Cloud Platform Lead

Location
Greater London, England, United Kingdom
London is hiring a Staff HPC Systems Software Engineer to define the technical direction of a core HPC platform. You will scope architecture for Slurm-based services, shaping how multiple teams build, automate, and run the platform in a cloud-native environment. You’ll work across engineering boundaries ...