5 of 5 Permanent NVIDIA Jobs in Gloucester

Platform Site Reliability Engineer

Location
Gloucester, England, United Kingdom
with Platform Engineering/Platform SRE to fully support both our infrastructure and platform stacks. Willingness to cross train with HPC Engineering, supported by NVIDIA to enhance our HPC supportability offering Requirements 5+ Years Proven experience in globally scaled, performance-intensive environments operating to a 24/7 support model ...

Infrastructure Site Reliability Engineer

Location
Gloucester, England, United Kingdom
with Platform Engineering/Platform SRE to fully support both our infrastructure and platform stacks. Willingness to cross train with HPC Engineering, supported by NVIDIA to enhance our HPC supportability offering What you bring 5+ Years Proven experience in globally scaled, performance-intensive environments operating to a 24/ ...

Cloud Infrastructure Support Engineer

Location
Gloucester, England, United Kingdom
most advanced high‐performance computing infrastructure. As a Cloud Support Engineer, you will support cutting‐edge GPU and CPU platforms — including the latest NVIDIA architectures — powering dense, large‐scale compute environments used for AI, machine learning, and next‐generation workloads. This is an opportunity to build expertise at the forefront ...

24/7 HPC Infra SRE for AI & GPU Compute

Location
Gloucester, England, United Kingdom
accelerated HPC infrastructure in a 24/7 production environment. You will work across network, storage, virtualization and orchestration with hands‐on Linux expertise, NVIDIA GPU ecosystems, RoCE/InfiniBand, and performance benchmarking. This role champions observability, automation and on‐call reliability, shaping next‐gen HPC platforms within a globally ...

HPC Infrastructure Site Reliability Engineer

Location
Gloucester, England, United Kingdom
generation GPU infrastructure at significant scale. You will bring strong breadth across bare metal, networking, storage, virtualisation, and orchestration, alongside deep HPC experience including NVIDIA GPU ecosystems, RDMA networking (RoCE and InfiniBand), and performance validation and benchmarking. Strong Linux and distributed systems expertise is essential. Alongside operational ownership, this … high‐performance computing infrastructure. As a HPC Infrastructure SRE, you’ll work hands‐on with cutting‐edge GPU and CPU platforms ‐ including the latest NVIDIA architectures ‐ powering dense, large‐scale compute environments used for AI, machine learning, and next‐generation workloads. This is an opportunity to build expertise ...