Principal - AI & HPC Data Centre Compute
- Location
- Greater London, England, United Kingdom
clients deliver performance and efficiency at scale. Responsibilities Advise hyperscalers, neoclouds and enterprises on compute strategies for AI and HPC, including power, cooling and network design Optimise AI training and inference workloads (LLMs, multimodal and scientific AI) across distributed clusters for cost, throughput and latency targets Lead … cluster design and orchestration using Slurm, Kubernetes and parallel processing models (MPI) Enhance GPU utilisation by addressing bottlenecks across compute, memory and data pipelines; collaborate with energy teams on power-aware scheduling Architect scalable AI platforms integrating MLOps frameworks, automation tools and reusable assets Develop repeatable AI/ ...