126 to 143 of 143 CUDA Jobs in England

Founding GPU Engineer

Hiring Organisation
Fuse Energy Supply
Location
London, United Kingdom
Salary
£ 80 K
data centre systems: low-level performance engineering for large-scale compute clusters, tying GPU workload behaviour to energy availability and grid demand. This puts CUDA/GPU performance engineering at the centre of how Fuse scales its compute infrastructure.ResponsibilitiesDesign, implement, and optimise CUDA kernels for high-throughput, latency … against CPU/GPU baselines and drive continuous performance improvementsContribute to internal libraries, documentation, and best practices for GPU performance engineeringRequirements4+ years writing production CUDA code, or equivalent strong project/industry experienceDeep understanding of GPU architecture (SMs, warps, memory hierarchy, occupancy)Proficiency in C++ and CUDA; experience ...

AI Research Engineer, Pre-Training

Hiring Organisation
Hudson River Trading
Location
London, United Kingdom
Salary
£ 70 K
impactful on the business, and it will be challenging: this is a field with no easy or obvious solutions.QualificationsStrong engineering skills, especially any of: CUDA/Triton/Pallas/CuTe DSL kernel development, lower-level PyTorch/JAX/XLA development, CUDA Graphs, FPGA/ASIC experienceMust ...

Campus ML Engineer: Build Scalable AI for Finance

Location
Greater London, England, United Kingdom
Jump Trading Group is seeking world-class engineers to collaborate with our research, trading and engineering teams to build state-of-the-art ML systems for quantitative finance. You will work on training pipelines, low ...

Founding GPU Engineer

Location
Greater London, England, United Kingdom
operate in significantly outpaces what we can currently build, meaning speed to power and reliability are critical to how we scale. This puts CUDA/GPU performance engineering at the center of how Fuse scales its compute infrastructure. Responsibilities Design, implement, and optimise CUDA kernels for high-throughput … baselines and drive continuous performance improvements. Contribute to internal libraries, documentation, and best practices for GPU performance engineering. 4+ years of experience writing production CUDA code, or equivalent strong project/industry experience. Deep understanding of GPU architecture (SMs, warps, memory hierarchy, occupancy). Proficiency in C++ and CUDA ...

CUDA Engineer

Location
Greater London, England, United Kingdom
optimising how power‐dense GPU workloads are scheduled, cooled, and balanced against grid conditions in real time. We're looking for a CUDA Engineer to write and optimise the low‐level GPU code that powers our inference workloads. You'll design custom CUDA kernels, tune performance across memory … operate in significantly outpaces what we can currently build, meaning speed to power and reliability are critical to how we scale. This puts CUDA/GPU performance engineering at the center of how Fuse scales its compute infrastructure. Responsibilities Write and optimise custom CUDA kernels for core transformer ...

CUDA Engineer

Hiring Organisation
Fuse Energy Supply
Location
London, United Kingdom
Salary
£ 80 K
sources of electricity demand, Fuse is expanding into high-performance compute infrastructure at the intersection of energy and AI. We're looking for a CUDA Engineer to write and optimise the low-level GPU code that powers our inference workloads: designing custom CUDA kernels, tuning performance across memory … squeezing maximum throughput out of every GPU in our fleet, working at the level of SMs, warps and memory hierarchies.ResponsibilitiesWrite and optimise custom CUDA kernels for core transformer inference operationsProfile kernels to identify and eliminate bottlenecks in occupancy, memory throughput and warp divergenceApply kernel fusion to reduce memory round ...

AI Inference Engineer

Location
Greater London, England, United Kingdom
demand, Fuse is expanding into high-performance compute infrastructure that sits at the intersection of energy and AI. We're building the GPU/CUDA performance layer and the inference serving layer at the same time, from scratch -- and we're looking for the founding engineer … Founding AI Inference Engineer to define and build how Fuse serves AI inference workloads at scale, reporting directly to the CTO. Where our CUDA and GPU engineering hires own kernel-level and hardware performance, this role owns the layer above it: how models actually get served, scaled, and delivered ...

AI Inference Engineer

Hiring Organisation
Fuse Energy Supply
Location
London, United Kingdom
Salary
£ 80 K
electricity demand, Fuse is expanding into high-performance compute infrastructure at the intersection of energy and AI. We're building the GPU/CUDA performance layer and the inference serving layer at the same time, from scratch, and we're looking for the founding engineer to own the latter. … serving, deciding where and how to apply quantisation, distillation, speculative decoding and similar techniques to improve throughput and cost per token, partnering with the CUDA/GPU engineersMake the core software architecture calls on serving frameworks and orchestration (e.g. vLLM, TensorRT-LLM, SGLang, Triton Inference Server or equivalents)Translate ...

Linux/RHEL Engineer - HPC

Location
Stevenage, England, United Kingdom
application support in Linux HPC environments MPI, compilers and scientific libraries Hardware, OS, scheduler and application troubleshooting ServiceNow or equivalent ITSM experience GPU/CUDA, Docker, Ansible and InfiniBand are desirable The role is based in Stevenage with a minimum of 3 days onsite each week. Desired Skills … Experience RHEL 7, RHEL 8, RHEL 9, Slurm, HPC Clusters, MPI, Scientific Applications, GPU Computing, CUDA, Ansible #J-18808-Ljbffr ...

Founding GPU Engineer — CUDA Performance for HPC

Location
Greater London, England, United Kingdom
Fuse Energy, LLC seeks an experienced CUDA performance engineer to design, optimize, and deploy high-throughput CUDA kernels across multi-GPU systems. You will profile hardware bottlenecks, develop tooling to monitor energy usage, and collaborate with ML engineers to integrate kernels into scalable training and inference pipelines. … role requires deep GPU architecture knowledge, strong C++/CUDA skills, and experience with Nsight, NCCL, and MPI. #J-18808-Ljbffr ...

LLM/GenAI MLOps Engineer - Platform Architect

Location
Sheffield, England, United Kingdom
慨正橡扯 is seeking an MLOps Engineer (LLM/GenAI) based in Sheffield, UK, to engineer production-grade infrastructure for modern AI. The ideal candidate will design scalable model hosting platforms, optimise inference performance, and build ...

Senior CUDA Engineer for GPU-Accelerated Crypto Protocols

Location
England, United Kingdom
Lawrence Harvey Search & Selection partners with a high-growth infrastructure team building GPU-accelerated cryptographic systems. We seek a Senior Software Engineer with deep CUDA expertise to own the performance layer and push GPU pipelines to the limit. You will work alongside researchers and infrastructure engineers to integrate CUDA ...

Machine Learning Performance Engineer

Location
Greater London, England, United Kingdom
about efficient large-scale training, low-latency inference in real-time systems and high-throughput inference in research. Part of this is improving straightforward CUDA, but the interesting part needs a whole-systems approach, including storage systems, networking and host- and GPU-level considerations. Zooming in, we also want … level GPU knowledge of PTX, SASS, warps, cooperative groups, Tensor Cores and the memory hierarchy Debugging and optimisation experience using tools like CUDA GDB, NSight Systems, NSight Computesight-systems and nsight-compute Library knowledge of Triton, CUTLASS, CUB, Thrust, cuDNN and cuBLAS Intuition about the latency and throughput characteristics ...

Machine Learning Performance Engineer

Location
Greater London, England, United Kingdom
about efficient large-scale training, low-latency inference in real-time systems and high-throughput inference in research. Part of this is improving straightforward CUDA, but the interesting part needs a whole-systems approach, including storage systems, networking and host- and GPU-level considerations. Zooming in, we also want … end. Low-level GPU knowledge of PTX, SASS, warps, cooperative groups, Tensor Cores and the memory hierarchy. Debugging and optimisation experience using tools like CUDA GDB, NSight Systems, NSight Compute-sight-systems and nsight-compute. Library knowledge of Triton, CUTLASS, CUB, Thrust, cuDNN and cuBLAS. Intuition about the latency ...

ML Systems Performance Engineer

Location
Greater London, England, United Kingdom
Quant Blueprint LLC is seeking an engineer proficient in low-level systems programming and optimization to enhance our machine learning team. This role centers on optimizing model performance, both for training and real-time inference. ...

Senior CUDA Engineer - High-Performance GPU Inference

Location
Greater London, England, United Kingdom
Fuse Energy is seeking a CUDA Engineer to design and optimize low-level GPU kernels powering our transformer inference workloads. You will write custom CUDA kernels, tune memory bandwidth, and push the throughput of our GPU fleet at the SM, warp, and memory hierarchy level. Collaborate with ...

Senior CUDA Engineer for High-Performance Inference

Location
Greater London, England, United Kingdom
Fuse Energy is seeking an experienced CUDA Engineer to design and optimise low‐level GPU code for high‐throughput inference workloads. You will work on custom CUDA kernels, memory hierarchies, and mixed‐precision arithmetic to maximise throughput and minimise latency across GPU fleets. The role focuses on profiling ...