51 to 75 of 88 CUDA Jobs in England

AI Research Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
plus; a strong background in mathematics is required. Alternatively, experience training models with strong math skills. Strong coding ability in Rust (preferred), CUDA, or C. Solid understanding of transformer architectures. A product mindset; you’ll build production‐ready products. Hybrid Remote with the London Office. What ...

Senior ML Systems Engineer, Frameworks & Tooling

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
libraries, or custom kernels/fused ops. Experience with multi-node cluster orchestration (Slurm, Ray, Kubernetes, or similar). Comfort debugging performance issues across CUDA/NCCL, networking, IO, and data pipelines. Experience working with containerized environments (Docker, Singularity/Apptainer). A track record of building tools that ...

ML Performance Engineer – Scale GPU/CPU Workloads

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
evolve the compute stack. The role shapes platform evolution and enables researchers to push the boundaries of machine learning. You will work with Python, CUDA, Kubernetes, and deep learning frameworks like PyTorch, applying #J-18808-Ljbffr ...

ML Engineering Manager, Industrial Vision & Robotics

Hiring Organisation
Jobleads-UK
Location
England, United Kingdom
systems, collaborating with software, product management and research teams. You will require BS+10y or MS+7y in CS/ML/Robotics, expert Python and CUDA/PyTorch skills, and a track record in computer vision and model fine-tuning. #J-18808-Ljbffr ...

ML Systems Engineer Intern — High-Perf Finance

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
apply advanced techniques, optimize training pipelines on HPC clusters, and integrate ML models into latency-sensitive production environments, using C/C++, Python, and CUDA across large-scale systems. #J-18808-Ljbffr ...

ML Systems Engineer for Quantitative Finance

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
systems for quantitative finance. You will optimize training pipelines, deploy models to low-latency production environments, and work across C/C++, Python, CUDA, and related GPU technologies. Join a fast-paced, collaborative team focused on impactful projects that push the boundaries of AI research and its applications ...

Member of Technical Staff (Infrastructure Engineer, Training and Inference Systems)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
front and centre. Strong candidates may also have Experience with a systems programming language like Rust or C++. Experience with writing and profiling CUDA kernels. A track record of building reliable research tools. Why this is interesting You’ll shape the core technical foundation of a frontier ...

Senior Applied Diffusion Engineer

Hiring Organisation
ASOS
Location
London, United Kingdom
Salary
£ 80 K
vision concepts including segmentation, pose estimation, and embeddings. Experience building production ML systems and inference pipelines. Strong Python engineering skills. Experience with GPU optimisation, CUDA fundamentals, TensorRT, or ONNX is advantageous. Experience in fashion, e-commerce, creative tooling, gaming avatars, or virtual try-on is highly desirable. Additional InformationBeneFITS ...

Research Engineer

Hiring Organisation
Jobleads-UK
Location
Oxford, England, United Kingdom
training (fine‐tuning, alignment, reward modelling). Publications at top‐tier AI conferences such as ICML, ICLR, NeurIPS, CVPR, etc. Experience using HPCs and CUDA for training large‐scale models. The ability to translate research into a product vision and carry it through to delivery. Nice to have Computer ...

Autonomous AI Research Engineer — Hybrid-Remote

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
practical product impact and close collaboration with product teams. The role requires a strong math background, expertise in PyTorch, and experience with Rust/CUDA for high-performance inference. This is a unique opportunity to shape the AI strategy for a fast-growing platform. #J-18808-Ljbffr ...

Research Engineer

Hiring Organisation
Cubiq Recruitment
Location
City of London, London, United Kingdom
purely academic backgrounds. Full-stack depth: distributed training frameworks such as Megatron-LM, DeepSpeed or FSDP, GPU performance work, and experience down to the CUDA, Triton or XLA level. Contributions to open-source ML frameworks or research codebases. They genuinely value this. Strong software engineering habits. Clean code, good ...

Senior ML Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
optimize the end-to-end ML stack: data pipelines, training loops, inference serving, and deployment. Design and implement GPU-accelerated components, including custom CUDA kernels where off-the-shelf libraries are not enough. Work closely with the founders to translate product requirements into concrete optimization goals and technical roadmaps. ...

Developer Experience Engineer New London

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
SDKs, APIs, or platforms that other engineers depend on. Experience developing for inference hardware and/or GPUs (e.g. one year or more using CUDA or ROCm). A good understanding of modern ML workloads and the practical challenges of deploying large language models to production. An instinct ...

Senior Machine Learning Engineer, AI Performance

Hiring Organisation
wayve
Location
London, United Kingdom
Salary
£ 80 K
deep learning models in PyTorch (not just using high-level tooling).Strong proficiency with at least one relevant stack/toolchain (e.g. TensorRT, CUDA, Qualcomm QNN, Triton, OpenCL) and confidence learning adjacent frameworks quickly.Comfort operating at multiple levels of abstraction — from high-level model behaviour down to low-level ...

Senior Machine Learning Engineer, AI Performance

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
deep learning models in PyTorch (not just using high‐level tooling). Strong proficiency with at least one relevant stack/toolchain (e.g. TensorRT, CUDA, Qualcomm QNN, Triton, OpenCL) and confidence learning adjacent frameworks quickly. Comfort operating at multiple levels of abstraction — from high-level model behaviour down ...

Product Manager - ENT AI Infrastructure

Hiring Organisation
SCC
Location
Birmingham, West Midlands (County), United Kingdom
Salary
£ 60 K
. Nice to have:• Product experience in datacentre hardware, cloud infrastructure, HPC, AI platforms, or security/compliance products. • Familiarity with GPU ecosystem concepts (CUDA awareness, GPU memory constraints, interconnects, training vs. inference economics). • Understanding of private AI solution patterns and components (RAG, vector databases, model serving, evaluation ...

Founding AI Inference Engineer – Scale & Serving Expert

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Inference Engineer to define and build how we serve AI workloads at scale, reporting to the CTO. You’ll own the layer above CUDA/GPU work, shaping inference delivery, throughput, and reliability as we scale data-centre compute for energy applications. With 4+ years in large-scale inference ...

GPU Systems Engineer

Hiring Organisation
Hudson River Trading
Location
London, United Kingdom
Salary
£ 80 K
system installation, performance tuning, and troubleshootingExpertise in troubleshooting distributed GPU workloadsDeep knowledge around GPU optimization and performance Proficiency in Python scripting and automation frameworks CUDA or C/C++ experience is a plusExperience with NVIDIA technologies beyond CUDA, such as NCCL, GPUDirect RDMA, and NVLink Familiarity with configuration ...

Campus ML Research Engineer (Intern)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
resources. Integrate ML models into production systems where latency matters. Work across a mix of programming languages: C/C++/Python/CUDA and other low-level GPU languages. Build large scale ML systems that are observable, performant, and flexible. Help improve productivity by reducing the iteration cycle … Proficiency in Pytorch, JAX, Tensorflow or other DL library. Ability to thrive in a collaborative, team-oriented environment Expertise in GPU or Accelerator programming (CUDA, Triton, SYCL, ROCm or equivalent) Experience building ML systems at large scale (hundreds of TBs of training data, low latency or high throughput inference ...

Campus ML Research Engineer (Full-Time)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
resources. Integrate ML models into production systems where latency matters. Work across a mix of programming languages: C/C++/Python/CUDA and other low-level GPU languages. Build large scale ML systems that are observable, performant, and flexible. Help improve productivity by reducing the iteration cycle … Proficiency in Pytorch, JAX, Tensorflow or other DL library. Ability to thrive in a collaborative, team-oriented environment Expertise in GPU or Accelerator programming (CUDA, Triton, SYCL, ROCm or equivalent) Experience building ML systems at large scale (hundreds of TBs of training data, low latency or high throughput inference ...

Research Engineer

Hiring Organisation
European Tech Recruit
Location
London Area, United Kingdom
Research Engineer Cambridge or London, UK (100% Onsite) Our client is a global leader in semiconductor innovation, developing advanced technologies that power billions of smart devices worldwide. With significant investment in artificial intelligence research, the ...

Research Software Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
About Mistral Mistral provides full-stack AI solutions: from frontier models to developer tools, applications, and compute. We partner with enterprises tackling the hardest problems—across high-stakes industries like finance, manufacturing, defense, healthcare, and ...

Senior Principal AI Infrastructure Architect

Hiring Organisation
NTT
Location
London, United Kingdom
Salary
£ 80 K
systems, storage, fabric, MLOps stack and managed services — to land service-led AI solutions. Lead integration of compute, storage, networking, the AI software stack (CUDA, ROCm, Triton, NIM, NVIDIA AI Enterprise, Run:ai, Slurm, Kubernetes/Kubeflow) and managed-service operating models across multiple domains, delivery units and geographies. …/NVSwitch fabrics, congestion control and fabric design for rail-optimised and fat-tree topologies. Working knowledge of the AI software and orchestration stack: CUDA, cuDNN, NCCL, ROCm, Triton Inference Server, NIM, vLLM, TensorRT-LLM, Slurm, Kubernetes (with GPU Operator), Kubeflow, Run:ai, MLflow and NVIDIA AI Enterprise. Familiarity ...

Machine Learning Performance Engineer

Hiring Organisation
G Research
Location
London, United Kingdom
Salary
£ 80 K
We tackle the most complex problems in quantitative finance, by bringing scientific clarity to financial complexity.From our London HQ, we unite world-class researchers and engineers in an environment that values deep exploration and methodical ...

Machine Learning Performance Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
We tackle the most complex problems in quantitative finance, by bringing scientific clarity to financial complexity. From our London HQ, we unite world-class researchers and engineers in an environment that values deep exploration and ...