26 to 50 of 104 CUDA Jobs in London

Senior Machine Learning Research Engineer

Location
Greater London, England, United Kingdom
infrastructure and containerised environments. A track record of taking research code from prototype to robust, reusable infrastructure that other people actually use. Bonus experience: CUDA/Triton kernel development; FlashAttention-style attention implementations; experience with foundation models for biology, vision, or language; contributions to open-source ML frameworks. Personally ...

Staff Software Engineer, Inference

Location
Greater London, England, United Kingdom
modern inference frameworks (e.g., vLLM, Triton, TensorRT‐LLM, Ray Serve, or TorchServe). Deep experience with GPU systems engineering and hardware performance optimisation (e.g., CUDA, NCCL, RDMA, NUMA, or GPU interconnects). Direct exposure to large‐scale AI/ML infrastructure or hyperscale cloud environments. Wondering ...

Senior Field Application Engineer - HPC (UK + Multiple European Locations)

Hiring Organisation
AMD
Location
London, UK
Employment Type
Full-time
audience Some Linux administration; understanding setup for HPC middleware. Nice to Haves:5+ years HPC application experienceExperience building and running HPC applications on GPU. CUDA or OpenACC or OpenMP paradigms. Experience running AI models on CPU or GPU.Any experience understanding/inspecting/writing assembly Understanding of memory ...

Machine Learning Engineer

Location
Greater London, England, United Kingdom
automation and orchestration systems for these platforms Proficiency in a low-level language such as C, C++, or Rust and in GPU frameworks like CUDA Competence in front-end web design to allow easy interfacing with large datasets Our Culture Follow the science. We prioritise rigorous scientific inquiry, relying ...

Contract C ++ Developer

Hiring Organisation
Hackajob Ltd
Location
South West London, London, United Kingdom
Employment Type
Contract
equivalent safety standards. Experience working with state-machine-driven supervisors and fault-management systems. Experience with NVIDIA Jetson, CUDA or GPU-accelerated processing . Experience with Yocto and embedded Linux board-support packages. Experience with systemd, sd_notify, D-Bus and journald. Experience with OpenSSL, TensorRT or similar external ...

Senior C++ Developer

Hiring Organisation
Hackajob Ltd
Location
South West London, London, United Kingdom
Employment Type
Contract
equivalent safety standards. Experience working with state-machine-driven supervisors and fault-management systems. Experience with NVIDIA Jetson, CUDA or GPU-accelerated processing. Experience with Yocto and embedded Linux board-support packages. Experience with systemd, sd_notify, D-Bus and journald. Experience with OpenSSL, TensorRT or similar external technologies. ...

Contract C + Developer

Location
Greater London, England, United Kingdom
equivalent safety standards. Experience working with state-machine-driven supervisors and fault-management systems. Experience with NVIDIA Jetson, CUDA or GPU-accelerated processing . Experience with Yocto and embedded Linux board-support packages. Experience with systemd, sd_notify, D-Bus and journald. Experience with OpenSSL, TensorRT or similar external ...

c++ developer for medical technology

Location
Greater London, England, United Kingdom
understanding of IEC 60601-1-8 or equivalent safety standards, experience with state-machine-driven supervisors and fault-management systems, experience with NVIDIA Jetson, CUDA or GPU-accelerated processing, experience with Yocto and embedded Linux board-support packages, experience with systemd, sd_notify, D-Bus and journald, experience with ...

Optical Systems Product Engineer (Photonic Quantum Computing)

Location
Greater London, England, United Kingdom
into existing AI and HPC infrastructure. Its current PT-2 system, a rack-mounted quantum computer, connects directly to GPU clusters via NVIDIA’s CUDA-Q platform, enabling accelerated AI workloads with significantly lower energy consumption than traditional silicon-based systems. Backed by a strong portfolio of patent families ...

Machine Learning Engineer

Hiring Organisation
Your Tech Future
Location
South West London, London, United Kingdom
Employment Type
Permanent
architectures and model optimisation Excellent software engineering and problem-solving abilities Experience working with complex datasets and real-world machine learning challenges Desirable Skills CUDA TensorRT Triton Quantised model training Edge AI deployment Computer Vision Robotics or autonomous systems Advanced C++ development Synthetic data generation What We're Looking ...

Machine Learning Researcher

Location
City Of London, England, United Kingdom
finance, trading, or quantitative research (not required). Publications, competition results (e.g., Kaggle, academic ML contests), or open-source contributions. Familiarity with C++, CUDA, or low-latency systems. Here is why you should join our dynamic team: Opportunity to work at one of the world's leading algorithmic trading ...

Senior Machine Learning Engineer - AV Core

Hiring Organisation
wayve
Location
London, UK
Employment Type
Full-time
time alerts. Experience with transformer-based and multimodal architectures, including vision-language models (VLM), vision-language-action models (VLA), or equivalent. Proficiency in C++, CUDA, distributed training, or performance optimization for production machine learning systems. This is a full-time role based in our office in London. At Wayve ...

Principal Biostatistician/ Sr Biostatistician (R/Rshiny - EMEA BASED)

Hiring Organisation
Syneos Health
Location
London, UK
Employment Type
Full-time
clinical data structures and programming with data expert in functional and object-oriented programming. Knowledgeable in Javascript/Typescript, HTML, WebGL, experience in CUDA/GPU-programing, cloud-computing, Github, web-hosting. Strong communication skills and ability to work both, independently and collaboratively, clear in the presentation of complex ...

Staff ML Performance Engineer (Compiler)

Location
Greater London, England, United Kingdom
with tight constraints (latency, memory, bandwidth, power/thermal, or cost). Strong proficiency with at least one relevant stack/toolchain (e.g. TensorRT, CUDA, Qualcomm QNN, Triton, OpenCL, MLIR, ONNX) and confidence learning adjacent frameworks quickly. Comfort operating at multiple levels of abstraction — from high-level model behaviour ...

Staff ML Performance Engineer (Inference Optimisation)

Hiring Organisation
wayve
Location
London, UK
Employment Type
Full-time
systems with tight constraints (latency, memory, bandwidth, power/thermal, or cost).Strong proficiency with at least one relevant stack/toolchain (e.g. TensorRT, CUDA, Qualcomm QNN, Triton, OpenCL) and confidence learning adjacent frameworks quickly. Comfort operating at multiple levels of abstraction — from high-level model behaviour down ...

GPU Infrastructure Lead - Systems Integrator

Location
Greater London, England, United Kingdom
deployments, monitoring and alerting stacks, and hardware acceptance and regression testing. Deep familiarity with the NVIDIA technology stack, including HGX platforms, NVLink/NVSwitch, CUDA-level debugging, and NCCL performance tuning. Strong leadership skills with the ability to set technical direction, solve complex problems, and guide engineering teams. Excellent ...

Research Scientist – Quantum Optics

Location
Greater London, England, United Kingdom
into existing AI and HPC infrastructure. Its current PT-2 system, a rack-mounted quantum computer, connects directly to GPU clusters via NVIDIA’s CUDA-Q platform, enabling accelerated AI workloads with significantly lower energy consumption than traditional silicon-based systems. Backed by a strong portfolio of patent families ...

AI Infrastructure Lead: GPUs, HPC & Distributed Systems

Location
Greater London, England, United Kingdom
PyTorch, Triton, and advanced scheduling tools like Kubernetes or Ray. Ideal candidates bring 3+ years in AI infrastructure, strong Python skills, and hands-on CUDA/C++ experience, with a degree in CS/EE/Applied Math. #J-18808-Ljbffr ...

Research Engineer (Inference & Serving)

Location
Greater London, England, United Kingdom
architectures Optional Bonus Research engagement: advanced degree with research output, top-tier publications (NeurIPS, ICML, MLSys, OSDI), or open-source contributions GPU kernel work - CUDA, Triton, or similar Experience with quantisation, speculative decoding, disaggregated inference, or KV-cache compression Shortlisted candidates will be contacted within 48 hours. #J ...

Software Engineer, Model Inference, DeepMind

Hiring Organisation
Google
Location
London, UK
Employment Type
Full-time
Experience with developing serving infrastructure. Experience programming hardware accelerators (GPUs, TPUs) via ML frameworks (e.g., JAX, PyTorch) or low-level programming models (e.g., Pallas, CUDA, OpenCL).Experience profiling software to identify performance bottlenecks. Experience with distributed ML systems optimization and parallelism (e.g., data, model, or pipeline parallelism).Familiarity with ...

Software Engineer

Hiring Organisation
Hackajob Ltd
Location
South West London, London, United Kingdom
Employment Type
Permanent
Experience with developing serving infrastructure. Experience programming hardware accelerators (GPUs, TPUs) via ML frameworks (e.g., JAX, PyTorch) or low-level programming models (e.g., Pallas, CUDA, OpenCL). Experience profiling software to identify performance bottlenecks. Experience with distributed ML systems optimization and parallelism (e.g., data, model, or pipeline parallelism). ...

Senior Software Engineer, Full-Stack

Location
Greater London, England, United Kingdom
these — as long as you're open to learning, Backend: Python Frontend: TypeScript and React Deployment: Kubernetes Infrastructure: GCP Machine learning: PyTorch, CUDA, Ray Why Encord Competitive salary, commission, and meaningful equity in a high-growth startup Strong in-person culture — most of the team works from our London ...

Software Engineer, Model Inference, DeepMind

Hiring Organisation
Hackajob Ltd
Location
South West London, London, United Kingdom
Employment Type
Permanent
Experience with developing serving infrastructure. Experience programming hardware accelerators (GPUs, TPUs) via ML frameworks (e.g., JAX, PyTorch) or low-level programming models (e.g., Pallas, CUDA, OpenCL). Experience profiling software to identify performance bottlenecks. Experience with distributed ML systems optimization and parallelism (e.g., data, model, or pipeline parallelism). ...

Product Engineer

Location
Greater London, England, United Kingdom
these — as long as you're open to learning. Backend: Python Frontend: TypeScript and React Deployment: Kubernetes Infrastructure: GCP Machine learning: PyTorch, CUDA, Ray Why Encord Competitive salary, commission, and meaningful equity in a high-growth startup Strong in-person culture — most of the team works from our London ...

Senior Software Engineer - Backend

Location
Greater London, England, United Kingdom
these. As long as you're open to learning, please apply. Backend: Python Frontend: TypeScript and React Deployment: Kubernetes Infrastructure: GCP Machine learning: PyTorch, CUDA, Ray Why Encord Competitive salary, commission, and meaningful equity in a high-growth startup Strong in-person culture: the team works from our London ...