6 of 6 OpenCL Jobs in London

Staff ML Performance Engineer (Compiler)

Location
Greater London, England, United Kingdom
memory, bandwidth, power/thermal, or cost). Strong proficiency with at least one relevant stack/toolchain (e.g. TensorRT, CUDA, Qualcomm QNN, Triton, OpenCL, MLIR, ONNX) and confidence learning adjacent frameworks quickly. Comfort operating at multiple levels of abstraction — from high-level model behaviour down to low-level ...

Software Engineer, Model Inference, DeepMind

Hiring Organisation
Hackajob Ltd
Location
South West London, London, United Kingdom
Employment Type
Permanent
with developing serving infrastructure. Experience programming hardware accelerators (GPUs, TPUs) via ML frameworks (e.g., JAX, PyTorch) or low-level programming models (e.g., Pallas, CUDA, OpenCL). Experience profiling software to identify performance bottlenecks. Experience with distributed ML systems optimization and parallelism (e.g., data, model, or pipeline parallelism). Familiarity ...

Staff ML Performance Engineer (Inference Optimisation)

Location
Greater London, England, United Kingdom
memory, bandwidth, power/thermal, or cost). Strong proficiency with at least one relevant stack/toolchain (e.g. TensorRT, CUDA, Qualcomm QNN, Triton, OpenCL) and confidence learning adjacent frameworks quickly. Comfort operating at multiple levels of abstraction — from high‐level model behaviour down to low‐level kernel/ ...

C++ Developer HPC

Hiring Organisation
Quant Capital
Location
London, United Kingdom
Developers MUST have: (the below also gives you an idea of stack)C++ 11Computer Science DegreeSingle Core and Distributed C++C++ Shared LibrariesVersion ControlIdeally OpenMP, OpenCL or TBBThe environment is that of Facebook or Google, relaxed open with time to think and make the right decisions. The atmosphere is calm ...

Senior Machine Learning Engineer, AI Performance London, United Kingdom

Location
Greater London, England, United Kingdom
PyTorch (not just using high‐level tooling). Strong proficiency with at least one relevant stack/toolchain (e.g. TensorRT, CUDA, Qualcomm QNN, Triton, OpenCL) and confidence learning adjacent frameworks quickly. Comfort operating at multiple levels of abstraction — from high‐level model behaviour down to low‐level kernel/ ...

Senior Machine Learning Engineer, AI Performance

Hiring Organisation
wayve
Location
London, UK
Employment Type
Full-time
PyTorch (not just using high-level tooling).Strong proficiency with at least one relevant stack/toolchain (e.g. TensorRT, CUDA, Qualcomm QNN, Triton, OpenCL) and confidence learning adjacent frameworks quickly. Comfort operating at multiple levels of abstraction — from high-level model behaviour down to low-level kernel/runtime ...