21 of 21 Remote/Hybrid CUDA Jobs in London

Senior Computer Vision Engineer - Deep Learning

Hiring Organisation
MFK Recruitment
Location
TW8, Brentford, Greater London, United Kingdom
Employment Type
Permanent
Salary
£70000 - £95000/annum
communication skills. Desirable experience Vision Transformer architectures. Object detection, tracking, segmentation or pose estimation. Multi-object localisation or multi-modal perception systems. GPU optimisation, CUDA, inference acceleration or edge AI. Real-time computer vision systems. Docker, Kubernetes, Kubeflow or MLOps pipelines. Aerial, satellite or ISR imagery. Synthetic data generation ...

Machine Learning Engineer - Spiking Neural Networks

Hiring Organisation
MFK Recruitment
Location
TW8, Brentford, Greater London, United Kingdom
Employment Type
Permanent
Salary
£75000 - £100000/annum
into reliable software. Experience developing maintainable code using version control, testing and CI/CD. Strong analytical, problem-solving and communication skills. Desirable experience CUDA and GPU programming. Python and PyTorch. CNNs, Transformers or broader deep learning architectures. Synthetic training imagery and simulation environments. ISR imagery or defence-related ...

Senior Software Engineer, Machine Learning Services

Location
Greater London, England, United Kingdom
lifecycle of models in a multi‐tenant, high‐availability system. Familiarity with building ML inference services, model serialization (e.g., ONNX), and GPU programming (CUDA). You've built or worked on custom storage or job‐queueing systems before and have the scars to prove it. Maybe ...

Senior Robotics Engineer

Location
Greater London, England, United Kingdom
share technical ideas clearly, support colleagues and work effectively across multidisciplinary teams.* Experience with areas such as DDS and ROS2 communication, edge acceleration, CUDA, TensorRT, cloud-connected robotics, telemetry or JavaScript-based integration would be valuable, but we don't expect one person to bring experience in every area. ...

AI Engineer (Fluent in Mandarin & English)

Location
Greater London, England, United Kingdom
systems, including vector database management and semantic search optimization. Preferred Qualifications Experience in the insurance or financial services sector. Deep knowledge of GPU architecture , CUDA, and hardware-level performance optimization. Familiarity with Document Intelligence frameworks (OCR, layout analysis, and multimodal extraction). MUST be fluent in Mandarin OR Cantonese ...

Senior Field Application Engineer

Hiring Organisation
Advanced Micro Devices
Location
Central London, London, United Kingdom
Employment Type
Permanent, Work From Home
Some Linux administration; understanding setup for HPC middleware. Nice to Haves: 5+ years HPC application experience Experience building and running HPC applications on GPU. CUDA or OpenACC or OpenMP paradigms. Experience running AI models on CPU or GPU. Any experience understanding/inspecting/writing assembly Understanding of memory ...

AI Integration Engineer

Location
Greater London, England, United Kingdom
complex APIs, third-party libraries, and scalable data pipelines. Desirables: Direct experience with FFmpeg or GStreamer. Familiarity with GPU-accelerated video/AI tech (CUDA, PyTorch, TensorRT, DeepStream). Experience with multimodal AI (VLMs, LLM APIs, Retrieval-Augmented Generation). Cloud and containerization skills (AWS, Docker). If this ...

Senior Machine Learning Engineer - AV Core London, United Kingdom

Location
Greater London, England, United Kingdom
time alerts. Experience with transformer‐based and multimodal architectures, including vision‐language models (VLM), vision‐language‐action models (VLA), or equivalent. Proficiency in C++, CUDA, distributed training, or performance optimization for production machine learning systems. This is a full‐time role based in our office in London. At Wayve ...

Senior Machine Learning Engineer - AV Core

Hiring Organisation
wayve
Location
London, UK
Employment Type
Full-time
time alerts. Experience with transformer-based and multimodal architectures, including vision-language models (VLM), vision-language-action models (VLA), or equivalent. Proficiency in C++, CUDA, distributed training, or performance optimization for production machine learning systems. This is a full-time role based in our office in London. At Wayve ...

Staff ML Performance Engineer (Inference Optimisation)

Location
Greater London, England, United Kingdom
with tight constraints (latency, memory, bandwidth, power/thermal, or cost). Strong proficiency with at least one relevant stack/toolchain (e.g. TensorRT, CUDA, Qualcomm QNN, Triton, OpenCL) and confidence learning adjacent frameworks quickly. Comfort operating at multiple levels of abstraction — from high‐level model behaviour down ...

Staff ML Performance Engineer (Inference Optimisation)

Hiring Organisation
wayve
Location
London, UK
Employment Type
Full-time
systems with tight constraints (latency, memory, bandwidth, power/thermal, or cost).Strong proficiency with at least one relevant stack/toolchain (e.g. TensorRT, CUDA, Qualcomm QNN, Triton, OpenCL) and confidence learning adjacent frameworks quickly. Comfort operating at multiple levels of abstraction — from high-level model behaviour down ...

Research Scientist / Senior Research Scientist

Location
Greater London, England, United Kingdom
into existing AI and HPC infrastructure. Its current PT-2 system, a rack-mounted quantum computer, connects directly to GPU clusters via NVIDIA’s CUDA-Q platform, enabling accelerated AI workloads with significantly lower energy consumption than traditional silicon-based systems. Backed by a strong portfolio of patent families ...

AI Research Engineer

Location
Greater London, England, United Kingdom
plus; a strong background in mathematics is required. Alternatively, experience training models with strong math skills. Strong coding ability in Rust (preferred), CUDA, or C. Solid understanding of transformer architectures. A product mindset; you’ll build production‐ready products. Hybrid Remote with the London Office. What ...

Senior ML Engineer

Location
Greater London, England, United Kingdom
optimize the end-to-end ML stack: data pipelines, training loops, inference serving, and deployment. Design and implement GPU-accelerated components, including custom CUDA kernels where off-the-shelf libraries are not enough. Work closely with the founders to translate product requirements into concrete optimization goals and technical roadmaps. ...

Research Engineer (LLM Performance), London

Hiring Organisation
Isomorphic Labs
Location
London, UK
Employment Type
Full-time
more important than writing kernels from scratchExcellent collaboration skills. Nice to have: Experience with general LLM serving stacks. Knowledge of XLA, Triton, Pallas, CUDA or similar accelerator DSLs/compilers. Experience with optimising ML accuracy using low-precision formats. Prior experience building, deploying and maintaining production systems on GCP.Interest ...

Senior GPU Software Engineer

Location
Greater London, England, United Kingdom
week. Experience for the Senior GPU Software Engineer includes: 5+ years of software engineering experience with significant GPU or high-performance computing experience Strong CUDA and GPU programming experience Excellent C++ programming skills Strong understanding of CPU/GPU interaction and data movement #J-18808-Ljbffr ...

Senior GPU Software Engineer

Hiring Organisation
Intellectual Capital Resources
Location
London, UK
Employment Type
Full-time
week. Experience for the Senior GPU Software Engineer includes: 5+ years of software engineering experience with significant GPU or high-performance computing experience Strong CUDA and GPU programming experience Excellent C++ programming skills Strong understanding of CPU/GPU interaction and data movement If you are a Senior ...

AI Infrastructure Engineer

Location
Greater London, England, United Kingdom
engineers who have: A track record of working on model training or model inference at scale , or on low‐level GPU coding (e.g. CUDA, Triton). Experience with one is great, multiple is even better. What will I be doing? As a Senior AI Infrastructure Engineer focused on model … Model training (especially transformers and LLMs). Model inference at scale (again, especially transformers and LLMs). Low‐level GPU work , such as writing CUDA or Triton kernels. Comfortable working in production environments at meaningful scale (traffic, data, or organizational). You communicate clearly, can explain complex technical topics ...

Staff Robotics Engineer, Localisation London, United Kingdom

Location
Greater London, England, United Kingdom
About us Founded in 2017, Wayve is the leading developer of Embodied AI technology. Our advanced AI software and foundation models enable vehicles to perceive, understand, and navigate any complex environment, enhancing the usability and ...

Staff Robotics Engineer, Localisation

Hiring Organisation
wayve
Location
London, UK
Employment Type
Full-time
About us Founded in 2017, Wayve is the leading developer of Embodied AI technology. Our advanced AI software and foundation models enable vehicles to perceive, understand, and navigate any complex environment, enhancing the usability and ...

Software Engineering Manager

Location
Greater London, England, United Kingdom
This role is based in our London (Kings Cross) office. Responsibilities: Manage and mentor a team of software engineers working across embedded C++/CUDA, GUI, and embedded linux, setting priorities, individual development plans, conducting performance reviews, and maintaining a high standard of engineering practice Own the technical roadmap … product software stack, from the C++/CUDA host application and real-time processing pipeline, aligning with product and system requirements Lead design and architecture reviews across the software stack, ensuring quality, consistency, and sound technical decisions Drive software from prototype through to production release, managing dependencies with ...