1 to 25 of 71 CUDA Jobs in London

Senior Computer Vision Engineer - Deep Learning

Hiring Organisation
MFK Recruitment
Location
TW8, Brentford, Greater London, United Kingdom
Employment Type
Permanent
Salary
£70000 - £95000/annum
communication skills. Desirable experience Vision Transformer architectures. Object detection, tracking, segmentation or pose estimation. Multi-object localisation or multi-modal perception systems. GPU optimisation, CUDA, inference acceleration or edge AI. Real-time computer vision systems. Docker, Kubernetes, Kubeflow or MLOps pipelines. Aerial, satellite or ISR imagery. Synthetic data generation ...

Platform Engineer

Location
Greater London, England, United Kingdom
building things properly the first time, even under early-stage constraints. Nice to Have Experience with GPU cluster management and ML training workloads (NVIDIA, CUDA, distributed training). Familiarity with MLOps tooling: Experiment tracking (MLflow, Weights & Biases). Workflow orchestration (Airflow, Prefect, Argo). Data versioning (DVC). Background ...

Machine Learning Engineer - Spiking Neural Networks

Hiring Organisation
MFK Recruitment
Location
TW8, Brentford, Greater London, United Kingdom
Employment Type
Permanent
Salary
£75000 - £100000/annum
into reliable software. Experience developing maintainable code using version control, testing and CI/CD. Strong analytical, problem-solving and communication skills. Desirable experience CUDA and GPU programming. Python and PyTorch. CNNs, Transformers or broader deep learning architectures. Synthetic training imagery and simulation environments. ISR imagery or defence-related ...

Senior MLOps Engineer

Location
City of Westminster, England, United Kingdom
Practical MLOps: experiment tracking and model versioning with MLflow, hyperparameter optimisation, and reproducible, configuration-driven training runs. Containerisation and infrastructure as code: Docker and CUDA-based GPU images, plus enough Terraform to ship your own model to an environment through code review rather than a console. Strong software engineering ...

Software Engineer, AI Video

Hiring Organisation
Enterprise Recruitment Ltd
Location
W2, Sheldon Square, Greater London, United Kingdom
Employment Type
Permanent
Salary
£60000 - £90000/annum
containerisation. AWS or cloud platform experience. FFmpeg or GStreamer. AI, computer vision or multimodal systems. Video curation or data preparation pipelines. GPU technologies including CUDA, PyTorch, TensorRT or Triton. Batch or real-time processing environments. ~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~ Position : Software Engineer Location : West London Salary : £50-90k Benefits: Pension, Private medical ...

Sr. Software Engineer, Inference

Location
Greater London, England, United Kingdom
managing capacity planning, autoscaling policies, and driving post-incident remediation. Preferred Experience with low-level systems and high-performance computing components (e.g., C++ development, CUDA kernels, NCCL/SHARP, RDMA/NUMA, or GPU interconnect topologies). Active open-source or production contributions to modern inference frameworks (e.g., vLLM ...

Campus AI Researcher, PhD/Postdoc (Full-Time)

Location
Greater London, England, United Kingdom
effectively with trading researchers Reliable and predictable availability required Bonus Points Experience with HPC and distributed large model training Experience with GPU performance optimisation (CUDA or ROCm) Experience with end-to-end model development Strong opinions on best practices in ML research, tooling, and/or infrastructure International Students ...

Senior Software Engineer, Machine Learning Services

Location
Greater London, England, United Kingdom
lifecycle of models in a multi‐tenant, high‐availability system. Familiarity with building ML inference services, model serialization (e.g., ONNX), and GPU programming (CUDA). You've built or worked on custom storage or job‐queueing systems before and have the scars to prove it. Maybe ...

Software Engineer Python

Hiring Organisation
Enterprise Recruitment Limited
Location
West London, London, United Kingdom
Employment Type
Permanent
Salary
£90,000
containerisation. AWS or cloud platform experience. FFmpeg or GStreamer. AI, computer vision or multimodal systems. Video curation or data preparation pipelines. GPU technologies including CUDA, PyTorch, TensorRT or Triton. Batch or real-time processing environments. ~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~~ Position : Software Engineer Location : West London Salary : £50-90k Benefits: Pension, Private medical ...

Machine Learning Engineer - Conversational AI & MLOps

Hiring Organisation
Robert Walters
Location
London, South East England, United Kingdom
Employment Type
Full-Time
Salary
£70,000 - £90,000 per annum
utilisation across both single-GPU and distributed multi-GPU environments . Improving Python and model inference performance using technologies including NumPy, Numba, Triton and CUDA-based libraries . Conducting load and stress testing to ensure AI services remain performant and stable under high levels of concurrent traffic. Optimising cloud ...

AI Research Scientist | Research & Development

Location
Greater London, England, United Kingdom
dependable working rhythm. Experience working with HPC environments or training large models in distributed settings, along with some exposure to GPU-level optimisation using CUDA or ROCm. A track record of taking models end-to-end, particularly in the context of LLMs, is valuable. Prior academic publications or meaningful ...

ML Research Engineer, London

Hiring Organisation
Isomorphic Labs
Location
London, UK
Employment Type
Full-time
models. Scale & Performance: Experience training models across distributed systems (multi-GPU/multi-node) and optimising training and inference performance (e.g., XLA, Triton, CUDA, Pallas).Domain Knowledge: A strong interest in, or knowledge of, biochemistry, computational biology, or drug discovery fundamentals. Industry Experience: Proven track record working in reputable ...

Senior AI Platform Engineer

Location
Greater London, England, United Kingdom
large-scale AI, machine learning, or distributed computing platforms in enterprise environments. Deep understanding of LLM architectures and their interaction with GPU infrastructure, including CUDA, cuDNN, NCCL, kernel-level acceleration libraries, and distributed training frameworks such as PyTorch. Strong knowledge of distributed training and inference strategies, including tensor, pipeline ...

Account Solution Architect

Location
Greater London, England, United Kingdom
Dutch, Swedish, Norwegian, Danish, or Finnish is a plus. Preferred: Familiarity with NVIDIA GPU architectures (H100, A100, H200) and the software stack around them: CUDA, NCCL, cuDNN. Working knowledge of high-performance networking concepts: InfiniBand, RDMA, RoCE, TCP/IP. Background working directly with AI labs, research institutions ...

Senior Machine Learning Engineer

Hiring Organisation
Hexwired Recruitment Limited
Location
London, United Kingdom
Employment Type
Contract
they are looking for a Senior Machine Learning Engineer to join their expanding team in London. Required Experience Strong C++ development skills. Experience with CUDA for GPU acceleration. Python and PyTorch experience. Knowledge of Transformers, computer vision or image analysis. Postgraduate or commercial experience within AI/Machine Learning ...

Senior Machine Learning Engineer

Location
City Of London, England, United Kingdom
Optimization: Maximize hardware utilization for single-GPU and distributed multi-GPU environments. Algorithm Acceleration: Optimize Python code execution using Numba, NumPy, and specialized CUDA libraries. Load Testing: Conduct rigorous load and stress testing to guarantee system stability under high concurrent traffic. Required Skills and Qualifications Core Programming & Frameworks Language ...

Junior Algorithmic Developer, Analyst, London

Hiring Organisation
Jefferies Financial Group
Location
London, UK
Employment Type
Full-time
Computer Science, Computer Engineering, Mathematics, or a highly quantitative field. Preferred Qualifications: Experience with networking protocols (TCP/IP, UDP) or hardware acceleration (CUDA/GPUs).Familiarity with CVS, git and Linux/Unix environments. Basic understanding of quantitative finance and algorithmic trading. Jefferies is a leading global, full ...

Senior Software Engineer, Full-Stack

Location
Greater London, England, United Kingdom
these — as long as you're open to learning, Backend: Python Frontend: TypeScript and React Deployment: Kubernetes Infrastructure: GCP Machine learning: PyTorch, CUDA, Ray Why Encord Competitive salary, commission, and meaningful equity in a high-growth startup Strong in-person culture — most of the team works from our London ...

Senior Field Application Engineer

Hiring Organisation
Advanced Micro Devices
Location
Central London, London, United Kingdom
Employment Type
Permanent, Work From Home
Some Linux administration; understanding setup for HPC middleware. Nice to Haves: 5+ years HPC application experience Experience building and running HPC applications on GPU. CUDA or OpenACC or OpenMP paradigms. Experience running AI models on CPU or GPU. Any experience understanding/inspecting/writing assembly Understanding of memory ...

Senior Machine Learning Research Engineer

Location
Greater London, England, United Kingdom
infrastructure and containerised environments. A track record of taking research code from prototype to robust, reusable infrastructure that other people actually use. Bonus experience: CUDA/Triton kernel development; FlashAttention-style attention implementations; experience with foundation models for biology, vision, or language; contributions to open-source ML frameworks. Personally ...

AI Integration Engineer

Hiring Organisation
Microtech Global Ltd
Location
London, United Kingdom
Employment Type
Permanent, Work From Home
complex APIs, third-party libraries, and scalable data pipelines. Desirables: Direct experience with FFmpeg or GStreamer. Familiarity with GPU-accelerated video/AI tech (CUDA, PyTorch, TensorRT, DeepStream). Experience with multimodal AI (VLMs, LLM APIs, Retrieval-Augmented Generation). Cloud and containerization skills (AWS, Docker). ...

Software Engineer

Hiring Organisation
Intellectual Capital Resources
Location
London, UK
Employment Type
Full-time
computational fluid dynamics (CFD), combustion, heat transfer, atmospheric dynamics, meteorology, weather forecasting, or stochastic simulation Desired: Experience with parallel computing, HPC, GPU acceleration or CUDA If you are a Software Engineer looking to apply your scientific software expertise to cutting-edge climate technology and help develop solutions that improve ...

Machine Learning Engineer

Hiring Organisation
Your Tech Future
Location
South West London, London, United Kingdom
Employment Type
Permanent
architectures and model optimisation Excellent software engineering and problem-solving abilities Experience working with complex datasets and real-world machine learning challenges Desirable Skills CUDA TensorRT Triton Quantised model training Edge AI deployment Computer Vision Robotics or autonomous systems Advanced C++ development Synthetic data generation What We're Looking ...

NLP Performance Engineer

Hiring Organisation
G Research
Location
London, UK
Employment Type
Full-time
PyTorch ecosystemExperience with inference optimisation techniques, including quantisation, speculative decoding and model parallelism across modern GPU architecturesStrong software engineering skills, including Python, CUDA and building reliable systems for machine learning workloadsStrong communication skills, with the ability to collaborate across research, infrastructure and engineering teamsWhy join us? Highly competitive compensation ...

Software Engineer, Model Inference, DeepMind

Hiring Organisation
Hackajob Ltd
Location
South West London, London, United Kingdom
Employment Type
Permanent
Experience with developing serving infrastructure. Experience programming hardware accelerators (GPUs, TPUs) via ML frameworks (e.g., JAX, PyTorch) or low-level programming models (e.g., Pallas, CUDA, OpenCL). Experience profiling software to identify performance bottlenecks. Experience with distributed ML systems optimization and parallelism (e.g., data, model, or pipeline parallelism). ...