26 to 50 of 98 CUDA Jobs in London

Senior Software Engineer, Full-Stack

Location
Greater London, England, United Kingdom
these — as long as you're open to learning, Backend: Python Frontend: TypeScript and React Deployment: Kubernetes Infrastructure: GCP Machine learning: PyTorch, CUDA, Ray Why Encord Competitive salary, commission, and meaningful equity in a high-growth startup Strong in-person culture — most of the team works from our London ...

Senior Field Application Engineer

Hiring Organisation
Advanced Micro Devices
Location
Central London, London, United Kingdom
Employment Type
Permanent, Work From Home
Some Linux administration; understanding setup for HPC middleware. Nice to Haves: 5+ years HPC application experience Experience building and running HPC applications on GPU. CUDA or OpenACC or OpenMP paradigms. Experience running AI models on CPU or GPU. Any experience understanding/inspecting/writing assembly Understanding of memory ...

Senior Machine Learning Research Engineer

Location
Greater London, England, United Kingdom
infrastructure and containerised environments. A track record of taking research code from prototype to robust, reusable infrastructure that other people actually use. Bonus experience: CUDA/Triton kernel development; FlashAttention-style attention implementations; experience with foundation models for biology, vision, or language; contributions to open-source ML frameworks. Personally ...

Machine Learning Engineer

Location
Greater London, England, United Kingdom
automation and orchestration systems for these platforms Proficiency in a low-level language such as C, C++, or Rust and in GPU frameworks like CUDA Competence in front-end web design to allow easy interfacing with large datasets Our Culture Follow the science. We prioritise rigorous scientific inquiry, relying ...

AI Integration Engineer

Hiring Organisation
Microtech Global Ltd
Location
London, United Kingdom
Employment Type
Permanent, Work From Home
complex APIs, third-party libraries, and scalable data pipelines. Desirables: Direct experience with FFmpeg or GStreamer. Familiarity with GPU-accelerated video/AI tech (CUDA, PyTorch, TensorRT, DeepStream). Experience with multimodal AI (VLMs, LLM APIs, Retrieval-Augmented Generation). Cloud and containerization skills (AWS, Docker). ...

Optical Systems Product Engineer (Photonic Quantum Computing)

Location
Greater London, England, United Kingdom
into existing AI and HPC infrastructure. Its current PT-2 system, a rack-mounted quantum computer, connects directly to GPU clusters via NVIDIA’s CUDA-Q platform, enabling accelerated AI workloads with significantly lower energy consumption than traditional silicon-based systems. Backed by a strong portfolio of patent families ...

Software Engineer

Hiring Organisation
Intellectual Capital Resources
Location
London, UK
Employment Type
Full-time
computational fluid dynamics (CFD), combustion, heat transfer, atmospheric dynamics, meteorology, weather forecasting, or stochastic simulation Desired: Experience with parallel computing, HPC, GPU acceleration or CUDA If you are a Software Engineer looking to apply your scientific software expertise to cutting-edge climate technology and help develop solutions that improve ...

Machine Learning Engineer

Hiring Organisation
Your Tech Future
Location
South West London, London, United Kingdom
Employment Type
Permanent
architectures and model optimisation Excellent software engineering and problem-solving abilities Experience working with complex datasets and real-world machine learning challenges Desirable Skills CUDA TensorRT Triton Quantised model training Edge AI deployment Computer Vision Robotics or autonomous systems Advanced C++ development Synthetic data generation What We're Looking ...

NLP Performance Engineer

Hiring Organisation
G Research
Location
London, UK
Employment Type
Full-time
PyTorch ecosystemExperience with inference optimisation techniques, including quantisation, speculative decoding and model parallelism across modern GPU architecturesStrong software engineering skills, including Python, CUDA and building reliable systems for machine learning workloadsStrong communication skills, with the ability to collaborate across research, infrastructure and engineering teamsWhy join us? Highly competitive compensation ...

Software Engineer, Model Inference, DeepMind

Hiring Organisation
Hackajob Ltd
Location
South West London, London, United Kingdom
Employment Type
Permanent
Experience with developing serving infrastructure. Experience programming hardware accelerators (GPUs, TPUs) via ML frameworks (e.g., JAX, PyTorch) or low-level programming models (e.g., Pallas, CUDA, OpenCL). Experience profiling software to identify performance bottlenecks. Experience with distributed ML systems optimization and parallelism (e.g., data, model, or pipeline parallelism). ...

NLP Performance Engineer

Location
Greater London, England, United Kingdom
PyTorch ecosystem Experience with inference optimisation techniques, including quantisation, speculative decoding and model parallelism across modern GPU architectures Strong software engineering skills, including Python, CUDA and building reliable systems for machine learning workloads Strong communication skills, with the ability to collaborate across research, infrastructure and engineering teams Why join ...

Product Engineer

Location
Greater London, England, United Kingdom
these — as long as you\'re open to learning, please apply. Backend: Python Frontend: TypeScript and React Deployment: Kubernetes Infrastructure: GCP Machine learning: PyTorch, CUDA, Ray Why Encord Competitive salary, commission, and meaningful equity in a high-growth startup Strong in-person culture — most of the team works from ...

AI Engineer (Fluent Portuguese & English)

Hiring Organisation
Appcast
Location
London, UK
Retrieval-Augmented Generation systems, including vector database management and semantic search optimization.Preferred QualificationsExperience in the insurance or financial services sector.Deep knowledge of GPU architecture, CUDA, and hardware-level performance optimization.Familiarity with Document Intelligence frameworks (OCR, layout analysis, and multimodal extraction).MUST be fluent in Portuguese and EnglishWe offer ...

Senior Software Engineer - Backend

Location
Greater London, England, United Kingdom
these. As long as you're open to learning, please apply. Backend: Python Frontend: TypeScript and React Deployment: Kubernetes Infrastructure: GCP Machine learning: PyTorch, CUDA, Ray Why Encord Competitive salary, commission, and meaningful equity in a high-growth startup Strong in-person culture: the team works from our London ...

AI Engineer (Fluent in Mandarin & English)

Location
Greater London, England, United Kingdom
systems, including vector database management and semantic search optimization. Preferred Qualifications Experience in the insurance or financial services sector. Deep knowledge of GPU architecture , CUDA, and hardware-level performance optimization. Familiarity with Document Intelligence frameworks (OCR, layout analysis, and multimodal extraction). MUST be fluent in Mandarin OR Cantonese ...

Principal Biostatistician/ Sr Biostatistician (R/Rshiny - EMEA BASED)

Hiring Organisation
Syneos Health
Location
London, UK
Employment Type
Full-time
clinical data structures and programming with data expert in functional and object-oriented programming. Knowledgeable in Javascript/Typescript, HTML, WebGL, experience in CUDA/GPU-programing, cloud-computing, Github, web-hosting. Strong communication skills and ability to work both, independently and collaboratively, clear in the presentation of complex ...

Staff ML Performance Engineer (Compiler)

Location
Greater London, England, United Kingdom
with tight constraints (latency, memory, bandwidth, power/thermal, or cost). Strong proficiency with at least one relevant stack/toolchain (e.g. TensorRT, CUDA, Qualcomm QNN, Triton, OpenCL, MLIR, ONNX) and confidence learning adjacent frameworks quickly. Comfort operating at multiple levels of abstraction — from high-level model behaviour ...

Staff ML Performance Engineer (Inference Optimisation)

Hiring Organisation
wayve
Location
London, UK
Employment Type
Full-time
systems with tight constraints (latency, memory, bandwidth, power/thermal, or cost).Strong proficiency with at least one relevant stack/toolchain (e.g. TensorRT, CUDA, Qualcomm QNN, Triton, OpenCL) and confidence learning adjacent frameworks quickly. Comfort operating at multiple levels of abstraction — from high-level model behaviour down ...

Research Engineer (Inference & Serving)

Location
Greater London, England, United Kingdom
architectures Optional Bonus Research engagement: advanced degree with research output, top-tier publications (NeurIPS, ICML, MLSys, OSDI), or open-source contributions GPU kernel work - CUDA, Triton, or similar Experience with quantisation, speculative decoding, disaggregated inference, or KV-cache compression Shortlisted candidates will be contacted within 48 hours. #J ...

Machine Learning Researcher

Location
City Of London, England, United Kingdom
finance, trading, or quantitative research (not required). Publications, competition results (e.g., Kaggle, academic ML contests), or open-source contributions. Familiarity with C++, CUDA, or low-latency systems. Here is why you should join our dynamic team: Opportunity to work at one of the world's leading algorithmic trading ...

Product Engineer, Physical AI

Location
Greater London, England, United Kingdom
these — as long as you're open to learning, please apply. Backend: Python Frontend: TypeScript and React Deployment: Kubernetes Infrastructure: GCP Machine learning: PyTorch, CUDA, Ray Why Encord Competitive salary, commission, and meaningful equity in a high-growth startup Strong in-person culture — most of the team works from ...

GPU Infrastructure Lead - Systems Integrator

Location
Greater London, England, United Kingdom
deployments, monitoring and alerting stacks, and hardware acceptance and regression testing. Deep familiarity with the NVIDIA technology stack, including HGX platforms, NVLink/NVSwitch, CUDA-level debugging, and NCCL performance tuning. Strong leadership skills with the ability to set technical direction, solve complex problems, and guide engineering teams. Excellent ...

Electronic Engineer

Location
Greater London, England, United Kingdom
into existing AI and HPC infrastructure. Its current PT-2 system, a rack-mounted quantum computer, connects directly to GPU clusters via NVIDIA’s CUDA-Q platform, enabling accelerated AI workloads with significantly lower energy consumption than traditional silicon-based systems. Backed by a strong portfolio of patent families ...

Software Engineer - Systems

Location
Greater London, England, United Kingdom
engineering Familiarity with tools like Apache Arrow, Parquet, DataFusion, Clickhouse, or DuckDB is a plus Understanding of cutting-edge ML infrastructure stack (e.g. PyTorch, CUDA) is also a plus Experience with Rust is a bonus Willingness to work in-person at our NYC or London office #J-18808-Ljbffr ...

Research Scientist, World Models, DeepMind

Hiring Organisation
Appcast
Location
London, UK
years of experience working in frontier AI research labs on pre-training or post-training teams.Experience writing TPU/GPU kernels (e.g., JAX, CUDA) to optimize model performance and real-time inference.Experience designing, training, and scaling generative pixel or real-time video architectures.Track record of cross-functional research collaboration ...