26 to 50 of 108 Permanent CUDA Jobs

Software Inference Deployment Engineer

Location
Oxford, England, United Kingdom
Strong Preference For Experience integrating accelerator hardware (GPUs, FPGAs, ASICs, NPUs, or novel architectures) into customer inference workflows Familiarity with the NVIDIA inference stack - CUDA, TensorRT, Triton Exposure to disaggregated inference architectures, prefill/decode separation, or KV cache management Compensation & Benefits Highly Competitive Salary: We are not saying ...

Account Solution Architect

Location
Greater London, England, United Kingdom
Dutch, Swedish, Norwegian, Danish, or Finnish is a plus. Preferred: Familiarity with NVIDIA GPU architectures (H100, A100, H200) and the software stack around them: CUDA, NCCL, cuDNN. Working knowledge of high-performance networking concepts: InfiniBand, RDMA, RoCE, TCP/IP. Background working directly with AI labs, research institutions ...

Junior SRE

Hiring Organisation
SQUAREPOINT CAPITAL
Location
London, United Kingdom
Salary
£ 70 K
distributed systems, HPC, or cloud platforms for CPU and GPU workloads.Experience with profiling, benchmarking, or performance analysis.Additional Qualifications (Nice to Have)Familiarity with CUDA or other GPU development frameworks.The minimum base salary for this role is $150,000 if located in New York. This expectation is based on available ...

Senior Machine Learning Engineer

Location
City Of London, England, United Kingdom
Optimization: Maximize hardware utilization for single-GPU and distributed multi-GPU environments. Algorithm Acceleration: Optimize Python code execution using Numba, NumPy, and specialized CUDA libraries. Load Testing: Conduct rigorous load and stress testing to guarantee system stability under high concurrent traffic. Required Skills and Qualifications Core Programming & Frameworks Language ...

Junior SRE

Hiring Organisation
SQUAREPOINT CAPITAL
Location
London, UK
Employment Type
Full-time
systems, HPC, or cloud platforms for CPU and GPU workloads. Experience with profiling, benchmarking, or performance analysis. Additional Qualifications (Nice to Have)Familiarity with CUDA or other GPU development frameworks. The minimum base salary for this role is $150,000 if located in New York. This expectation is based ...

Machine Learning Engineer ()

Hiring Organisation
Placement Services USA, Inc
Location
San Mateo, California, United States
Employment Type
Any
Salary
USD 195,000 Annual
fraud risk scoring systems by applying patterns from end-to-end ID verification and KYC fraud models, utilizing Python, TensorFlow, AWS, OpenCV, and CUDA, and machine learning libraries such as NumPy, Scikit-learn, Keras, and Pandas to develop, validate, and optimize data quality and risk-scoring mechanisms. Experience must ...

AI Engineer - Portuguese & English, Perm, Hybrid, London, 65-80k

Hiring Organisation
Bangura Solutions
Location
London, United Kingdom
Salary
£ 70 K
Retrieval-Augmented Generation systems, including vector database management and semantic search optimization.Preferred QualificationsExperience in the insurance or financial services sector.Deep knowledge of GPU architecture, CUDA, and hardware-level performance optimization.Familiarity with Document Intelligence frameworks (OCR, layout analysis, and multimodal extraction).MUST be fluent in Portuguese and EnglishMinorities, women, LGBTQ+ ...

AI Engineer - Portuguese & English, Perm, Hybrid, London, 65-80k

Hiring Organisation
Bangura Solutions
Location
London, UK
Employment Type
Full-time
Generation systems, including vector database management and semantic search optimization. Preferred QualificationsExperience in the insurance or financial services sector. Deep knowledge of GPU architecture, CUDA, and hardware-level performance optimization. Familiarity with Document Intelligence frameworks (OCR, layout analysis, and multimodal extraction).MUST be fluent in Portuguese and EnglishMinorities, women ...

NLP Performance Engineer

Hiring Organisation
G Research
Location
London, United Kingdom
Salary
£ 80 K
PyTorch ecosystemExperience with inference optimisation techniques, including quantisation, speculative decoding and model parallelism across modern GPU architecturesStrong software engineering skills, including Python, CUDA and building reliable systems for machine learning workloadsStrong communication skills, with the ability to collaborate across research, infrastructure and engineering teamsWhy join us Highly competitive compensation ...

Software Engineer, Model Inference, DeepMind

Hiring Organisation
Hackajob Ltd
Location
South West London, London, United Kingdom
Employment Type
Permanent
Experience with developing serving infrastructure. Experience programming hardware accelerators (GPUs, TPUs) via ML frameworks (e.g., JAX, PyTorch) or low-level programming models (e.g., Pallas, CUDA, OpenCL). Experience profiling software to identify performance bottlenecks. Experience with distributed ML systems optimization and parallelism (e.g., data, model, or pipeline parallelism). ...

NLP Performance Engineer

Location
Greater London, England, United Kingdom
PyTorch ecosystem Experience with inference optimisation techniques, including quantisation, speculative decoding and model parallelism across modern GPU architectures Strong software engineering skills, including Python, CUDA and building reliable systems for machine learning workloads Strong communication skills, with the ability to collaborate across research, infrastructure and engineering teams Why join ...

Software Engineer, Model Inference, DeepMind

Location
Westminster, West End, United Kingdom
Experience with developing serving infrastructure. Experience programming hardware accelerators (GPUs, TPUs) via ML frameworks (e.g., JAX, PyTorch) or low-level programming models (e.g., Pallas, CUDA, OpenCL). Experience profiling software to identify performance bottlenecks. Experience with distributed ML systems optimization and parallelism (e.g., data, model, or pipeline parallelism). ...

Software Engineer, Model Inference, DeepMind

Hiring Organisation
Google
Location
London, United Kingdom
Salary
£ 80 K
qualifications:Experience with developing serving infrastructure.Experience programming hardware accelerators (GPUs, TPUs) via ML frameworks (e.g., JAX, PyTorch) or low-level programming models (e.g., Pallas, CUDA, OpenCL).Experience profiling software to identify performance bottlenecks.Experience with distributed ML systems optimization and parallelism (e.g., data, model, or pipeline parallelism).Familiarity with writing ...

Product Engineer

Location
Greater London, England, United Kingdom
these — as long as you're open to learning. Backend: Python Frontend: TypeScript and React Deployment: Kubernetes Infrastructure: GCP Machine learning: PyTorch, CUDA, Ray Why Encord Competitive salary, commission, and meaningful equity in a high-growth startup Strong in-person culture — most of the team works from our London ...

Senior Software Engineer - Backend

Location
Greater London, England, United Kingdom
Encord and are not looking for experience across all of these. Backend: Python Frontend: TypeScript and React Deployment: Kubernetes Infrastructure: GCP Machine learning: PyTorch, CUDA, Ray Why Encord Competitive salary, commission, and meaningful equity in a high-growth startup Strong in-person culture: the team works from our London ...

AI Engineer (Fluent Portuguese & English)

Hiring Organisation
Chubb
Location
London, UK
Employment Type
Full-time
Generation systems, including vector database management and semantic search optimization. Preferred QualificationsExperience in the insurance or financial services sector. Deep knowledge of GPU architecture, CUDA, and hardware-level performance optimization. Familiarity with Document Intelligence frameworks (OCR, layout analysis, and multimodal extraction).MUST be fluent in Portuguese and EnglishWe offer ...

AI Engineer (Fluent in Mandarin & English)

Location
Greater London, England, United Kingdom
systems, including vector database management and semantic search optimization. Preferred Qualifications Experience in the insurance or financial services sector. Deep knowledge of GPU architecture , CUDA, and hardware-level performance optimization. Familiarity with Document Intelligence frameworks (OCR, layout analysis, and multimodal extraction). MUST be fluent in Mandarin OR Cantonese ...

Senior Field Application Engineer - HPC (UK + Multiple European Locations)

Hiring Organisation
AMD
Location
London, United Kingdom
Salary
£ 80 K
audience Some Linux administration; understanding setup for HPC middleware. Nice to Haves:5+ years HPC application experienceExperience building and running HPC applications on GPU. CUDA or OpenACC or OpenMP paradigms.Experience running AI models on CPU or GPU.Any experience understanding/inspecting/writing assembly Understanding of memory and cache ...

Software Engineer, Model Inference, DeepMind

Location
Greater London, England, United Kingdom
Experience with developing serving infrastructure. Experience programming hardware accelerators (GPUs, TPUs) via ML frameworks (e.g., JAX, PyTorch) or low-level programming models (e.g., Pallas, CUDA, OpenCL). Experience profiling software to identify performance bottlenecks. Experience with distributed ML systems optimization and parallelism (e.g., data, model, or pipeline parallelism). ...

MLOps Engineer

Hiring Organisation
AECOM
Location
London, United Kingdom
Salary
£ 70 K
reinforcement learning, or generative AI (a plus, not required) Identifying and resolving bottlenecks in distributed machine learning workloads (knowledge of low-level languages and CUDA library is a plus) Additional InformationOur Hiring Process25-minute screening callTake-home challenge: A hands-on task to assess your problem-solving and technical ...

Machine Learning Engineer

Hiring Organisation
Intellectual Capital Resources
Location
Oxford, Oxfordshire, United Kingdom
Salary
£ 80 K
model deployment Clear communication - writing specs, documenting APIs, presenting to customers Familiarity with circuit simulation or EDA tools (nice to have) C++/CUDA experience for performance‐critical components (nice to have) Startup or fast‐moving team experience Why consider it: Competitive salary, bonus, and meaningful equity Flexible remote ...

Machine Learning Engineer

Hiring Organisation
Your Tech Future
Location
South West London, London, United Kingdom
Employment Type
Permanent
architectures and model optimisation Excellent software engineering and problem-solving abilities Experience working with complex datasets and real-world machine learning challenges Desirable Skills CUDA TensorRT Triton Quantised model training Edge AI deployment Computer Vision Robotics or autonomous systems Advanced C++ development Synthetic data generation What We're Looking ...

Senior Machine Learning Researcher

Location
Greater London, England, United Kingdom
finance, trading, or quantitative research (not required). Publications, competition results (e.g., Kaggle, academic ML contests), or open-source contributions. Familiarity with C++, CUDA, or low-latency systems. Benefits Opportunity to work at one of the world's leading algorithmic trading firms Engaging projects offering accelerated responsibilities and ownership ...

Product Engineer, Physical AI

Location
Greater London, England, United Kingdom
these — as long as you're open to learning, please apply. Backend: Python Frontend: TypeScript and React Deployment: Kubernetes Infrastructure: GCP Machine learning: PyTorch, CUDA, Ray Why Encord Competitive salary, commission, and meaningful equity in a high-growth startup Strong in-person culture — most of the team works from ...

Senior Machine Learning Research Engineer

Location
Greater London, England, United Kingdom
infrastructure and containerised environments. A track record of taking research code from prototype to robust, reusable infrastructure that other people actually use. Bonus experience: CUDA/Triton kernel development; FlashAttention-style attention implementations; experience with foundation models for biology, vision, or language; contributions to open-source ML frameworks. Personally ...