1 to 25 of 243 NVIDIA Jobs in the UK

AI Platform Engineer- Senior Consultant-AI and Digital Factory

Location
Manchester, England, United Kingdom
LLMOps• Hands-on with MLOps platforms (Azure ML, Databricks, SageMaker) and vector/retrieval databases (Pinecone, Milvus, pgvector)• Experience with GPU-accelerated infrastructure and NVIDIA AI Enterprise or equivalent stacks• Exposure to fine-tuning, RLHF, or SLM distillation is a strong plusCloud-Native & Infrastructure• Deep expertise in Kubernetes and container ...

ML Ops Engineer

Location
Greater London, England, United Kingdom
utilisation, reliability, and cost-efficiency. Your Impact Provision and manage cloud-native AI/ML infrastructure utilising Kubernetes, Docker, and GPU orchestration frameworks (e.g., NVIDIA GPU Operator, Slurm, or Ray). Automate core platform infrastructure using Infrastructure as Code (IaC) tools like Terraform, Helm, and Ansible. Optimise GPU compute workloads ...

Machine Learning Engineer (Mid to Principal)

Location
West of England, England, United Kingdom
/ML/CS or related field. Beneficial knowledge General tooling and platforms: Databricks, AWS, GCP, GitHub, Docker/Kubernetes, MLflow, Jira. Edge deployments: Nvidia Jetson (e.g. AGX Orin), Raspberry Pi, or other embedded accelerators. Distributed model training & infra: Pytorch DDP, FDSP and TorchTitan, Megatron, Slurm, Run:ai, DeepSpeed, Kubernetes ...

Senior Research HPC Engineer

Location
Greater London, England, United Kingdom
dive when necessary Desirable Exposure to computational research involving GPU, including AI/ML methods applied to biomedical problems Experience supporting GPU‐accelerated workloads, NVIDIA tooling, CUDA‐aware environments, and/or AI/ML and bioinformatics workloads on shared compute platforms. Familiarity with bioinformatics or scientific workflow frameworks (e.g. ...

Senior Data Centre Solutions Architect

Hiring Organisation
Shaw Daniels Solutions
Location
London, UK
Employment Type
Full-time
Storage or Dell EMCExperience with Cisco Nexus and ACIExperience with VMware Cloud FoundationExperience with Terraform, Ansible, or similar automation technologiesDesirable: CCNP or equivalent certificationDesirable: NVIDIA certifications or AI infrastructure experienceDesirable: DevOps, automation, and orchestration toolsDesirable: Python, REST APIs, and GitDesirable: Docker and KubernetesDesirable: Knowledge of modern CPU, GPU, and infrastructure ...

Senior Computer Vision Engineer - Deep Learning

Hiring Organisation
MFK Recruitment
Location
TW8, Brentford, Greater London, United Kingdom
Employment Type
Permanent
Salary
£70000 - £95000/annum
edge AI. Real-time computer vision systems. Docker, Kubernetes, Kubeflow or MLOps pipelines. Aerial, satellite or ISR imagery. Synthetic data generation, Unreal Engine or NVIDIA Omniverse. Distributed training and large-scale model optimisation. Defence, government or national security projects. An MSc or PhD in Computer Vision, Artificial Intelligence, Machine Learning ...

Embedded Software Engineer - Contractor

Location
Retford, England, United Kingdom
architectures, specifically implementing Intel Secure Boot and Arm TrustZone. Engineer reliable software connectivity and data pipelines between Intel host boards and Eizo (ARM/NVIDIA) GPGPUs. Design and integrate Intelligent Platform Management Interface (IPMI) solutions for robust hardware management. Implement Ubuntu and relevant driver infrastructure Technical Operations Author and maintain ...

Senior Software Engineer-DV Security Cleared

Hiring Organisation
Morson Edge
Location
Milton Keynes, Buckinghamshire, South East, United Kingdom
Employment Type
Contract
Contract Rate
£550 - 600 per day
related STEM discipline. Senior Software Engineer- DV Security Cleared- Desirable Technical Experience Experience with some of the following technologies would be highly beneficial: NVIDIA CUDA/GPU-accelerated computing Machine learning and data analytics NumPy/SciPy, PostgreSQL/SQLite, Docker/containerised environments Ansible, Cython, Celery/Redis JavaScript ...

Platform Site Reliability Engineer

Location
Gloucester, England, United Kingdom
with Platform Engineering/Platform SRE to fully support both our infrastructure and platform stacks. Willingness to cross train with HPC Engineering, supported by NVIDIA to enhance our HPC supportability offering Requirements 5+ Years Proven experience in globally scaled, performance-intensive environments operating to a 24/7 support model ...

Infrastructure Site Reliability Engineer

Location
Gloucester, England, United Kingdom
with Platform Engineering/Platform SRE to fully support both our infrastructure and platform stacks. Willingness to cross train with HPC Engineering, supported by NVIDIA to enhance our HPC supportability offering What you bring 5+ Years Proven experience in globally scaled, performance-intensive environments operating to a 24/ ...

HPC/AI Benchmarker

Location
United Kingdom
NCCL, and GPU-accelerated applications. Experience benchmarking and tuning scientific, engineering, AI, and data analytics workloads on large-scale HPC platforms. Strong understanding of NVIDIA GPU technologies and architectures, including H100, H200, B200, and GB200/GB300. Hands-on experience with high-performance networking technologies such as InfiniBand and RoCEv2. ...

Enterprise Architect - AI

Location
Greater London, England, United Kingdom
record leading technical delivery (not just advisory) on enterprise-scale AI or HPC infrastructure programmes. AI Hardware & Data Center Infrastructure GPU/accelerator architectures: NVIDIA/AMD, including multi-node scale-out design. Accelerator interconnects: NVLink, NVSwitch High-performance networking: InfiniBand and RoCEv2 fabric design, 400G/800G Ethernet, rail … frameworks: PyTorch and TensorFlow at a working, hands-on level. Distributed training: Horovod, DeepSpeed, Megatron-LM, or equivalent multi-node training frameworks. Inference & serving: NVIDIA Triton, vLLM, TensorRT-LLM, or equivalent high-throughput serving platforms. MLOps/LLMOps: Kubeflow, MLflow, and at least one hyperscaler ML platform (SageMaker, Azure ...

Senior Observability Engineer

Location
Greater London, England, United Kingdom
have Experience building and operating large‐scale observability systems across multiple data centers. Experience monitoring GPU clusters and AI/ML workloads, including NVIDIA DCGM and training or inference performance metrics. Deeper experience with hardware and network telemetry, such as IPMI/Redfish, NetFlow/sFlow, SNMP, gNMI, or eBPF. ...

Lead Software Engineer - LLM Ops Platform Reliability

Location
Paisley, Scotland, United Kingdom
agents using frameworks such as LangChain, CrewAI, LangGraph, or similar orchestration platforms Experience operating or integrating serving platforms such as KServe, Ray Serve, NVIDIA Triton Inference Server, Text Generation Inference (TGI), alongside vLLM/llm-d Familiarity with Amazon SageMaker JumpStart, SageMaker Endpoints, and Amazon Bedrock for managed model hosting ...

Senior Robotics Engineer

Location
Greater London, England, United Kingdom
expect one person to bring experience in every area.* You may also bring experience with emerging technologies such as NeRFs, Gaussian Splatting, NVIDIA Omniverse, Unity or Unreal, or knowledge of model explainability, open-source licensing or CI/CD-based testing of robotics and machine learning pipelines. These are helpful ...

Senior Robotics Engineer

Location
Epsom, England, United Kingdom
expect one person to bring experience in every area. You may also bring experience with emerging technologies such as NeRFs, Gaussian Splatting, NVIDIA Omniverse, Unity or Unreal, or knowledge of model explainability, open-source licensing or CI/CD-based testing of robotics and machine learning pipelines. These are helpful ...

Senior Infrastructure Operations Engineer (EMEA) New London, England, United Kingdom

Location
Greater London, England, United Kingdom
Haves Experience troubleshooting, provisioning, or operating bare-metal server infrastructure. Experience operating or troubleshooting GPU, HPC, or other high-performance compute infrastructure. Familiarity with NVIDIA GPUs, DCGM, InfiniBand, RoCE/RDMA, NVLink, or high-speed data center networking. Experience with hardware management and provisioning technologies such as PXE, BMC, IPMI ...

Senior Robotics Engineer

Location
Epsom, England, United Kingdom
expect one person to bring experience in every area. You may also bring experience with emerging technologies such as NeRFs, Gaussian Splatting, NVIDIA Omniverse, Unity or Unreal, or knowledge of model explainability, open‐source licensing or CI/CD‐based testing of robotics and machine learning pipelines. These are helpful ...

Principal Network Engineer

Location
Greater London, England, United Kingdom
networking for AI/HPC workloads, including InfiniBand and/or RoCE. Experience with subnet managers and fabric orchestration technologies such as OpenSM or NVIDIA UFM. Expert-level understanding of modern data centre routing and control planes, including BGP, EVPN-VXLAN, and Clos/spine-leaf architectures. Production experience with ...

Enterprise Architect - AI

Hiring Organisation
World Wide Technology
Location
London, UK
Employment Type
Full-time
DescriptionCertificationsNVIDIA certifications (NCP-AI Infrastructure, or NVIDIA Deep Learning Institute credentials) — strongly preferred. Cloud AI/ML certification: AWS Certified Machine Learning – Specialty, Microsoft Certified: Azure AI Engineer Associate, or Google Professional Machine Learning Engineer — at least one preferred. Kubernetes: CKA or CKAD — preferred. TOGAF 9/10 or equivalent … ambiguity in a fast-moving technology space; makes sound architectural calls with incomplete information. Collaborates effectively across sales, pre-sales, delivery, and partner (NVIDIA, hyperscaler, ISV) teams. Education & ExperienceBachelor's degree in Computer Science, Computer Engineering, or a related technical field, or equivalent demonstrable experience. Advanced degree ...

Software Engineer, AI Libraries

Hiring Organisation
wayve
Location
London, UK
Employment Type
Full-time
Argo Workflows. Experience with containerisation and infrastructure tooling such as Docker, Kubernetes, or Terraform. Experience profiling or optimising ML systems, for example using NVIDIA Nsight. Understanding of ML workflows and researcher experience, even if you are not focused on model development. What we're not looking forThis is not primarily ...

Software Engineer London, United Kingdom

Location
Greater London, England, United Kingdom
Metaflow, or Argo Workflows. Experience with containerisation and infrastructure tooling such as Docker, Kubernetes, or Terraform.Experience profiling or optimising ML systems, for example using NVIDIA Nsight. Understanding of ML workflows and researcher experience, even if you are not focused on model development. What we’re not looking for This ...

Senior Lead Software Engineer - LLM Ops Platform Reliability

Hiring Organisation
JP Morgan Chase
Location
Glasgow, UK
Employment Type
Full-time
agents using orchestration frameworks such as LangChain, LangGraph, CrewAI, or similar platformsExperience operating or integrating model serving platforms such as KServe, Ray Serve, or NVIDIA Triton Inference Server alongside other large language model serving stacksFamiliarity with Amazon SageMaker JumpStart, SageMaker Endpoints, and Amazon Bedrock for managed model hostingExperience with online ...

ML/AI Engineer

Location
Manchester, England, United Kingdom
safe in production. Deploy and tune GPU‐backed inference services (e.g., A100), optimise CUDA environments, and leverage TensorRT where appropriate. Operate scalable serving frameworks (NVIDIA Triton, TorchServe) with attention to latency, efficiency, resilience, and cost. Implement end‐to‐end observability for models and pipelines: drift, data quality, fairness signals, latency ...

Account Solution Architect

Location
Greater London, England, United Kingdom
concurrent opportunities across a geographic territory. Fluency in English required; proficiency in Dutch, Swedish, Norwegian, Danish, or Finnish is a plus. Preferred Familiarity with NVIDIA GPU architectures (H100, A100, H200) and the software stack around them: CUDA, NCCL, cuDNN. Working knowledge of high-performance networking concepts: InfiniBand, RDMA, RoCE ...