26 to 31 of 31 AI Model Optimisation Jobs in London

Software Engineer, Model Inference, DeepMind

Hiring Organisation
Hackajob Ltd
Location
South West London, London, United Kingdom
Employment Type
Permanent
mission is the complex task of measuring the intelligence of our prototypes. As a Software Engineer, you will be working with the cutting edge AI agents developed by our exceptional team of Machine Learning and Neuroscience research scientists. Your responsibilities will include everything from creating systems for agent testing … working on a wide range of challenging problems within a mission-driven team. In this role, you will be at the forefront of bringing AI research to life. You'll work directly with researchers and engineers to optimize and deploy large language models (LLMs) like Gemini onto Google ...

Software Engineer, Model Inference, DeepMind

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
production environment. Experience in profiling, configuring, or executing ML workloads directly on hardware accelerators (e.g., GPU or TPU). Experience designing, building, or optimizing model serving infrastructure or inference backends. Preferred qualifications: Experience with developing serving infrastructure. Experience programming hardware accelerators (GPUs, TPUs) via ML frameworks (e.g., JAX, PyTorch … programming models (e.g., Pallas, CUDA, OpenCL). Experience profiling software to identify performance bottlenecks. Experience with distributed ML systems optimization and parallelism (e.g., data, model, or pipeline parallelism). Familiarity with writing performance-optimized kernels. Understanding of LLM architecture and inference performance dynamics (e.g., Transformer models, memory bandwidth ...

Machine Learning Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
learning technologies. Past projects have included: Evaluating alternative accelerators for ML workloads Multi-node distributed training to understand trade‐offs in networking technology Optimising model inference to minimise latency or maximise throughput Understanding and optimising different storage technology to maximise bandwidth Evaluating the latest hardware and software … data‐science competitions, such as Kaggle) Strong object‐oriented engineering skills, with experience in Python, PyTorch and NumPy desirable The ability to apply advanced optimisation methods, modern ML techniques, HPC, profiling or model‐inference expertise; you do not need to have all of the above A passion ...

AI Systems Researcher/Engineer — Privacy ML (Hybrid UK)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Brave is seeking an AI Systems Researcher/Research Engineer in London (Hybrid) to advance on-device ML and privacy-focused technologies. The role requires a deep background in AI systems research, ML model optimization, and experience with modern ML frameworks. You will contribute to high-value ...

Principal Technical Product Manager, Prime Video Commerce Personalization and Member Experience

Hiring Organisation
Appcast
Location
London, UK
including offer propensity, churn prediction, subscription ranking, and cross-offer optimization — serve decisions across every Prime Video commerce surface. We are evolving from siloed model optimization to a unified system that coordinates across the full customer journey. We are a product development team that partners closely with cross-functional … software engineering teams- Experience in designing experiments and statistical analysis of results- Experience owning and delivering end-to-end product strategy for ML/AI-powered consumer products.Preferred qualification - Experience driving direction and alignment with cross-functional teams- Experience with subscription or commerce business models and LTV-based decision ...

LLM Serving Platform Engineer — Build Scalable AI Infra

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
environments and customers. The role emphasizes backend system design, end-to-end ownership, and collaboration with researchers. You will contribute to fault-tolerant services, model optimization, observability, and platform discovery, while staying #J-18808-Ljbffr ...