Founding AI Inference Engineer – Scale & Serving Expert
- Hiring Organisation
- Jobleads-UK
- Location
- Greater London, England, United Kingdom
CTO. You’ll own the layer above CUDA/GPU work, shaping inference delivery, throughput, and reliability as we scale data-centre compute for energy applications. With 4+ years in large-scale inference systems, you’ll collaborate with GPU engineers on model-level optimisations and architecture decisions ...