GenAI Infra Engineer: Real-Time GPU Serving & Fine-Tuning
- Hiring Organisation
- Jobleads-UK
- Location
- Greater London, England, United Kingdom
serving, high-throughput batch inference, and model fine-tuning for open-weight platforms. You will collaborate across model serving, inference engines, training pipelines, and observability, pushing cost/performance frontiers while meeting latency and reliability targets. #J-18808-Ljbffr ...