Senior Software Engineer, Inference Platform
- Hiring Organisation
- Jobleads-UK
- Location
- Greater London, England, United Kingdom
engineer who has built and scaled high-performance inference systems for AI/ML workloads. You understand the complexities of serving models at scale latency optimization, resource orchestration, autoscaling dynamics, and production reliability. You’ve designed distributed systems that handle thousands of requests per second while maintaining sub‐second … think about the end‐to‐end user experience. You’re a team player comfortable wearing multiple hats one day you’re optimizing inference latency, the next you’re joining customer calls to understand their deployment challenges, and the day after you’re helping with UI/UX, customer success ...