Staff ML Performance Engineer (Compiler)
- Location
- Greater London, England, United Kingdom
pinpoint bottlenecks across the full inference stack (model graph, compiler/runtime, kernel execution, memory movement) and deliver measurable improvements. Build robust benchmarking and regression testing to ensure performance improvements hold across models, devices, and software releases. Develop and optimise for multiple target platforms (e.g. NVIDIA Orin/… multiple levels of abstraction — from high-level model behaviour down to low-level kernel/runtime execution. Strong software engineering fundamentals (debugging, profiling, testing, and maintainable code). Clear communicator and collaborative teammate; able to align multiple stakeholders on performance trade‐offs and priorities. Desirable Experience with compute graph ...