ML/AI Engineer
- Location
- Manchester, England, United Kingdom
leverage TensorRT where appropriate. Operate scalable serving frameworks (NVIDIA Triton, TorchServe) with attention to latency, efficiency, resilience, and cost. Implement end‐to‐end observability for models and pipelines: drift, data quality, fairness signals, latency, GPU utilisation, error budgets, and SLOs/SLIs via Prometheus, Grafana, and Dynatrace. Establish actionable alerting ...