cloud engineer in insurance
- Location
- Greater London, England, United Kingdom
Drive improvements in reliability, resilience, performance, and efficiency through SRE practices Own and enhance observability and monitoring, including meaningful alerting, operational dashboards, and rapid incident diagnosis Improve operational maturity through incident management, root-cause analysis, and continuous improvement Ensure workloads are designed, deployed, and operated according … such as Terraform, CI/CD pipelines, and automation tooling Hands-on expertise in SRE and operational excellence, including monitoring, alerting, reliability improvement, and incident response Proven ability to lead technical initiatives end-to-end while balancing robustness, efficiency, and developer usability Experience operating and supporting Kubernetes-based workloads ...