1 of 1 Permanent Distributed Systems Jobs in Halifax

Lead Cloud Site Reliability Engineer

Location
Halifax, England, United Kingdom
comparable cloud environments. Site Reliability Engineering (SRE), Platform Engineering, Infrastructure Engineering, Cloud Engineering or Production Operations. Observability and monitoring practices, including metrics, logging and distributed tracing. Incident management, problem management and service reliability improvement. Service Level Objectives (SLOs), Service Level Indicators (SLIs) and error budgets. Automation and reducing operational … both technical and non-technical stakeholders. Working collaboratively across multiple teams and disciplines. Desirable Experience Azure and GCP platform technologies. Cloud-native architectures and distributed systems. Infrastructure as Code (IaC) and platform automation. Large-scale enterprise or regulated technology environments. Operational resilience and availability engineering practices. What ...