Lead Site Reliability Engineer (Kubernetes Required) - Hybrid
- Location
- Greater London, England, United Kingdom
reliability, scalability, and performance of our systems and services. You will work closely with development and operations teams to build and maintain robust infrastructure, automate processes, and drive engineering best practices. **Key Responsibilities*** Monitor, maintain, and improve the reliability and availability of production systems* Respond to and resolve incidents … Technical Skills*** **Cloud Platforms:** *(e.g. AWS, GCP, Azure)** **CI/CD Tooling:** *(e.g. GitHub Actions, ArgoCD, Harness)** **Monitoring & Observability:** *(e.g. Prometheus, Grafana, Coralogix, OpenTelemetry)** **Infrastructure as Code:** *(e.g. Terraform, Pulumi)** **Config Management**: *(e.g. Ansible, Puppet, Chef)** **Programming/Scripting:** *(e.g. Python, Go, Bash)* **Soft Skills ...