Principal Site Reliability Engineer
- Location
- Greater London, England, United Kingdom
Amazon Web Services - multi-region, multi-account, with a broad range of managed services Containers and orchestration - Kubernetes (GKE and EKS) Infrastructure-as-Code - Terraform, Terragrunt, and Atlantis CI/CD - GitLab Pipelines, ArgoCD, Octopus Deploy Data - ElasticSearch hosted with Kubernetes Operator, PostgreSQL, SQL Server, BigQuery Monitoring and Security - Splunk … cross-cloud problems - networking, identity, data movement, and the operational patterns that hold across environments Drive automation of infrastructure provisioning and configuration management using Terraform and related IaC tooling Establish and maintain comprehensive monitoring, alerting, and observability practices Cross-train and mentor other engineers, with the explicit goal of broadening ...