SRE Technical Lead
Location: Hybrid (UK - office, client site and home-based working)
Salary: Up to £100,000 + 5% bonus
We're looking for an experienced SRE Technical Lead to take ownership of the reliability, availability and operational excellence of critical platforms within complex, multi-vendor environments.
This is a senior technical leadership role where you'll act as the technical authority for Site Reliability Engineering, working closely with stakeholders, engineering teams and delivery partners to drive reliability across large-scale cloud platforms.
Essential Security Requirements
To be considered for this role, you must:
- Hold active SC (Security Check) clearance.
- Be a sole UK national.
Unfortunately, candidates who do not meet both of these essential requirements cannot be considered.
The Role
As the SRE Technical Lead, you will:
- Define and drive the SRE strategy, standards, SLAs, SLOs and error budgets.
- Embed reliability engineering principles into platform and service design.
- Lead the adoption of core SRE practices including reliability reviews, operational readiness and toil reduction.
- Drive automation across monitoring, incident response, recovery and remediation.
- Govern reliability-focused Infrastructure as Code, CI/CD pipelines and operational tooling.
- Identify and remove systemic causes of operational overhead while improving scalability, resilience and operability.
- Act as the senior technical escalation point for major incidents and high-risk releases.
- Lead blameless post-incident reviews and ensure measurable service improvements.
- Define and oversee observability, monitoring and capacity management practices.
- Ensure SRE approaches align with security, governance and compliance requirements.
- Mentor and coach senior engineers, helping to improve SRE maturity across engineering teams.
This role provides technical leadership through influence rather than direct line management and works closely with Cloud, Platform, Security and Operations teams.
About You
You'll have strong technical expertise gained within enterprise-scale environments, including:
- Deep knowledge of Kubernetes and OpenShift.
- Experience designing and supporting hybrid and multi-cloud platforms.
- Experience with service mesh technologies such as Istio.
- Strong hands-on experience with observability tooling including Prometheus, Grafana, Loki, Tempo and OpenTelemetry.
- Infrastructure as Code and GitOps expertise using tools such as Helm, Kustomize, ArgoCD and Tekton.
- Experience building and improving CI/CD pipelines with a focus on reliability engineering.
- Familiarity with Red Hat ACM/ACS, Submariner networking and enterprise databases such as PostgreSQL.
- A proven track record of providing technical leadership across complex, multi-vendor environments.
What's on Offer
- Salary up to £100,000
- 5% annual bonus
- Hybrid working model
- Opportunity to lead reliability engineering across large-scale, business-critical platforms
- Exposure to modern cloud-native technologies and complex enterprise environments
Please note: Active SC clearance and sole UK nationality are mandatory requirements for this position. Applications from candidates who do not meet these criteria cannot be progressed.