10 of 10 Permanent Site Reliability Engineering Jobs in Surrey

DevOps Engineer

Location
Reigate, England, United Kingdom
parameterized templates for Azure infrastructure. Automate environment provisioning across development and production. Manage backend state, pipelines, and state-change detection integrations. Platform Engineering & SRE Own and improve reliability, observability, and performance of the platform. Implement SLOs, alerting, dashboards, and auto remediation where possible. Troubleshoot cluster level, networking … Azure DevOps: repos, pipelines, artifacts. Solid understanding of Azure networking, DNS/private endpoints, Certificate/Secret management etc Strong debugging and operational experience (SRE mindset). Solid experience of DevSecOps architecture, processes & tooling Solid understanding of Observability Process & Tooling Logging, metrics, traces, dashboards Other highly desirable, but not essential ...

Senior Data Solutions Engineer

Location
Woking, England, United Kingdom
leading-edge products within CI. You will work collaboratively to solve interesting and challenging problems in a cross functional team with other engineering disciplines including Software Engineers, Platform Engineers, Test Engineers, Site Reliability Engineers and Systems Engineers. The analytics team is missioned to build operational insight … with the expectation of a minimum of two days a week in the office, which should be co‐ordinated with other members of the Engineering team. Basic Qualifications We are looking for an applicant who has: Bachelor's degree in a numerate discipline: Maths, Physics, Computer Science, Statistics, Engineering. ...

Senior Platform Engineer

Location
Reigate, England, United Kingdom
Description As a Senior Software Engineer (SaaS), you will join a substantial global engineering organisation and provide leadership to highly skilled engineers in the software platform area. You will be part of a team that follows Agile methodologies to deliver market-leading insurance solutions. This is a new position … access management using Azure AD/Entra ID, B2B, RBAC, OAuth2, OIDC, JWT and claims‐based authorisation. Experience in platform engineering or SRE roles: building internal platforms, defining SLIs/SLOs, managing error budgets, and implementing observability (centralised logging, metrics, distributed tracing). Strong awareness of emerging cloud ...

Senior Data Solutions Engineer: Cloud & Edge Analytics Lead

Location
Woking, England, United Kingdom
Data Platform, delivering scalable data solutions across cloud and edge environments. You will work in a cross-functional team with engineers across software, platform, SRE and testing, applying Python and modern data tooling to build streaming pipelines and analytics apps. Hybrid role based in Woking with two days ...

Senior DevOps Engineer: Cloud Platform & SRE (Hybrid)

Location
Reigate, England, United Kingdom
DevOps Engineer to join our global SaaS platform team. You will help evolve Radar Live SaaS, applying IaC, CI/CD, cloud automation, and SRE practices to ensure reliability and security at scale. You’ll work across Dev, Ops, Security and Architecture in an Agile environment, building automated deployments ...

Principal Platform Engineer || Identity Platform (AuthN/AuthZ)

Location
Staines-upon-Thames, England, United Kingdom
provisioning workflows Security governance, IAM audit, policy authoring or architecture-only work Kubernetes RBAC and cloud IAM policies as part of a DevOps or SRE role All valuable work. None of it is this job. It is a fit if you have personally run an identity provider in production. Installed … with Keycloak estates migrating onto it. You will architect and build that, own it in production, and set the identity patterns the rest of engineering follows. This is a hands‐on engineering role. You will write Go. What We Need To See Authorisation Fine‐grained authorisation systems ...

Principal Platform Engineers

Location
Staines-upon-Thames, England, United Kingdom
with Keycloak estates migrating onto it. You will architect and build that, own it in production, and set the identity patterns the rest of engineering follows. This is a hands-on engineering role. You will write Go. Fine-grained authorisation systems you have built and run at production … provisioning workflows Security governance, IAM audit, policy authoring or architecture-only work Kubernetes RBAC and cloud IAM policies as part of a DevOps or SRE role All valuable work. None of it is this job. It is a fit if you have personally run an identity provider in production. Installed ...

Operations Team Lead (Production & Reliability)

Location
Guildford, England, United Kingdom
improving. This is a hands‐on role. You’ll shape process, lead incidents, build the team, and move us from reactive firefighting to proactive reliability engineering. What You’ll Own Production Stability and availability of all live systems Operational readiness for new releases Safe production access and change coordination … Raise the bar on operational discipline You’re responsible for both system performance and team performance. What We’re Looking For Strong experience in SRE, DevOps, Infrastructure, or Production Engineering Prior experience leading technical teams Deep hands‐on incident management experience Strong observability and reliability mindset Calm under ...

Senior Project Manager

Location
Woking, England, United Kingdom
safe, secure and prosperous. Work on leading edge technology solutions in the following disciplines: AI & Data Science, Cyber, Cloud, Big Data, Software Development, DevOps, SRE, Platform Engineering. Role We are looking for an experienced, tenacious and self‐motivated Project Manager to join our National Security (NS) Business Unit as part … effective problem solving ability supported by pro‐active and adaptive team leadership skills Must have a strong appreciation of technology and trends, and associated engineering methods Must be flexible and willing to contribute where needed in the business cycle to further the team’s success Able to operate ...

Operations Team Lead: Scale Reliable Production

Location
Guildford, England, United Kingdom
will drive reliability, observability, and continuous improvement, shaping processes, leading incidents, and building a high-performing team. This hands-on role requires deep SRE/DevOps expertise, leadership, and a calm, clear communicator in a fast-paced environment. You will be responsible for incident lifecycle management, defining runbooks ...