26 to 34 of 34 Site Reliability Engineering Jobs in the East of England

Operations Team Lead (Production & Reliability)

Location
Cambridge, England, United Kingdom
improving. This is a hands‐on role. You’ll shape process, lead incidents, build the team, and move us from reactive firefighting to proactive reliability engineering. What You’ll Own Production Stability and availability of all live systems Operational readiness for new releases Safe production access and change coordination … Raise the bar on operational discipline You’re responsible for both system performance and team performance. What We’re Looking For Strong experience in SRE, DevOps, Infrastructure, or Production Engineering Prior experience leading technical teams Deep hands‐on incident management experience Strong observability and reliability mindset Calm under ...

Operations Team Lead (Production & Reliability)

Location
Watford, England, United Kingdom
improving. This is a hands‐on role. You’ll shape process, lead incidents, build the team, and move us from reactive firefighting to proactive reliability engineering. What You’ll Own Production Stability and availability of all live systems Operational readiness for new releases Safe production access and change coordination … Raise the bar on operational discipline You’re responsible for both system performance and team performance. What We’re Looking For Strong experience in SRE, DevOps, Infrastructure, or Production Engineering Prior experience leading technical teams Deep hands‐on incident management experience Strong observability and reliability mindset Calm under ...

Service Design Specialist

Hiring Organisation
ARM
Location
Cambridge, Cambridgeshire, UK
Employment Type
Full-time
Overview: We are building a modern Service Management capability combining ITIL 4, SRE, automation and operational governance to enable fast, reliable delivery. This role translates business and technical requirements into practical, end-to-end service designs, creating the models, documentation and readiness evidence needed to ensure services are supportable, resilient … Skills and Experience: Experience with ServiceNow or a comparable ITSM platform, particularly service catalogue, CMDB, CSDM, service mapping, knowledge or workflow capabilities! Understanding of SRE concepts such as SLAs, SLOs, SLIs, error budgets, service health, reliability and observability. Experience in DevOps, CI/CD, cloud, SaaS, PaaS, platform engineering ...

Performance and Monitoring Engineer

Hiring Organisation
Solus Accident Repair Centres
Location
Birchanger, Hertfordshire, United Kingdom
Employment Type
Permanent
Salary
GBP 40,000 - 50,000 Annual
Aviva family, is growing our Technology capability and we're looking for a talented Performance and Monitoring Engineer to help us strengthen the stability, reliability and performance of our systems. If you're passionate about monitoring, observability and using data to proactively improve service health, this is a great … qualifications Microsoft certifications (AZ-900, AZ-104, AZ-305, AZ-500) or similar Experience with LogicMonitor admin, Grafana or other observability tools Familiarity with SRE concepts (SLIs, SLOs, error budgets) Understanding of ITIL processes Who are Solus? Solus, who are owned by Aviva, are one of the UK leaders ...

Performance and Monitoring Engineer

Hiring Organisation
Solus Accident Repair Centres
Location
Stansted, Essex, South East, United Kingdom
Employment Type
Permanent
Salary
£50,000
Aviva family, is growing our Technology capability and we're looking for a talented Performance and Monitoring Engineer to help us strengthen the stability, reliability and performance of our systems. If you're passionate about monitoring, observability and using data to proactively improve service health, this is a great … qualifications Microsoft certifications (AZ-900, AZ-104, AZ-305, AZ-500) or similar Experience with LogicMonitor admin, Grafana or other observability tools Familiarity with SRE concepts (SLIs, SLOs, error budgets) Understanding of ITIL processes Who are Solus? Solus, who are owned by Aviva, are one of the UK leaders ...

Operations Team Lead — Production Reliability & Scale

Location
Cambridge, England, United Kingdom
Cambridge is seeking an Operations Team Lead to own production and scale systems. You will lead operational excellence across live customer-facing platforms, ensuring reliability, observability, and proactive improvements. This hands … role involves shaping processes, guiding incidents, building the team, and moving from firefighting to sustainable reliability engineering. The ideal candidate has strong SRE/DevOps background and a calm, clear communicator. #J-18808-Ljbffr ...

Production Reliability Lead

Location
Norwich, England, United Kingdom
Complexio in the United Kingdom is seeking an Operations Team Lead to own production across live customer-facing systems and to build a scalable reliability engine. … will shape processes, lead incidents, and grow the team to move from firefighting to proactive reliability engineering. This hands-on role requires leading SRE/DevOps practices, defining SLIs/SLOs, improving observability, and ensuring clear ownership with blameless accountability. #J-18808-Ljbffr ...

Production Reliability Team Lead

Location
Hemel Hempstead, England, United Kingdom
Complexio is seeking an Operations Team Lead to own production. Not just keep it running, but build a system that scales, improves reliability, and delivers predictable performance for our enterprise customers. You’ll lead incident management, define monitoring … drive on-call rotations, and mentor the team to ship steady, observable, and blameless production practices. This hands-on role requires strong experience in SRE, DevOps, and production engineering, plus clear communication and a #J-18808-Ljbffr ...

Operations Team Lead: Drive Reliable Production

Location
Basildon, England, United Kingdom
Complexio is seeking an Operations Team Lead to own production and build scalable reliability across live … customer-facing systems. You will lead incidents, define runbooks, and shape processes to move from firefighting to proactive reliability engineering. We expect strong SRE/DevOps experience, hands-on leadership, and a calm, communicative approach under pressure. This is a demanding, high-discipline role focused on measurable improvements ...