10 of 10 Site Reliability Engineering Jobs in Hertfordshire

Senior / Lead Site Reliability Engineer

Hiring Organisation
Jobleads-UK
Location
Watford, England, United Kingdom
customer-facing systems during both normal operation and peak lottery events. The role combines hands‐on engineering, incident leadership, and ownership of the SRE improvement backlog and reporting, working across platform, product, and operational teams. Objectives of the role Own reliability outcomes across services using SLOs, SLIs … performance optimisation: Latency reduction Throughput scaling Cost efficiency (AWS utilisation and associated log costs, observability license consumption) Backlog ownership & reporting Own and prioritise the SRE backlog, balancing: Reliability improvements Technical debt Automation opportunities to reduce/offload toil Produce structured reporting covering: SLO performance Incident trends and MTTR Platform ...

Senior / Lead Site Reliability Engineer

Hiring Organisation
Jobleads-UK
Location
Watford, England, United Kingdom
customer-facing systems during both normal operation and peak lottery events. The role combines hands‐on engineering, incident leadership, and ownership of the SRE improvement backlog and reporting, working across platform, product, and operational teams. Objectives of the role Own reliability outcomes across services using SLOs, SLIs … performance optimisation: Latency reduction Throughput scaling Cost efficiency (AWS utilisation and associated log costs, observability license consumption) Backlog ownership & reporting Own and prioritise the SRE backlog, balancing: Reliability improvements Technical debt Automation opportunities to reduce/offload toil Produce structured reporting covering: SLO performance Incident trends and MTTR Platform ...

Site Reliability Engineer

Hiring Organisation
Jobleads-UK
Location
Watford, England, United Kingdom
capacity planning Collaboration Work closely with engineers to: + Improve service reliability + Support releases and production readiness Contribute to adoption of SRE practices within teams What experience we’re looking for Technical Experience in cloud environments (AWS preferred) Working knowledge of: + Containers (ECS; exposure to Kubernetes … Bash, etc.) Troubleshooting & operations Ability to diagnose issues in distributed systems Familiarity with monitoring and logging tools Understanding of Linux systems and networking fundamentals SRE fundamentals Understanding of: + Monitoring and alerting concepts + Reliability principles + Incident response processes Willingness to be part of an on-call rotation ...

Principal/Senior Site Reliability Engineer

Hiring Organisation
Jobleads-UK
Location
Welwyn, England, United Kingdom
global regions. Design for resilience, building disaster recovery and failover plans with auto‐scaling and load balancing to keep critical systems available worldwide. Strengthen reliability through chaos engineering experiments that validate systems and surface weaknesses before incidents. Build deep observability with monitoring, logging, and alerting frameworks such … through influence, with strong communication, mentoring, and problem‐solving abilities. Degree in Computer Science or related technical field, or equivalent experience in software and site reliability engineering. Preferred experience with distributed ML frameworks such as Horovod or TensorFlow Distributed, familiarity with data engineering pipelines such as Apache ...

Site Reliability Engineer - Scale, Observe, Automate

Hiring Organisation
Jobleads-UK
Location
Watford, England, United Kingdom
Allwyn is seeking a Site Reliability Engineer to support the reliability and performance of our digital services. You will work with production systems, automation, and observability to keep services stable, scalable and well-instrumented across peak events. You will collaborate with senior SREs and engineering teams … drive automation, incident response, and improvement of runbooks, while contributing to platform reliability and user experience signals. #J-18808-Ljbffr ...

Senior Site Reliability Engineer - Lead Resilient Platform

Hiring Organisation
Jobleads-UK
Location
Watford, England, United Kingdom
Allwyn in the UK is seeking a Senior/Lead Site Reliability Engineer to provide technical leadership for reliability across the digital estate, ensuring high availability, performance, and resilience of customer-facing systems during normal operation and peak lottery events. The role blends hands-on engineering ...

Senior SRE: Lead Reliability for Scalable Platforms

Hiring Organisation
Jobleads-UK
Location
Watford, England, United Kingdom
Allwyn UK is seeking a Senior/Lead Site Reliability Engineer to provide technical leadership for reliability across the digital estate, ensuring high availability and performance during normal operations and peak lottery events. You will own SLO/SLI definitions, incident leadership, and a roadmap for platform ...

Senior Platform Engineer

Hiring Organisation
Jobleads-UK
Location
Welwyn Garden City, England, United Kingdom
Senior Platform Engineer, you will lead the design, evolution, and reliability of the core platform that underpins our engineering ecosystem. You will set technical direction, define standards, and drive best practices that enable product teams to deliver securely, efficiently, and at scale. Your role goes beyond implementation … tools, technologies, and approaches to keep the platform modern, efficient, and competitive. What we would like from you Strong experience in platform engineering, SRE, or DevOps within a distributed cloud environment. Deep expertise in Kubernetes and containerised workloads, ideally in managed environments such as AKS. Proven experience designing ...

Cloud Native DevOps Engineer- SC Cleared

Hiring Organisation
17918
Location
Watford, Hertfordshire, United Kingdom
CLEARED CLOUD NATIVE DEVOPS ENGINEER - Permanent opportunity for a Cloud Native DevOps Engineer with SC Clearance. - Salary up to £95,000 DOE - On-site opportunity with Gloucester based offices - To apply, please call Laura Jackson on 02038540120, or email with an up-to-date CV. … process and submit (subject to required skills) your application to our client in conjunction with this vacancy only. KEY SKILLS DEVOPS ENGINEER, DEVOPS, SITE RELIABILITY ENGINEER, PLATFROM ENGINEER, PLATFROM, CLOUD, CLOUD NATIVE, AWS, AWS CLOUD, AWS ENGINEER, ANSIBLE, TERRAFORM, CLOUD SUPPORT, CLOUD INFRASTRUCTURE, DEFENCE, NATIONAL SECURITY, DV CLEARED ...

Performance and Monitoring Engineer

Hiring Organisation
Solus Accident Repair Centres
Location
Birchanger, Hertfordshire, United Kingdom
Employment Type
Permanent
Salary
GBP 40,000 - 50,000 Annual
Aviva family, is growing our Technology capability and we're looking for a talented Performance and Monitoring Engineer to help us strengthen the stability, reliability and performance of our systems. If you're passionate about monitoring, observability and using data to proactively improve service health, this is a great … qualifications Microsoft certifications (AZ-900, AZ-104, AZ-305, AZ-500) or similar Experience with LogicMonitor admin, Grafana or other observability tools Familiarity with SRE concepts (SLIs, SLOs, error budgets) Understanding of ITIL processes Who are Solus? Solus, who are owned by Aviva, are one of the UK leaders ...