1 to 25 of 32 Site Reliability Engineering Jobs in the East of England

Site Reliability Engineering Professional

Location
Ipswich, England, United Kingdom
role BT International is transforming the way we operate, evolving from traditional network operations to a modern Site Reliability Engineering (SRE) and Platform Operations model. As a Site Reliability Engineering Professional, you will play a key role in the operational management of BT International … International's next-generation platforms, collaborating with engineering, product, supplier, and operational teams to improve service reliability through automation, observability, and SRE best practices. What you'll be doing Support the 24x7 operation of BT International's core network and platform services, ensuring high availability and performance. Proactively ...

Site Reliability Engineer

Location
Cambridge, England, United Kingdom
learn more, visit http://www.darktrace.com. **Job D****escription****:**## **About the Role**We’re looking for a **Site Reliability Engineer (SRE)** to bring deep expertise in a key reliability domain and help shape the future of our platform reliability strategy.SRE sits at the heart … your area of specialism**, working across teams to embed best practices, solve complex reliability challenges, and improve system resilience at scale.Unlike a generalist SRE, this role focuses on a **core domain of expertise**—such as **observability, performance engineering, data infrastructure reliability, security-focused SRE, or network reliability ...

Site Reliability Engineer

Location
Cambridge, England, United Kingdom
world’s largest content providers, including Amazon, Google and Microsoft trust Bango technology to reach subscribers everywhere. Bango, where people subscribe. Role As a Site Reliability Engineer at Bango, you own the reliability, performance and continuous improvement of the Bango Platform end-to-end — from the infrastructure … constructive, detailed feedback. Call out areas where Bango can improve service or reduce cost. Essentials 3+ years' experience in a Cloud, Platform, DevOps or SRE role in a commercial environment. Production experience with a major cloud provider (AWS, Azure or GCP). Strong Linux administration and troubleshooting (process management, memory ...

Senior Site Reliability Engineer

Location
Cambridge, England, United Kingdom
greater scope, complexity and influence, and you take responsibility for lifting the capability of the engineering function around you. Like the SRE role, you combine three things that have historically sat in separate teams: platform and cloud infrastructure engineering, automation and delivery pipeline ownership, and proactive/reactive … trusted presence in the most serious incidents and the most contentious design decisions. You take deliberate responsibility for the strength and depth of the SRE function, not just your own output — spotting capability gaps and key-person risks before they bite, growing engineers into harder work, spreading concentrated knowledge ...

Senior / Lead Site Reliability Engineer

Location
Watford, England, United Kingdom
customer-facing systems during both normal operation and peak lottery events. The role combines hands-on engineering, incident leadership, and ownership of the SRE improvement backlog and reporting, working across platform, product, and operational teams. Objectives of the role Own reliability outcomes across services using SLOs, SLIs … performance optimisation: Latency reduction Throughput scaling Cost efficiency (AWS utilisation and associated log costs, observability license consumption) Backlog ownership & reporting Own and prioritise the SRE backlog, balancing: Reliability improvements Technical debt Automation opportunities to reduce/offload toil Produce structured reporting covering: SLO performance Incident trends and MTTR Platform ...

Site Reliability Engineer: Observability & Platform Resilience

Location
Cambridge, England, United Kingdom
Darktrace Ltd is looking for a Site Reliability Engineer in Cambridge, UK, to enhance their platform reliability strategy. In this pivotal role, you will work closely with Platform Engineering and DevSecOps to implement best practices and solve complex reliability challenges. The ideal candidate will have … substantial experience in Site Reliability Engineering or a related field, with a strong background in programming, cloud platforms, and performance engineering. The role offers a dynamic work environment with competitive benefits including private medical insurance, life insurance, and additional holiday days. #J-18808-Ljbffr ...

Cloud Platform Tech Lead

Hiring Organisation
Danaher
Location
Cambridgeshire, United Kingdom
Employment Type
Full Time
Azure supporting specific business and technology requirements. Working closely with Enterprise Architecture, Cyber Security, Product Engineering, and Site Reliability Engineering (SRE), the Platform Technology Lead delivers cloud platform capabilities, infrastructure automation, self-service services, and operational improvements that enable faster, safer, and more reliable technology delivery. … driving record It would be a plus if you also possess previous experience in: Microsoft Azure and Azure Landing Zones. Platform Engineering and SRE operating models. GitHub Actions and Azure DevOps. Observability and monitoring platforms. Abcam, a Danaher operating company, offers a broad array of comprehensive, competitive benefit programs ...

Site Reliability Engineer

Location
Hemel Hempstead, England, United Kingdom
Site Reliability Engineer****Home-based (Remote-first, with occasional travel to our Hemel Hempstead office and off-site events)****Permanent | Full Time****Competitive salary + bonus and benefits****About the role**Are you a hands-on Site Reliability Engineer who thrives … intersection of cloud, code and reliability? If so, we want to hear from you! Haven is looking for a technically strong SRE to join our Product Technology function and help raise the bar for how our digital platforms perform, scale and recover.You'll work shoulder to shoulder with engineering ...

Senior Software Development Engineer (SRE)

Location
Cambridge, England, United Kingdom
patterns that standardize and elevate the resilience of our SaaS products across multiple regions and environments. Cultivate a shared responsibility model where the SRE team collaborates with and educates engineering teams on reliability best practices. Contribute to incident response and management, ensuring rapid resolution, clear stakeholder communication … improve operational efficiency and scalability. Champion Service-Oriented Organization (SOO) principles to ensure accountability and clarity in service ownership. Qualifications 6+ years in SRE, DevOps or related role in a large-scale environment Software development experience(ideally working with and as a .NET developer) Strong understanding of SDLC, microservice ...

Network Engineer

Hiring Organisation
Third Nexus Group Limited
Location
Cambridge, Cambridgeshire, United Kingdom
Employment Type
Contract
Contract Rate
£375 - £400/annum
governance, ITSM integration, automation expansion, documentation standards, platform optimisation, network baselining and development of a strategic roadmap towards Site Reliability Engineering (SRE) practices. Key Responsibilities Review the existing network data landscape, including current data sources, data quality, and data flows between network management, monitoring, automation … automation workflows, operational runbooks and controlled implementation processes. Implement new platform features and reduce repetitive manual activity. Define and develop a roadmap towards Network SRE practices. Propose service reliability measures, including relevant KPIs, SLIs and SLOs. Identify opportunities for proactive monitoring, fault prevention and self-healing automation. Work with ...

Senior Site Reliability Engineer – Global Platform Ops

Location
Ipswich, England, United Kingdom
Group is seeking an Site Reliability Engineering Professional to help operate BT International's global core platforms. You will ensure reliability, security, and scalability while driving incident resolution and continual service improvement across international networks and digital platforms. You will collaborate with engineering, product, suppliers ...

Site Reliability Engineer – Platform & Automation

Location
Cambridge, England, United Kingdom
Bango plc is seeking a Site Reliability Engineer to own reliability, performance and continuous improvement of the Bango Platform end-to-end. You’ll merge platform engineering, automation and delivery ownership with proactive incident response and customer impact management, shaping the automation and security posture across … stack. As a member of the Managed Services & Support team, you’ll collaborate with NOC, Partner Support, TSM and other engineering groups to ensure secure, #J-18808-Ljbffr ...

Senior Site Reliability Engineer: Lead Resilience & Incidents

Location
Watford, England, United Kingdom
Allwyn UK in Watford seeks a Senior/Lead Site Reliability Engineer to drive reliability across the digital estate, ensuring high availability and performance of customer-facing systems during normal operation and peak lottery events. You will own SLOs/SLIs, push automation with Terraform, mentor engineers ...

Senior SRE: Observability, Automation & Reliability

Location
Cambridge, England, United Kingdom
Altium is seeking a Senior Site Reliability Engineer to ensure the reliability, availability, and performance of large-scale SaaS platforms across regions. You will automate operations, improve observability, and contribute to incident management in collaboration with DevOps and engineering teams. Join a team that champions … principles, and scalable deployments, driving reliability improvements and faster incident resolution across the cloud platform. #J-18808-Ljbffr ...

DevOps Engineer (Platform Engineering & Security0

Hiring Organisation
Sanderson Recruitment
Location
Peterborough, Cambridgeshire, East Anglia, United Kingdom
Employment Type
Contract
Contract Rate
£500 - £550 per day
DevOps Engineer (Platform Engineering & Security) Peterborough - 3 days per week onsite £500-£520 (Inside IR35) 6 Months Initial We're supporting a large enterprise organisation as they continue to invest in platform engineering, security automation and software delivery modernisation. This is an opportunity for an experienced Senior DevOps … scalable platform solutions. Contribute to platform standardisation, engineering best practices, and continuous improvement initiatives. Experience Required Strong background in DevOps, Platform Engineering, SRE, or Infrastructure Engineering. Experience building and supporting enterprise-scale CI/CD and automation platforms. Solid understanding of DevSecOps principles and security-focused engineering ...

Senior Platform Engineer

Location
Welwyn Garden City, England, United Kingdom
Senior Platform Engineer, you will lead the design, evolution, and reliability of the core platform that underpins our engineering ecosystem. You will set technical direction, define standards, and drive best practices that enable product teams to deliver securely, efficiently, and at scale. Your role goes beyond implementation … tools, technologies, and approaches to keep the platform modern, efficient, and competitive. What we would like from you Strong experience in platform engineering, SRE, or DevOps within a distributed cloud environment. Deep expertise in Kubernetes and containerised workloads, ideally in managed environments such as AKS. Proven experience designing ...

Remote SRE: Cloud Reliability & AI-Driven Ops

Location
Hemel Hempstead, England, United Kingdom
Haven is seeking a hands-on Site Reliability Engineer to join our Product Technology team. This remote-first role involves shaping CI/CD, observability, and incident response while collaborating with engineers and tech leads to ensure reliable, scalable platforms for guests and colleagues. You’ll tackle infrastructure ...

Staff DevOps Engineer

Location
Cambridge, England, United Kingdom
Mentor team members and drive knowledge sharing. Participate in project management and on‐going strategic planning. Qualifications 6+ years of experience in DevOps/SRE environment. 6+ years of experience working with Linux and Windows systems. Strong understanding and knowledge of Kubernetes. Strong understanding and knowledge of cloud infrastructure components ...

Head of Cloud

Location
Norwich, England, United Kingdom
embedded finance solutions, trusted by 90,000+ businesses worldwide. We build modern, scalable platforms that power payments, data, and AI-driven experiences. Join an engineering culture that values ownership, clean and testable code, and continuous improvement. The Opportunity We are seeking a Head of Cloud to serve … infrastructure-as-code mindset. Strategic Thinker: Ability to influence directors and align technical initiatives with commercial objectives. Reliability & Resilience: Expert understanding of SRE principles, building for failure, and high-availability systems. Nice to Have Experience with service mesh and advanced traffic management. Expertise in data persistence (RDS, Aurora, DynamoDB ...

Staff Private Cloud Engineer

Location
Cambridge, England, United Kingdom
lead the design, build and operation of a greenfield multi-tenant private cloud platform based on OpenStack, delivering scalable and reliable infrastructure services for engineering teams. The platform underpins large-scale engineering workloads and is central to the organisation’s infrastructure strategy. The platform is to be built … storage or compute. Hands-on GitOps driven and CI/CD pipelines (Jenkins, ArgoCD, FluxCD, etc). Experience of working in a DevOps/SRE environment. “Nice To Have” Skills and Experience: Experience managing hybrid platforms across multiple data centres and clouds. Hands-On experience with Kubernetes and cloud-native ...

Lead Private Cloud Engineer — OpenStack Platform

Location
Cambridge, England, United Kingdom
experience across compute, networking and storage layers. The role involves leading incident response, improving scalability and performance, and mentoring engineers in a DevOps/SRE culture within a hybrid working framework. #J-18808-Ljbffr ...

Senior Network Engineer

Hiring Organisation
Infoplus Technologies UK Limited
Location
Cambridge, Cambridgeshire, United Kingdom
Employment Type
Full-Time
Salary
£450.00 - £500.00 per day
with enterprise platforms such as ServiceNow . The successful candidate will also contribute towards expanding network automation and developing a strategic roadmap towards Network SRE practices . Key Responsibilities Review and improve the existing network data landscape, data quality and integrations. Develop and enhance NetBox for enterprise network inventory … documentation. Perform network device baselining and configuration validation. Improve network monitoring, data quality and operational reliability. Define KPIs, SLIs and SLOs to support Network SRE practices. Identify opportunities for proactive monitoring, fault prevention and self-healing automation. Work closely with Operations, Security, Architecture, Platform and ITSM teams. Required Technical Skills ...

Lead Engineer

Location
Hemel Hempstead, England, United Kingdom
visits to Hemel Hempstead or on Park****Permanent | Full Time****Competitive salary + bonus and benefits****About the role**Are you an experienced software engineering leader passionate about AI-enabled ways of working? If so, we want to hear from you! Haven is looking for a Lead Engineer … tool selection, systems architecture, and solution definition* Solid working knowledge of Agile, Lean, and OKRs practices and principles* Good working knowledge of DevOps and SRE (CI/CD, logging, monitoring, alerting), plus AWS or Azure* Good working knowledge of relevant security concerns (e.g. OWASP Top 10)* Strong coaching, communication ...

Operations Team Lead (Production & Reliability)

Location
Watford, England, United Kingdom
improving. This is a hands‐on role. You’ll shape process, lead incidents, build the team, and move us from reactive firefighting to proactive reliability engineering. What You’ll Own Production Stability and availability of all live systems Operational readiness for new releases Safe production access and change coordination … Raise the bar on operational discipline You’re responsible for both system performance and team performance. What We’re Looking For Strong experience in SRE, DevOps, Infrastructure, or Production Engineering Prior experience leading technical teams Deep hands‐on incident management experience Strong observability and reliability mindset Calm under ...

Operations Team Lead (Production & Reliability)

Location
Cambridge, England, United Kingdom
improving. This is a hands‐on role. You’ll shape process, lead incidents, build the team, and move us from reactive firefighting to proactive reliability engineering. What You’ll Own Production Stability and availability of all live systems Operational readiness for new releases Safe production access and change coordination … Raise the bar on operational discipline You’re responsible for both system performance and team performance. What We’re Looking For Strong experience in SRE, DevOps, Infrastructure, or Production Engineering Prior experience leading technical teams Deep hands‐on incident management experience Strong observability and reliability mindset Calm under ...