126 to 150 of 279 Site Reliability Engineer Jobs

Hybrid AI-Driven SRE & Reliability Engineer

Location
United Kingdom
bet365 Group is seeking a Site Reliability Engineer to enhance system reliability, observability, and performance. You will treat reliability as a software problem, protecting uptime and driving improvements across critical systems. Responsibilities include building tools, dashboards, and automation, contributing to live incident resolution and post ...

Senior Site Reliability Engineer – AI-Driven Production Resilience (Hybrid)

Location
Greater London, England, United Kingdom
London seeks a Site Reliability Engineer/Senior Engineer to join Corporate Bank Production. The role focuses on embedding SRE principles, improving production stability, and partnering with global teams to manage risks and changes across platforms. You will define and monitor SLOs, enhance observability, and explore ...

Site Reliability Engineer – Platform & Automation

Location
Cambridge, England, United Kingdom
Bango plc is seeking a Site Reliability Engineer to own reliability, performance and continuous improvement of the Bango Platform end-to-end. You’ll merge platform engineering, automation and delivery ownership with proactive incident response and customer impact management, shaping the automation and security posture across ...

Site Reliability Engineer (Edv) - National Security

Location
Cheltenham, England, United Kingdom
Site Reliability Engineer (eDV) - National Security Location: Key National Security hubs Salary: £65,000-£90,000+ depending on experience Clearance: Active eDV required Level: Mid-level through to Senior … spending more time firefighting than improving reliability, this one should resonate. I'm supporting a National Security engineering team looking for an SRE who wants to work closer to the platform, improve how services behave in production and take real ownership of resilience rather than simply reacting to incidents. ...

Senior Site Reliability Engineer

Location
Reading, England, United Kingdom
principles, operational knowledge, security, and automation to work towards platform/service production excellence from an angle of infrastructure, reliability, and security. The SRE team owns the foundation of AI Platform’s Core platform - the services and infrastructure that let us deploy to a multitude of public cloud providers … platform in close partnership with our lead/backend/staff engineers. Who you are (must-haves) 5+ years in infrastructure engineering, DevOps, or SRE, operating large-scale, high-availability production systems using Kubernetes Production Operational experience - a live cluster under real load, not a lab. Fluent with Helm ...

Vice President - Site Reliability Engineering (SRE) - The Core Engineering - Birmingham Birmingham · United Kingdom · Vice President

Location
Birmingham, England, United Kingdom
Vice President - Site Reliability Engineering (SRE) - The Core Engineering - Birmingham location_on Birmingham, West Midlands, England, United Kingdom WHAT WE DO Site Reliability Engineering at Goldman Sachs sits at the intersection of software engineering, systems design, and production excellence. In this VP role, you will help … This role is for software engineers who enjoy solving complex distributed system problems, building tools and platforms that make teams more effective, and championing SRE principles (such as SLOs, error budgets, and blameless post-mortems) across a large engineering organization. Key Responsibilities Partner with engineering leadership to establish service level ...

Senior Site Reliability Engineer: Lead Resilience & Incidents

Location
Watford, England, United Kingdom
Allwyn UK in Watford seeks a Senior/Lead Site Reliability Engineer to drive reliability across the digital estate, ensuring high availability and performance of customer-facing systems during normal operation and peak lottery events. You will own SLOs/SLIs, push automation with Terraform, mentor ...

Senior Site Reliability Engineer - AI-Driven Reliability

Location
Greater London, England, United Kingdom
JPMorgan Chase is seeking a Site Reliability Engineer to join the International Consumer Bank group in the UK. The role focuses on building reliable, scalable digital banking services and leading initiatives to reduce operational toil through automation. You will collaborate with product and platform teams to implement ...

Site Reliability Engineer – Hybrid, Scale & Resilience

Location
Leeds, England, United Kingdom
William Hill PLC, Leeds, is seeking a Site Reliability Engineer to join our SRE team. You’ll partner with engineering and operations to improve platform reliability, scalability and customer experience across betting and gaming products. You’ll lead major incidents, perform postmortems and drive improvements. ...

Site Reliability Engineer - Data Platform & Scale

Location
East Hagbourne, England, United Kingdom
Open Cosmos Ltd is seeking a Site Reliability Engineer to enhance the reliability and scalability of our data platform. You will be responsible for monitoring, troubleshooting, and improving our systems while collaborating with engineering teams to design robust infrastructure. The ideal candidate will have expertise ...

Service Manager, Site Reliability Engineering

Location
Belfast City District, Northern Ireland, United Kingdom
pricing sophistication, telematics, and, more recently, device and identity protection.**Your role in the team**The Service Manager - Site Reliability Engineering (SRE) is responsible for ensuring the reliability, availability, observability, and operational excellence of technology services while maintaining strong alignment with business objectives. This role serves … this vacancy.* A minimum of 4 years of experience supporting, or improving enterprise technology services, infrastructure environments, platform operations, Site Reliability Engineering (SRE), IT Operations, or Service Management disciplines. (Or Equivalent).* A minimum of 2 years leading and mentoring teams within service reliability, availability, performance ...

Senior Systems Reliability Engineer (SRE), Edge

Location
Greater London, England, United Kingdom
Senior Systems Reliability Engineer (SRE), Edge Cloudflare ·United States, London, United Kingdom Job Description About Us At Cloudflare, we are on a mission to help build a better Internet. Today the company runs one of the world’s largest networks that powers millions of websites and other Internet … capabilities. We own a wide portfolio of applications and services, running a tight feedback loop of developer and operator patterns. The ideal SRE candidate has a passionate curiosity about how the Internet fundamentally works and has a strong knowledge of networking, Linux and TLS along with coding ability ...

Lead Site Reliability Engineer - Observability & Resilience

Location
Glasgow, Scotland, United Kingdom
JPMorgan Chase & Co. seeks a Lead Site Reliability Engineer to define the future of reliability for a global firm. You will lead critical resiliency design reviews, break complex problems into actionable work, and mentor engineers across large-scale OpenTelemetry pipelines in hybrid environments. You will guide ...

Manufacturing Site Reliability Engineer — Drive Lasting Improvements

Location
United Kingdom
Whitworths is seeking a Site Reliability Engineer to drive long-term reliability of manufacturing assets. You’ll lead RCA, analyze downtime data, and develop practical, lasting solutions that reduce repeat failures and improve maintenance strategies. You’ll work with Engineering and Operations to implement medium ...

Data Platform Infra Site Reliability Engineer

Location
Greater London, England, United Kingdom
Selection changes the language of the page/content Data Platform Infra Site Reliability Engineer London, England, United Kingdom Software and Services At Apple, we believe that innovation flourishes in an environment where ideas are challenged, collaboration is encouraged, and technology is pushed to its limits. This … inspire innovation in everything we do. Imagine what you could accomplish here! Join Apple and help us make the world a better place.As an SRE on our team, you'll own the reliability, performance, and scale of the distributed storage and data platform systems that power Apple's services. ...

Cloud-Scale Site Reliability Engineer: Automation & Resilience

Location
Birmingham, England, United Kingdom
prominent technology firm in Birmingham is seeking a Site Reliability Engineer tasked with ensuring the reliability, performance, and scalability of critical enterprise systems. This role combines software and systems engineering to enhance automation and support cloud transformations. The ideal candidate will have expertise in Unix, Windows ...

Site Reliability Engineer / Production Support

Hiring Organisation
Hackajob Ltd
Location
London, United Kingdom
Employment Type
Permanent
fastest growing fintech in 2025. The momentum is real. THE OPPORTUNITY Monuments production environment is the heartbeat of a licensed bank, and the SRE role is the single point of ownership when incidents occur. You will directly oversee the offshore Production Support team, run on-call and incident response … eliminate it. Quality-driven - you care about alert quality, observability standards, and reliability patterns that prevent problems at source. WHAT YOU BRING Strong SRE or production support experience with accountability for incident response in a production environment. Deep understanding of observability tools, alerting, logging, and distributed systems debugging. Experience ...

Site Reliability Engineer - Banking & Finance

Location
Greater London, England, United Kingdom
invests heavily in modern platform engineering practices, enabling teams to build reliable, scalable, and highly automated production environments. This opportunity is ideal for a Site Reliability Engineer looking to work at the intersection of software engineering and infrastructure. You'll develop internal platforms, tooling, and automation across … Have: Strong experience programming with Python, Go and/or C++ Strong Linux knowledge and understanding of distributed systems. Experience with monitoring, observability or SRE practices. Experience with CI/CD pipelines, Git and infrastructure automation. Familiarity with Kubernetes and containerised workloads. Strong analytical and troubleshooting skills. Benefits: Build ...

Azure SRE & Reliability Engineer – Multi-Region

Location
Belfast City District, Northern Ireland, United Kingdom
Delinea is seeking an experienced Site Reliability Engineer with deep Azure expertise to maintain the availability, performance, and reliability of our SaaS applications. You will own automation, monitoring, and incident response across a multi-cloud, multi-region environment. You will work with senior engineers and cross … functional teams to improve reliability, observability, and best practices, while contributing to disaster recovery and on-call incident management. #J-18808-Ljbffr ...

Site Reliability Engineer — Production & Incident Response

Location
Greater London, England, United Kingdom
慨正橡扯 is looking for a Site Reliability Engineer to manage incident response and oversee the offshore Production Support team … London. This hybrid position will require you to ensure system reliability and maintain high observability standards. Ideal candidates will bring significant experience in SRE, embracing automation tools to enhance performance. You will be instrumental in protecting client interests and contributing to the company during a vital stage of growth. ...

Site Reliability Engineer

Location
Birmingham, England, United Kingdom
Overview Our client is seeking a high-impact Site Reliability Engineer to join a team responsible for ensuring the reliability, performance, and scalability of critical enterprise systems. This role blends software and systems engineering to drive automation, prevent service-impacting incidents, and support transformative cloud initiatives. … work for any US Employer without sponsorship. Benefits & Extras Work on cutting-edge distributed systems and cloud transformations Solve challenging performance, scalability, and reliability problems Collaborate with teams driving automation and monitoring initiatives Exposure to enterprise-scale network and fault-tolerant architectures High-impact role with visibility across technical ...

Site Reliability Engineer - Hybrid, Observability

Location
United Kingdom
bet365 Group is seeking a Site Reliability Engineer to strengthen the stability … performance and resilience of the systems behind every click and live change across our global product. In this full-time role, you will apply SRE principles, develop automated tooling, and work with OpenTelemetry, Grafana, Splunk, New Relic and Cloudflare edge services to boost reliability and observability. The position supports ...

Lead Site Reliability Engineer

Hiring Organisation
Hackajob Ltd
Location
Glasgow, Lanarkshire, United Kingdom
Employment Type
Permanent
Salary
GBP Annual
defining the future of a globally recognized firm and have a direct and significant effect in a realm tailored for top achievers in site reliability. As a Lead Site Reliability Engineer at JPMorgan Chase within the Infrastructure Platforms team, you hold a leadership role in your ...

Lead Site Reliability Engineer - Chief Technology Office

Hiring Organisation
Hackajob Ltd
Location
Glasgow, Lanarkshire, United Kingdom
Employment Type
Permanent
Salary
GBP Annual
defining the future of a globally recognized firm and have a direct and significant effect in a realm tailored for top achievers in site reliability. As a Lead Site Reliability Engineer at JPMorgan Chase within the Chief Technology Office, youwill solve complex and broad business prob ...

Site Reliability Engineer — Scale, Automation & Observability

Location
Greater London, England, United Kingdom
Apple Inc. in London is seeking a Site Reliability Engineer to help manage and optimize scalable services that power Apple’s media and services. You will own the reliability of distributed systems across data centers, collaborating with software engineers to improve performance and resilience. The role ...