1 to 25 of 1,177 Site Reliability Engineering Jobs in the UK

Director, Site Reliability Engineering

Location
Manchester, England, United Kingdom
Operations, the Director of Site Reliability Engineering will establish Omnicell’s enterprise reliability engineering strategy, build a globally distributed SRE organization, and partner closely with Cloud Platform Engineering, Site Reliability Operations, Cloud Security, Product Engineering, and Enterprise Architecture to ensure reliabilityReliability Engineering Organization Build and scale Omnicell’s global Site Reliability Engineering organization. You Will Recruit, mentor, and develop SRE Managers, Principal Engineers, and senior technical talent. Build a high‐performing engineering organization focused on reliability, resilience, and production engineering. Define engineering ...

Head of Site Reliability Engineering – SRE

Location
West of England, England, United Kingdom
automation, and continuous improvement across our technology landscape. Work closely with Engineering, Infrastructure, Security, and Technology Operations teams to establish and embed modern SRE practices that enable highly reliable, scalable, and resilient services while fostering a culture of shared ownership and continuous learning. Drive adoption of SRE principles (SLOs … operations. Improve incident and problem management maturity. Partner with software and infrastructure engineering teams to embed reliability into the product lifecycle. Establish SRE governance, standards, and operating model. Requirements Proven experience building, leading, and developing Site Reliability Engineering or Production Engineering teams, with ...

AI Platform & Site Reliability Engineering Managing Consultant

Location
Manchester, England, United Kingdom
build and scale secure, reliable and operationally effective AI platforms. You will combine expertise in platform engineering, Site Reliability Engineering (SRE), observability and intelligent operations to help organisations move from isolated AI experimentation to production‐grade, enterprise‐scale AI services. You will work with technology, engineering … observability, platform automation and operational guardrails. Enable reliable and repeatable delivery of AI services from experimentation through to production. Reliability Engineering & SRE: Establish SRE practices including SLIs, SLOs, error budgets, capacity planning, resilience engineering and reliability governance. Help clients shift from reactive operations to data‐driven ...

AI Platform & Site Reliability Engineering Managing Consultant

Location
United Kingdom
build and scale secure, reliable and operationally effective AI platforms. You will combine expertise in platform engineering, Site Reliability Engineering (SRE), observability and intelligent operations to help organisations move from isolated AI experimentation to production-grade, enterprise-scale AI services. You will work with technology, engineering … observability, platform automation and operational guardrails. Enable reliable and repeatable delivery of AI services from experimentation through to production. Reliability Engineering & SRE: Establish SRE practices including SLIs, SLOs, error budgets, capacity planning, resilience engineering and reliability governance. Help clients shift from reactive operations to data-driven ...

The Core Engineering - Site Reliability Engineering - Associate - Birmingham

Location
Birmingham, England, United Kingdom
Core Engineering - Site Reliability Engineering - Associate - Birmingham location_on Birmingham, West Midlands, England, United Kingdom What We Do Core Engineering is a global team of more than 2,500 engineers and scientists focused on solving complex, mission-critical problems across the firm. We build … enable faster delivery of new capabilities, reduce downtime, and eliminate repetitive operational work through automation. This role sits within Core Engineering and applies SRE practices to services that support compliance, risk, and other critical business functions. Responsibilities Proactively manage production services by measuring and monitoring availability, capacity, latency ...

Systems Engineering Manager, Site Reliability Engineering, ML Compute

Location
Greater London, England, United Kingdom
. Track record of mentoring technical leads. Proven success leading and influencing multiple technical teams. About the job Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed, fault-tolerant systems. SRE ensures that Google's services—both … internally critical and our externally-visible systems—have reliability, uptime appropriate to users' needs and a fast rate of improvement. Additionally SRE’s will keep an ever-watchful eye on our systems capacity and performance. Much of our software development focuses on optimizing existing systems, building infrastructure and eliminating ...

Senior Manager, Core Infrastructure Engineering

Hiring Organisation
Oracle Corporation
Location
United Kingdom, UK
Employment Type
Full-time
– OCI Core InfrastructureCore Infrastructure Engineering within Oracle Cloud Infrastructure (OCI) is seeking an experienced Senior Manager, Site Reliability Engineering (SRE) to lead a team responsible for the reliability, availability, performance, and operational excellence of critical database and storage services that form the backbone … Reliability Engineering, cloud infrastructure, distributed systems, storage, database, or related infrastructure technologies. Proven experience managing and developing experienced engineers, ideally within SRE, infrastructure, platform, or cloud engineering teams. Strong technical foundation in operating large-scale, highly distributed production systems on cloud platforms such ...

Lead Cloud Site Reliability Engineer

Location
West of England, England, United Kingdom
Lead Site Reliability Engineer - Public Cloud Platform Location: Manchester or Bristol Salary: £92,701- £109,043 Working Pattern: Hybrid (2 days in office per week) About this opportunity At Lloyds Banking Group, our purpose is to Help Britain Prosper. As we continue our technology transformation, we're investing … following areas: Designing, building or operating large-scale cloud platforms within Azure, GCP or comparable cloud environments. Site Reliability Engineering (SRE), Platform Engineering, Infrastructure Engineering, Cloud Engineering or Production Operations. Observability and monitoring practices, including metrics, logging and distributed tracing. Incident management, problem management ...

Cloud Operating Model - Managing Consultant

Location
Greater London, England, United Kingdom
build and scale secure, reliable and operationally effective AI platforms. You will combine expertise in platform engineering, Site Reliability Engineering (SRE), observability and intelligent operations to help organisations move from isolated AI experimentation to production-grade, enterprise-scale AI services.You will work with technology, engineering … observability, platform automation and operational guardrails. Enable reliable and repeatable delivery of AI services from experimentation through to production.• Reliability Engineering & SRE: Establish SRE practices including SLIs, SLOs, error budgets, capacity planning, resilience engineering and reliability governance. Help clients shift from reactive operations to data-driven ...

Head Of Infrastructure and Cloud - Internal Applicants Only

Location
Greater London, England, United Kingdom
transition from traditional infrastructure management to a platform-centric, product-led operating model, integrating platform engineering, DevOps, Site Reliability Engineering (SRE), and Network Operations (NOC) to enable scalable, automated, and resilient technology services. To place the interests of customers at the centre of all activities … YBIYRI) model with shared accountability for service delivery and operational outcomes. Establish and integrate Site Reliability Engineering (SRE) practices, defining and managing service‐level objectives (SLOs), error budgets, and proactive reliability engineering across critical services. Ensure end‐to‐end service reliability and resilience, including ...

Head Of Infrastructure and Cloud

Hiring Organisation
Arbuthnot Latham
Location
London, UK
Employment Type
Full-time
transition from traditional infrastructure management to a platform-centric, product-led operating model, integrating platform engineering, DevOps, Site Reliability Engineering (SRE), and Network Operations (NOC) to enable scalable, automated, and resilient technology services. To place the interests of customers at the centre of all activities … YBIYRI) model with shared accountability for service delivery and operational outcomes. Establish and integrate Site Reliability Engineering (SRE) practices, defining and managing service-level objectives (SLOs), error budgets, and proactive reliability engineering across critical services. Ensure end-to-end service reliability and resilience, including ...

Lead Cloud Site Reliability Engineer

Location
Manchester, England, United Kingdom
Summary End Date Monday 21 September 2026 Salary Range £92,701 - £109,060 Flexible Working Options Hybrid Working, Job Share Lead Site Reliability Engineer – Public Cloud Platform Location: Manchester or Bristol Salary: £92,701- £109,043 Working Pattern: Hybrid (2 days in office per week) About this opportunity … Skills and Experience Designing, building or operating large-scale cloud platforms within Azure, GCP or comparable cloud environments. Site Reliability Engineering (SRE), Platform Engineering, Infrastructure Engineering, Cloud Engineering or Production Operations. Observability and monitoring practices, including metrics, logging and distributed tracing. Incident management, problem ...

Lead Cloud Site Reliability Engineer

Location
Halifax, England, United Kingdom
Summary End Date Monday 21 September 2026 Salary Range £92,701 - £109,060 Flexible Working Options Hybrid Working, Job Share Lead Site Reliability Engineer – Public Cloud Platform Location: Manchester or Bristol Salary: £92,701- £109,043 Working Pattern: Hybrid (2 days in office per week) About this opportunity … Skills and Experience Designing, building or operating large-scale cloud platforms within Azure, GCP or comparable cloud environments. Site Reliability Engineering (SRE), Platform Engineering, Infrastructure Engineering, Cloud Engineering or Production Operations. Observability and monitoring practices, including metrics, logging and distributed tracing. Incident management, problem ...

The Core Engineering - Site Reliability Engineering - Associate - Birmingham

Location
Birmingham, England, United Kingdom
What We Do Core Engineering is a global team of more than 2,500 engineers and scientists focused on solving complex, mission-critical problems across the firm. We build and operate platforms and applications that produce metrics, analyze risk, curate financial reports, enable people processes, support budgeting and financial … enable faster delivery of new capabilities, reduce downtime, and eliminate repetitive operational work through automation. This role sits within Core Engineering and applies SRE practices to services that support compliance, risk, and other critical business functions. Responsibilities Proactively manage production services by measuring and monitoring availability, capacity, latency ...

Senior Site Reliability Engineer

Hiring Organisation
Malvern Panalytical
Location
United Kingdom, UK
Employment Type
Full-time
delivered rapidly without compromising stability, security, or customer experience. As a senior technical specialist, you will champion Site Reliability Engineering (SRE) best practices, improve platform resilience, and contribute to the continuous evolution of our cloud services and operational excellence. This is an exciting opportunity for an experienced … SRE professional who enjoys solving complex technical challenges, improving system reliability, and influencing the future of cloud platform operations within a global technology organization. What you'll be doing: Monitor the performance, availability, reliability, and overall health of Azure cloud applications and services using telemetry, metrics, logging ...

Site Reliability Engineer (SRE) – Cloud Engineer

Location
Glasgow, Scotland, United Kingdom
Glasgow Contract Our client is undergoing a major cloud transformation programme and are seeking an experienced Site Reliability Engineer (SRE)/Cloud Engineer to join a high-performing engineering team. You will play a key role in enhancing cloud reliability, scalability, automation and operational excellence across … potential extension) Start Date: Immediate Our client is undergoing a major cloud transformation programme and are seeking an experienced Site Reliability Engineer (SRE)/Cloud Engineer to join a high-performing engineering team. You will play a key role in enhancing cloud reliability, scalability, automation ...

Senior Site Reliability Engineer

Hiring Organisation
VIQU IT
Location
United Kingdom, Morley, West Yorkshire
Employment Type
Permanent
Salary
£60000 - £66000/annum 15% Bonus
Senior Site Reliability Engineer (AWS CDK) Remote UK | Permanent, Full Time £60,000 - £66,000 + 15% bonus VIQU has partnered with a leading UK technology organisation investing significantly in its cloud platform and engineering capability. They are looking for a Senior Site Reliability Engineer … incident management. Improve cloud performance, resilience and operational efficiency. Work closely with engineering, architecture and delivery teams. Mentor and support engineers, promoting strong SRE and DevOps practices. Contribute to technical direction, architecture discussions and continuous improvement. Key Requirements of the Senior Site Reliability Engineer Strong commercial ...

Site Reliability Engineer

Location
Manchester, England, United Kingdom
Security Office, our Security Data & AI Lab is building intelligent, AI-driven security capabilities that help keep our customers and colleagues safe! As a Site Reliability Engineer, you'll join a collaborative team focused on the resilient operation of cloud-hosted applications and services. You'll work alongside … DevOps and Site Reliability Engineering practices, including continuous improvement, operational excellence and collaborative delivery. The SBO Skills Library identifies DevOps and SRE & Service Engineering as core engineering skill areas. Familiarity with Infrastructure as Code concepts and tools such as Terraform, alongside source control and software ...

Lead SRE- Azure & GCP

Location
Glasgow, Scotland, United Kingdom
have a Lead Site Reliability Engineer (SRE) opportunity within our Google Cloud Site Reliability Engineering team. As a Lead Site Reliability Engineer at JPMorgan Chase within the Infrastructure Platform - Cloud Foundational Services SRE organization, you will join our Google Cloud Site Reliability Engineering team operating within a global follow-the-sun support model. Job Responsibilities: Lead and Implement SRE frameworks to support global google cloud environments and ensure the highest level of SLOs through operational excellence Mastery of application, data, infrastructure, and Agentic AI disciplines Keen understanding of financial control ...

Site Reliability Engineer

Location
Swindon, England, United Kingdom
kaikille hakijoillemme yhtäläiset mahdollisuudet sukupuolesta ja sukupuolen ilmaisusta, vammaisuudesta, alkuperästä, uskonnollisesta vakaumuksesta ja seksuaalisesta suuntautumisesta tai muista kriteereistä riippumatta.**Site Reliability Engineer (SRE)****Location: Swindon (Hybrid)**Join Edenred and help shape resilient, future-ready infrastructure that supports millions of users worldwide.At Edenred, we’re looking for a Site Reliability Engineer (SRE) to join our Infrastructure Engineering team. This is an exciting opportunity for someone who is passionate about reliability, automation, operational excellence, and continuous improvement.**About the role**As a Site Reliability Engineer, you will play a key role in ensuring ...

Sr. Manager, Site Reliability

Location
Manchester, England, United Kingdom
7. The Site Reliability Engineering function is the reliability engine of that organization, and this role is the first senior SRE hire — the person who will design the practice, set the standards, and then run the plays themselves until the team is large enough to delegate. … 7. The Site Reliability Engineering function is the reliability engine of that organization, and this role is the first senior SRE hire — the person who will design the practice, set the standards, and then run the plays themselves until the team is large enough to delegate. ...

Site Reliability Engineer

Location
Cambridge, England, United Kingdom
learn more, visit http://www.darktrace.com. **Job D****escription****:**## **About the Role**We’re looking for a **Site Reliability Engineer (SRE)** to bring deep expertise in a key reliability domain and help shape the future of our platform reliability strategy.SRE sits at the heart … your area of specialism**, working across teams to embed best practices, solve complex reliability challenges, and improve system resilience at scale.Unlike a generalist SRE, this role focuses on a **core domain of expertise**—such as **observability, performance engineering, data infrastructure reliability, security-focused SRE, or network reliability ...

Site Reliability Engineer - SC Cleared

Location
Gloucester, England, United Kingdom
SITE RELIABILITY ENGINEER- SC CLEARED SITE RELIABILITY ENGINEER Permanent opportunity for a Site Reliability Engineer with SC Clearance. Salary up to £65,000 DOE Hybrid opportunity with Gloucester based offices WHO ARE WE? We're hiring for Site Reliability Engineers to join … mission-critical projects, shaping solutions that make a real difference. Due to the sensitive nature of the work, active SC Clearance is required. THE SITE RELIABILITY ENGINEER Active SC Clearance and DV Clearance eligibility Gloucester Based or ability to travel to Gloucester. Experience as in a Site ...

Vice President - Site Reliability Engineering (SRE) - The Core Engineering - Birmingham Birmingham · United Kingdom · Vice President

Location
Birmingham, England, United Kingdom
Vice President - Site Reliability Engineering (SRE) - The Core Engineering - Birmingham location_on Birmingham, West Midlands, England, United Kingdom WHAT WE DO Site Reliability Engineering at Goldman Sachs sits at the intersection of software engineering, systems design, and production excellence. In this … This role is for software engineers who enjoy solving complex distributed system problems, building tools and platforms that make teams more effective, and championing SRE principles (such as SLOs, error budgets, and blameless post-mortems) across a large engineering organization. Key Responsibilities Partner with engineering leadership to establish ...

Software Engineer III, Site Reliability Engineering, GCE AI

Location
Greater London, England, United Kingdom
Science or Engineering. 2 years of experience designing, analyzing, and troubleshooting large-scale distributed systems. About the job Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed, fault-tolerant systems. SRE ensures that Google Cloud's services—both … internally critical and our externally-visible systems—have reliability, uptime appropriate to customer's needs and a fast rate of improvement. Additionally SRE’s will keep an ever-watchful eye on our systems capacity and performance. Much of our software development focuses on optimizing existing systems, building infrastructure ...