1 to 25 of 853 Permanent Site Reliability Engineering Jobs in England

Director, Site Reliability Engineering

Location
Manchester, England, United Kingdom
Operations, the Director of Site Reliability Engineering will establish Omnicell’s enterprise reliability engineering strategy, build a globally distributed SRE organization, and partner closely with Cloud Platform Engineering, Site Reliability Operations, Cloud Security, Product Engineering, and Enterprise Architecture to ensure reliabilityReliability Engineering Organization Build and scale Omnicell’s global Site Reliability Engineering organization. You Will Recruit, mentor, and develop SRE Managers, Principal Engineers, and senior technical talent. Build a high‐performing engineering organization focused on reliability, resilience, and production engineering. Define engineering ...

Head of Site Reliability Engineering – SRE

Location
West of England, England, United Kingdom
automation, and continuous improvement across our technology landscape. Work closely with Engineering, Infrastructure, Security, and Technology Operations teams to establish and embed modern SRE practices that enable highly reliable, scalable, and resilient services while fostering a culture of shared ownership and continuous learning. Drive adoption of SRE principles (SLOs … operations. Improve incident and problem management maturity. Partner with software and infrastructure engineering teams to embed reliability into the product lifecycle. Establish SRE governance, standards, and operating model. Requirements Proven experience building, leading, and developing Site Reliability Engineering or Production Engineering teams, with ...

AI Platform & Site Reliability Engineering Managing Consultant

Location
Manchester, England, United Kingdom
build and scale secure, reliable and operationally effective AI platforms. You will combine expertise in platform engineering, Site Reliability Engineering (SRE), observability and intelligent operations to help organisations move from isolated AI experimentation to production‐grade, enterprise‐scale AI services. You will work with technology, engineering … observability, platform automation and operational guardrails. Enable reliable and repeatable delivery of AI services from experimentation through to production. Reliability Engineering & SRE: Establish SRE practices including SLIs, SLOs, error budgets, capacity planning, resilience engineering and reliability governance. Help clients shift from reactive operations to data‐driven ...

The Core Engineering - Site Reliability Engineering - Associate - Birmingham

Location
Birmingham, England, United Kingdom
Core Engineering - Site Reliability Engineering - Associate - Birmingham location_on Birmingham, West Midlands, England, United Kingdom What We Do Core Engineering is a global team of more than 2,500 engineers and scientists focused on solving complex, mission-critical problems across the firm. We build … enable faster delivery of new capabilities, reduce downtime, and eliminate repetitive operational work through automation. This role sits within Core Engineering and applies SRE practices to services that support compliance, risk, and other critical business functions. Responsibilities Proactively manage production services by measuring and monitoring availability, capacity, latency ...

Lead Cloud Site Reliability Engineer

Location
West of England, England, United Kingdom
Lead Site Reliability Engineer - Public Cloud Platform Location: Manchester or Bristol Salary: £92,701- £109,043 Working Pattern: Hybrid (2 days in office per week) About this opportunity At Lloyds Banking Group, our purpose is to Help Britain Prosper. As we continue our technology transformation, we're investing … following areas: Designing, building or operating large-scale cloud platforms within Azure, GCP or comparable cloud environments. Site Reliability Engineering (SRE), Platform Engineering, Infrastructure Engineering, Cloud Engineering or Production Operations. Observability and monitoring practices, including metrics, logging and distributed tracing. Incident management, problem management ...

Cloud Operating Model - Managing Consultant

Location
Greater London, England, United Kingdom
build and scale secure, reliable and operationally effective AI platforms. You will combine expertise in platform engineering, Site Reliability Engineering (SRE), observability and intelligent operations to help organisations move from isolated AI experimentation to production-grade, enterprise-scale AI services.You will work with technology, engineering … observability, platform automation and operational guardrails. Enable reliable and repeatable delivery of AI services from experimentation through to production.• Reliability Engineering & SRE: Establish SRE practices including SLIs, SLOs, error budgets, capacity planning, resilience engineering and reliability governance. Help clients shift from reactive operations to data-driven ...

Head Of Infrastructure and Cloud - Internal Applicants Only

Location
Greater London, England, United Kingdom
transition from traditional infrastructure management to a platform-centric, product-led operating model, integrating platform engineering, DevOps, Site Reliability Engineering (SRE), and Network Operations (NOC) to enable scalable, automated, and resilient technology services. To place the interests of customers at the centre of all activities … YBIYRI) model with shared accountability for service delivery and operational outcomes. Establish and integrate Site Reliability Engineering (SRE) practices, defining and managing service‐level objectives (SLOs), error budgets, and proactive reliability engineering across critical services. Ensure end‐to‐end service reliability and resilience, including ...

Head Of Infrastructure and Cloud

Hiring Organisation
Arbuthnot Latham
Location
London, UK
Employment Type
Full-time
transition from traditional infrastructure management to a platform-centric, product-led operating model, integrating platform engineering, DevOps, Site Reliability Engineering (SRE), and Network Operations (NOC) to enable scalable, automated, and resilient technology services. To place the interests of customers at the centre of all activities … YBIYRI) model with shared accountability for service delivery and operational outcomes. Establish and integrate Site Reliability Engineering (SRE) practices, defining and managing service-level objectives (SLOs), error budgets, and proactive reliability engineering across critical services. Ensure end-to-end service reliability and resilience, including ...

Lead Cloud Site Reliability Engineer

Location
Manchester, England, United Kingdom
Summary End Date Monday 21 September 2026 Salary Range £92,701 - £109,060 Flexible Working Options Hybrid Working, Job Share Lead Site Reliability Engineer – Public Cloud Platform Location: Manchester or Bristol Salary: £92,701- £109,043 Working Pattern: Hybrid (2 days in office per week) About this opportunity … Skills and Experience Designing, building or operating large-scale cloud platforms within Azure, GCP or comparable cloud environments. Site Reliability Engineering (SRE), Platform Engineering, Infrastructure Engineering, Cloud Engineering or Production Operations. Observability and monitoring practices, including metrics, logging and distributed tracing. Incident management, problem ...

Lead Cloud Site Reliability Engineer

Location
Halifax, England, United Kingdom
Summary End Date Monday 21 September 2026 Salary Range £92,701 - £109,060 Flexible Working Options Hybrid Working, Job Share Lead Site Reliability Engineer – Public Cloud Platform Location: Manchester or Bristol Salary: £92,701- £109,043 Working Pattern: Hybrid (2 days in office per week) About this opportunity … Skills and Experience Designing, building or operating large-scale cloud platforms within Azure, GCP or comparable cloud environments. Site Reliability Engineering (SRE), Platform Engineering, Infrastructure Engineering, Cloud Engineering or Production Operations. Observability and monitoring practices, including metrics, logging and distributed tracing. Incident management, problem ...

The Core Engineering - Site Reliability Engineering - Associate - Birmingham

Location
Birmingham, England, United Kingdom
What We Do Core Engineering is a global team of more than 2,500 engineers and scientists focused on solving complex, mission-critical problems across the firm. We build and operate platforms and applications that produce metrics, analyze risk, curate financial reports, enable people processes, support budgeting and financial … enable faster delivery of new capabilities, reduce downtime, and eliminate repetitive operational work through automation. This role sits within Core Engineering and applies SRE practices to services that support compliance, risk, and other critical business functions. Responsibilities Proactively manage production services by measuring and monitoring availability, capacity, latency ...

Senior Site Reliability Engineer

Hiring Organisation
VIQU IT
Location
United Kingdom, Morley, West Yorkshire
Employment Type
Permanent
Salary
£60000 - £66000/annum 15% Bonus
Senior Site Reliability Engineer (AWS CDK) Remote UK | Permanent, Full Time £60,000 - £66,000 + 15% bonus VIQU has partnered with a leading UK technology organisation investing significantly in its cloud platform and engineering capability. They are looking for a Senior Site Reliability Engineer … incident management. Improve cloud performance, resilience and operational efficiency. Work closely with engineering, architecture and delivery teams. Mentor and support engineers, promoting strong SRE and DevOps practices. Contribute to technical direction, architecture discussions and continuous improvement. Key Requirements of the Senior Site Reliability Engineer Strong commercial ...

Site Reliability Engineer

Location
Manchester, England, United Kingdom
Security Office, our Security Data & AI Lab is building intelligent, AI-driven security capabilities that help keep our customers and colleagues safe! As a Site Reliability Engineer, you'll join a collaborative team focused on the resilient operation of cloud-hosted applications and services. You'll work alongside … DevOps and Site Reliability Engineering practices, including continuous improvement, operational excellence and collaborative delivery. The SBO Skills Library identifies DevOps and SRE & Service Engineering as core engineering skill areas. Familiarity with Infrastructure as Code concepts and tools such as Terraform, alongside source control and software ...

Site Reliability Engineer

Location
Swindon, England, United Kingdom
kaikille hakijoillemme yhtäläiset mahdollisuudet sukupuolesta ja sukupuolen ilmaisusta, vammaisuudesta, alkuperästä, uskonnollisesta vakaumuksesta ja seksuaalisesta suuntautumisesta tai muista kriteereistä riippumatta.**Site Reliability Engineer (SRE)****Location: Swindon (Hybrid)**Join Edenred and help shape resilient, future-ready infrastructure that supports millions of users worldwide.At Edenred, we’re looking for a Site Reliability Engineer (SRE) to join our Infrastructure Engineering team. This is an exciting opportunity for someone who is passionate about reliability, automation, operational excellence, and continuous improvement.**About the role**As a Site Reliability Engineer, you will play a key role in ensuring ...

Systems Engineering Manager, Site Reliability Engineering, ML Compute

Location
City of Westminster, England, United Kingdom
. Track record of mentoring technical leads. Proven success leading and influencing multiple technical teams. About the job Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed, fault-tolerant systems. SRE ensures that Google's services—both … internally critical and our externally-visible systems—have reliability, uptime appropriate to users' needs and a fast rate of improvement. Additionally SRE’s will keep an ever-watchful eye on our systems capacity and performance. Much of our software development focuses on optimizing existing systems, building infrastructure and eliminating ...

Sr. Manager, Site Reliability

Location
Manchester, England, United Kingdom
7. The Site Reliability Engineering function is the reliability engine of that organization, and this role is the first senior SRE hire — the person who will design the practice, set the standards, and then run the plays themselves until the team is large enough to delegate. … 7. The Site Reliability Engineering function is the reliability engine of that organization, and this role is the first senior SRE hire — the person who will design the practice, set the standards, and then run the plays themselves until the team is large enough to delegate. ...

Site Reliability Engineer

Location
Cambridge, England, United Kingdom
learn more, visit http://www.darktrace.com. **Job D****escription****:**## **About the Role**We’re looking for a **Site Reliability Engineer (SRE)** to bring deep expertise in a key reliability domain and help shape the future of our platform reliability strategy.SRE sits at the heart … your area of specialism**, working across teams to embed best practices, solve complex reliability challenges, and improve system resilience at scale.Unlike a generalist SRE, this role focuses on a **core domain of expertise**—such as **observability, performance engineering, data infrastructure reliability, security-focused SRE, or network reliability ...

Site Reliability Engineer - SC Cleared

Location
Gloucester, England, United Kingdom
SITE RELIABILITY ENGINEER- SC CLEARED SITE RELIABILITY ENGINEER Permanent opportunity for a Site Reliability Engineer with SC Clearance. Salary up to £65,000 DOE Hybrid opportunity with Gloucester based offices WHO ARE WE? We're hiring for Site Reliability Engineers to join … mission-critical projects, shaping solutions that make a real difference. Due to the sensitive nature of the work, active SC Clearance is required. THE SITE RELIABILITY ENGINEER Active SC Clearance and DV Clearance eligibility Gloucester Based or ability to travel to Gloucester. Experience as in a Site ...

Vice President - Site Reliability Engineering (SRE) - The Core Engineering - Birmingham Birmingham · United Kingdom · Vice President

Location
Birmingham, England, United Kingdom
Vice President - Site Reliability Engineering (SRE) - The Core Engineering - Birmingham location_on Birmingham, West Midlands, England, United Kingdom WHAT WE DO Site Reliability Engineering at Goldman Sachs sits at the intersection of software engineering, systems design, and production excellence. In this … This role is for software engineers who enjoy solving complex distributed system problems, building tools and platforms that make teams more effective, and championing SRE principles (such as SLOs, error budgets, and blameless post-mortems) across a large engineering organization. Key Responsibilities Partner with engineering leadership to establish ...

Software Engineer III, Site Reliability Engineering, GCE AI

Location
City of Westminster, England, United Kingdom
Science or Engineering. 2 years of experience designing, analyzing, and troubleshooting large-scale distributed systems. About the job Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed, fault-tolerant systems. SRE ensures that Google Cloud's services—both … internally critical and our externally-visible systems—have reliability, uptime appropriate to customer's needs and a fast rate of improvement. Additionally SRE’s will keep an ever-watchful eye on our systems capacity and performance. Much of our software development focuses on optimizing existing systems, building infrastructure ...

Site Reliability Engineer

Location
Swindon, England, United Kingdom
opportunities regardless of their gender and gender expression, disability, origin, religious belief and sexual orientation or any other criteria.**Site Reliability Engineer (SRE)****Location: Swindon (Hybrid)**Join Edenred and help shape resilient, future-ready infrastructure that supports millions of users worldwide.At Edenred, we’re looking for a Site Reliability Engineer (SRE) to join our Infrastructure Engineering team. This is an exciting opportunity for someone who is passionate about reliability, automation, operational excellence, and continuous improvement.**About the role**As a Site Reliability Engineer, you will play a key role in ensuring ...

Senior SRE Engineer

Hiring Organisation
SF Partners
Location
Birmingham, West Midlands (County), United Kingdom
Employment Type
Permanent
Salary
£100000 - £110000/annum great training and progression opp
some of the UK's largest and most complex enterprise platforms. Working within a highly skilled, multi-disciplinary engineering team, this Senior SRE Engineer will take technical ownership of observability and reliability across large-scale AWS environments, helping engineering teams improve platform performance, resilience and availability. … mindset, using automation, Infrastructure as Code and cloud-native tooling to improve reliability and reduce operational overhead. This role would suit a Senior SRE DevOps or Cloud Engineer who combines hands-on AWS expertise with genuine Dynatrace implementation experience, rather than simply having used Dynatrace as an end user. ...

Engineer - Site Reliability

Location
Greater London, England, United Kingdom
early‐session coverage from London that ensures continuous, high‐availability operations across Cboe's real‐time low‐latency trading platforms. The London‐based SRE provides technical support to Cboe Trade Desk and Operations Support Center staff across time zones, and works closely with Software Engineering, Systems Engineering … spoken — is required. This role demands clear, precise, and unambiguous communication at all times. As the operational bridge between Cboe's European and APAC SRE teams and its US‐based leadership, the ability to communicate with clarity across time zones, cultures, and technical disciplines is fundamental to the success ...

Vice President - Site Reliability Engineering (SRE) - The Core Engineering - Birmingham

Location
West Midlands, England, United Kingdom
This role is for software engineers who enjoy solving complex distributed system problems, building tools and platforms that make teams more effective, and championing SRE principles (such as SLOs, error budgets, and blameless post-mortems) across a large engineering organization. Key Responsibilities Partner with engineering leadership to establish … understand how individual components interact under load. Strong interpersonal skills to collaborate with product developers, influence architectural decisions, prioritize toil reduction, and drive SRE adoption without direct authority. Ability to translate complex technical issues into clear, actionable insights for both technical and non-technical stakeholders. Highly motivated, pro-active ...

Vice President, Site Reliability Engineering

Hiring Organisation
The Bank of New York Mellon
Location
London, UK
Employment Type
Full-time
seeking a future team member for the role of Vice President - Site Reliability Engineer to join our team. This role is located in London. Role SummaryBNY is seeking a Vice President - Site Reliability Engineer to design, build, deploy, and scale resilient, automated, and centrally managed engineering … driven automation, and modern software delivery practices. Experience supporting distributed systems, cloud-native platforms, or container-based architectures. Knowledge of Agile, DevOps, and SRE operating models, including continuous improvement and blameless post-incident practices. Ability to influence engineering standards and drive adoption of common tooling and automation patterns across ...