26 to 50 of 1,094 Site Reliability Engineering Jobs in the UK

Software Engineer III, Site Reliability Engineering, GCE AI

Location
Greater London, England, United Kingdom
Science or Engineering. 2 years of experience designing, analyzing, and troubleshooting large-scale distributed systems. About the job Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed, fault-tolerant systems. SRE ensures that Google Cloud's services—both … internally critical and our externally-visible systems—have reliability, uptime appropriate to customer's needs and a fast rate of improvement. Additionally SRE’s will keep an ever-watchful eye on our systems capacity and performance. Much of our software development focuses on optimizing existing systems, building infrastructure ...

Tech Lead, Site Reliability Engineering

Hiring Organisation
London Stock Exchange Group
Location
Nottingham, United Kingdom
seeking an accomplished and forward-thinking technical leader to join our Risk Intelligence organisation as the Tech Lead – Site Reliability Engineering (SRE) for the Risk Screening Product Line. Based in Nottingham, UK, this role will lead the reliability engineering function for critical Screening applications supporting … observability, ITSM, and incident communication.Ensure DR readiness, runbook quality, and resilience patterns are consistently applied across services.People LeadershipLead and mentor a team of SRE engineers, fostering a culture of ownership, learning, and engineering excellence.Drive career development, performance management, and technical capability growth across the team.Collaborate with HR and Talent ...

Site Reliability Engineer

Location
Swindon, England, United Kingdom
opportunities regardless of their gender and gender expression, disability, origin, religious belief and sexual orientation or any other criteria.**Site Reliability Engineer (SRE)****Location: Swindon (Hybrid)**Join Edenred and help shape resilient, future-ready infrastructure that supports millions of users worldwide.At Edenred, we’re looking for a Site Reliability Engineer (SRE) to join our Infrastructure Engineering team. This is an exciting opportunity for someone who is passionate about reliability, automation, operational excellence, and continuous improvement.**About the role**As a Site Reliability Engineer, you will play a key role in ensuring ...

Senior SRE Engineer

Hiring Organisation
SF Partners
Location
Birmingham, West Midlands (County), United Kingdom
Employment Type
Permanent
Salary
£100000 - £110000/annum great training and progression opp
some of the UK's largest and most complex enterprise platforms. Working within a highly skilled, multi-disciplinary engineering team, this Senior SRE Engineer will take technical ownership of observability and reliability across large-scale AWS environments, helping engineering teams improve platform performance, resilience and availability. … mindset, using automation, Infrastructure as Code and cloud-native tooling to improve reliability and reduce operational overhead. This role would suit a Senior SRE DevOps or Cloud Engineer who combines hands-on AWS expertise with genuine Dynatrace implementation experience, rather than simply having used Dynatrace as an end user. ...

Engineer - Site Reliability

Location
Greater London, England, United Kingdom
early‐session coverage from London that ensures continuous, high‐availability operations across Cboe's real‐time low‐latency trading platforms. The London‐based SRE provides technical support to Cboe Trade Desk and Operations Support Center staff across time zones, and works closely with Software Engineering, Systems Engineering … spoken — is required. This role demands clear, precise, and unambiguous communication at all times. As the operational bridge between Cboe's European and APAC SRE teams and its US‐based leadership, the ability to communicate with clarity across time zones, cultures, and technical disciplines is fundamental to the success ...

Vice President - Site Reliability Engineering (SRE) - The Core Engineering - Birmingham

Location
West Midlands, England, United Kingdom
This role is for software engineers who enjoy solving complex distributed system problems, building tools and platforms that make teams more effective, and championing SRE principles (such as SLOs, error budgets, and blameless post-mortems) across a large engineering organization. Key Responsibilities Partner with engineering leadership to establish … understand how individual components interact under load. Strong interpersonal skills to collaborate with product developers, influence architectural decisions, prioritize toil reduction, and drive SRE adoption without direct authority. Ability to translate complex technical issues into clear, actionable insights for both technical and non-technical stakeholders. Highly motivated, pro-active ...

Vice President, Site Reliability Engineering

Hiring Organisation
Hackajob Ltd
Location
South West London, London, United Kingdom
Employment Type
Permanent
Were seeking a future team member for the role of Vice President - Site Reliability Engineer to join our team. This role is located in London. Role Summary BNY is seeking a Vice President - Site Reliability Engineer to design, build, deploy, and scale resilient, automated, and centrally … driven automation, and modern software delivery practices. Experience supporting distributed systems, cloud-native platforms, or container-based architectures. Knowledge of Agile, DevOps, and SRE operating models, including continuous improvement and blameless post-incident practices. Ability to influence engineering standards and drive adoption of common tooling and automation patterns across ...

Vice President, Site Reliability Engineering

Hiring Organisation
The Bank of New York Mellon
Location
London, UK
Employment Type
Full-time
seeking a future team member for the role of Vice President - Site Reliability Engineer to join our team. This role is located in London. Role SummaryBNY is seeking a Vice President - Site Reliability Engineer to design, build, deploy, and scale resilient, automated, and centrally managed engineering … driven automation, and modern software delivery practices. Experience supporting distributed systems, cloud-native platforms, or container-based architectures. Knowledge of Agile, DevOps, and SRE operating models, including continuous improvement and blameless post-incident practices. Ability to influence engineering standards and drive adoption of common tooling and automation patterns across ...

Director of Site Reliability Engineering

Hiring Organisation
Hackajob Ltd
Location
Glasgow, Lanarkshire, Scotland, United Kingdom
Employment Type
Permanent
your profession to new heights by contributing to revolutionary projects. You've discovered the perfect environment to have a major impact. As a Principal Site Reliability Engineer at JPMorgan Chase within the Corporate Technology and Enterprise Technology Team, you draw upon your advanced knowledge to identify new opportunities … lifecycle of software development for the firm. You will have the opportunity to manage, design, and implement infrastructure components to improve reliability and ensure operational efficiency. Job Responsibilities Identifies and solves problems of high complexity and drives improvements as outcomes Uses enterprise-authorized AI capabilities within the work environment ...

Azure Infrastructure / SRE Engineer

Hiring Organisation
Teksystems
Location
Central London, London, United Kingdom
Employment Type
Contract
Azure tenant/infrastructure. Required Skills The successful candidate will be a student of the Google Site Reliability Engineering (SRE) philosophy as applied to managing large-scale cloud infrastructure, possess skills and xp within one or more of the following areas, and demonstrate a willingness to learn … additional skills via certification and/or on-the-job learning where required. Programming, Software & Network Principles xp with SRE and Azure DevOps Ability to script (Bash/PowerShell, Azure CLI), code (Python, C#, Java), query (SQL, Kusto query language) coupled with xp with software versioning control systems (e.g., GitHub ...

Systems Engineering Manager, Site Reliability Engineering, ML Compute

Location
Westminster, West End, United Kingdom
Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed, fault-tolerant systems. SRE ensures that Google's services both our internally critical and our externally-visible systems have reliability, uptime appropriate to users' needs and a fast … rate of improvement. Additionally SRE's will keep an ever-watchful eye on our systems capacity and performance. Much of our software development focuses on optimizing existing systems, building infrastructure and eliminating work through automation. On the SRE team, you'll have the opportunity to manage the complex challenges ...

Sr. Manager, Site Reliability

Location
Manchester, England, United Kingdom
7. The Site Reliability Engineering function is the reliability engine of that organization, and this role is the first senior SRE hire — the person who will design the practice, set the standards, and then run the plays themselves until the team is large enough to delegate. … investment is prioritized against feature velocity. You will make those calls in partnership with the VP of Global Cloud Operations and an Engineer III SRE you will coach and grow. The environment is hybrid. Some of our products are still hardware in hospitals communicating with cloud services; others are fully ...

Senior DevSecOps Engineer

Location
Greater London, England, United Kingdom
repeatable, and secure delivery of autonomy and mission software • Embed security throughout the software development lifecycle, integrating security controls, testing, evidence, and assurance into engineering workflows • Apply UK MOD Secure by Design principles and help engineering teams meet cyber security and technical assurance responsibilities • Collaborate with software, autonomy … controls including static analysis, dependency scanning, container scanning, secrets detection, software composition analysis, vulnerability management, and policy enforcement • Champion a DevSecOps culture emphasizing security, reliability, deployability, and operational performance ownership Requirements BS or MS in Computer Science, Software Engineering, Cyber Security, Electrical Engineering, Systems Engineering ...

Systems Engineering Manager, Site Reliability Engineering, ML Compute

Hiring Organisation
Hackajob Ltd
Location
South West London, London, United Kingdom
Employment Type
Permanent
Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed, fault-tolerant systems. SRE ensures that Google's servicesboth our internally critical and our externally-visible systemshave reliability, uptime appropriate to users' needs and a fast rate … systems capacity and performance. Much of our software development focuses on optimizing existing systems, building infrastructure and eliminating work through automation. On the SRE team, youll have the opportunity to manage the complex challenges of scale which are unique to Google, while using your expertise in coding, algorithms, complexity analysis ...

Senior Specialist Engineer (Specialist Site Reliability Engineer SRE)

Hiring Organisation
UK Health Security Agency
Location
Birmingham, Chilton, Leeds, Liverpool, London, Porton, E14 4PU, United Kingdom
Salary
£41983.00 to £52113.00
summary An SRE engineer will apply engineering principles to remediate infrastructure and operational problems. The primary focus will be on automation and CI/CD; ensuring our services run reliably, are scalable, and perform optimally in production environments. The role will monitor and manage these aspects while taking responsibility … operational service improvements and performance improvements to meet and exceed SLOs (Service Level Objectives). Main duties of the job Working with the HPC & SRE Team to: Ensure services are stable, scalable, performant and automated Respond to incidents, troubleshooting issues, and restoring services as quickly as possible Prioritise operational service ...

Technical Lead - Site Reliability Engineering

Location
Greater London, England, United Kingdom
Reliability Engineering capabilities to strengthen reliability, observability, security, and operational excellence across our Markets and Risk Intelligence division.As a **Technical Lead SRE**, you will be a senior hands‐on technical person help shape the foundations of reliability across both new and existing platforms. You will collaborate … person who is passionate about reliability engineering and who bring a continuous improvement approach to everything they do!Lead the establishment of SRE foundations for new projects building environments, monitoring, alerting, and ensuring operational readiness from day one.Collaborate with Architecture and Engineering teams to embed reliability ...

Site Reliability Engineer

Location
Cambridge, England, United Kingdom
world’s largest content providers, including Amazon, Google and Microsoft trust Bango technology to reach subscribers everywhere. Bango, where people subscribe. Role As a Site Reliability Engineer at Bango, you own the reliability, performance and continuous improvement of the Bango Platform end-to-end — from the infrastructure … constructive, detailed feedback. Call out areas where Bango can improve service or reduce cost. Essentials 3+ years' experience in a Cloud, Platform, DevOps or SRE role in a commercial environment. Production experience with a major cloud provider (AWS, Azure or GCP). Strong Linux administration and troubleshooting (process management, memory ...

Site Reliability Engineer

Hiring Organisation
Twinstream Limited
Location
Cheltenham, Gloucestershire, South West, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£95,000
Site Reliability Engineer | Up to £95,000 DOE | Fully Remote Initially | Cheltenham | Future Hybrid Working Build resilient systems. Solve complex challenges. Make a real impact. Are you a Site Reliability Engineer or DevOps professional who enjoys getting stuck into complex infrastructure challenges? Do you want … Government residency and right-to-work requirements associated with the required level of clearance. Ready for your next challenge? If you're an SRE or DevOps professional who enjoys solving difficult problems, improving infrastructure and working on projects that genuinely matter, we'd love to hear from you. Apply today ...

Director of Site Reliability Engineering

Location
Greater London, England, United Kingdom
influence engineering standards, enhance operational frameworks, and foster a culture of continuous improvement across mission‐critical environments. Responsibilities Lead and scale a global SRE organization, focusing on engineering excellence and team empowerment Collaborate with product, platform, operations, and security teams to embed reliability within SDLC practices Define … deliver systemic improvements across production environments Establish observability strategies with standardized tooling for metrics, logs, and tracing to support distributed systems Adopt and enforce SRE practices, including SLIs, SLOs, SLAs, and error budgets across services Drive resilience strategies with highly available architectures and disaster recovery readiness Champion an automation‐first ...

AI Platform & Site Reliability Engineering Consultant

Hiring Organisation
Akkodis
Location
City of London, London, United Kingdom
Employment Type
Permanent
Salary
£88000 - £96000/annum
SRE Managing Consultant Cloud Operating Model & Reliability Transformation Security Clearance: SC eligible (UK residency required) Shape the Future of Cloud Reliability Are you passionate about building resilient, scalable cloud platforms that truly support the business? Do you thrive at the intersection of engineering excellence, operating models … senior stakeholder advisory? We're looking for a Managing Consultant in Site Reliability Engineering (SRE) to help organisations shift from reactive operations to measurable, product-aligned reliability - embedding SRE as a core engineering discipline across cloud and hybrid environments. You'll work with senior leaders ...

AI Platform & Site Reliability Engineering Consultant

Hiring Organisation
Akkodis
Location
London, South East England, United Kingdom
Employment Type
Full-Time
Salary
£88,000 - £96,000 per annum
SRE Managing Consultant Cloud Operating Model & Reliability Transformation Security Clearance: SC eligible (UK residency required) Shape the Future of Cloud Reliability Are you passionate about building resilient, scalable cloud platforms that truly support the business? Do you thrive at the intersection of engineering excellence, operating models … senior stakeholder advisory? We're looking for a Managing Consultant in Site Reliability Engineering (SRE) to help organisations shift from reactive operations to measurable, product-aligned reliability - embedding SRE as a core engineering discipline across cloud and hybrid environments. You'll work with senior leaders ...

Custody Support Applications Support - Assistant Vice President

Location
Belfast City District, Northern Ireland, United Kingdom
modern observability platforms supporting securities processing and settlement functions. This position combines traditional application support responsibilities with Site Reliability Engineering (SRE) principles, automation, resiliency engineering, and operational risk management. The ideal candidate will be passionate about improving platform reliability, reducing operational toil, and driving continuous … enterprise-scale distributed applications within a financial services or mission-critical technology environment. Proven experience in Production Support, Site Reliability Engineering (SRE), Platform Support, or similar technical operations roles. Demonstrated experience managing or coordinating major incident responses. Strong analytical and problem-solving capabilities with a structured approach ...

Lead SRE - AWS Platform

Hiring Organisation
Hackajob Ltd
Location
Glasgow, Lanarkshire, Scotland, United Kingdom
Employment Type
Permanent
role in defining the future of a globally recognized firm and make a direct, meaningful impact in a space built for top achievers in site reliability engineering. As a Lead Site Reliability Engineer at JPMorganChase within Infrastructure Platforms, you bring deep technical expertise across multiple domains … play a key role in driving reliability outcomes for your team. You will conduct resiliency design reviews, break complex problems into digestible work, act as a technical authority for medium to large-sized products, and share your knowledge and experience with peers to raise the bar across the team. ...

Lead SRE - AWS Platform

Location
Glasgow, Scotland, United Kingdom
role in defining the future of a globally recognized firm and make a direct, meaningful impact in a space built for top achievers in site reliability engineering. As a Lead Site Reliability Engineer at JPMorganChase within Infrastructure Platforms, you bring deep technical expertise across multiple domains … play a key role in driving reliability outcomes for your team. You will conduct resiliency design reviews, break complex problems into digestible work, act as a technical authority for medium to large-sized products, and share your knowledge and experience with peers to raise the bar across the team. ...

Lead SRE- Azure & GCP

Hiring Organisation
Hackajob Ltd
Location
Glasgow, Lanarkshire, Scotland, United Kingdom
Employment Type
Permanent
DESCRIPTION We have a Lead Site Reliability Engineer (SRE) opportunity within our Google Cloud Site Reliability Engineering team. As a Lead Site Reliability Engineer at JPMorgan Chase within the Infrastructure Platform - Cloud Foundational Services SRE organization, you will join our Google Cloud Site Reliability Engineering team operating within a global follow-the-sun support model. Job Responsibilities: Lead and Implement SRE frameworks to support global google cloud environments and ensure the highest level of SLOs through operational excellence Mastery of application, data, infrastructure, and Agentic AI disciplines Keen understanding ...