1 to 25 of 412 Remote/Hybrid Permanent Site Reliability Engineering Jobs

AI Platform & Site Reliability Engineering Managing Consultant

Location
United Kingdom
build and scale secure, reliable and operationally effective AI platforms. You will combine expertise in platform engineering, Site Reliability Engineering (SRE), observability and intelligent operations to help organisations move from isolated AI experimentation to production-grade, enterprise-scale AI services. You will work with technology, engineering … observability, platform automation and operational guardrails. Enable reliable and repeatable delivery of AI services from experimentation through to production. Reliability Engineering & SRE: Establish SRE practices including SLIs, SLOs, error budgets, capacity planning, resilience engineering and reliability governance. Help clients shift from reactive operations to data-driven ...

Head of Site Reliability Engineering (SRE)

Hiring Organisation
Computershare
Location
Bristol, Gloucestershire, United Kingdom
Salary
£ 70 K
flex.We give you a world of potential Computershare have an up-and-coming opportunity for a Head of Site Reliability Engineering (SRE) to join our global technology team at a time when transforming our organisations toward an SRE operating model is a key focus.Reporting directly … improve service reliability, resilience, and performance of critical platforms.A role you will love We are seeking an experienced and visionary Head of SRE to define, lead, and evolve our global reliability strategy. This is a senior leadership role responsible for driving operational excellence, service reliability, observability, automation ...

The Core Engineering - Site Reliability Engineering - Associate - Birmingham

Location
Birmingham, England, United Kingdom
Core Engineering - Site Reliability Engineering - Associate - Birmingham location_on Birmingham, West Midlands, England, United Kingdom What We Do Core Engineering is a global team of more than 2,500 engineers and scientists focused on solving complex, mission-critical problems across the firm. We build … enable faster delivery of new capabilities, reduce downtime, and eliminate repetitive operational work through automation. This role sits within Core Engineering and applies SRE practices to services that support compliance, risk, and other critical business functions. Responsibilities Proactively manage production services by measuring and monitoring availability, capacity, latency ...

Cloud Operating Model - Managing Consultant

Location
Greater London, England, United Kingdom
build and scale secure, reliable and operationally effective AI platforms. You will combine expertise in platform engineering, Site Reliability Engineering (SRE), observability and intelligent operations to help organisations move from isolated AI experimentation to production-grade, enterprise-scale AI services.You will work with technology, engineering … observability, platform automation and operational guardrails. Enable reliable and repeatable delivery of AI services from experimentation through to production.• Reliability Engineering & SRE: Establish SRE practices including SLIs, SLOs, error budgets, capacity planning, resilience engineering and reliability governance. Help clients shift from reactive operations to data-driven ...

Head Of Infrastructure and Cloud

Hiring Organisation
Arbuthnot Latham
Location
London, United Kingdom
Salary
£ 100 K
transition from traditional infrastructure management to a platform-centric, product-led operating model, integrating platform engineering, DevOps, Site Reliability Engineering (SRE), and Network Operations (NOC) to enable scalable, automated, and resilient technology services.To place the interests of customers at the centre of all activities … Build It, You Run It” (YBIYRI) model with shared accountability for service delivery and operational outcomes.Establish and integrate Site Reliability Engineering (SRE) practices, defining and managing service-level objectives (SLOs), error budgets, and proactive reliability engineering across critical services.Ensure end-to-end service reliability ...

Head Of Infrastructure and Cloud - Internal Applicants Only

Location
Greater London, England, United Kingdom
transition from traditional infrastructure management to a platform-centric, product-led operating model, integrating platform engineering, DevOps, Site Reliability Engineering (SRE), and Network Operations (NOC) to enable scalable, automated, and resilient technology services. To place the interests of customers at the centre of all activities … YBIYRI) model with shared accountability for service delivery and operational outcomes. Establish and integrate Site Reliability Engineering (SRE) practices, defining and managing service‐level objectives (SLOs), error budgets, and proactive reliability engineering across critical services. Ensure end‐to‐end service reliability and resilience, including ...

Head Of Infrastructure and Cloud

Hiring Organisation
Arbuthnot Latham
Location
London, UK
Employment Type
Full-time
transition from traditional infrastructure management to a platform-centric, product-led operating model, integrating platform engineering, DevOps, Site Reliability Engineering (SRE), and Network Operations (NOC) to enable scalable, automated, and resilient technology services. To place the interests of customers at the centre of all activities … YBIYRI) model with shared accountability for service delivery and operational outcomes. Establish and integrate Site Reliability Engineering (SRE) practices, defining and managing service-level objectives (SLOs), error budgets, and proactive reliability engineering across critical services. Ensure end-to-end service reliability and resilience, including ...

Lead Cloud Site Reliability Engineer

Location
Manchester, England, United Kingdom
Summary End Date Monday 21 September 2026 Salary Range £92,701 - £109,060 Flexible Working Options Hybrid Working, Job Share Lead Site Reliability Engineer – Public Cloud Platform Location: Manchester or Bristol Salary: £92,701- £109,043 Working Pattern: Hybrid (2 days in office per week) About this opportunity … Skills and Experience Designing, building or operating large-scale cloud platforms within Azure, GCP or comparable cloud environments. Site Reliability Engineering (SRE), Platform Engineering, Infrastructure Engineering, Cloud Engineering or Production Operations. Observability and monitoring practices, including metrics, logging and distributed tracing. Incident management, problem ...

Lead Cloud Site Reliability Engineer

Location
Halifax, England, United Kingdom
Summary End Date Monday 21 September 2026 Salary Range £92,701 - £109,060 Flexible Working Options Hybrid Working, Job Share Lead Site Reliability Engineer – Public Cloud Platform Location: Manchester or Bristol Salary: £92,701- £109,043 Working Pattern: Hybrid (2 days in office per week) About this opportunity … Skills and Experience Designing, building or operating large-scale cloud platforms within Azure, GCP or comparable cloud environments. Site Reliability Engineering (SRE), Platform Engineering, Infrastructure Engineering, Cloud Engineering or Production Operations. Observability and monitoring practices, including metrics, logging and distributed tracing. Incident management, problem ...

Senior Site Reliability Engineer

Hiring Organisation
Malvern Panalytical
Location
United Kingdom
Salary
£ 55 K
delivered rapidly without compromising stability, security, or customer experience. As a senior technical specialist, you will champion Site Reliability Engineering (SRE) best practices, improve platform resilience, and contribute to the continuous evolution of our cloud services and operational excellence.This is an exciting opportunity for an experienced SRE … experience operating and supporting production workloads on Microsoft Azure• Expertise in monitoring, alerting, observability platforms, and cloud-native operational practices• Deep understanding of SRE principles including SLOs, SLIs, SLAs, error budgets, and reliability engineering methodologies• Experience with automation and development using technologies such as C#, Python, PowerShell ...

Site Reliability Engineer

Location
Swindon, England, United Kingdom
kaikille hakijoillemme yhtäläiset mahdollisuudet sukupuolesta ja sukupuolen ilmaisusta, vammaisuudesta, alkuperästä, uskonnollisesta vakaumuksesta ja seksuaalisesta suuntautumisesta tai muista kriteereistä riippumatta.**Site Reliability Engineer (SRE)****Location: Swindon (Hybrid)**Join Edenred and help shape resilient, future-ready infrastructure that supports millions of users worldwide.At Edenred, we’re looking for a Site Reliability Engineer (SRE) to join our Infrastructure Engineering team. This is an exciting opportunity for someone who is passionate about reliability, automation, operational excellence, and continuous improvement.**About the role**As a Site Reliability Engineer, you will play a key role in ensuring ...

Vice President - Site Reliability Engineering (SRE) - The Core Engineering - Birmingham Birmingham · United Kingdom · Vice President

Location
Birmingham, England, United Kingdom
Vice President - Site Reliability Engineering (SRE) - The Core Engineering - Birmingham location_on Birmingham, West Midlands, England, United Kingdom WHAT WE DO Site Reliability Engineering at Goldman Sachs sits at the intersection of software engineering, systems design, and production excellence. In this … This role is for software engineers who enjoy solving complex distributed system problems, building tools and platforms that make teams more effective, and championing SRE principles (such as SLOs, error budgets, and blameless post-mortems) across a large engineering organization. Key Responsibilities Partner with engineering leadership to establish ...

Site Reliability Engineer

Location
Hove, England, United Kingdom
Title : Site Reliability Engineer Job Location : Hove, UK (Hybrid 3 days office) Job Type : FTE Job Description: SRE will play a pivotal role in driving the modernization of IT operations by implementing observability practices and automating toil. This position requires a deep understanding of Site Reliability Engineering (SRE) principles, modern observability tools, and automation techniques to ensure scalability, reliability, and efficiency in IT systems. This role requires a strategic thinker with hands-on expertise who can lead modernization efforts while fostering a culture of reliability and innovation. Primary Responsibilities: Work closely with ...

Site Reliability Engineer

Location
Swindon, England, United Kingdom
opportunities regardless of their gender and gender expression, disability, origin, religious belief and sexual orientation or any other criteria.**Site Reliability Engineer (SRE)****Location: Swindon (Hybrid)**Join Edenred and help shape resilient, future-ready infrastructure that supports millions of users worldwide.At Edenred, we’re looking for a Site Reliability Engineer (SRE) to join our Infrastructure Engineering team. This is an exciting opportunity for someone who is passionate about reliability, automation, operational excellence, and continuous improvement.**About the role**As a Site Reliability Engineer, you will play a key role in ensuring ...

Senior SRE Engineer

Hiring Organisation
SF Partners
Location
Birmingham, West Midlands (County), United Kingdom
Employment Type
Permanent
Salary
£100000 - £110000/annum great training and progression opp
some of the UK's largest and most complex enterprise platforms. Working within a highly skilled, multi-disciplinary engineering team, this Senior SRE Engineer will take technical ownership of observability and reliability across large-scale AWS environments, helping engineering teams improve platform performance, resilience and availability. … mindset, using automation, Infrastructure as Code and cloud-native tooling to improve reliability and reduce operational overhead. This role would suit a Senior SRE DevOps or Cloud Engineer who combines hands-on AWS expertise with genuine Dynatrace implementation experience, rather than simply having used Dynatrace as an end user. ...

Site Reliability Engineer

Hiring Organisation
Hackajob Ltd
Location
Swindon, Wiltshire, South West, United Kingdom
Employment Type
Permanent
Salary
£85,000
opportunities regardless of their gender and gender expression, disability, origin, religious belief and sexual orientation or any other criteria. Site Reliability Engineer (SRE) Location: Swindon (Hybrid) Join Edenred and help shape resilient, future-ready infrastructure that supports millions of users worldwide. At Edenred, we're looking … Site Reliability Engineer (SRE) to join our Infrastructure Engineering team. This is an exciting opportunity for someone who is passionate about reliability, automation, operational excellence, and continuous improvement. About the role As a Site Reliability Engineer, you will play a key role in ensuring ...

Site Reliability Engineer

Hiring Organisation
Twinstream Limited
Location
Cheltenham, Gloucestershire, South West, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£95,000
Site Reliability Engineer | Up to £95,000 DOE | Fully Remote Initially | Cheltenham | Future Hybrid Working Build resilient systems. Solve complex challenges. Make a real impact. Are you a Site Reliability Engineer or DevOps professional who enjoys getting stuck into complex infrastructure challenges? Do you want … Government residency and right-to-work requirements associated with the required level of clearance. Ready for your next challenge? If you're an SRE or DevOps professional who enjoys solving difficult problems, improving infrastructure and working on projects that genuinely matter, we'd love to hear from you. Apply today ...

Director of Site Reliability Engineering

Location
Greater London, England, United Kingdom
influence engineering standards, enhance operational frameworks, and foster a culture of continuous improvement across mission‐critical environments. Responsibilities Lead and scale a global SRE organization, focusing on engineering excellence and team empowerment Collaborate with product, platform, operations, and security teams to embed reliability within SDLC practices Define … deliver systemic improvements across production environments Establish observability strategies with standardized tooling for metrics, logs, and tracing to support distributed systems Adopt and enforce SRE practices, including SLIs, SLOs, SLAs, and error budgets across services Drive resilience strategies with highly available architectures and disaster recovery readiness Champion an automation‐first ...

Custody Support Applications Support - Assistant Vice President

Location
Belfast City District, Northern Ireland, United Kingdom
modern observability platforms supporting securities processing and settlement functions. This position combines traditional application support responsibilities with Site Reliability Engineering (SRE) principles, automation, resiliency engineering, and operational risk management. The ideal candidate will be passionate about improving platform reliability, reducing operational toil, and driving continuous … enterprise-scale distributed applications within a financial services or mission-critical technology environment. Proven experience in Production Support, Site Reliability Engineering (SRE), Platform Support, or similar technical operations roles. Demonstrated experience managing or coordinating major incident responses. Strong analytical and problem-solving capabilities with a structured approach ...

Platform Engineering Manager (SRE)

Location
Bracknell, England, United Kingdom
cloud products fast, reliable, and trusted by some of the world's largest SAP-run businesses. We are hiring a Platform Engineering Manager (SRE) to own reliability and platform engineering for Klario and our Private Cloud estate. This is a hands-on leadership role and a genuine … developer tooling. Drive standardisation and reduce engineering friction through automation and self-service. Site Reliability Engineering Introduce and embed SRE practices across Engineering. Improve reliability, resilience, recoverability, and operational readiness. Stand up monitoring, alerting, logging, and service health capabilities, including public-facing uptime and status ...

Senior Site Reliability Engineer

Hiring Organisation
GCS
Location
Glasgow, City of Glasgow, United Kingdom
Employment Type
Permanent
Salary
£75000 - £95000/annum Bonus
experienced Senior Site Reliability Engineer to drive reliability, scalability and performance across critical banking systems. This role combines hands-on SRE engineering with technical leadership, with a strong focus on observability, automation, continuous improvement and optimisation. Responsibilities: * Build and maintain reliable, scalable and secure infrastructure platforms … solutions. * Apply SRE and software engineering practices to improve reliability, availability and performance. * Monitor systems, manage incidents and lead complex troubleshooting and root cause analysis. * Develop automation using programming and scripting to reduce manual intervention and improve efficiency. * Develop and improve observability, monitoring, instrumentation and performance capabilities. ...

SRE

Location
Hove, England, United Kingdom
Role: SRE Location: Hove, UK Is it Permanent/Contract: Open for both Perm/Contract Is it Onsite/Remote/Hybrid: 2 days per week from office No. of Positions: 1 We are seeking an experienced Site Reliability Engineer (SRE) to drive the modernization … driven alerting and proactive anomaly detection to reduce Mean Time to Detect (MTTD) and Mean Time to Resolve (MTTR). Establish and enforce SRE best practices, including Service Level Indicators (SLIs), Service Level Objectives (SLOs), and Error Budgets. Define and implement an AIOps roadmap to enhance operational intelligence and automation. ...

Lead Site Reliability Engineer

Hiring Organisation
Inspire People
Location
Edinburgh, Midlothian, United Kingdom
Employment Type
Full-Time
Salary
£63,824 - £80,158 per annum
support economic growth across the UK. The Department for Business, Innovation, Science and Trade (BIST), in partnership with Inspire People, is seeking a Senior SRE Squad Lead with experience leading and developing engineers, strong DevOps and Site Reliability Engineering expertise, cloud platform experience, infrastructure-as-code capability … UK. BIST's Digital, Data and Technology (DDaT) directorate develops and operates the tools and services that enable this mission. As a Senior SRE Squad Lead, you will play a key role in leading engineers while remaining hands-on in the design, delivery and continuous improvement of reliable, secure ...

Lead Site Reliability Engineer

Hiring Organisation
Inspire People
Location
Salford, Lancashire, United Kingdom
Employment Type
Full-Time
Salary
£63,824 - £80,158 per annum
support economic growth across the UK. The Department for Business, Innovation, Science and Trade (BIST), in partnership with Inspire People, is seeking a Senior SRE Squad Lead with experience leading and developing engineers, strong DevOps and Site Reliability Engineering expertise, cloud platform experience, infrastructure-as-code capability … UK. BIST's Digital, Data and Technology (DDaT) directorate develops and operates the tools and services that enable this mission. As a Senior SRE Squad Lead, you will play a key role in leading engineers while remaining hands-on in the design, delivery and continuous improvement of reliable, secure ...

Lead Site Reliability Engineer

Hiring Organisation
Inspire People
Location
City of London, London, United Kingdom
Employment Type
Permanent, Part Time, Work From Home
Salary
£80,000
support economic growth across the UK. The Department for Business, Innovation, Science and Trade (BIST), in partnership with Inspire People, is seeking a Senior SRE Squad Lead with experience leading and developing engineers, strong DevOps and Site Reliability Engineering expertise, cloud platform experience, infrastructure-as-code capability … UK. BIST's Digital, Data and Technology (DDaT) directorate develops and operates the tools and services that enable this mission. As a Senior SRE Squad Lead, you will play a key role in leading engineers while remaining hands-on in the design, delivery and continuous improvement of reliable, secure ...