1 to 25 of 146 Remote/Hybrid Site Reliability Engineering Jobs in London

Cloud Operating Model - Managing Consultant

Location
Greater London, England, United Kingdom
build and scale secure, reliable and operationally effective AI platforms. You will combine expertise in platform engineering, Site Reliability Engineering (SRE), observability and intelligent operations to help organisations move from isolated AI experimentation to production-grade, enterprise-scale AI services.You will work with technology, engineering … observability, platform automation and operational guardrails. Enable reliable and repeatable delivery of AI services from experimentation through to production.• Reliability Engineering & SRE: Establish SRE practices including SLIs, SLOs, error budgets, capacity planning, resilience engineering and reliability governance. Help clients shift from reactive operations to data-driven ...

Head Of Infrastructure and Cloud - Internal Applicants Only

Location
Greater London, England, United Kingdom
transition from traditional infrastructure management to a platform-centric, product-led operating model, integrating platform engineering, DevOps, Site Reliability Engineering (SRE), and Network Operations (NOC) to enable scalable, automated, and resilient technology services. To place the interests of customers at the centre of all activities … YBIYRI) model with shared accountability for service delivery and operational outcomes. Establish and integrate Site Reliability Engineering (SRE) practices, defining and managing service‐level objectives (SLOs), error budgets, and proactive reliability engineering across critical services. Ensure end‐to‐end service reliability and resilience, including ...

Head Of Infrastructure and Cloud

Hiring Organisation
Arbuthnot Latham
Location
London, UK
Employment Type
Full-time
transition from traditional infrastructure management to a platform-centric, product-led operating model, integrating platform engineering, DevOps, Site Reliability Engineering (SRE), and Network Operations (NOC) to enable scalable, automated, and resilient technology services. To place the interests of customers at the centre of all activities … YBIYRI) model with shared accountability for service delivery and operational outcomes. Establish and integrate Site Reliability Engineering (SRE) practices, defining and managing service-level objectives (SLOs), error budgets, and proactive reliability engineering across critical services. Ensure end-to-end service reliability and resilience, including ...

Director of Site Reliability Engineering

Location
Greater London, England, United Kingdom
influence engineering standards, enhance operational frameworks, and foster a culture of continuous improvement across mission‐critical environments. Responsibilities Lead and scale a global SRE organization, focusing on engineering excellence and team empowerment Collaborate with product, platform, operations, and security teams to embed reliability within SDLC practices Define … deliver systemic improvements across production environments Establish observability strategies with standardized tooling for metrics, logs, and tracing to support distributed systems Adopt and enforce SRE practices, including SLIs, SLOs, SLAs, and error budgets across services Drive resilience strategies with highly available architectures and disaster recovery readiness Champion an automation‐first ...

Lead Site Reliability Engineer

Hiring Organisation
Inspire People
Location
City of London, London, United Kingdom
Employment Type
Permanent, Part Time, Work From Home
Salary
£80,000
support economic growth across the UK. The Department for Business, Innovation, Science and Trade (BIST), in partnership with Inspire People, is seeking a Senior SRE Squad Lead with experience leading and developing engineers, strong DevOps and Site Reliability Engineering expertise, cloud platform experience, infrastructure-as-code capability … UK. BIST's Digital, Data and Technology (DDaT) directorate develops and operates the tools and services that enable this mission. As a Senior SRE Squad Lead, you will play a key role in leading engineers while remaining hands-on in the design, delivery and continuous improvement of reliable, secure ...

Lead Site Reliability Engineer

Hiring Organisation
Inspire People
Location
London, South East England, United Kingdom
Employment Type
Full-Time
Salary
£67,547 - £83,778 per annum
support economic growth across the UK. The Department for Business, Innovation, Science and Trade (BIST), in partnership with Inspire People, is seeking a Senior SRE Squad Lead with experience leading and developing engineers, strong DevOps and Site Reliability Engineering expertise, cloud platform experience, infrastructure-as-code capability … UK. BIST's Digital, Data and Technology (DDaT) directorate develops and operates the tools and services that enable this mission. As a Senior SRE Squad Lead, you will play a key role in leading engineers while remaining hands-on in the design, delivery and continuous improvement of reliable, secure ...

Site Reliability Engineer

Hiring Organisation
REVYBE IT RECRUITMENT LIMITED
Location
City of London, London, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£85,000
period of growth and investing heavily in its engineering and platform capabilities. They're looking for an experienced Site Reliability Engineer (SRE) to join the team and play a key role in building highly reliable, scalable, and observable infrastructure. This is a hands-on role focused … experience Help improve platform resilience, scalability, and disaster recovery capabilities Contribute to capacity planning and performance optimisation as the platform scales Establish and champion SRE best practices across the wider engineering function What We're Looking For Proven commercial experience working as an SRE, DevOps Engineer, Platform Engineer ...

Site Reliability Engineer

Hiring Organisation
Bristow Holland Ltd
Location
London, United Kingdom
Employment Type
Permanent
Salary
£55000 - £60000/annum - Offering 100% Work from home
exciting global technology organisation is looking for a Site Reliability Engineer (SRE) to join its growing engineering team. This is a fully remote position, offering the opportunity to work on large-scale, business-critical platforms used by customers around the world. The role would suit an experienced … Site Reliability, DevOps, Platform or Cloud Engineer with strong hands-on experience across Kubernetes and Microsoft Azure who enjoys solving complex production problems, improving reliability and automating manual processes. You will work closely with Development and DevOps teams, helping to design, build, operate and scale highly available ...

principal engineer- international technology & Starbucks digital solutions

Location
Greater London, England, United Kingdom
technical excellence across EMEA while aligning to global technology strategy and leading the Starbucks Digital Solutions technical direction for International markets. It will set engineering direction, raise standards and guide decisions across internally developed and third-party platforms that matter most to our business, customers, partners, baristas and shareholders.As … Management organisations in a product-led operating model.• Knowledge of modern engineering practices including Platform Engineering, Site Reliability Engineering (SRE), AI-assisted development and Developer Experience (DevEx).What else should you know?• We have a flexible working policy. Meaning 50% of the time we collaborate ...

Site Reliability Engineer, Studios

Location
Uxbridge, England, United Kingdom
rotations, to support live operations and critical systems. Occasional travel may be required depending on project and client needs. IMG is looking for a Site Reliability Engineer to help design, build, operate, and continuously improve resilient, secure, and highly available platforms that underpin our digital, cloud, and broadcast … adjacent services. This role is suited to someone who combines strong infrastructure and software engineering capability with an operational mindset, and who can help embed reliability engineering practices across systems that support live, business‐critical environments. The successful candidate will play a key role in improving service ...

Site Reliability Engineer, Studios

Location
Greater London, England, United Kingdom
rotations, to support live operations and critical systems. Occasional travel may be required depending on project and client needs. IMG is looking for a Site Reliability Engineer to help design, build, operate, and continuously improve resilient, secure, and highly available platforms that underpin our digital, cloud, and broadcast … adjacent services. This role is suited to someone who combines strong infrastructure and software engineering capability with an operational mindset, and who can help embed reliability engineering practices across systems that support live, business-critical environments. The successful candidate will play a key role in improving service ...

Site Relaibility Engineer

Hiring Organisation
Bristow Holland
Location
London, South East England, United Kingdom
Employment Type
Full-Time
Salary
£55,000 - £60,000 per annum
exciting global technology organisation is looking for a Site Reliability Engineer (SRE) to join its growing engineering team. This is a fully remote position, offering the opportunity to work on large-scale, business-critical platforms used by customers around the world. The role would suit an experienced … Site Reliability, DevOps, Platform or Cloud Engineer with strong hands-on experience across Kubernetes and Microsoft Azure who enjoys solving complex production problems, improving reliability and automating manual processes. You will work closely with Development and DevOps teams, helping to design, build, operate and scale highly available ...

Site Reliability Manager - Environment Strategy

Location
Greater London, England, United Kingdom
looking for a Site Reliability Manager to join our team in London, United Kingdom in a hybrid working mode. In this role, you will lead a team focused on environment strategy, automation, patch governance and operational reliability for AWS-based platforms. Your responsibilities include setting roadmaps, driving … while ensuring strong technical standards, compliance and resilience across all production and non-production systems. Responsibilities Define and own the vision and roadmap for site reliability and environment strategy Lead, mentor and develop a team of DevOps and environment engineers Set and enforce standards for environment provisioning, lifecycle ...

Site Reliability Engineer

Location
Greater London, England, United Kingdom
TITLE: Site Reliability Engineer SALARY: Up to £80,780 (Outside London), £93,390 (London) LOCATION: Remote (UK Based) HOURS: Full Time (35 Hours) WORKING PATTERN: Our work style is hybrid, which involves spending at least two days per week, or 40% of our time … with workplace adjustments including hybrid working expectations in line with our Flexibility Works policy. What you\'ll be doing We\'re looking for a Site Reliability Engineer to help build, scale and operate the platforms that power Curve\'s products and services. Working closely with engineering teams ...

Strategic DevSecOps Consultant

Hiring Organisation
CloudBees
Location
London, UK
Employment Type
Full-time
transform how software is built, secured, and delivered. As a member of the Services team, you will work at the intersection of DevSecOps, Platform Engineering, AI, and software delivery innovation, helping customers accelerate outcomes and unlock new levels of engineering productivity. You will collaborate with some … customer obsession. What You'll BringRequired:5+ years of experience in consulting, solutions architecture, platform engineering, DevOps, Site Reliability Engineering (SRE), or related customer-facing technical roles. Proven experience advising customers on DevSecOps, software delivery modernization, platform engineering, or cloud transformation initiatives. Strong understanding ...

Security Engineer (Site Reliability Engineering) - SC Cleared

Hiring Organisation
Sanderson Government and Defence
Location
London, United Kingdom
Employment Type
Contract
Contract Rate
£545 - £590 per day + Inside IR35
Length: 6-18 Months Rate: £545-£590 per day (Inside IR35) Positions Available: 2 About the Role We are seeking two experienced Security Engineers (SRE) to join a specialist consultancy delivering cyber security services across a portfolio of government projects and digital transformation programmes. This role is ideal for security … best practices. Drive continual improvements across security engineering and platform security functions. Essential Skills & Experience Strong experience as a Security Engineer, Security-focused SRE, Platform Security Engineer, or similar role. Advanced knowledge of Enterprise Security Architecture principles. Hands-on experiencewithSplunk, including: Security monitoring Dashboard development Alerting and reporting ...

Observability SME/Architect/Consultant

Hiring Organisation
Hays Specialist Recruitment Limited
Location
London, South East England, United Kingdom
Employment Type
Full-Time
Salary
Salary negotiable
Dependency Mapping Impact Tolerances Scenario Testing Resilience Risk Assessment Operational Resilience Frameworks Business Continuity Principles Service Management & Operations Site Reliability Engineering (SRE) Incident Management Problem Management Major Incident Support IT Operations Service Reliability & Availability Management ITIL Frameworks and Best Practice Consulting & Leadership Stakeholder Management Communication & Presentation … , IT Operations, Service Management, Operational Resilience, or Platform Engineering environments. Desirable ITIL Foundation or Advanced ITIL Certifications TOGAF Dynatrace Certification Splunk Certification SRE Foundation PRINCE2, Agile, Scrum, or PMP AWS, Azure, or GCP Certifications Hays Specialist Recruitment Limited acts as an employment agency for permanent recruitment and employment ...

Senior Backend Engineer

Hiring Organisation
Inspire People
Location
South West London, London, United Kingdom
Employment Type
Permanent, Part Time, Work From Home
Salary
£80,000
platform services that underpin critical digital products, helping development teams improve observability, monitoring, CI/CD and service resilience. As a Senior Backend Engineer (Site Reliability), you will: * Design, build and operate reliable, secure and scalable cloud platform services supporting critical digital products. * Develop and maintain platform tooling … automation, observability, monitoring and CI/CD capabilities. * Build software solutions using Python and modern engineering practices. * Write clean, maintainable code and infrastructure-as-code solutions to support service delivery. * Embed Site Reliability Engineering principles including SLIs, SLOs, error budgets and continual service improvement. * Support live ...

Senior Backend Engineer

Hiring Organisation
Inspire People
Location
London, South East England, United Kingdom
Employment Type
Full-Time
Salary
£63,824 - £80,158 per annum
platform services that underpin critical digital products, helping development teams improve observability, monitoring, CI/CD and service resilience. As a Senior Backend Engineer (Site Reliability), you will: * Design, build and operate reliable, secure and scalable cloud platform services supporting critical digital products. * Develop and maintain platform tooling … automation, observability, monitoring and CI/CD capabilities. * Build software solutions using Python and modern engineering practices. * Write clean, maintainable code and infrastructure-as-code solutions to support service delivery. * Embed Site Reliability Engineering principles including SLIs, SLOs, error budgets and continual service improvement. * Support live ...

Site Reliability Engineer / Production Support

Hiring Organisation
Hackajob Ltd
Location
London, United Kingdom
Employment Type
Permanent
fastest growing fintech in 2025. The momentum is real. THE OPPORTUNITY Monuments production environment is the heartbeat of a licensed bank, and the SRE role is the single point of ownership when incidents occur. You will directly oversee the offshore Production Support team, run on-call and incident response … eliminate it. Quality-driven - you care about alert quality, observability standards, and reliability patterns that prevent problems at source. WHAT YOU BRING Strong SRE or production support experience with accountability for incident response in a production environment. Deep understanding of observability tools, alerting, logging, and distributed systems debugging. Experience ...

Azure DevOps Engineer

Hiring Organisation
Anson Mccade
Location
South East London, London, United Kingdom
Employment Type
Contract, Work From Home
Contract Rate
From £500 to £700 per day Inside IR35
programme work with strong extension potential High-impact role with significant technical ownership Enterprise-scale Azure cloud environment Opportunity to shape DevSecOps and platform engineering best practices Exposure to cloud security and infrastructure automation initiatives Work alongside experienced cloud, security and engineering teams Flexible hybrid working arrangements Opportunity … Azure DevOps Engineer Microsoft Certifications such as AZ-104, AZ-305 or AZ-400 Platform Engineering practices Site Reliability Engineering (SRE) Agile and Scrum methodologies Enterprise change and release management Mission-critical cloud platforms Disaster recovery and resilience planning Apply Today If you're an experienced ...

AI Technical Platform Leader

Location
Greater London, England, United Kingdom
business stakeholders to mold and implement strategy. The role applies broad technical knowledge with depth in generative AI, agentic systems, enterprise platforms and engineering governance to ensure that AI platforms at WTW are optimally configured to achieve company vision and imperatives. It has end-to-end ownership of implementation … operations in a large, global enterprise: Software or platform engineering for global, enterprise-scaled solutions Management of Site reliability engineering (SRE) programs for mission‐critical systems Creation and management of DevOps practices for automated, consistent, and secure solution deployment in regulated environments. Literacy in global compliance ...

Senior Software Engineer

Location
Greater London, England, United Kingdom
contribute to modern digital capabilities, drive continuous improvement, and support the delivery of future-ready solutions. This is an opportunity to shape high-quality engineering outcomes, embrace innovation and AI-enabled ways of working, and create lasting value in a complex, enterprise-scale environment.Hybrid working:The places that … principles, secure software development practices, and security-focused engineering approaches.Experience working within regulated Financial Services environments.Understanding of Site Reliability Engineering (SRE) concepts, operational resilience, and service reliability practices.Relevant cloud, DevOps, or engineering certifications.We are a Disability Confident Employer:Capgemini is proud ...

AWS Cloud Architect, Technology Consulting

Location
Greater London, England, United Kingdom
security and automation technologies. You will help design scalable, secure and resilient cloud solutions for our clients, enabling digital transformation through modern architecture, platform engineering and DevOps practices. You should bring practical experience of cloud technologies and infrastructure modernisation, alongside an understanding of how AI-enabled tooling and modern … capabilities Strong understanding of cloud networking, identity, security and resilience patterns Experience supporting cloud engineering, DevOps or Site Reliability Engineering (SRE) teams Cloud architecture certifications in additional public cloud platforms (Azure or GCP) Experience using AI or agentic techniques within the role through GitHub Copilot, Claude ...

Site Reliability Engineer

Location
Greater London, England, United Kingdom
Site Reliability Engineer Reports to: Labs Team Lead Looper Insights Remote-first, with regular visits to our Byfleet and Hounslow data centres The company Looper Insights builds analytics products that help the world’s leading media and entertainment companies understand how their content is performing across digital platforms … collaborate closely to produce a valuable service for an industry about which we are all passionate. The role We’re looking for a Site Reliability Engineer to keep the global LooperBox fleet running, the physical backbone behind every piece of data Looper Insights produces. LooperBoxes sit in front ...