1 to 25 of 572 Site Reliability Engineering Jobs in London

Site Reliability Engineering Lead

Location
City Of London, England, United Kingdom
lead a team of SREs responsible for ensuring the reliability, scalability, security, performance, and operational excellence of mission-critical applications and platforms. The SRE Lead will partner closely with engineering, architecture, security, operations, and business stakeholders to drive cloud modernization, operational maturity, observability excellence, automation, and continuous improvement. … mentor a team of Site Reliability Engineers, fostering a culture of ownership, operational excellence, collaboration, and continuous learning. Define and drive SRE strategy, standards, best practices, and operational frameworks across engineering organizations. Partner with product and platform teams to improve application reliability, scalability, security, performance ...

Head Of Infrastructure and Cloud - Internal Applicants Only

Location
Greater London, England, United Kingdom
transition from traditional infrastructure management to a platform-centric, product-led operating model, integrating platform engineering, DevOps, Site Reliability Engineering (SRE), and Network Operations (NOC) to enable scalable, automated, and resilient technology services. To place the interests of customers at the centre of all activities … YBIYRI) model with shared accountability for service delivery and operational outcomes. Establish and integrate Site Reliability Engineering (SRE) practices, defining and managing service‐level objectives (SLOs), error budgets, and proactive reliability engineering across critical services. Ensure end‐to‐end service reliability and resilience, including ...

Head Of Infrastructure and Cloud

Hiring Organisation
Arbuthnot Latham
Location
London, UK
Employment Type
Full-time
transition from traditional infrastructure management to a platform-centric, product-led operating model, integrating platform engineering, DevOps, Site Reliability Engineering (SRE), and Network Operations (NOC) to enable scalable, automated, and resilient technology services. To place the interests of customers at the centre of all activities … YBIYRI) model with shared accountability for service delivery and operational outcomes. Establish and integrate Site Reliability Engineering (SRE) practices, defining and managing service-level objectives (SLOs), error budgets, and proactive reliability engineering across critical services. Ensure end-to-end service reliability and resilience, including ...

Distinguished Engineer - Head of Service Reliability Engineering

Hiring Organisation
Global Resourcing
Location
London, UK
Employment Type
Full-time
reliability, operability, and observability integral to engineering from the outset. With enterprise-wide reach and influence, you will set the direction for SRE, raise service maturity, and build a lasting capability that enables teams to innovate safely and operate with confidence. This executive-level role has accountability … , with the credibility to influence architecture, set standards, and deliver reliability outcomes in complex, regulated environments. Site Reliability Engineering (SRE): Leads the adoption of SRE practices, using SLOs, error budgets, and reliability metrics to drive measurable improvements in service performance and operational decision-making. ...

Engineer - Site Reliability

Location
Greater London, England, United Kingdom
early‐session coverage from London that ensures continuous, high‐availability operations across Cboe's real‐time low‐latency trading platforms. The London‐based SRE provides technical support to Cboe Trade Desk and Operations Support Center staff across time zones, and works closely with Software Engineering, Systems Engineering … spoken — is required. This role demands clear, precise, and unambiguous communication at all times. As the operational bridge between Cboe's European and APAC SRE teams and its US‐based leadership, the ability to communicate with clarity across time zones, cultures, and technical disciplines is fundamental to the success ...

Vice President, Site Reliability Engineering

Hiring Organisation
The Bank of New York Mellon
Location
London, UK
Employment Type
Full-time
seeking a future team member for the role of Vice President - Site Reliability Engineer to join our team. This role is located in London. Role SummaryBNY is seeking a Vice President - Site Reliability Engineer to design, build, deploy, and scale resilient, automated, and centrally managed engineering … driven automation, and modern software delivery practices. Experience supporting distributed systems, cloud-native platforms, or container-based architectures. Knowledge of Agile, DevOps, and SRE operating models, including continuous improvement and blameless post-incident practices. Ability to influence engineering standards and drive adoption of common tooling and automation patterns across ...

Senior Site Reliability Engineer

Hiring Organisation
Pathfinder Business Solutions Ltd
Location
City of London, London, United Kingdom
Employment Type
Permanent
Salary
£90,000
Senior Site Reliability Engineer (SRE) London, hybrid, one day a week onsite Up to £90,000 plus 10% cash allowance, 10% non contributory pension and discretionary bonus Were looking for a Senior Site Reliability Engineer with hands on Azure, Kubernetes and Terraform experience to join … service health and help improve how systems perform, scale and recover. As the teams Senior Site Reliability Engineer, youll also represent SRE across the wider business, working with engineering and delivery colleagues to establish reliable ways of working. The Role Keep systems healthy, reliable and ready ...

Senior DevSecOps Engineer

Location
Greater London, England, United Kingdom
repeatable, and secure delivery of autonomy and mission software • Embed security throughout the software development lifecycle, integrating security controls, testing, evidence, and assurance into engineering workflows • Apply UK MOD Secure by Design principles and help engineering teams meet cyber security and technical assurance responsibilities • Collaborate with software, autonomy … controls including static analysis, dependency scanning, container scanning, secrets detection, software composition analysis, vulnerability management, and policy enforcement • Champion a DevSecOps culture emphasizing security, reliability, deployability, and operational performance ownership Requirements BS or MS in Computer Science, Software Engineering, Cyber Security, Electrical Engineering, Systems Engineering ...

Technical Lead - Site Reliability Engineering

Location
Greater London, England, United Kingdom
Reliability Engineering capabilities to strengthen reliability, observability, security, and operational excellence across our Markets and Risk Intelligence division.As a **Technical Lead SRE**, you will be a senior hands‐on technical person help shape the foundations of reliability across both new and existing platforms. You will collaborate … person who is passionate about reliability engineering and who bring a continuous improvement approach to everything they do!Lead the establishment of SRE foundations for new projects building environments, monitoring, alerting, and ensuring operational readiness from day one.Collaborate with Architecture and Engineering teams to embed reliability ...

Foundation Engineering - SRE Platforms - Site Reliability Engineer – Associate - London London · United Kingdom · Associate

Location
Greater London, England, United Kingdom
Foundation Engineering - SRE Platforms - Site Reliability Engineer – Associate - London location_on London, Greater London, England, United Kingdom Role Overview Goldman Sachs has embarked on one of its most ambitious engineering programs: theConsolidated Trade Ledger (CTL), a ground-up reimagining of the front-to-back architecture that … underpins every trade the firm executes. CTL is a flagship initiative jointly sponsored by Global Markets and Engineering leadership, and it sits at the very heart of the firm's core technology strategy. This new, cloud-native platform will deliver the capacity, extensibility, scalability, and innovation capabilities to power ...

Director of Site Reliability Engineering

Location
Greater London, England, United Kingdom
influence engineering standards, enhance operational frameworks, and foster a culture of continuous improvement across mission‐critical environments. Responsibilities Lead and scale a global SRE organization, focusing on engineering excellence and team empowerment Collaborate with product, platform, operations, and security teams to embed reliability within SDLC practices Define … deliver systemic improvements across production environments Establish observability strategies with standardized tooling for metrics, logs, and tracing to support distributed systems Adopt and enforce SRE practices, including SLIs, SLOs, SLAs, and error budgets across services Drive resilience strategies with highly available architectures and disaster recovery readiness Champion an automation‐first ...

AI Platform & Site Reliability Engineering Consultant

Hiring Organisation
Akkodis
Location
City of London, London, United Kingdom
Employment Type
Permanent
Salary
£88000 - £96000/annum
SRE Managing Consultant Cloud Operating Model & Reliability Transformation Security Clearance: SC eligible (UK residency required) Shape the Future of Cloud Reliability Are you passionate about building resilient, scalable cloud platforms that truly support the business? Do you thrive at the intersection of engineering excellence, operating models … senior stakeholder advisory? We're looking for a Managing Consultant in Site Reliability Engineering (SRE) to help organisations shift from reactive operations to measurable, product-aligned reliability - embedding SRE as a core engineering discipline across cloud and hybrid environments. You'll work with senior leaders ...

Lead Site Reliability Engineer

Hiring Organisation
Inspire People
Location
Clerkenwell, Greater London, UK
related job information below. The Department for Business, Innovation, Science and Trade (BIST), in partnership with Inspire People, is seeking a Senior SRE Squad Lead with experience leading and developing engineers, strong DevOps and Site Reliability Engineering expertise, cloud platform experience, infrastructure-as-code capability … UK. BIST's Digital, Data and Technology (DDaT) directorate develops and operates the tools and services that enable this mission. As a Senior SRE Squad Lead, you will play a key role in leading engineers while remaining hands-on in the design, delivery and continuous improvement of reliable, secure ...

Cloud and Platform Engineer-Consultant-AI and Digital Factory

Location
Greater London, England, United Kingdom
more of the following disciplines: cloud engineering, platform engineering, infrastructure engineering, DevOps, and Site Reliability Engineering (SRE). As an engineer in our Cloud Team, you will design, innovate, optimize, build, and run the infrastructure and platforms our clients depend on. Depending on your … service tooling that improve developer experience• Work with clients and internal teams to shape new engineering opportunities and grow a strong DevOps/SRE/platform culture• Provide operational support: monitoring, alerting, troubleshooting, and production issue resolution• Conduct systems tests for security, performance, resilience, and availability• Share your knowledge ...

Site Reliability Engineer

Hiring Organisation
REVYBE IT RECRUITMENT LIMITED
Location
City of London, London, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£85,000
period of growth and investing heavily in its engineering and platform capabilities. They're looking for an experienced Site Reliability Engineer (SRE) to join the team and play a key role in building highly reliable, scalable, and observable infrastructure. This is a hands-on role focused … experience Help improve platform resilience, scalability, and disaster recovery capabilities Contribute to capacity planning and performance optimisation as the platform scales Establish and champion SRE best practices across the wider engineering function What We're Looking For Proven commercial experience working as an SRE, DevOps Engineer, Platform Engineer ...

Lead Site Reliability Engineer

Hiring Organisation
JP Morgan Chase
Location
London, UK
Employment Type
Full-time
defining the future of a globally recognized firm and have a direct and significant effect in a realm tailored for top achievers in site reliability. As a Lead Site Reliability Engineer at JPMorgan Chase within the Infrastructure Platforms team, you hold a leadership role in your team … ability to expand and collaborate across different levels and stakeholder groupsDemonstrated experience using enterprise-authorized AI capabilities within the work environment to improve SRE workflows (e.g., incident investigation support and knowledge capture) with strong validation habits and awareness of data sensitivity. Ability to evaluate AI-assisted operational recommendations for correctness ...

Lead Site Reliability Engineer, Athena Core

Location
Greater London, England, United Kingdom
defining the future of a globally recognized firm and have a direct and significant effect in a realm tailored for top achievers in site reliability. As a Lead Site Reliability Engineer at JPMorgan Chase within the Commercial and Investment Banking, Markets Technology – Athena Core team, you hold … programming languages such as: Python, Java/Spring Boot, .Net Demonstrated experience using enterprise-authorized AI capabilities within the work environment to improve SRE workflows (e.g., incident investigation support and knowledge capture) with strong validation habits and awareness of data sensitivity. Ability to evaluate AI-assisted operational recommendations for correctness ...

Lead Site Reliability Engineer, Athena Core

Location
Greater London, England, United Kingdom
defining the future of a globally recognized firm and have a direct and significant effect in a realm tailored for top achievers in site reliability. As a Lead Site Reliability Engineer at JPMorgan Chase within the Commercial and Investment Banking, Markets Technology - Athena Core team, you hold … programming languages such as: Python, Java/Spring Boot, .Net Demonstrated experience using enterprise-authorized AI capabilities within the work environment to improve SRE workflows (e.g., incident investigation support and knowledge capture) with strong validation habits and awareness of data sensitivity. Ability to evaluate AI-assisted operational recommendations for correctness ...

Head of Infrastructure and Cyber Security

Location
Greater London, England, United Kingdom
systems through modern tech paradigms such as Infrastructure as Code (IaC), automated CI/CD pipelines, FinOps, and Site Reliability Engineering (SRE). You will establish clear technical roadmaps, professional standards and investment priorities that support Hackney’s wider digital transformation. Cyber security will be central … improved commercial arrangements. – Strong supplier, procurement, budget and risk‐management experience. – Experience with modern engineering practices including Site Reliability Engineering (SRE), Infrastructure as Code (IaC), DevSecOps, and cloud FinOps optimization. This is an opportunity to lead a strategically important service, strengthen Hackney’s cyber resilience ...

Site Reliability Engineer

Hiring Organisation
Morgan McKinley
Location
London, South East England, United Kingdom
Employment Type
Full-Time
Salary
Salary negotiable
About the Role We are seeking a Senior Level 3 Wintel Site Reliability Engineer (SRE) to join our infrastructure engineering function in London. This is a hands-on "builder" position focused on modernizing, automating, and maintaining our enterprise Windows Server ecosystem and hybrid cloud integrations. Moving away … point for complex Windows Server OS internals, Active Directory forests, Group Policies, Kerberos authentication, and core domain services. Site Reliability Engineering (SRE): Implement proactive observability, monitoring, automated recovery workflows, and performance optimization to enforce continuous service availability and operational resilience. Architecture & Modernization: Partner with engineering, cloud ...

Azure CloudOps Engineer

Location
Greater London, England, United Kingdom
Infrastructure as Code, enhancing observability and AIOps capabilities, and driving automation across both application and infrastructure lifecycles. This role combines Cloud Engineering, DevOps, SRE, and AIOps practices, leveraging automation, AI-assisted operations, and self-healing capabilities to improve platform reliability, operational efficiency, and service availability. Key Responsibilities: Design … cloud and hybrid environments. Essential Skills & Experience 3+ years' experience in Cloud Engineering, DevOps, Platform Engineering, Site Reliability Engineering (SRE), or Cloud Operations roles. Strong experience with Microsoft Azure services and cloud infrastructure environments. Strong experience with Terraform is preferred. Experience with Helm, CloudFormation ...

Junior Software Engineer (with SRE Influences)

Location
Greater London, England, United Kingdom
Junior Software Engineer (with SRE Influences) About Fogsphere Fogsphere is a London‐based innovator focused on transforming workplace and urban safety through advanced AI, Computer Vision, and Industrial IoT. Built on a principled “Edge‐to‐Fog‐to‐Cloud” architecture, our platform turns passive CCTV cameras and sensors into proactive hazard … reliability of Fogsphere’s AI platform . This role blends software engineering with elements of site reliability engineering (SRE) — giving you exposure not only to building robust software but also to ensuring its smooth, automated, and scalable deployment in production environments. We are looking ...

Senior Site Reliability Engineer

Location
Greater London, England, United Kingdom
Senior Site Reliability EngineerApplylocations: London (82)time type: Full timeposted on: Posted Todayjob requisition id: JR101516**Senior Site Reliability Engineer (SRE) - GCP/Kubernetes****About the Role**We are seeking an experienced and highly motivated Senior Site Reliability Engineer (SRE) to join our small … agile engineering team. This role offers the unique opportunity to drive the reliability, scalability, and performance of our core platform with a high degree of autonomy and ownership.The successful candidate will split their time between providing expert operational support for our critical systems and leading exciting new infrastructure ...

Site Reliability Engineer

Hiring Organisation
Quant Capital
Location
London, UK
Employment Type
Full-time
Site Reliability Engineer – FintechQuant Capital is urgently looking for a Site Reliability Engineer to join or well known Fintech50 client who produces software disrupting the wealth management market. My client is a market leading SAAS provider to financial advisory business nationwide. They are currently in growth … stable solutions. Familiarity with development languages, such as .NET, Java or PythonRedisDocker/KubernetesDatabase experiencesThis role suits a senior Engineer from a DevOps or SRE background who is real technologist interested in the latets tooling and technologies that support software development and infrasturtcure. The firm has a corporate feel ...

Senior Platform / Site Reliability Engineer LAMP & AWS

Hiring Organisation
Bilge Adam Technologies UK Ltd
Location
London, United Kingdom
Employment Type
Permanent
Salary
£75,000
enior Platform/Site Reliability Engineer LAMP & AWS Location: London, UK Hybrid Occasional Travel: Basingstoke Employment Type: Permanent/FTE with BGTS Start Date: ASAP Salary: Competitive/Best Salary About the Role We are looking for a Senior Platform/Site Reliability Engineer to join … team and take ownership of the reliability, operational support and gradual modernization of established platforms. The role is primarily focused on maintaining and supporting a legacy Linux/Apache/MySQL/PHP (LAMP) environment , including troubleshooting, break-fix and operational support. Alongside this, you will play an important ...