76 to 100 of 157 Site Reliability Engineer Jobs in London

Site Reliability Engineer - Live Ops & Cloud Resilience

Location
Greater London, England, United Kingdom
seeking an experienced Site Reliability Engineer to design, build, and operate resilient, secure platforms underpinning our digital and live operations. You’ll focus on reliability, observability, automation, and disaster recovery across hybrid environments, collaborating with engineering, operations, and project stakeholders. The role emphasizes improving service availability ...

Site Reliability Engineer: Automation & Observability

Location
Greater London, England, United Kingdom
Apple Inc. is seeking a Site Reliability Engineer in London to join the Apple Services … Engineering team. You will help sustain large-scale services powering the App Store, Apple Music, TV, Podcasts and Books for users worldwide. As an SRE, you’ll work with Linux, open source tools and internal software to manage configuration, deployment, logging and monitoring. You’ll collaborate with development teams ...

Remote Site Reliability Engineer - Cloud-Native Infra

Location
City of Westminster, England, United Kingdom
Lloyds Bank plc is seeking a Site Reliability Engineer to help build, scale and operate the platforms powering Curve's products and services. This role focuses on reliability, security, and observability across our cloud-native stack, with a hybrid work model allowing UK-based … remote work and partial office attendance. You will collaborate with engineers, product managers and stakeholders to drive automation, incident post-mortems, and best-practice SRE culture while supporting millions of #J-18808-Ljbffr ...

Technical Site Reliability Engineer

Hiring Organisation
Anduril Industries
Location
London, United Kingdom
Salary
£ 70 K
current, and trustworthy. A failed scenario run or a silent regression after a software release costs operators and engineers’ real time.As our founding Site Reliability Engineer, you will design, build, and operate the infrastructure that makes this possible. You'll work at the intersection of hardware, simulation … them to permanent resolution, and implement the guardrails, monitoring, or process changes that prevent recurrence.Partner with development teams and stakeholders - review upcoming changes for reliability risk, surface concerns early, and implement mitigation strategies before releases land in the simulation environment.Monitor overall system health - instrument and watch the environment, triage ...

Senior Site Reliability Engineer — Hybrid Kubernetes & GitOps

Location
Greater London, England, United Kingdom
Cisco’s Webex Engineering Group in London is seeking a Senior Site Reliability Engineer to own the design, deployment, and operation of Kubernetes-based microservices, delivering reliability and scalable deployments in a hybrid work environment. You will drive GitOps workflows with Argo CD, use Helm ...

Site Reliability Engineer: Scale Global Services

Location
Greater London, England, United Kingdom
Apple Services Engineering is seeking a Site Reliability Engineer to own and improve the massive-scale infrastructure powering App Store, Apple TV, Apple Music, and more. You’ll work across Linux-based systems, build automated tooling, and collaborate with developers to deliver reliable services. Responsibilities include troubleshooting ...

Hybrid Site Reliability Engineer — Healthcare Infra & Automation

Location
Greater London, England, United Kingdom
Cranial Technologies is seeking a Site Reliability Engineer to help build and maintain scalable, secure healthcare technology systems. The role focuses on automation, observability, and reducing manual operations to keep clinical and business processes reliable. You will support production apps, APIs, and integrations, while collaborating with ...

Lead Site Reliability Engineer: Architect Resilience & AI Ops

Location
City of Westminster, England, United Kingdom
JPMorgan Chase & Co. in London seeks a Lead Site Reliability Engineer to shape the future for a globally recognized firm. You will lead resiliency reviews, break complex problems into actionable work for engineers, and serve as technical lead for medium to large products. As part … Infrastructure Platforms team, you will guide incident response, mentor peers, and drive AI-enabled reliability workflows across the SDLC while ensuring security and traceability throughout. #J-18808-Ljbffr ...

Site Reliability Engineer

Location
Greater London, England, United Kingdom
iGaming company based in London. Estimated Benefits Health Insurance Pension Stock Options Benefits estimated based on industry standards We’re hiring a Site Reliability Engineer to join our London team This is a fantastic opportunity for someone passionate about reliability, scalability and automation. ...

Site Reliability Engineer – Scalable Trading Infra

Location
City Of London, England, United Kingdom
Selby Jennings partners with a world-class hedge fund in London to hire a Site Reliability Engineer who bridges software engineering and infrastructure. You will design scalable, automated platforms, improve observability, and ensure production systems achieve top-tier availability for trading and research environments. Working with software ...

Lead Engineer, Site Reliability Engineering

Location
Greater London, England, United Kingdom
## Our TeamWe are evolving our Reliability Engineering team to move beyond support and operations. As a Senior Engineer in Site Reliability, you will be part of a diverse and inclusive organization that has full ownership of the availability, performance, and scalability … purpose. Write automation to scale systems sustainably, prevent service issues, or when they occur, quickly recover service. Partner with development teams to improve system reliability, observability, and release velocity. Participate in on-call rotations, incident response, postmortems, and root cause analysis and resolution. Be a vocal advocate of strong ...

SRE Engineer – FinTech Reliability, Observability & Cloud

Location
Greater London, England, United Kingdom
Hamilton Barnes Associates Limited is seeking a Site Reliability Engineer to work at the intersection of software engineering and infrastructure. You'll develop internal platforms, tooling, and automation across Linux, distributed systems, and cloud-native technologies to improve reliability and operational efficiency in a global production ...

Site Reliability Engineer , Cryptography, Access and Identity Services

Hiring Organisation
AmazonWebServices
Location
London, United Kingdom
Salary
£ 80 K
with the latest cloud computing technologies and becoming a core part of the largest cloud infrastructure on the planet AWS is seeking a Systems Engineer to build and operate services for our customers, to automate service operations and deployment methods, and provide mentoring for junior team members. You will … feel supported in the workplace and at home, there’s nothing we can’t achieve. Basic qualifications- Experience in site reliability engineering (SRE), systems engineering, systems administration, DevOps, security administration, or network administration- Experience working with Linux- Experience in systems engineering- Experience in any of the following: Python ...

Senior SRE Engineer: Reliability, Cloud & Automation

Location
Greater London, England, United Kingdom
London Stock Exchange Group is looking for a Senior Engineer in Site Reliability who will join a driven team focused on system availability, performance, and scalability. Responsibilities include maintaining service level objectives, writing automation for system resilience, and partnering with development teams. Required qualifications include a Bachelor … computer science, experience in Object Oriented programming and cloud systems, and DevOps familiarity. The role is pivotal in ensuring 24/7 system reliability and promoting engineering best practices. #J-18808-Ljbffr ...

Site Reliability Engineer

Hiring Organisation
DeepL
Location
London, United Kingdom
Salary
£ 80 K
well-being. Discover more about life at DeepL onLinkedIn,Instagram, and our Blog.Meet the team behind this journeyWe currently have two teams in the SRE track, and some SRE peers in other parts of the business. The in-track teams work closely, often collaborating.The first team SRE: Excellence focussed … making it easier to use our systems and provide tools and services to help run and monitor our products.The second team SRE: Accelerate works closely with our product development teams to use these as effectively as possible and embed better practice in teams. Both teams help ensure the services remain ...

Senior Site Reliability Engineer

Location
Greater London, England, United Kingdom
more about life at DeepL on LinkedIn, Instagram, and our Blog. Meet the team behind this journey We currently have two teams in the SRE track, and some SRE peers in other parts of the business. The in-track teams work closely, often collaborating. The first team SRE: Excellence focussed … making it easier to use our systems and provide tools and services to help run and monitor our products. The second team SRE: Accelerate works closely with our product development teams to use these as effectively as possible and embed better practice in teams. Both teams help ensure the services ...

Oracle-Focused SRE: Secure Cloud Reliability Engineer

Location
Greater London, England, United Kingdom
partners with a leading technology and engineering organisation delivering secure, innovative solutions to the UK Government and Defence sector. We are seeking an experienced Site Reliability Engineer with strong Oracle experience to join a skilled team responsible for the reliability and performance of production systems. … will work across application support, infrastructure, automation and reliability, building and maintaining CI/CD pipelines, and #J-18808-Ljbffr ...

Director of Site Reliability Engineering

Location
Greater London, England, United Kingdom
will influence engineering standards, enhance operational frameworks, and foster a culture of continuous improvement across mission‐critical environments. Responsibilities Lead and scale a global SRE organization, focusing on engineering excellence and team empowerment Collaborate with product, platform, operations, and security teams to embed reliability within SDLC practices Define … deliver systemic improvements across production environments Establish observability strategies with standardized tooling for metrics, logs, and tracing to support distributed systems Adopt and enforce SRE practices, including SLIs, SLOs, SLAs, and error budgets across services Drive resilience strategies with highly available architectures and disaster recovery readiness Champion an automation‐first ...

Entry Level - Site Reliability Engineer - (Remote - United Kingdom)

Location
Greater London, England, United Kingdom
growing as engineers, and having fun in a collaborative environment. Do you want to build and manage scaleable, self-healing, globally-distributed systems? Our Site Reliability engineers keep Yelp fast, available, and growing, connecting users to great local businesses. No matter how many times we get searched, scraped … solutions don’t work at our scale and contribute upstream to open source projects. Participate in light on-call rotations - we have geographically distributed SRE teams for follow-the-sun support, which means nobody needs to be on-call 24h a day! What It Takes To Succeed Familiarity with Linux ...

Site Reliability Engineer III: Scale & Automation

Location
Greater London, England, United Kingdom
Google London, UK, is seeking a Mid-level Software Engineer III in Site Reliability Engineering for the GCE AI team. You will help build and run large‐scale, fault‐tolerant systems, with emphasis on reliability, uptime, and performance. Expect collaboration across software and systems teams, mentoring ...

site reliability engineer

Location
Greater London, England, United Kingdom
digital engineering, cloud, and AI-enabled transformation services, focusing on complex software product development and digital platform engineering. Задачи Lead and scale a global SRE organization, focusing on engineering excellence and team empowerment Collaborate with product, platform, operations, and security teams to embed reliability within SDLC practices Define … deliver systemic improvements across production environments Establish observability strategies with standardized tooling for metrics, logs, and tracing to support distributed systems Adopt and enforce SRE practices, including SLIs, SLOs, SLAs, and error budgets across services Drive resilience strategies with highly available architectures and disaster recovery readiness Champion an automation-first ...

Senior Site Reliability Engineer

Hiring Organisation
Carta
Location
London, United Kingdom
Salary
£ 80 K
equity ownership for more people in more places. We believe that the problems we solve today unlock the opportunities of tomorrow. As a SeniorSite Reliability Engineer, you’ll work to: Build and scale our internal platform offerings (compute, storage and networking services) to ensure the reliability … powers the entire company. We’re looking for strong communicators who enjoy collaborating to solve complex problems. Familiarity with infrastructure best practices on performance, reliability and security and their associated tools is appreciated. Our stack is Python, Java, Terraform, gRPC, Docker, Kubernetes, Postgres, running on AWS. Come join ...

Software Engineer III, Site Reliability Engineering, GCE AI

Location
Greater London, England, United Kingdom
Computer Science or Engineering. 2 years of experience designing, analyzing, and troubleshooting large-scale distributed systems. About The Job Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed, fault-tolerant systems. SRE ensures that Google Cloud's services—both … internally critical and our externally-visible systems—have reliability, uptime appropriate to customer's needs and a fast rate of improvement. Additionally SRE’s will keep an ever-watchful eye on our systems capacity and performance. Much of our software development focuses on optimizing existing systems, building infrastructure ...

Software Engineer III, Site Reliability Engineering, GCE AI

Hiring Organisation
Google
Location
London, United Kingdom
Salary
£ 80 K
degree in Computer Science or Engineering. 2 years of experience designing, analyzing, and troubleshooting large-scale distributed systems. Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed, fault-tolerant systems. SRE ensures that Google Cloud's services—both our internally … critical and our externally-visible systems—have reliability, uptime appropriate to customer's needs and a fast rate of improvement. Additionally SRE’s will keep an ever-watchful eye on our systems capacity and performance. Much of our software development focuses on optimizing existing systems, building infrastructure and eliminating ...

AVP, Observability & SRE Engineer

Location
Greater London, England, United Kingdom
Citi is seeking a Site Reliability Engineer - Assistant Vice President in London to drive end‐to‐end observability, migrate legacy monitoring to Google Cloud Observability and Grafana, and implement OpenTelemetry instrumentation across OpenShift/Kubernetes environments. The role emphasizes hands‐on deployment, automation (Ansible/Terraform ...