51 to 75 of 106 Site Reliability Engineer Jobs in London

Senior Site Reliability Engineer — Hybrid Kubernetes & GitOps

Location
Greater London, England, United Kingdom
Cisco’s Webex Engineering Group in London is seeking a Senior Site Reliability Engineer to own the design, deployment, and operation of Kubernetes-based microservices, delivering reliability and scalable deployments in a hybrid work environment. You will drive GitOps workflows with Argo CD, use Helm ...

Hybrid Site Reliability Engineer — Healthcare Infra & Automation

Location
Greater London, England, United Kingdom
Cranial Technologies is seeking a Site Reliability Engineer to help build and maintain scalable, secure healthcare technology systems. The role focuses on automation, observability, and reducing manual operations to keep clinical and business processes reliable. You will support production apps, APIs, and integrations, while collaborating with ...

Lead Site Reliability Engineer: Architect Resilience & AI Ops

Location
City of Westminster, England, United Kingdom
JPMorgan Chase & Co. in London seeks a Lead Site Reliability Engineer to shape the future for a globally recognized firm. You will lead resiliency reviews, break complex problems into actionable work for engineers, and serve as technical lead for medium to large products. As part … Infrastructure Platforms team, you will guide incident response, mentor peers, and drive AI-enabled reliability workflows across the SDLC while ensuring security and traceability throughout. #J-18808-Ljbffr ...

Site Reliability Engineer

Location
Greater London, England, United Kingdom
iGaming company based in London. Estimated Benefits Health Insurance Pension Stock Options Benefits estimated based on industry standards We’re hiring a Site Reliability Engineer to join our London team This is a fantastic opportunity for someone passionate about reliability, scalability and automation. ...

IAM Secrets Management Engineering - SRE Platform Engineer - VP - London London · United Kingdo[...]

Location
Greater London, England, United Kingdom
Secrets Management Engineering - SRE Platform Engineer - VP - London location_on London, Greater London, England, United Kingdom Who We Are Goldman Sachs is a leading global investment banking, securities and investment management firm that provides a wide range of services worldwide to a substantial and diversified client base that includes … institutions, governments and high net‐worth individuals. The Role We are seeking a skilled and experienced Lead Site Reliability Platform Engineer (SRE) to join our team. The ideal candidate will be responsible for ensuring the reliability, performance, and scalability of mission‐critical, high‐availability, high‐throughput ...

Engineer - Site Reliability Engineering

Location
Greater London, England, United Kingdom
TeamWe are evolving our Reliability Engineering team to move beyond support and operations. As a Senior Engineer in Site Reliability, you will be part of a diverse and inclusive organization that has full ownership of the availability, performance, and scalability of one of the most critical … purpose.* Write automation to scale systems sustainably, prevent service issues, or when they occur, quickly recover service.* Partner with development teams to improve system reliability, observability, and release velocity.* Participate in on-call rotations, incident response, postmortems, and root cause analysis and resolution.* Be a vocal advocate of strong ...

SRE Engineer

Location
London, United Kingdom
preferred supplier to one of our biggest Clients, I am seeking for a SRE Engineer (Datadog) for a position in York (UK). Key Responsibilities: Experience: 10+ years of hands-on experience in Site Reliability Engineering (SRE), DevOps, or Systems Architecture, with at least 3+ years specializing … maintain compliance postures and mitigate runtime threats. Collaboration & Enablement Cross-Functional Mentorship: Act as the go-to escalation point and technical mentor for DevOps, SRE, and Software Engineering teams regarding troubleshooting and instrumentation. Training & Documentation: Create internal documentation, runbooks, and training modules to elevate organizational proficiency in observability. Vendor Management ...

SRE Engineer – FinTech Reliability, Observability & Cloud

Location
Greater London, England, United Kingdom
Hamilton Barnes Associates Limited is seeking a Site Reliability Engineer to work at the intersection of software engineering and infrastructure. You'll develop internal platforms, tooling, and automation across Linux, distributed systems, and cloud-native technologies to improve reliability and operational efficiency in a global production ...

Site Reliability Engineer , Cryptography, Access and Identity Services

Hiring Organisation
AmazonWebServices
Location
London, UK
Employment Type
Full-time
with the latest cloud computing technologies and becoming a core part of the largest cloud infrastructure on the planet? AWS is seeking a Systems Engineer to build and operate services for our customers, to automate service operations and deployment methods, and provide mentoring for junior team members. You will … feel supported in the workplace and at home, there's nothing we can't achieve. Basic qualifications- Experience in site reliability engineering (SRE), systems engineering, systems administration, DevOps, security administration, or network administration- Experience working with Linux- Experience in systems engineering- Experience in any of the following: Python ...

Senior SRE Engineer: Reliability, Cloud & Automation

Location
Greater London, England, United Kingdom
London Stock Exchange Group is looking for a Senior Engineer in Site Reliability who will join a driven team focused on system availability, performance, and scalability. Responsibilities include maintaining service level objectives, writing automation for system resilience, and partnering with development teams. Required qualifications include a Bachelor … computer science, experience in Object Oriented programming and cloud systems, and DevOps familiarity. The role is pivotal in ensuring 24/7 system reliability and promoting engineering best practices. #J-18808-Ljbffr ...

Senior Site Reliability Engineer

Location
Greater London, England, United Kingdom
more about life at DeepL on LinkedIn, Instagram, and our Blog. Meet the team behind this journey We currently have two teams in the SRE track, and some SRE peers in other parts of the business. The in‐track teams work closely, often collaborating. The first team SRE: Excellence focussed … making it easier to use our systems and provide tools and services to help run and monitor our products. The second team SRE: Accelerate works closely with our product development teams to use these as effectively as possible and embed better practice in teams. Both teams help ensure the services ...

Site Reliability Engineer (SRE)

Hiring Organisation
Virtusa
Location
London, UK
Employment Type
Full-time
Description"SRE Observability EngineerKey Responsibilities The Monitoring and Observability team is responsible for managing Operating with a global footprint. Collaborating across various organizations within Citi to understand and develop observabilitysolutions for enterprise-wide deployment at scale. Managing the legacy monitoring stack across the Production Management organization withinCiti. Driving the strategic ...

Head of Site Reliability Engineering (SRE)

Hiring Organisation
Blockchain
Location
London, UK
Employment Type
Full-time
than 40 million verified users, facilitating over $1 trillion in crypto transactions. We are looking for a Head of Site Reliability Engineering (SRE) who will serve as the principal leader in developing and executing our infrastructure reliability strategy that scales with the company as it continues … uptime and safety of Blockchain.com's needs. WHAT YOU WILL DOEstablish a multi-year platform and reliability strategy, defining the SRE roadmap and driving engineering standards for observability, automation, and production readinessBe a strong business enabler, allowing us to deliver products quickly whilst maintaining standards-compliance and security through ...

site reliability engineer

Location
Greater London, England, United Kingdom
digital engineering, cloud, and AI-enabled transformation services, focusing on complex software product development and digital platform engineering. Задачи Lead and scale a global SRE organization, focusing on engineering excellence and team empowerment Collaborate with product, platform, operations, and security teams to embed reliability within SDLC practices Define … deliver systemic improvements across production environments Establish observability strategies with standardized tooling for metrics, logs, and tracing to support distributed systems Adopt and enforce SRE practices, including SLIs, SLOs, SLAs, and error budgets across services Drive resilience strategies with highly available architectures and disaster recovery readiness Champion an automation-first ...

Senior Site Reliability Engineer

Hiring Organisation
Carta
Location
London, UK
Employment Type
Full-time
equity ownership for more people in more places. We believe that the problems we solve today unlock the opportunities of tomorrow. As a SeniorSite Reliability Engineer, you'll work to: Build and scale our internal platform offerings (compute, storage and networking services) to ensure the reliability … powers the entire company. We're looking for strong communicators who enjoy collaborating to solve complex problems. Familiarity with infrastructure best practices on performance, reliability and security and their associated tools is appreciated. Our stack is Python, Java, Terraform, gRPC, Docker, Kubernetes, Postgres, running on AWS. Come join ...

AVP, Observability & SRE Engineer

Location
Greater London, England, United Kingdom
Citi is seeking a Site Reliability Engineer - Assistant Vice President in London to drive end‐to‐end observability, migrate legacy monitoring to Google Cloud Observability and Grafana, and implement OpenTelemetry instrumentation across OpenShift/Kubernetes environments. The role emphasizes hands‐on deployment, automation (Ansible/Terraform ...

Site Reliability / Infrastructure Engineer - Defence & Government

Hiring Organisation
JLA Resourcing Ltd
Location
London, United Kingdom
Employment Type
Permanent, Work From Home
engineering foundation who enjoys building systems, solving complex production problems and improving infrastructure through code and automation. This would particularly suit someone from an SRE, Software Engineering or Infrastructure Engineering background who wants to work on technically challenging, high-impact projects. The Role You'll work across software and infrastructure … purely operational DevOps experience. Ideally, you'll have: Active UK SC clearance - this is a key requirement Around 1-3+ years' experience within SRE, Infrastructure, Platform or Software Engineering A Computer Science degree or similarly strong technical/software engineering background Hands-on Kubernetes experience AWS/cloud infrastructure ...

Site Reliability Engineer, Infrastructure - ThousandEyes

Location
Greater London, England, United Kingdom
deeply integrated across the Cisco technology portfolio, delivering AI‐powered assurance insights within Cisco’s Networking, Security, Collaboration, and Observability portfolios. Our distributed Site Reliability Engineering team of approximately nine engineers owns the availability, latency, performance, efficiency, monitoring, emergency response, and capacity planning of the platform while partnering … call rotation. Hands‐on experience with infrastructure‐as‐code tooling and codebases, preferably Terraform. Hands‐on experience leveraging AI as a force multiplier of SRE activities, such as automating toil away and improving operational efficiency. Professional experience administering and troubleshooting GNU/Linux systems, including system libraries, file systems, networking ...

Site Reliability Engineer, Infrastructure - ThousandEyes

Location
City Of London, England, United Kingdom
deeply integrated across the Cisco technology portfolio, delivering AI-powered assurance insights within Cisco’s Networking, Security, Collaboration, and Observability portfolios. Our distributed Site Reliability Engineering team of approximately nine engineers owns the availability, latency, performance, efficiency, monitoring, emergency response, and capacity planning of the platform while partnering … operational on-call rotation. Hands-on experience with infrastructure-as-code tooling and codebases, preferably Terraform. Hands-on experienceleveraging AIas a force multiplier of SRE activities, such as automati ng toil away and improving operational efficiency. Professional experience administering and troubleshooting GNU/Linux systems, including system libraries, file systems ...

Site Reliability Engineer - Private Cloud Compute

Location
Greater London, England, United Kingdom
approach to cloud intelligence, extending the security and privacy of Apple devices into the cloud to unlock even more intelligence for our users. This SRE team is responsible for the availability and automation of the critical systems and services that enable PCC to deliver cloud intelligence without compromising user privacy. … future of privacy-preserving cloud infrastructure at scale, this is the opportunity for you! Description We're looking for a hardworking and passionate SRE Engineer to join this amazing team. You will be an accomplished builder and problem-solver, eager to tackle challenging technical problems. You have a deep ...

Staff Platform Site Reliability Engineer

Hiring Organisation
Index Exchange
Location
London, UK
Employment Type
Full-time
users—it just works. Mentor and raise the bar. Coach engineers, foster a culture of engineering excellence, and collaborate across Cloud Platform Operations, SRE, Network, Security, and Software Engineering teams. What You BringYou've built and shipped platform infrastructure at scale—not just operated someone else's. You think … care as much about the developer experience of your platform as you do about its architecture. Must Have8+ years in platform engineering, SRE, infrastructure engineering, or DevOps. Deep experience with Linux internals: kernel tuning, network stack, system observability, security. Strong Kubernetes expertise: cluster lifecycle, networking, storage, RBAC, multi-cluster—across ...

Site Reliability Engineer – Privacy‐First Cloud at Scale

Location
Greater London, England, United Kingdom
Apple Inc. in London, England invites a hardworking SRE Engineer to join the Private Cloud Compute team. You will tackle challenging problems, own responsibilities for high-availability systems and contribute to global-scale infrastructure with privacy-centric principles. This role emphasizes automation, performance optimization and secure, scalable service delivery. ...

Technical Lead - Site Reliability Engineering

Location
Greater London, England, United Kingdom
Reliability Engineering capabilities to strengthen reliability, observability, security, and operational excellence across our Markets and Risk Intelligence division.As a **Technical Lead SRE**, you will be a senior hands‐on technical person help shape the foundations of reliability across both new and existing platforms. You will collaborate … person who is passionate about reliability engineering and who bring a continuous improvement approach to everything they do!Lead the establishment of SRE foundations for new projects building environments, monitoring, alerting, and ensuring operational readiness from day one.Collaborate with Architecture and Engineering teams to embed reliability, scalability, security ...

Principal Platform Engineer (SRE/Cloud)

Location
Greater London, England, United Kingdom
honesty ensuring our workforce is able to bring their full selves to work. ABOUT THE ROLE Principal Platform Engineers at Beamery solve the toughest reliability, scalability and infrastructure problems with the highest impact. Together they collaborate to set the standards for how Engineering will build, run and operate services … whole engineering organisation WHO ARE WE LOOKING FOR? We are seeking a hands-on technical leader with deep Site Reliability Engineering (SRE) and Cloud expertise who can set direction across the engineering organisation. Key skills/experience: A proven track record of designing and delivering scalable, reliable cloud ...

Lead Site Reliability Engineer

Hiring Organisation
Hackajob Ltd
Location
London, United Kingdom
Employment Type
Permanent
Salary
GBP 100,000 Annual
trading technology stack is undergoing a multi year convergence and modernization journey. You will play a pivotal role in shaping our next generation SRE patterns, reliability frameworks, observability strategy, and performance engineering capabilities across globally distributed systems click apply for full job details ...