251 to 275 of 597 Site Reliability Engineering Jobs in London

Senior SRE Engineer: Reliability, Cloud & Automation

Location
Greater London, England, United Kingdom
London Stock Exchange Group is looking for a Senior Engineer in Site Reliability who will join a driven team focused on system availability, performance, and scalability. Responsibilities include maintaining service level objectives, writing automation for system resilience, and partnering with development teams. Required qualifications include a Bachelor … computer science, experience in Object Oriented programming and cloud systems, and DevOps familiarity. The role is pivotal in ensuring 24/7 system reliability and promoting engineering best practices. #J-18808-Ljbffr ...

AI-Powered Production Intelligence Tech Lead

Location
Greater London, England, United Kingdom
Cisco is seeking a Technical Leader to drive architectural vision for an AI-powered Production Intelligence platform. You will blend Site Reliability Engineering with agentic AI to improve monitoring, diagnosis, and auto-remediation across global SaaS infrastructure. You will lead architecture, mentor engineers, and partner across teams ...

Principal Platform Engineer

Location
Greater London, England, United Kingdom
impact***** ****Build and maintain infrastructure-as-code (Terraform or similar), CI/CD pipelines, and Kubernetes platforms as products consumed by engineering teams********SRE & reliability***** ****Define and drive SLOs, error budgets, and observability standards (metrics, logging, tracing) across the platform***** ****Take part in post-incident reviews and help … roadmap for the platform function, balancing reliability investment against delivery****## ****What we're looking for********Must have***** ****A background in Operations or SRE running highly available, redundant production platforms — you understand failure domains, graceful degradation, and what "five nines" costs***** ****Deep hands-on experience with AWS (VPC design ...

Senior SRE: GCP, Kubernetes & Automation Leader

Location
Greater London, England, United Kingdom
Brevan Howard CFD LTD is seeking a Senior Site Reliability Engineer (SRE) to enhance the reliability, scalability, and performance of its core platform. The successful candidate will provide operational support and lead infrastructure projects. This role requires hands-on experience with Google Cloud Platform (GCP) and Kubernetes ...

Site Reliability Engineer- Spacetime UK

Location
Greater London, England, United Kingdom
Role Overview This isn't a "keep the lights on" SRE role. This is a strategic, high-impact opportunity to build the nervous system for a platform that transforms how networks of satellites, ground stations, and fleets are interconnected and orchestrated. You will be building the core observability stack that … cloud-native tools to a robust, scalable, and insightful platform built on best-in-class technologies (Prometheus, OpenTelemetry, etc.). If you are an SRE who thrives on platform-building challenges and wants to be relied upon to build a production-grade observability stack from the ground up, this role ...

Senior Data Platform SRE — Hybrid, Obs & Reliability Leader

Location
Greater London, England, United Kingdom
Outerlimit is seeking a Senior Site Reliability Engineer to embed within the Data Engineering team. You will own reliability, performance, and operability … data platform, building observability from the ground up and leading incident response for data outages. You'll balance feature velocity with system stability, applying SRE practices and capacity planning to help the team ship more reliably and efficiently. #J-18808-Ljbffr ...

Lead Software Engineer – DevSecOps & Platform Automation

Location
Greater London, England, United Kingdom
seeking a Lead Software Engineer (DevSecOps) to drive the hands‐on delivery of automated, secure, and reliable software delivery capabilities across our product engineering organization. This role is responsible for building, operating, and continuously improving CI/CD pipelines, infrastructure automation, and embedded security controls that enable product teams … checks, and automated rollback signals; Implement and report on DORA and deployment reliability metrics (deployment frequency, lead time, change failure rate, MTTR); Support SRE‐aligned practices such as error budgets, runbooks, and post‐incident reviews; Participate in incident response and on‐call rotation for delivery and platform services, driving ...

Site Reliability Engineer

Hiring Organisation
GoCardless
Location
London, UK
Employment Type
Full-time
follow us on LinkedIn @GoCardless. Mollie is the leading payments and financial services partner for business, rooted in Europe, with global reach. Platform Engineering at GoCardlessThe Platform Engineering team is a unified, globally distributed team. We are currently located in London, Riga, and Lisbon. We work with … other engineering teams to empower them to build, release, run and scale their products. We operate with a strong focus on both strategic project delivery and operational excellence. Our work is a balance of: Project Delivery: Building new platform components, from initial design to deployment. Operational Support: Ensuring ...

Senior SRE

Hiring Organisation
Pigment
Location
London, UK
Employment Type
Full-time
innovation and ready to make an impact at scale, we'd love to hear from you. The opportunityWe are looking for a Senior SRE profile who will design and implement the Pigment infrastructure for tomorrow. Pigment is a technically challenging platform. It calculates and synchronizes large datasets that must … aggregated on-demand through our formula engine while rendering live on our front end. Here are the challenges to be tackled as an SRE :Define and build the infrastructure needed to answer our performance challenges, automate it, and make it scalable. In particular, ensure that the infrastructure scales ...

Senior DevOps Platform Engineer

Hiring Organisation
LEAP29
Location
London, UK
Employment Type
Full-time
Senior DevOps Platform EngineerLondon We are currently supporting a leading organisation looking to hire an experienced Senior DevOps Platform Engineer to join their cloud engineering function. This is an excellent opportunity for a highly skilled DevOps professional to work on complex cloud-native environments, designing and automating scalable platforms … Kubernetes-based architecturesImplement GitOps delivery models using tools such as ArgoCD and HelmDevelop reusable automation frameworks for infrastructure provisioning, configuration management and deploymentsImprove platform reliability, scalability and security through automation and engineering best practicesSupport cloud migrations from traditional data centre environments into modern cloud-native platformsImplement monitoring, observability ...

Senior AWS Site Reliability Engineer

Hiring Organisation
Spectrum IT Recruitment
Location
City of London, London, United Kingdom
Employment Type
Permanent
Salary
£60000 - £70000/annum Bonus, Medical Care
credentials Do You Have What It Takes? 3-6 years of hands-on experience in a similar role, with a strong emphasis on systems engineering, automation, and service reliability Proficient in at least one programming language such as Python, Go, Java, or C#, along with scripting skills … PowerShell Solid grasp of cloud platforms like AWS, including an understanding of how core services like EC2, ECS, Lambda, and DynamoDB operate under reliability constraints Practical experience using infrastructure-as-code tools like CloudFormation or Terraform In-depth knowledge of CI/CD principles and hands-on experience with ...

SRE Engineer – Hybrid Cloud Reliability & Automation

Location
Greater London, England, United Kingdom
Site Reliability Engineer to drive reliability, performance and operational excellence across hybrid cloud and on‐prem environments. You will shape SRE practices, support incidents, embed automation and ensure ISO 27001 security alignment. This hands‐on role has significant influence across engineering, security and operations teams, with ...

Engineering Environments Lead

Hiring Organisation
Adecco
Location
London, United Kingdom
Employment Type
Contract
environment estate. Work closely with Engineering, Product, Platform and Infrastructure teams to support delivery at scale. Champion modern engineering practices including DevOps, SRE and Environment as Code. Identify opportunities for continuous improvement and operational excellence. Experience Required Proven experience leading engineering or enterprise environment functions within … strategy, governance and operational management. Experience delivering automation, Infrastructure as Code and self-service initiatives. Good understanding of cloud platforms, platform engineering and SRE principles. Experience driving resilience, disaster recovery and service continuity programmes. Strong stakeholder management skills with the ability to engage and influence senior leaders. Track record ...

AWS DevOps Engineer

Hiring Organisation
Opus Recruitment Solutions
Location
London, United Kingdom
Employment Type
Contract
Contract Rate
£500/day Inside IR35
growing technology team focused on supporting and improving cloud infrastructure and production applications. This is a hands-on role with a strong DevOps/SRE focus, working across AWS environments, automation, operational stability and continuous improvement, with some exposure to C#/.NET development . What you'll do: Support … application issues Contribute to C#/.NET development when required What we're looking for: 5+ years' experience in DevOps, Cloud Engineering or SRE Strong hands-on AWS expertise Experience supporting production environments Automation and CI/CD experience C#/.NET experience desirable but not essential A passion ...

Principal Platform Security Engineer

Location
Greater London, England, United Kingdom
Type:**Permanent**Build a brilliant future with Hiscox****Position**: Principal Platform Security Engineer, London MarketAs a senior and influential member of the London Platform Engineering Chapter, the Principal Platform Security Engineer will set direction and lead by example in maturing our platform security practices. You will guide multiple Innovation … squads and Engineering Chapters, driving cloud-first adoption and championing secure-by-design initiatives.You will be an integral member of a Chapter spanning Platform, DevOps, and Site Reliability Engineers, and a core contributor to a Platform Engineering squad focused on continuous improvement across our cloud ...

Principal Platform Security Engineer

Hiring Organisation
Hiscox AG
Location
London, UK
Employment Type
Full-time
Type: PermanentBuild a brilliant future with HiscoxPosition: Principal Platform Security Engineer, London MarketAs a senior and influential member of the London Platform Engineering Chapter, the Principal Platform Security Engineer will set direction and lead by example in maturing our platform security practices. You will guide multiple Innovation squads … Engineering Chapters, driving cloud-first adoption and championing secure-by-design initiatives. You will be an integral member of a Chapter spanning Platform, DevOps, and Site Reliability Engineers, and a core contributor to a Platform Engineering squad focused on continuous improvement across our cloud ...

Operations Engineering Lead

Hiring Organisation
Willis Towers Watson
Location
London, UK
Employment Type
Full-time
Ready to shape the infrastructure that powers engineering at scale? As our Engineering Lead - Infrastructure Operations, you'll lead the platforms, tooling, and operations that keep our teams shipping safely and efficiently. Here at Cushon by WTW, we like to do things a bit differently... our mission … practices enabling teams to support anything they build in their domains. Building FinOps frameworks and monitoring, focussing on cost optimizationTechnical experienceExperience in Operations, DevSecOps, SRE, or Cloud Infrastructure roles in fast-growing, regulated environmentsProven experience leading or acting as technical lead for a platform or operations teamStrong AWS infrastructure experience ...

Front-Office SRE Lead: Observability & AI-Driven Reliability

Location
Greater London, England, United Kingdom
JPMorgan Chase & Co. in London is seeking a Lead Site Reliability Engineer to shape next‐gen SRE patterns, observability, and reliability across globally distributed trading systems. You will partner with front‐office traders, contribute production code (Java/Python/Kotlin), drive incident response, and lead … assisted reliability initiatives while collaborating with infrastructure, cloud, and security teams. #J-18808-Ljbffr ...

Staff Software Engineer, Infrastructure - Python & Kubernetes

Location
Greater London, England, United Kingdom
headquartered in the United Kingdom, with offices in London, New York, and Singapore and an expanding presence in the Bay Area. Staff Software Engineer – SRE and Core Infra London (Hybrid) | Engineering | Full Time The Role PhysicsX is growing rapidly and so is the infrastructure that underpins our platform. … complex engineering workloads, the foundational infrastructure layer becomes ever more critical. We are looking for a Senior Software Engineer to join our Platform SRE Core Infrastructure team. This role is responsible for the design, provisioning and operation of the shared infrastructure that the entire PhysicsX platform depends on. ...

SRE Engineer – FinTech Reliability, Observability & Cloud

Location
Greater London, England, United Kingdom
Hamilton Barnes Associates Limited is seeking a Site Reliability Engineer to work at the intersection of software engineering and infrastructure. You'll develop internal platforms, tooling, and automation across Linux, distributed systems, and cloud-native technologies to improve reliability and operational efficiency in a global production ...

Senior Site Reliability Engineer

Location
Greater London, England, United Kingdom
production environment by monitoring availability and taking a holistic view of system health Build software and systems to manage platform infrastructure and applications Improve reliability, quality, and time-to-market of our suite of software solutions Measure and optimize system performance, with an eye toward pushing our capabilities forward … getting ahead of customer needs, and innovating to continually improve Provide primary operational support and engineering for multiple large distributed software applications How will you make an impact? Gather and analyze metrics from both operating systems and applications to assist in performance tuning and fault finding Partner with development ...

Golang Lead Engineer

Location
City Of London, England, United Kingdom
lead the technical development and delivery of backend services that are reliable, secure, scalable and easy to operate! This role sits within Software Engineering and offers the opportunity to shape the Group's Mobile Security Core and wider backend capabilities. You'll combine hands-on engineering with technical … skills, with the ability to communicate technical trade-offs clearly and influence outcomes. The Skills Taxonomy identifies relevant skills including Software Engineering, DevOps, SRE & Service Engineering, Technology Leadership, Critical Thinking, Decision Making, Coaching & Feedback and Stakeholder Management. And any experience of this would be really useful. Experience Experience ...

SC Cleared DevOps Engineer

Hiring Organisation
Sanderson Recruitment
Location
London, United Kingdom
Employment Type
Contract
Contract Rate
Up to £570 per day + Inside IR-35
must have been used within the last 12 months, with a minimum of 3 years remaining before expiry. This is a hands-on platform engineering role focused on creating secure, resilient and highly automated AWS environments that enable teams to develop, test and deploy software efficiently while maintaining strong … platform engineering best practices. What We're Looking For * Strong AWS cloud engineering and platform experience. * Background in DevOps, Platform Engineering, SRE or Cloud Infrastructure roles. * Experience with Infrastructure as Code, ideally Terraform, CloudFormation or AWS CDK. * Strong CI/CD pipeline experience. * Experience with containerisation technologies ...

Senior Nework Programmer - Algo trading

Hiring Organisation
Quant Capital
Location
London, UK
Employment Type
Full-time
Network SRE – 250,000-300,000 total compensation – 4 days in officeQuant Capital is urgently looking Network SRE for our high profile client. Our client is a leading quantitative trading company and liquidity provider. Their focus on technology has allowed them to deeply penetrate the market and gain market share. … Shared Engineering team that focuses on designing, developing, and maintaining infrastructure and tools. The team requires a Network Site Reliability Engineer (SRE) with strong network fundamentals, problem-solving skills, and a keen interest in diverse tools and techniques. The role involves collaborative work across various teams, exploring ...

MLOps Engineer

Location
Greater London, England, United Kingdom
2025. Learn more at www.coreweave.com. We're proud to be a Living Wage accredited Employer. What You'll Do CoreWeave’s Physical AI Platform Engineering team builds and scales the data and workflow backbone powering advanced engineering simulation and AI workflows. Our ambition is to become the super … engineers on production‐grade ML practices. Who You Are 5–6+ years of professional experience in MLOps, ML platform engineering, ML infrastructure, or SRE/DevOps for production machine learning systems. Proven experience building, operating, and automating production ML pipelines covering experiment tracking, model registries, artifact versioning, dataset management ...