8 of 8 Observability Jobs in the City of Westminster

Principal/Senior Site Reliability Engineer

Hiring Organisation
Jobleads-UK
Location
City of Westminster, England, United Kingdom
systems available worldwide. You strengthen reliability through chaos engineering running experiments that validate systems and surface weaknesses before they become incidents. You build deep observability with monitoring, logging, and alerting frameworks such as Prometheus, Grafana, Datadog, and ELK. You provide technical leadership to a team of engineers, fostering collaboration, innovation ...

Staff Software Engineer- Loyalty

Hiring Organisation
Jobleads-UK
Location
City of Westminster, England, United Kingdom
Delivery and Engineering teams to solve complex business challenges and deliver customer-focused outcomes. Owning operational stability by promoting a production-first mindset, improving observability, maintaining service reliability and advancing engineering roadmaps. Who you are Strong software engineering experience across multiple layers of the technology stack, including back-end systems ...

Principal Software Engineer - International

Hiring Organisation
Jobleads-UK
Location
City of Westminster, England, United Kingdom
objectives, engineering standards, governance requirements and long‐term architectural direction. Deep knowledge of cloud and platform engineering, including cloud‐native architectures, Kubernetes, networking, security, observability, reliability engineering and automation practices. Strong understanding of quality engineering, test strategies, testing quadrants, testing pyramids, continuous integration, continuous delivery and DevOps ownership models. Excellent ...

Software Engineering Manager - Store Operations

Hiring Organisation
Jobleads-UK
Location
City of Westminster, England, United Kingdom
desirbale Tech Stack M&S uses a variety of technologies including; Java, Spring, SpringBOOT, Micronaut React, Next.js, Typescript, Angular Azure Cloud, Kubernetes, Dynatrace (observability) SQL Server, MongoDB Ignite, Redis What’s In It For You Working at M&S means being part of something bigger — helping to deliver quality, value ...

Senior Site Reliability Engineer – Cloud, MLOps & HPC

Hiring Organisation
Jobleads-UK
Location
City of Westminster, England, United Kingdom
systems and accelerate scientific discovery. You will architect IaC using Terraform, Pulumi or CloudFormation, ensure DR with auto-scaling, run chaos experiments, and build observability with Prometheus, Grafana, Datadog and the ELK stack. #J-18808-Ljbffr ...

AI Platform Architect – Internal Dev Platform Lead

Hiring Organisation
Jobleads-UK
Location
City of Westminster, England, United Kingdom
build, deploy, operate and improve software safely at scale, creating self-service capabilities and governance guardrails. You will design compute, Kubernetes, networking, secrets, observability, and IaC, while driving reliability with SLOs and incident learnings. #J-18808-Ljbffr ...

Principal AI Platform Engineer

Hiring Organisation
Jobleads-UK
Location
City of Westminster, England, United Kingdom
reusable golden paths, and standard service templates to simplify service provisioning and operations.Contribute to cloud-native platform architecture, including compute, Kubernetes, networking, secrets management, observability, CI/CD, and infrastructure as code.Integrate AI-assisted engineering workflows to support faster delivery, improved code quality, automation, and data-driven operational decisions.Establish governance … security, and compliance guardrails through policy-as-code and auditable platform patterns.Improve platform reliability using SLOs, observability practices, resilience engineering, and insights from incidents.Collaborate with product, engineering, security, and architecture teams to align platform capabilities with business priorities and user needs.Drive efficiency and sustainability through automation, standardisation, and FinOps-informed ...

Site Reliability Engineering Manager

Hiring Organisation
Jobleads-UK
Location
City of Westminster, England, United Kingdom
Reliability Engineers. Shape and deliver our Site Reliability Engineering roadmap alongside the Head of Platform. Champion modern engineering practices including SLIs, SLOs, error budgets, observability and automation. Improve the reliability, scalability and performance of our cloud platforms and digital services. Partner with Engineering, Security, Data and Product teams to embed … operational excellence from design through to production. Drive the adoption of our observability platform, helping teams gain deeper insight into the health and performance of their services. Lead incident learning, continuous improvement and automation initiatives that reduce operational toil. Provide technical leadership across AWS, Kubernetes, Infrastructure as Code, CI/ ...