5 of 5 Application Performance Monitoring Jobs in London

Observability Architect - 12 Month FTC

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
observability strategy for the co-location migration programme. Ensure logging, metrics, tracing, alerting, and operational dashboards provide comprehensive visibility across the new infrastructure and application estate, enabling the successful migration of the tightly-coupled monolithic platform with minimal operational risk. Identify gaps in the existing observability capability and recommend … Responsibilities Observability Assessment & Strategy Review the current observability architecture across infrastructure, networks, middleware, databases, and applications. Assess existing logging, metrics, distributed tracing, and monitoring capabilities to determine readiness for the co-location migration. Develop an observability strategy that supports both migration activities and long-term operational support. Recommend enhancements ...

Site Reliability Engineering (SRE) / Observability Technical Lead

Hiring Organisation
NTT
Location
London, United Kingdom
Salary
£ 80 K
team and drive the strategy and execution of observability and reliability projects across our clients. The ideal candidate will have deep expertise in Application Performance Monitoring (APM), Infrastructure as Code (IaC), automation, and distributed tracing using OpenTelemetry. As a lead, you will guide the design, implementation … continuous improvement of observability solutions, ensuring system reliability, performance, and scalability while fostering best practices in SRE and DevOps. What you'll be doing: Lead the strategic development and management of observability and reliability frameworks across the organization, ensuring alignment with business goals and technical requirements.Design and implementation ...

Site Reliability Engineer

Hiring Organisation
Inspire People
Location
South West London, London, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£55,000
team that ensures BIST's digital services work as users expect, working with development teams giving them the tools for their job, including application performance monitoring, exception, log and metrics aggregation, dashboards, and declarative CD/CI pipelines. £45,835 to £60,993 (including allowances) London … Belfast. As a Site Reliability Engineer, you will pro-actively engage development teams and use initiative to develop the tools for their job, including application performance monitoring, exception, log and metrics aggregation, dashboards, and declarative CI/CD (continuous integration/continuous delivery) pipelines. You'll work ...

Sr Software Engineer I

Hiring Organisation
American Express
Location
London, United Kingdom
Salary
£ 80 K
GitHub Actions, Cloud Build, Jenkins, and Docker/Kubernetes-based CI/CD pipelines. Design resilient, secure, and observable distributed systems using OpenTelemetry, Cloud Monitoring, Prometheus, Grafana, Datadog, and centralized logging solutions. Collaborate closely with Product, UX, Architecture, and Platform Engineering teams to deliver scalable, secure, and high-quality … such as Jest, React Testing Library, Playwright, Cypress, Go testing, and Postman. Experience implementing observability and operational excellence using OpenTelemetry, Prometheus, Grafana, Google Cloud Monitoring, Cloud Logging, and application performance monitoring tools. Strong understanding of application security, cloud security best practices, OWASP principles, secure software ...

Senior Platform Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
this is the right place. About the Role We are seeking highly skilled and experienced Platform Product Engineers to join our Security, Infrastructure and Performance team. This is a crucial, dual‐faceted role that combines high‐level engineering strategy with hands‐on operational excellence. The successful candidates will … Engineering (SRE) mindset. The successful candidates will be instrumental in defining and upholding Service Level Objectives (SLOs) and Service Level Indicators (SLIs), implementing effective monitoring and alerting strategies, and leading operational incident response processes. What you'll do The core responsibility is to implement, maintain, and continuously improve ...