176 to 200 of 223 Incident Management Jobs in London

Senior Backend Engineer | AI Platform

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
similar. Experience working with cloud platforms such as Google Cloud Platform (preferred), AWS, or Azure. Strong understanding of system reliability, observability, monitoring, and incident management. Experience with Infrastructure as Code and cloud-native architectures. Previous experience working within a Platform Engineering team is a strong plus. Key responsibilities: Design … methodologies. Build platform capabilities that enable teams to safely deploy, monitor, and iterate on AI-powered applications. Configure and maintain tracing, monitoring, observability, and incident management solutions to ensure platform reliability. Partner closely with product and engineering teams to understand their needs and provide scalable platform solutions. Continuously ...

SRE Architect: Cloud & Data Reliability Leader

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Engineering practice, transforming operations toward proactive, engineering-led reliability. You will define NFRs with FMEA-based methods, champion observability, self-healing automation, and automated incident management, while ensuring cost efficiency and continuous improvement. Responsibilities include defining NFRs, building self-healing automation, and leading DB automation with UK release ...

Senior SRE: Cloud Reliability (Azure/AWS, Terraform, K8s)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
. You will partner with global development teams to define patterns, standardise practices, and drive reliable, scalable systems. The role emphasizes hands‐on design, incident management, documentation, and mentoring, with a bias toward Terraform, Kubernetes, and modern CI/CD tooling. Hybrid work and comprehensive benefits are offered. ...

Interim VP of Engineering

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
scalability, reliability, and cost efficiency at scale Delivery & Operations Drive delivery throughput, predictability, and platform reliability across all teams — CI/CD, on-call, incident management, disaster recovery Ensure security and regulatory compliance across R&D and production platforms — including ISMS implementation and applicable industry standards Manage relationships ...

Head of Engineering

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
direction, data capabilities and system reliability Own the reliability and performance of core systems, including uptime, latency, SLOs/SLAs and continuous improvement through incident management Ensure security, data protection, and regulatory compliance, while leading and developing a high‐performing, accountable engineering organisation Requirements Significant experience of commercial ...

Salesforce Portal Owner / Developer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
order - GDPR/UK GDPR and FCA obligations (including Consumer Duty), data‐subject requests, retention and access control - and lead GDPR, security and regulatory incident management (containment, remediation and reporting) with the DPO and Compliance. Build and maintain the portal hands‐on in Experience Cloud - Lightning Web Components ...

System Engineer Linux (f/m/d)

Hiring Organisation
Visa
Location
London, England, United Kingdom
Client globally to all Core Systems and services at Visa Data Centers. The team is responsible for all technical support, Change Execution, Problem and Incident management. The administrator is responsible to ensure high availability and drive better systems monitoring and performance investigations. This includes working in partnership with other ...

Financial Markets Infrastructure Services - Technical Analyst

Hiring Organisation
Standard Chartered Bank
Location
Greater London, United Kingdom
Employment Type
Full Time
Directory, email, and collaboration tools. Support trader voice and telecommunications systems, including handsets, headsets, turrets, and voice recording solutions. Assist with user administration, profile management, and access provisioning. Trading & Market Data Systems Support market data platforms such as Bloomberg and Refinitiv. Develop knowledge of trading applications and financial market … technologies. Help maintain availability and performance of critical business systems used by trading and dealing teams. Assist with transaction systems and associated infrastructure. Service Management & Continuous Improvement Manage incidents, problems, service requests, and changes using ServiceNow. Contribute to root cause analysis and high-severity incident management. Monitor service ...

Principal Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
supporting a foundation of tools and common services - including our managed Kubernetes clusters, observability tooling, and much more - while setting best practices around incident management and cost controls and empowering teams to build it, run it. The Principal Engineers work together as a unified team with a common … will deliver on Beamery's strategy. Engage with other Principal Engineers in setting and advocating company-wide standards for operational excellence, observability, reliability and incident response. Take a whole-company view of major incidents, identifying recurring themes and turning them into company-level investments. Own and evolve the platform ...

Digital Product Delivery Manager

Hiring Organisation
Boux Avenue
Location
Wimbledon, England, United Kingdom
User Acceptance Testing (UAT) processes. You will also serve as the primary technical point of contact for day-to-day site health and incident management. You are accountable for turning agreed priorities into working, tested and released digital products on time and to high quality. Main Duties & Responsibilities Requirements … Structure and lead User Acceptance Testing sessions with business stakeholders, ensuring that deliverables meet the original business requirements and are defect-free before deployment. Incident Management & Triage: Act as the central triage point for critical site bugs. Investigate issues to isolate whether they are user error, edge cases ...

Director, Technology, Cyber & Resilience Risk

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
meets regulatory expectations. Drives risk-informed engineering delivery, embedding robust controls, resilience practices, and data-led assurance across platforms. Reports to Head of Business Management, Markets & Risk Intelligence Engineering.**Core Accountabilities** 1. Risk & Control Ownership* Own the first-line technology risk profile, ensuring alignment to divisional risk appetite. … govern KRIs, KPIs and control effectiveness metrics (KCIs).* Ensure availability of accurate, decision-ready risk data.* Drive adoption of data-led risk management across engineering teams. 7. Leadership & Operating Model* Lead and develop a high-performing technology risk team.* Define clear roles, responsibilities, and RACI across first ...

Platform & Cloud Operations Director (FinOps)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
cloud environments. The role demands expertise in managing cloud FinOps outcomes and building operational maturity while reducing team size. Key responsibilities include vendor management, internal IT operations, and security coordination. The ideal candidate will have significant experience in Azure operations, incident management, and strong vendor relationships. Benefits ...

Architect/Staff Embedded Software Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Linux BSP: Yocto/OpenEmbedded, U‐Boot or UEFI, device driver development, and device tree authoring for custom silicon platforms. Deep expertise in server management protocols IPMI/IPMB, Redfish, MCTP, PLDM — and low‐speed peripheral integration (I2C, SPI, UART, CPLD) in a rack‐scale hardware context. Experience with … shipped across multiple generations in large and fast‐moving organizations. Track record driving technical outcomes in organizations with high reliability expectations, including robust observability, incident management, and close collaboration with hardware and silicon teams on field issues. Outstanding technical communicator. You can articulate architectural decisions and their consequences ...

Architect/Staff Embedded Software Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Linux BSP : Yocto/OpenEmbedded, U‐Boot or UEFI, device driver development, and device tree authoring for custom silicon platforms. Deep expertise in server management protocols IPMI/IPMB, Redfish, MCTP, PLDM – and low‐speed peripheral integration (I2C, SPI, UART, CPLD) in a rack‐scale hardware context. Experience with … shipped, across multiple generations in large and fast‐moving organizations. Track record driving technical outcomes in organisations with high reliability expectations, including robust observability, incident management, and close collaboration with hardware and silicon teams on field issues. Outstanding technical communicator. You can articulate architectural decisions and their consequences ...

Senior Program Manager, Payments

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
days a week) in our London office. You will interact with financial partners directly and drive efforts to build payment programs, improve operations and incident management procedures, and use metrics to track payments and partner performance. You will partner with product, engineers, operations, data analytics and other program … working with other stakeholders and financial partners to ensure delivery Help to define and implement metrics to measure payments and partner health Assist with incident resolution when needed, working with Engineering and Product teams to drive partner communications and escalations You Have: 8+ years of relevant experience in program ...

Information Security Manager

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
call,assessing new CVEs,vulnerabilitiesand supply-chain risk against our real infrastructure, not just a checklist. Running hands-on audits covering pen testing, vulnerability management, supply-chain security, bringing technical judgement, not just policy references. Owning our governance, risk and compliance frameworks, and keeping them honest, turning policy into … security management. CISSP or CISM (or equivalent depth of experience). We care more about the substance than the letters. A track record in incident management and governance, risk and control frameworks, ideally somewhere that also demanded technical credibility, not just process. The confidence to say "that policy ...

Head of Engineering, SAP/Oracle Platform & Transformation (HR and Finance)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
platforms supporting HR and Finance domains Lead engineering teams responsible for Oracle HCM configuration, extensions, integrations and reporting, setting standards for code quality, release management and platform integrity Own technical decision making across configuration versus extension in Oracle, including Redwood adoption, BIP reporting, Oracle Integration Cloud, REST APIs … security compliance Support wider platform modernisation initiatives, including alignment with cloud and emerging technologies Support operational excellence across platforms, including monitoring, resilience and incident management Contribute to wider platform modernisation initiatives and alignment with cloud and emerging technologies. Experience and Skills 15+ years of experience working with enterprise ...

Senior Azure SRE & Cloud Reliability Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
reliability targets, and drive resilience across multi-region deployments. The role requires deep Azure knowledge, Terraform/IaC discipline, and strong DevOps practices, including incident management, runbooks, and proactive capacity planning. #J-18808-Ljbffr ...

Hybrid SRE Engineer — Observability & Cloud (London)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
toward an SRE model while working in a hybrid setup, visiting the London office twice weekly. The role focuses on observability, high availability and incident management, with collaboration across Product Engineering and Infrastructure teams. Strong AWS, Terraform, Python and Kubernetes skills are valued. #J-18808-Ljbffr ...

Senior C++ Trading Systems Engineer — Low-Latency

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
performance and resilience using C++ on Linux in a front-office setting. The role emphasizes building scalable, highly available production systems with strong observability, incident management and modern DevOps practices, including cloud adoption and automation, in a London-based #J-18808-Ljbffr ...

Senior Site Reliability Engineer, Scalable Infra

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
will work across AWS, Kubernetes, and infrastructure‐as‐code, applying AI to automate toil and improve reliability. Collaboration with software engineers and on‐call incident management are core parts of the role. Ideal candidates have 5+ years in SRE/DevOps, strong coding skills in Python ...

Director, Technology, Cyber & Resilience Risk

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
control frameworks aligned to recognised standards (NIST, ISO, COBIT). Strong track record in risk governance and remediation of systemic issues. Operational resilience and incident management expertise. Experience engaging with regulators and executive stakeholders. Cloud and third-party risk oversight. Degree in Computer Science/Engineering or equivalent ...

Engineering Leader - Hybrid | Build High-Impact Teams

Hiring Organisation
Jobleads-UK
Location
Wimbledon, England, United Kingdom
meet strategic priorities. You will coach engineers and testers, build high-performing teams, and drive continuous improvement while supporting live production systems and incident management. A strong Agile background is essential. #J-18808-Ljbffr ...

Onsite Service Assistant

Hiring Organisation
Apogee Corporation**
Location
City of London, London, United Kingdom
Employment Type
Permanent
Print Service, this role is to support and maintain devices within a specified client account based on site. Utilising both remote monitoring software and incident management applications the Onsite Service Assistant will provide first level service support (hardware and software). Acting as Apogees front-line representative … service requests Mapping Desktop Print Queues and Troubleshooting Attend to basic break/fix for printer jams and/or other minor issues Consumable management/Re-stock (Ink/Toner Order and Installation) including recycling Monitor a large fleet of print devices Ensuring client devises are operating correctly ...

Privacy and Commercial Counsel & DPO (Maternity cover)

Hiring Organisation
LAW Absolute
Location
City of London, London, United Kingdom
international colleagues on cross-border privacy matters. Advising on privacy initiatives including: Data Protection Impact Assessments (DPIAs) International data transfer arrangements Privacy by design Incident management Governance frameworks, policies and procedures Monitoring legal and regulatory developments and providing practical advice on emerging privacy and legal risks. Advising … other business functions to deliver pragmatic legal solutions. Building strong relationships with senior leaders and acting as a trusted legal adviser. Supporting contract management and broader legal team initiatives. About You You will be: A qualified lawyer in England & Wales. Experienced working in-house and/or within ...