401 to 425 of 481 Incident Management Jobs in London

Production Engineering Manager

Location
City of Westminster, England, United Kingdom
this role, you will manage a team of production engineers who own the full lifecycle of systems — from capacity planning and performance optimization to incident response and automation. You will drive technical strategy, champion AI-augmented workflows, and partner closely with software engineering, infrastructure, and product teams to ensure … team, sharing learnings and best practices with the broader production engineering organizationContribute hands-on to technical work including code, system design reviews, and incident response, using AI tooling to expand personal and team reach across disciplinesPartner cross-functionally with software engineering, data science, and product teams to unblock dependencies ...

AWS Network Engineer (Palo Alto, Firewalls, Python) – Amsterdam

Location
Greater London, England, United Kingdom
managing and evolving large-scale AWS networking and cloud security infrastructure. This is a hands‐on engineering role focused on AWS networking, firewall management, troubleshooting, automation, and operational excellence. You will work closely with platform, security, and engineering teams to design, build, and operate resilient cloud network solutions supporting … services, including: VPCs Transit Gateway Gateway Load Balancer Route Tables NAT Gateways VPN Connectivity VPC Peering Own and operate cloud firewall infrastructure, including policy management, upgrades, and security controls. Collaborate with stakeholders to gather requirements and deliver scalable networking solutions. Investigate and resolve complex network and security incidents. Support ...

Interim Senior Engineering Manager, IT Solutions

Location
Greater London, England, United Kingdom
consistently deliver against ambitious commercial goals. Key Responsibilities Lead multiple engineering teams, supporting engineers and managers in their growth through coaching, feedback, and performance management Own end-to-end delivery across the white label and DPAYG products, ensuring initiatives are executed effectively and deliver measurable impact Partner closely with … coordinated and transparent Champion engineering best practices that promote high-quality, scalable, and resilient systems Drive operational excellence, ensuring strong technical health, monitoring, and incident management across all teams Required Qualifications & Experience Significant experience leading engineering teams and delivering complex, multi-team initiatives in a B2B or platform ...

Senior Consultant | Cybersecurity - Incident Response

Location
Greater London, England, United Kingdom
Senior Consultant | Cybersecurity - Incident ResponseSkip to main contentWe use cookies to provide website functionality, to analyze our traffic, to personalize content and to enable social media functionality. For further information, please see our Cookie Policy. To enhance your experience, we use an AI assistant, Olivia, to help you explore … roles and learn more about FTI Consulting. For access to Olivia, please accept or decline.#Senior Consultant | Cybersecurity - Incident Response page is loaded## Senior Consultant | Cybersecurity - Incident ResponseApplyremote type: Hybridlocations: London, United Kingdomtime type: Full timeposted on: Posted Todayjob requisition id: JR2520U-TEE**Who We Are**FTI Consulting ...

Director, Technology, Cyber & Resilience Risk

Location
Greater London, England, United Kingdom
meets regulatory expectations. Drives risk-informed engineering delivery, embedding robust controls, resilience practices, and data-led assurance across platforms. Reports to Head of Business Management, Markets & Risk Intelligence Engineering.**Core Accountabilities** 1. Risk & Control Ownership* Own the first-line technology risk profile, ensuring alignment to divisional risk appetite. … govern KRIs, KPIs and control effectiveness metrics (KCIs).* Ensure availability of accurate, decision-ready risk data.* Drive adoption of data-led risk management across engineering teams. 7. Leadership & Operating Model* Lead and develop a high-performing technology risk team.* Define clear roles, responsibilities, and RACI across first ...

Platform & Cloud Operations Director (FinOps)

Location
Greater London, England, United Kingdom
cloud environments. The role demands expertise in managing cloud FinOps outcomes and building operational maturity while reducing team size. Key responsibilities include vendor management, internal IT operations, and security coordination. The ideal candidate will have significant experience in Azure operations, incident management, and strong vendor relationships. Benefits ...

Architect/Staff Embedded Software Engineer

Location
Greater London, England, United Kingdom
Linux BSP : Yocto/OpenEmbedded, U‐Boot or UEFI, device driver development, and device tree authoring for custom silicon platforms. Deep expertise in server management protocols IPMI/IPMB, Redfish, MCTP, PLDM – and low‐speed peripheral integration (I2C, SPI, UART, CPLD) in a rack‐scale hardware context. Experience with … shipped, across multiple generations in large and fast‐moving organizations. Track record driving technical outcomes in organisations with high reliability expectations, including robust observability, incident management, and close collaboration with hardware and silicon teams on field issues. Outstanding technical communicator. You can articulate architectural decisions and their consequences ...

Architect/Staff Embedded Software Engineer London, UK

Location
Greater London, England, United Kingdom
Linux BSP : Yocto/OpenEmbedded, U-Boot or UEFI, device driver development, and device tree authoring for custom silicon platforms. Deep expertise in server management protocols IPMI/IPMB, Redfish, MCTP, PLDM - and low‐speed peripheral integration (I2C, SPI, UART, CPLD) in a rack-scale hardware context. Experience with … shipped, across multiple generations in large and fast-moving organizations. Track record driving technical outcomes in organisations with high reliability expectations, including robust observability, incident management, and close collaboration with hardware and silicon teams on field issues. Outstanding technical communicator. You can articulate architectural decisions and their consequences ...

Software Engineer II - Backend (Ruby)

Location
Greater London, England, United Kingdom
impact on merchant revenue* Partners closely with Product to evolve checkout and payments capabilities while balancing quality, security (PCI), and speed of delivery* Owns incident management and on-call responsibilities for checkout services due to their critical natureThe Commerce Engineering organisation is on a mission to build … functionality is production-grade, with the appropriate test coverage on all layers, instrumentation in place, and documentation* Troubleshoot production issues and contribute to post-incident reviews to improve reliability* Produce bullet-proof code that is robust, efficient, and maintainable* Work on challenging problems such as query optimization and performance ...

Head of Engineering, SAP/Oracle Platform & Transformation (HR and Finance)

Location
Greater London, England, United Kingdom
platforms supporting HR and Finance domains Lead engineering teams responsible for Oracle HCM configuration, extensions, integrations and reporting, setting standards for code quality, release management and platform integrity Own technical decision making across configuration versus extension in Oracle, including Redwood adoption, BIP reporting, Oracle Integration Cloud, REST APIs … security compliance Support wider platform modernisation initiatives, including alignment with cloud and emerging technologies Support operational excellence across platforms, including monitoring, resilience and incident management Contribute to wider platform modernisation initiatives and alignment with cloud and emerging technologies. Experience and Skills 15+ years of experience working with enterprise ...

Global Systems Dev Manager - Automation & 24/7 Ops

Location
Greater London, England, United Kingdom
London. You will lead a global support engineering team, guiding it from reactive troubleshooting to proactive automation and engineering excellence, with a focus on incident management at scale and permanent code fixes. You will champion automation, define follow-the-sun operations, and collaborate with engineering groups to align ...

Platform Operations Leader: Scale, Reliability & Automation

Location
Greater London, England, United Kingdom
technical expertise and leadership skills to ensure operational excellence and continuous improvement. Your responsibilities will include managing a team of engineers, leading major incident management, and liaising with business and engineering teams. Familiarity with DevOps tools and methodologies is essential. You'll contribute to a high-performing culture ...

Senior Product Manager (SaaS)

Hiring Organisation
LinuxRecruit
Location
London, UK
Employment Type
Full-time
logs, metrics, traces, and security events while saving customers serious money. They are looking for several technical Product Managers to oversee their infrastructure and incident management platforms. These roles are about turning customer needs into smart roadmaps and products people actually enjoy using. It means working with design ...

Systematic Trading Production Engineer

Location
Greater London, England, United Kingdom
time or trading environments. Experience supporting business‐critical production systems with a strong sense of urgency and ownership. Strong understanding of monitoring, observability, alerting, incident management, and root cause analysis. Experience supporting front‐office users in a fast‐paced environment. Strong troubleshooting skills across application and system behavior ...

BI Engineer

Location
Greater London, England, United Kingdom
support adoption of self-service analytics Work with platform teams on architecture, scalability, and performance Follow engineering standards and contribute to continuous improvement Support incident management and ongoing maintenance of BI solutions Skills & Experience Experience with AWS data services (e.g. Redshift, Athena, S3, Glue, Lambda) Strong SQL skills ...

java developer for cloud-native microservices

Location
Greater London, England, United Kingdom
using appropriate approaches; Ensure systems are reliable and easy to operate; Continuously update technologies and patterns; Support products throughout their lifecycle, including production and incident management; Drive adoption and governance of approved AI-assisted engineering practices across teams; Establish measurable validation standards for secure coding, peer review ...

Lead Software Engineer - Backend Engineer - Chase UK

Location
City Of London, England, United Kingdom
date by continuously updating our technologies and patterns Support the products you've built through their entire lifecycle, including in production and during incident management Requirements Formal training or certification on software engineering concepts and applied experience Recent hands-on professional experience as a back-end software engineer ...

Senior Full Stack Payments Product Engineer

Location
Greater London, England, United Kingdom
Laminas and Mezzio frameworks AI is woven into how we work, using Claude Code and agentic development workflows Linear for project management, Notion for docs Cloudflare for CDN, queueing and workers Github and Github Actions for source control and CI Incident.io for incident management Datadog for logging ...

Platform Support Engineer (Eng I)

Location
Greater London, England, United Kingdom
platform ops role in a high-scale environment Experience with GitOps workflows, understanding how infrastructure changes flow from code to production Familiarity with incident management tooling (PagerDuty, incident.io, Rootly, or similar), not just receiving alerts but owning the response flow Exposure to a client-facing or stakeholder-heavy ...

Senior IT Infrastructure Administrator, Networking

Hiring Organisation
Confluence
Location
London, UK
Employment Type
Full-time
including BGP, OSPF and EIGRP across multi-site environments. Proven ability to design, implement and transition infrastructure solutions into production. Hands-on experience in incident management, root cause analysis and issue resolution. Strong troubleshooting mindset with the ability to resolve complex networking issues independently. Understanding of how networking ...

Python Software Engineer

Hiring Organisation
Cboe Exchange
Location
London, UK
Employment Type
Full-time
Providing operational support for Cboe Europe's trading systems by participating in a production support rota, responding to incidents in line with Cboe's Incident Management and Response processes, and contributing to post-mortem analyses and follow-up actions. Requirements Solid Python knowledge. A commitment to writing testable ...

Global Platform Operations Leader | SRE & CI/CD Excellence

Location
Greater London, England, United Kingdom
Trading Technologies seeks an SVP, Platform Operations to lead the end-to-end deployment pipeline, production observability, and incident management. You will build a global leadership layer and drive platform integration for acquisitions, with a focus on reliability and automated delivery across TT's engineering hubs. You will ...

Senior Data Protection Manager, M&A

Location
Greater London, England, United Kingdom
throughout the M&A lifecycle. Conducting due diligence and creating integration plans. Advising on privacy considerations for transitional service agreements. Supporting privacy compliance and incident management. Collaborating with teams across Legal, IT, HR, and more. About You You’re a privacy professional with experience in operational or consultative roles. ...

Senior CNS Engineer: AI-Driven Incident & Data Safeguards

Location
Greater London, England, United Kingdom
Cisco Systems, Inc. is hiring a Senior Network Support Engineer to lead incident management for UK Data Sovereignty cases. You will analyse critical network data, coordinate with global TAC, and route complex issues to teams with AI-assisted tooling and de-sensitisation processes. You will mentor peers, drive ...

Principal Python Software Engineer

Hiring Organisation
Cboe Exchange
Location
London, UK
Employment Type
Full-time
building highly reliable, highly testable Python systems to support Cboe Europe's trading operations. Leading complex projects including: Meet regularly with team members and management to discuss project progress and operational correctness/efficiency. Manage the involvement of developers across multiple Cboe systems and/or other developers both … Providing operational support for Cboe Europe's trading systems by participating in a production support rota, responding to incidents in line with Cboe's Incident Management and Response processes, and contributing to post-mortem analyses and follow-up actions. The ideal candidate has :Expert Python knowledge A commitment ...