426 to 450 of 481 Incident Management Jobs in London

Group IT Operations and Security Engineer

Location
Greater London, England, United Kingdom
system performance, availability, and capacity; implement improvements where needed. Support deployment, configuration, and patching of operating systems, applications, and infrastructure. Administer identity and access management (IAM) platforms, where supported by the CII. Automate recurring operational tasks using scripts or configuration‐management tools. Oversee vendors who manage backups, disaster … recovery processes, and business continuity plans. Troubleshoot and resolve hardware, software, and network issues in the event of a major incident, assist and oversee vendors. Plan, manage and deliver IS operational projects for Infrastructure, Cyber and Disaster recovery to ensure systems are maintained to current standards. Maintain service management ...

Senior Manager/ Lead - Trading Application Support OMS/EMS

Hiring Organisation
Robert Walters
Location
London, South East England, United Kingdom
Employment Type
Full-Time
Salary
£70,000 - £90,000 per annum
broker connectivity. Coordinate with global Application Support teams across Pune, the US and Australia. Manage incidents, perform root cause analysis and contribute to post-incident reviews. Maintain operational procedures, support documentation and knowledge repositories. Identify opportunities for automation and continuous service improvement. What you'll bring: Strong experience supporting … connectivity and/or exchange connectivity. Knowledge of Cash Equities and ideally exposure to Swaps and/or Futures & Options. Strong production troubleshooting and incident management experience. Comfortable interacting directly with traders and other front-office stakeholders. Able to work effectively within a global, follow-the-sun support ...

Senior Data Engineer | AWS | Azure | SAP (Contract)

Location
Greater London, England, United Kingdom
staff. Actively delivers the roll-out and embedding of DataFoundation initiatives in support of the key business programmers. Coordinate the change management process, incidentmanagement and problem management process. Present reports and findings to key stakeholders andbe the subject matter expert on Data Analysis &Design. Drive implementation efficiency … initiatives likeCoE, CoP. Mandatory skills: ELT - Master Data Modeling - Master Data Integration & Ingestion - Skill Data Manipulation and Processing - Skill Optional skills: Experience in project management, running a scrumteam. Experience working with BPC, Planning. Exposure to working with external technical ecosystem. Project Description: High Level Objective the business aims ...

InfoSec Engineer (FinTech) - IAM & DevSecOps

Location
Greater London, England, United Kingdom
role offers flexible, hybrid working and focuses on authentication, IAM, and DevSecOps across security tooling and platforms. You will contribute to monitoring, automation, incident management and secure architecture, working with Kubernetes, Microsoft Entra, SailPoint and CyberArk in a fast‐paced environment. #J-18808-Ljbffr ...

Azure Data Platform Operations Lead | Resilience & FinOps

Location
Greater London, England, United Kingdom
Operations Lead to manage the performance of their Azure-based data platform. The ideal candidate should have a strong background in operational leadership and incident management, along with hands-on experience with Azure tools. This role requires solving operational issues, leading stakeholder communication, and ensuring platform reliability ...

Remote Principal SRE - Healthcare Platform Reliability Lead

Location
Greater London, England, United Kingdom
ensure the reliability of healthcare platforms. The candidate will lead efforts in automating operations and improving service availability. With a focus on troubleshooting and incident management, applicants should have 7+ years of experience in enterprise applications and strong knowledge of Azure cloud environments. The role offers a competitive ...

Hybrid IT Application Delivery Analyst – Legal Tech

Location
Greater London, England, United Kingdom
line support for legal applications across the firm, collaborating with IT teams to ensure reliable operation of software and services. The role focuses on incident management, patching, testing, deployment, and documentation, with hybrid work arrangement and opportunities to influence the software lifecycle within a leading international law firm. ...

Principal SRE: Cloud Native (Terraform, Kubernetes)

Location
Greater London, England, United Kingdom
scalable, resilient systems and mentor engineers across disciplines. You will collaborate with Architecture, Product Owners and development teams to migrate workloads to cloud, drive incident management, and advance automation and documentation in a fast-paced hybrid environment. #J-18808-Ljbffr ...

EMEA Data Center Operations Leader

Location
Greater London, England, United Kingdom
into regional plans and represent EMEA Operations with the Executive Leadership, the Board, investors, and key customers. You will drive uptime and safety, lead incident management, chair governance on standards and KPIs, guide new site startups and staffing, and partner with Finance for OpEx/CapEx planning while ...

Global SVP, Platform Operations & AI-Driven Reliability

Location
Greater London, England, United Kingdom
Platform Operations to consolidate and lead the production platform across SRE, deployment, and platform engineering. You will own the CI/CD pipeline, incident management, and the SLO/SLI framework to raise reliability for critical trading services worldwide. You will build a leadership layer, expand global engineering ...

Senior Software Engineer II — Reliability & Observability

Location
Greater London, England, United Kingdom
build automated reliability and self-healing systems at scale, delivering platform tooling that engineers across the company adopt for their services. You will own incident management tooling, evolve observability infrastructure with SLOs and real-time signals, and contribute to AI-driven automation that reduces toil and speeds delivery. ...

Operational Resilience Lead — UK Financial Services

Location
Greater London, England, United Kingdom
tests with cross-functional teams, and contribute to risk committees. This role requires strong regulatory knowledge and hands-on experience in resilience, continuity, and incident management within UK financial services. #J-18808-Ljbffr ...

Backend Engineer - Cloud-Native FinTech (Java/Kotlin)

Location
Greater London, England, United Kingdom
services, money transfers, and core banking, collaborating with distributed teams to rapidly ship features. The role emphasizes software craftsmanship, CI/CD practices, and incident management, with strong communication in #J-18808-Ljbffr ...

Senior IT Service Desk Engineer – AI-Driven Support

Location
City Of London, England, United Kingdom
Europe, and global needs. You will act as a technical escalation point, troubleshoot Windows/macOS, identity and access, endpoint lifecycle, and incident management, while driving automation and knowledge base improvements. You will partner with security, networks, CRM, facilities, and regional IT teams, and support on-call rotations ...

Senior Software Engineer (Node & TypeScript)

Location
Greater London, England, United Kingdom
databases, ideallyMongoDB Experience mentoring engineers Experience leading technical delivery in a build-and-run engineering model Experience in a structured on-call process and incident management Experience working in regulated industries, ideally financial services or FinTech. #J-18808-Ljbffr ...

Senior Software Engineer — Infra Agent Systems UK Together AI London

Location
Greater London, England, United Kingdom
agents. Develop fleet intelligence systems that combine telemetry, infrastructure state, operational knowledge, and historical incidents to help agents make better decisions. Integrate with observability, incident management, ticketing, fleet inventory, source control, chat, and internal infrastructure systems through well-designed APIs. Own services end to end, including architecture, implementation ...

Senior Full-Stack Engineer: Cloud, On-Call & Incident Mgmt

Location
Greater London, England, United Kingdom
Burendo is seeking a Senior Full Stack Engineer to join our team, supporting a diverse range of applications and ecosystems. You will operate within an L2/L3 support function, ensuring the reliability of business ...

Senior SRE: Global Edge Network & Automation Lead

Location
Greater London, England, United Kingdom
Senior SRE to manage and enhance their global edge cloud platform infrastructure. You will build and maintain the network, ensuring reliability and incident management while innovating tools for team processes. The ideal candidate will have substantial experience with internet protocols and cloud services, as well as a strong ...

Head of Cloud Platform Foundations

Location
Greater London, England, United Kingdom
property technology company in Greater London seeks a Platform Engineering Manager to lead their Cloud Foundations team. Responsibilities include overseeing secure, reliable cloud infrastructure, incident management, and team leadership. Candidates should have a strong understanding of cloud operations, automation, and governance frameworks. The role offers hybrid working ...

Technology Risk & AI Governance Director

Location
Greater London, England, United Kingdom
with regulatory compliance and the firm’s objectives. You will lead cross‐team collaborations across infrastructure, cyber and risk stakeholders, guiding policy development and incident management while advancing operational resilience and governance. #J-18808-Ljbffr ...

Senior Identity Engineering Lead — IAM & SSO (Hybrid)

Location
Greater London, England, United Kingdom
enable seamless access across digital products. You will collaborate with security, compliance, and product teams, drive scalable APIs, and ensure resilience through strong incident management and platform strategy. #J-18808-Ljbffr ...

24/7 Network Reliability Engineer

Location
Greater London, England, United Kingdom
/7 eyes-on-glass monitoring, triage incidents, and perform first-line resolution to protect service availability and improve customer experience. Responsibilities include incident management, working with engineering and partner teams, and maintaining clear updates and knowledge articles while supporting planned changes and service validation. #J-18808-Ljbffr ...

Network Engineer: Secure, Automated Enterprise Networking

Location
Greater London, England, United Kingdom
services. You will own end-to-end networking, from routing and switching to wireless, WAN, and cloud connectivity, while balancing platform improvements with responsive incident management and change control. This role offers hybrid work (3 days in the office weekly), relocation assistance, private medical insurance ...

Principal Data Engineer — Remote Data Platform Lead

Location
Greater London, England, United Kingdom
resilient, scalable data platforms aligned with long-term business goals. You will lead multi-year technology roadmaps, drive large-scale data architectures, and own incident management, governance, and risk controls while mentoring teams and promoting data excellence across the organization. #J-18808-Ljbffr ...

Cloud-Native Senior Lead Engineer for FinTech Platform

Location
Greater London, England, United Kingdom
team, applying modern cloud-native approaches to banking products. Responsibilities span end-to-end cloud-native microservices, domain modeling, secure coding, performance optimization, and incident management. You will mentor peers on quality, AI-assisted development, and responsible AI use, shaping scalable, reliable software used #J-18808-Ljbffr ...