24 of 24 Permanent Incident Management Jobs in the City of London

Service Owner

Location
City Of London, England, United Kingdom
risk, manage supplier relationships and ensure the service continues to meet the needs of its users and business stakeholders. This is a senior service management role requiring someone comfortable operating across live services, ITIL processes, suppliers, technical teams and senior stakeholders within a complex and highly secure environment. … together performance, risks, roadmap priorities, service improvements and readiness for live service Maintain oversight of incidents and problems affecting the service, working closely with Incident and Problem Management teams to drive resolution and address root causes Act as an escalation point for significant operational issues Identify opportunities ...

Chanel Client Data Manager

Location
City Of London, England, United Kingdom
fosters cross-functional collaboration around a shared client-centric vision. As part of the Global PMO, you will lead key global client data management initiatives, partnering with Business, IT, Legal, Security, Data Governance and external partners to drive scalable, compliant and client-focused solutions. You will play a central … governance, data quality, privacy and client data capabilities while ensuring alignment across a complex global organisation. This role requires strong expertise in client data management, stakeholder engagement, governance and project delivery within an international matrix environment. Reports to: Head of Global PMO - Pioneer Key Responsibilities Client Data Management ...

Head of Trading Venue Operations, EMEA

Location
City Of London, England, United Kingdom
that provides operational support for TP ICAP's broker and client-facing trading platforms across the EMEA region. Reporting to the Head of Production Management, the role manages a team of 21 support professionals, including four asset class-aligned Team Leads, and works closely with Operations, Technology, Business Management and IT Service Management teams to ensure the effective delivery of support services. The team provides day-to-day support across a range of activities, including order and trade management, user administration, connectivity support, client onboarding, platform configuration and incident resolution. The successful candidate will ensure ...

Service Desk Team Lead

Location
City Of London, England, United Kingdom
Service Desk Team Lead Department: Service Management Employment Type: Permanent - Full Time Location: Hybrid Compensation: £40,000 - £45,000/year Description Hello. Welcome to Wanstor! At Wanstor, we’ve been delivering award-winning IT solutions for over 22 years, and we’re proud to keep growing year after … thrive and grow, you’ll feel right at home here! We’re looking for a Service Desk Team Leader to join our talented Service Management team. This role is crucial to ensuring customer requirements are met in terms of communication, prioritising, escalating and resolving incidents and requests. The Service ...

IT Systems Analyst Jessica McCormack Permanent contract London, GB IT Consulting Information Sy[...]

Location
City Of London, England, United Kingdom
cyber security best practices across the business. Administer security tools, including multi-factor authentication (MFA), mail filtering and Anti-Virus software. Participate in vulnerability management, incident response coordination, and system security reviews. Help maintain system documentation, asset registers, and compliance records in line with IT policies and data … Process & Documentation Maintain up‐to‐date documentation of IT systems, configurations, procedures, and security protocols. Support the continuous improvement of IT operational processes and incident management workflows. Skills & Experience Required Proven experience in a Systems Analyst, IT Support Analyst, or Desktop Support role, ideally within a retail, luxury ...

Site Reliability Engineer, Infrastructure - ThousandEyes

Location
City Of London, England, United Kingdom
recovery measures and addressing potential single points of failure. Troubleshoot complex issues across infrastructure and platform services,participate in the on-call rotation and incident-management process, and turn root-cause findings into lasting reliability improvements. Collaborate with application development teams and other stakeholders to meet internal Service … knowledge. Experience designing and implementing scalable, resilient, and well-tested distributed systems. Experience with Service Level Objectives, Service Level Agreements, monitoring, alerting, capacity planning, incident management, or disaster-recovery testing. Experience building automation that reduces repetitive work, improves release safety, or increases infrastructure efficiency. Strong communication and documentation ...

DBA Lead

Location
City Of London, England, United Kingdom
resilient and scalable data services that power business-critical applications. About the Role This candidate works with team on daily activities involving the design, management, maintenance, and utilization of databases. Provide leadership and drive strategy, standards, and processes for applicable database environments. Responsibilities Ensure the maintenance, stability, and operational … timely communication with the line manager and relevant stakeholders regarding progress, issues, and risks. Perform all other duties as assigned. Leadership and People Management Assign personnel to projects and direct their activities to meet business and technology objectives. Review and evaluate work, prepare performance reports, and support employee development. ...

Sr Backend Engineer

Location
City Of London, England, United Kingdom
asynchronous processing solutions Ensure operational excellence through comprehensive testing strategies (unit, component, integration, end-to-end, and performance testing) and production support Participate in incident management and system reliability initiatives Collaborate with cross-functional teams across multiple locations What you'll bring to the role: Essential Requirements: Extensive … experience building and supporting cloud-native applications in production Demonstrated experience developing RESTful APIs, messaging frameworks, and asynchronous processing Strong production support experience, including incident management and CI/CD pipelines/automation Comprehensive testing capability across unit, component, integration, end-to-end, and performance testing Desirable Requirements ...

Site Reliability Engineering Lead

Location
City Of London, England, United Kingdom
improve application reliability, scalability, security, performance, and resilience. Establish and maintain service level objectives (SLOs), service level indicators (SLIs), and error budgets. Lead major incident management, root cause analysis, problem management, and post‐incident review processes. Drive cloud modernization initiatives and support application migrations to Azure … security and regulatory requirements. Lead architecture reviews and provide guidance on cloud‐native and highly resilient application designs. Drive capacity planning, performance optimization, cost management, and operational efficiency initiatives. Establish engineering guardrails, governance controls, and deployment standards for production environments. Support organizational transformation toward DevOps and SRE practices. Manage ...

Data Center Engineering Manager

Location
City Of London, England, United Kingdom
internal stakeholders to maintain site availability, improve operational resilience and support the continued development of the data centre. Key Responsibilities Engineering Team Leadership and Management Lead, manage and develop a team of four M&E Shift Engineers providing 24/7 operational coverage of the data centre. Provide clear … data to identify recurring issues and opportunities to improve system performance and reliability. Support capacity planning, infrastructure upgrades, customer installations and expansion projects. Maintenance Management Ensure all critical infrastructure is maintained in accordance with manufacturer recommendations, statutory requirements, contractual obligations and recognised engineering standards. Plan and coordinate maintenance activities ...

Azure Data Support Engineer

Location
City Of London, England, United Kingdom
solutions. Develop and implement self‐healing automation for recurring failures. Service Operations & Support (Managed Services) Provide L2/L3 support aligned with ITIL practices (incident, problem, change management). Participate in on‐call rotations and handle critical incident response. Maintain detailed SOPs, runbooks, knowledge‐base articles … debugging, data validation, and optimization; experience with Azure SQL DB or SQL Server; familiarity with data modeling concepts and warehouse performance tuning. Support & Incident Management – Strong troubleshooting and analytical skills for root cause analysis; exposure to ITSM tools such as ServiceNow and Jira. Preferred Qualifications Microsoft Certifications (e.g. ...

Production Engineer

Location
City Of London, England, United Kingdom
trade flow, and post-trade issues. Lead the ‘follow the sun’ support model coordination, working closely with global teams in APAC and US. Drive incident management practices, trend identification, and post-mortem analysis. Collaborate with development teams to influence platform improvements based on support insights. Occasional weekend work … Advanced experience in troubleshooting network problems: firewalls, routing, DNS, load balancers, and connectivity issues Working knowledge of multiple buy-side or sell-side Order Management Systems Strong understanding of international equity market mechanics, flows, and settlement processes Proven ability to lead multiple projects simultaneously, delivering to deadlines Expert-level ...

AI Platform & Site Reliability Engineering Consultant

Hiring Organisation
Akkodis
Location
City of London, London, United Kingdom
Employment Type
Permanent
Salary
£88000 - £96000/annum
modern engineering and traditional ITSM/ITIL practices Establish SLIs, SLOs, and Error Budgets Shape observability strategies using metrics, logs, and traces Design incident response models and post-incident learning loops Reduce toil through automation and engineering excellence Deliver SRE capability assessments and roadmaps Act as a trusted … Looking For Extensive experience in SRE, cloud operations, or DevOps Proven consulting or advisory background Experience with AWS, Azure, or GCP Strong observability and incident management expertise Ability to obtain UK SC clearance Modis International Ltd acts as an employment agency for permanent recruitment and an employment business ...

Implementation Manager - Mainframe Product Migration

Location
City Of London, England, United Kingdom
dependencies, and critical path. Lead cutover planning : detailed runbooks, rehearsal plans, roles/responsibilities, communications, checkpoints, and contingency/rollback. Coordinate release and change management activities (CAB submissions, implementation approvals, change records, stakeholder signoffs). Drive environment and operational readiness : capacity, access, monitoring, scheduling/batch windows, support model … validation steps, reconciliation, and operational controls. Orchestrate dress rehearsals/mock cutovers , capture lessons learned and harden runbooks. Lead implementation governance : readiness reviews, RAID management, decision logs, and executive reporting. Coordinate hyper care/post-go-live : incident triage, defect backlog prioritisation, stabilisation metrics, and handover to BAU. ...

Senior Security Consultant

Location
City Of London, England, United Kingdom
technological partner for the core business operations of its clients worldwide. It stands at the forefront of key sectors including Transport, Defence, Air Traffic Management and Space, alongside advanced Information Technology services delivered through Minsait, and cutting-edge capabilities in Sovereign AI, Cybersecurity, and Cyberdefence via IndraMind. The company … global benchmark in innovative transportation and mobility solutions. It is recognised as one of the world's top three companies in public transportation management systems. Indra's technology supports the daily journeys of over 78 million people, helping to reduce more than 10 million tonnes of CO2 emissions annually ...

Lead GCP Engineer

Location
City Of London, England, United Kingdom
platform direction across the product lifecycle. Remain hands‐on in the delivery of complex engineering work rather than operating solely in an architectural or management capacity. Lead solution design activities, code reviews and technical assurance across the platform. Cloud Architecture & Engineering Design, implement and operate secure, scalable, cloud‐native … platform components and services. Ensure solutions align with organisational governance standards, security policies and regulatory obligations. Balance innovation and rapid delivery with appropriate risk management and control frameworks. Design secure data integration patterns incorporating authentication, authorisation and access management technologies. Support delivery of solutions that are suitable ...

Application Support Analyst | Enterprise Technology

Location
City Of London, England, United Kingdom
Enterprise platforms offered by Marex to both internal and external client base. Support business users offering second- and third-line support. Provide incident management per ITIL standards. Support applications and infrastructure hosted within AWS. Support REST APIs, API gateways, microservices and distributed services. Support application releases and deployments … through CI/CD pipelines, ensuring changes are managed and approved in accordance with Marex change management and release governance processes. Manage new system analysis and implementation. Liaison between technology departments to communicate system changes. Manage process and system documentation in existing template; produce and regularly maintain ...

Enterprise Applications Manager

Hiring Organisation
Gold Group
Location
City, London, United Kingdom
Employment Type
Permanent
Salary
GBP 85,000 - 95,000 Annual
application lifecycle, including selection, implementation, upgrades, optimisation, and retirement Ensure reliability, availability, and performance of critical systems, driving best-practice support and incident management Build, lead and develop a high-performing team of application specialists, fostering innovation and collaboration Partner with business units, IT service lines and external … security, and compliance standards across all enterprise applications What experience you need to be the successful Enterprise Applications Manager: Proven experience in senior applications management or IT leadership roles Strong technical knowledge across enterprise ecosystems such as ERP, PLM, BI, SaaS platforms, integration technologies, Atlassian tools, Microsoft 365, SharePoint ...

Site Reliability Engineer

Hiring Organisation
REVYBE IT RECRUITMENT LIMITED
Location
City of London, London, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£85,000
alerting, and dashboards Define and improve SLIs, SLOs, and reliability metrics Proactively identify and resolve performance, availability, and reliability issues Lead and contribute to incident response, troubleshooting, and root cause analysis Automate operational processes and eliminate repetitive manual tasks Work closely with software engineers to improve deployment processes, system … understanding of metrics, logging, tracing, alerting, and system health Experience troubleshooting complex production environments Understanding of SLIs, SLOs, SLAs, and error budgets Experience with incident management and root cause analysis Good understanding of cloud networking, security, and infrastructure fundamentals Strong scripting/automation skills A strong understanding ...

Senior Software Engineer (Full-Stack - TypeScript/Node)

Location
City Of London, England, United Kingdom
pairing using tools like Git and GitHub. Experience of operationally managing software components once live, including; observability, logging, metrics, error reporting, debugging and live incident management. Experience of working with sensitive personal data. Competencies Experience working in/with cross-functional teams consisting ofe.g.engineers, product,UXand non-technical stakeholders ...

Senior AWS Site Reliability Engineer

Hiring Organisation
Spectrum IT Recruitment
Location
City of London, London, United Kingdom
Employment Type
Permanent
Salary
£60000 - £70000/annum Bonus, Medical Care
environment by ensuring system availability and maintaining a comprehensive perspective on overall health. You'll develop tools and software to support and streamline the management of platform infrastructure and key applications. A major focus will be enhancing the dependability, performance, and delivery speed of our software products. … Mimir, and Tempo Background in administering or developing with popular monitoring and automation tools such as Splunk, Datadog, PagerDuty, or Rundeck Experience using configuration management platforms like Ansible, Puppet, or Chef Professional certifications in cloud DevOps, such as AWS Certified DevOps Engineer or Google Cloud Professional DevOps Engineer ...

Business Analyst / PM (Payments Systems)

Location
City Of London, England, United Kingdom
Ensure seamless interaction between payment, loyalty, and customer engagement solutions. Drive enhancements to improve customer experience and business value. Lead the implementation and ongoing management of point-to-point encryption (P2PE) across the estate. Support security audits, assessments, and remediation activities. Maintain awareness of emerging threats and regulatory developments. … success delivering technology and business change projects. Experience translating business requirements into technical solutions. Experience managing mobile applications and customer-facing digital platforms. Strong incident management and problem-solving capabilities. Experience operating within PCI DSS-regulated environments. Desirable Experience P2PE implementation experience. Knowledge of loyalty platforms and customer ...

Platform Engineering Lead

Hiring Organisation
Hays
Location
City of London, London, United Kingdom
Employment Type
Permanent
resilience across global load balancing, disaster recovery and observability. Act as a senior escalation point for complex platform incidents. Contribute to operational excellence through incident management and on-call participation. What We're Looking For Proven experience leading platform, cloud, infrastructure or DevOps engineering teams. Strong hands … cloud-native production platforms. Understanding of modern security practices, including Zero Trust approaches. Experience implementing automation, GitOps and platform engineering best practices. Strong stakeholder management and technical leadership skills. Highly Desirable Experience within fintech, SaaS or high-transaction digital platforms. Experience with Cloud SQL, Spanner and cloud-native databases. ...

Senior IT Service Desk Engineer – AI-Driven Support

Location
City Of London, England, United Kingdom
Europe, and global needs. You will act as a technical escalation point, troubleshoot Windows/macOS, identity and access, endpoint lifecycle, and incident management, while driving automation and knowledge base improvements. You will partner with security, networks, CRM, facilities, and regional IT teams, and support on-call rotations ...