126 to 150 of 224 Incident Management Jobs in London

Payments Manager

Hiring Organisation
Jobleads-UK
Location
City Of London, England, United Kingdom
execution of payment processes, ensuring efficiency, reliability, and security. This role requires a blend of strategic thinking and hands‐on operational management, working closely with various internal teams to drive continuous improvement and support Verto's rapid growth. What You’ll Be Doing Oversee and improve daily payment operations … managing payment operation projects, and driving process improvements. A strong understanding of cross‐border payment systems, multi‐currency processes, and FX operations. Experience with incident management principles and best practices in a real‐time payments context. Card and/or safeguarding experience is a plus. Exceptional analytical ...

Java Developer

Hiring Organisation
Global
Location
Greater London, United Kingdom
Employment Type
Full Time
Global, you will: Key Responsibilities Feature development (35%) : Build and enhance features on our ad server platform to support direct and programmatic campaign lifecycle management, from booking to fulfilment, tracking and reporting. Own features end-to-end, from design through to production release, ensuring quality and timely delivery. Collaboration … implement changes to enhance the platform and tooling. Production support and reliability (10%) : Help ensure a stable production environment by monitoring system health, supporting incident resolution and contributing to preventative measures and improvements. What You'll Love About This Role Think Big : Work on complex, high-scale systems powering ...

Site Reliability Engineer (SRE) / Platform Engineer

Hiring Organisation
Adecco
Location
City of London, London, United Kingdom
Employment Type
Contract
enable efficient and reliable software delivery. Work closely with development teams to improve platform resilience, scalability, and operational performance. Implement monitoring, alerting, and incident response processes to support production systems. Drive automation initiatives to reduce operational overhead and improve system reliability. We're Looking For: Proven experience … using FastAPI . Strong knowledge of Terraform and Infrastructure as Code principles. Experience designing highly available, resilient, and secure cloud environments. Strong troubleshooting and incident management skills within production environments. Experience with CI/CD pipelines, Git, GitHub, and DevOps best practices . Strong understanding of cloud networking ...

AWS Solution Architect

Hiring Organisation
Anson Mccade
Location
City of London, London, United Kingdom
Employment Type
Permanent, Part Time, Work From Home
Salary
£90,000
communicate deployment patterns for migration Support engineering teams in optimising the migration path Act as an escalation point for engineering teams on troubleshooting and incident management Drive the adoption of automation across the platform, using common methods with other platforms Document deployed systems, scripts and working practices … Professional), or AWS Speciality certifications (Security, Advanced Networking) A scaled agile framework certification, such as SAFe Experience managing stakeholders, including users and senior management Experience mentoring junior engineers What's in it for you Flexible and hybrid working, with core hours and part-time options available 25 days' holiday ...

Principal Data Engineer (14 Month FTC)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
design, build, test and deployment of secure and scalable data pipelines Maintain and operate data products, ensuring availability, performance and resilience through monitoring and incident management Provide assurance on code quality and automation used to manage data flows Lead service reviews, health checks and root cause analysis … understanding of cloud‐based data ingestion, processing and storage techniques Expert knowledge of data processing, testing and monitoring practices Strong understanding of enterprise data management, optimisation and security at scale Confidence influencing technical direction and engineering standards Strong stakeholder engagement skills across business and technical teams Ability to break ...

AWS Solution Architect

Hiring Organisation
Hackajob Ltd
Location
South West London, London, United Kingdom
Employment Type
Permanent, Part Time
automate, publish and communicate deployment patterns for migration Support engineering teams in optimizing the migration path Escalation point for engineering teams in troubleshooting and incident management Driving the adoption of automation in the platform, using common methods with other platforms Documentation of deployed systems, scripts and other working … Professional, AWS Speciality domains (Security, Advanced Networking) A scaled agile framework certification, such as SAFe or Scrum@Scale Managing stakeholders, including users and management Mentoring junior engineers and nurturing their passion for engineering Security Clearance is required for this vacancy. If you are not currently Security Cleared, you will ...

AWS Solution Architect

Hiring Organisation
BAE Systems
Location
Greater London, United Kingdom
Employment Type
Full Time
automate, publish and communicate deployment patterns for migration Support engineering teams in optimizing the migration path Escalation point for engineering teams in troubleshooting and incident management Driving the adoption of automation in the platform, using common methods with other platforms Documentation of deployed systems, scripts and other working … Professional, AWS Speciality domains (Security, Advanced Networking) A scaled agile framework certification, such as SAFe or Scrum@Scale Managing stakeholders, including users and management Mentoring junior engineers and nurturing their passion for engineering Security Clearance is required for this vacancy. If you are not currently Security Cleared, you will ...

Chase UK Product Director - Platform

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
looking for a Product Director in our platform team to own the platform product strategy and product management lifecycle from start to finish. Partnering with engineering leads and other product managers, you will define the platform roadmap, KPIs and product management best practices across multiple teams. By taking … manage stakeholders at all levels of seniority in the organisation. This is an exciting opportunity for an experienced leader to implement a technical product management discipline and key processes across the product team. You will define what good platform product management looks like and promote a customer-first ...

DevOps AI Platform Architecture - London, UK

Hiring Organisation
Capgemini
Location
Greater London, United Kingdom
Employment Type
Full Time
best practices and automation. Collaborating with data scientists, developers, security teams, and business stakeholders to deliver enterprise-grade AI solutions. Implementing observability, monitoring, and incident management processes to maintain platform performance and availability. Evaluating and integrating emerging technologies in Generative AI, platform engineering, and cloud-native ecosystems ...

AWS Solution Architect

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
automate, publish and communicate deployment patterns for migration. Support engineering teams in optimizing the migration path. Escalation point for engineering teams in troubleshooting and incident management. Driving the adoption of automation in the platform, using common methods with other platforms. Documentation of deployed systems, scripts and other working practices ...

Principal Data Engineer

Hiring Organisation
Hackajob Ltd
Location
Hounslow, London, United Kingdom
Employment Type
Permanent
design, build, test and deployment of secure and scalable data pipelines Maintain and operate data products, ensuring availability, performance and resilience through monitoring and incident management Provide assurance on code quality and automation used to manage data flows Lead service reviews, health checks and root cause analysis … understanding of cloud based data ingestion, processing and storage techniques Expert knowledge of data processing, testing and monitoring practices Strong understanding of enterprise data management, optimisation and security at scale Confidence influencing technical direction and engineering standards Strong stakeholder engagement skills across business and technical teams Ability to break ...

Principal Data Engineer - 14 Month FTC

Hiring Organisation
Hackajob Ltd
Location
Hounslow, London, United Kingdom
Employment Type
Permanent
Salary
£90,000
design, build, test and deployment of secure and scalable data pipelines Maintain and operate data products, ensuring availability, performance and resilience through monitoring and incident management Provide assurance on code quality and automation used to manage data flows Lead service reviews, health checks and root cause analysis … understanding of cloud based data ingestion, processing and storage techniques Expert knowledge of data processing, testing and monitoring practices Strong understanding of enterprise data management, optimisation and security at scale Confidence influencing technical direction and engineering standards Strong stakeholder engagement skills across business and technical teams Ability to break ...

Principal Data Engineer (14 Month FTC)

Hiring Organisation
Jobleads-UK
Location
City Of London, England, United Kingdom
design, build, test and deployment of secure and scalable data pipelines Maintain and operate data products, ensuring availability, performance and resilience through monitoring and incident management Provide assurance on code quality and automation used to manage data flows Lead service reviews, health checks and root cause analysis … understanding of cloud‐based data ingestion, processing and storage techniques Expert knowledge of data processing, testing and monitoring practices Strong understanding of enterprise data management, optimisation and security at scale Confidence influencing technical direction and engineering standards Strong stakeholder engagement skills across business and technical teams Ability to break ...

AVP, Automation & AI Delivery

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
determine feasibility Own, Track and report on delivery scope, progress, risks and dependencies Ensure delivery standards for requirements, documentation, testing, releases, monitoring and incident management are met Implement feedback loops and metrics to track and report on adoption, trust, and impact to SA leadership and senior stakeholders Collaborate … permitted use, privacy, security, model governance and third‐party risk). Role Requirements & Skills Skills/Competencies Highly organised with strong programme/project management skills; able to manage multiple initiatives concurrently and deliver results. Exceptional stakeholder management with the ability to build trusting relationships and drive consensus ...

Operational Resilience Manager

Hiring Organisation
Hanson Lee
Location
London Area, United Kingdom
strengthen the clients operational resilience framework by identifying, assessing, testing, and mitigating risks across the organisation. You will take ownership of third-party risk management, business continuity planning, incident management, IT and cyber resilience, disaster recovery, and AI-related initiatives into a coordinated function along with owning ...

Senior Manager – Technology Resilience & Risk Improvement

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
effective execution of core resilience activities (e.g. exercises, scenario testing, plan validation, recovery readiness) Coordinate (not lead) regional response in alignment with global event management structure. Act as resilience SME within incident response. Maintain oversight of resilience capability health, including tracking risks, gaps and remediation Partner with Business … prioritise improvement initiatives across technology risk and resilience Deliver initiatives to strengthen control effectiveness, resilience testing, recovery outcomes, and automation Translate regulatory, audit, and incident findings into actionable improvement plans Track and report measurable improvements in resilience maturity, control effectiveness, and risk reduction Support broader Technology Risk Office ...

Senior Cloud Engineer

Hiring Organisation
4Recruitment Services
Location
Hackney, Hackney Central, Greater London, United Kingdom
Employment Type
Contract
Contract Rate
£500/day
Infrastructure as Code, automation, CI/CD and systems integration. Strong background in cloud architecture, networking and solution design. Experience with monitoring, performance optimisation, incident management and service availability. Solid software development and scripting skills using modern development practices. Experience working in Agile, cross-functional teams with … user-focused approach. Strong stakeholder management, leadership and mentoring skills, with excellent communication abilities. Relevant cloud certification (e.g. AWS, Azure) or equivalent practical experience preferred. To find out more information please contact Abbie at (url removed) Recruitment is done in line with safe recruitment practices. We are an equal ...

Developer Enablement, Technical Architect – Release on Demand (SVP)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
extend your sphere of influence across various related up and downstream platforms which is highly encouraged.**Release on Demand (RoD)** is our strategic release management platform and the primary focus of this role. Built internally approximately three years ago, the platform has quickly become our strategic release generation … domain specific RESTful services & APIs* Proficiency with relational and/or NoSQL databases: PostgreSQL, MongoDB or Couchbase* Demonstrated SRE or platform engineering experience — SLOs, incident management, reliability engineering at scale* Experience defining and implementing observability strategies: distributed tracing, structured logging, metrics and alerting* Proven experience leading technical projects ...

Senior SRE / Platform Engineer- Global Prime Brokerage & Financing Platform

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
large-scale distributed systems using Amazon CodeBuild, GitHub Actions & Terraform Enterprise Implement and maintain observability solutions to establish real-time monitoring and proactive incident response using Datadog and AWS CloudWatch Ensure high availability and performance of relational and time-series databases (PostgreSQL/QuestDB), including replication, failover strategies … experience with observability and monitoring tools (e.g. Datadog, ELK, CloudWatch) Proven track record managing infrastructure with Terraform or AWS CDK Strong background in incident management, system reliability, and operational leadership, with experience conducting root cause analysis & implementing reliability improvements Benefits & Incentives Build and scale mission-critical systems that ...

Senior Azure Platform Engineer

Hiring Organisation
Opus Recruitment Solutions Ltd
Location
London, South East, England, United Kingdom
Employment Type
Contractor
Contract Rate
£500 - £600 per day
strong focus on designing, building and supporting enterprise-scale Azure cloud environments. Proven experience working across the Microsoft Azure ecosystem , including the deployment, management and optimisation of secure, highly available and scalable cloud infrastructure. Strong expertise in Azure Monitor , including monitoring strategy, log analytics, alerting, performance tuning and proactive … incident management across cloud platforms. Extensive experience with Azure DevOps , delivering CI/CD pipelines, release automation, source control management and infrastructure deployment best practices. Advanced Infrastructure as Code (IaC) skills using Bicep or Terraform , with a track record of developing reusable, maintainable and automated Azure infrastructure ...

Director of Platform Engineering

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Pulumi, Kubernetes, Docker Strong knowledge of DevOps principles, CI/CD pipelines, and engineering best practices Experience driving operational excellence including reliability, monitoring, and incident management Ability to influence and communicate effectively with senior stakeholders and non-technical audiences Commercial awareness with experience managing budgets and optimising platform ...

Cloud Data Engineer

Hiring Organisation
Capgemini
Location
Greater London, United Kingdom
Employment Type
Full Time
automate deployment processes. Collaborate with architects, business analysts, QA teams, and stakeholders. Ensure data security, governance, compliance, and access controls. Support production deployment, monitoring, incident management, and troubleshooting. Prepare technical documentation and operational runbooks. Your Skills Cloud & Data Engineering Strong hands-on experience with AWS cloud platform. Experience ...

Salesforce Technical Lead

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
align with the broader Salesforce ecosystem strategy Design and implement Salesforce best practices, focusing on scalable architecture, data integrity, security, and performance optimisation Support incident management and processes directly within Salesforce through custom development and troubleshooting Collaborate with stakeholders to gather complex requirements, translate them into technical specifications ...

Senior Backend Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
scaleCommunicate clearly with stakeholders across engineering, product, and design, translating technical complexity into shared understanding and actionable plansEnd-to-end application support, including production incident managementEmbrace agile methodologies and user-centred thinkingEngage in a culture of continuous improvement by attending events such as blameless post-mortems, architecture reviews ...

Senior Engineering Manager - Enterprise Trust & Reliability

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
clean, well-tested APIs and integrations that hold up under enterprise security reviews. Stability & Operations (S&O) keeps the platform dependable as it scales: incident management and the severity model, service-level objectives (SLOs) and error budgets, observability and DORA (DevOps Research and Assessment) delivery metrics … tighter discovery-to-build-to-review loops, and you modelling it in how you run the teams. Own the operational backbone. On-call health, incident practice, and clear, proactive communication when things go wrong, so reliability is a habit rather than a fire drill. What we’re looking ...