551 to 575 of 688 Incident Response Jobs in London

Senior Product Manager - Corporate Experience

Location
Greater London, England, United Kingdom
proportionate and well understood. Own outcomes in production Take responsibility for how the corporate platform and payments behave in the real world. Support incident response, learn from issues, and drive changes that reduce repeat problems over time. Set direction and priorities Define clear goals and success measures ...

Payments-Digital and Design-FX Product Manager-Vice President-LONDON

Hiring Organisation
Hackajob Ltd
Location
South West London, London, United Kingdom
Employment Type
Permanent
framework and tracks the product's key success metrics such as cost, feature and functionality, risk posture, and reliability Plays a critical role in incident response; facilitates Product communication deliverable during production outages Demonstrates superior judgment to mitigate risk; fosters an environment where risk/control issues ...

Senior Software Engineer (Product, AI Portal)

Location
Greater London, England, United Kingdom
systems. Comfort working across the stack including infrastructure, backend services, APIs and frontend development, although the role is backend leaning. Experience with on-call, incident response or support rotation for a production system. Strong communication skills and the ability to work directly with stakeholders. Product-minded: you care ...

Software Engineer, Machine Learning Infrastructure

Hiring Organisation
Deliveroo
Location
London, UK
Employment Type
Full-time
distributed systems. Experience building production services, APIs, data pipelines, or ML infrastructure at scale. Experience operating systems in production, including observability, debugging, reliability, incident response, and performance/cost optimization. Hands-on experience with LLM inference and/or fine-tuning of open-weight models in production — serving ...

Senior Software Engineer, GenAI Platform

Location
Greater London, England, United Kingdom
record of designing and owning production services, APIs, data pipelines, or ML infrastructure at scale. Experience operating systems in production, including observability, debugging, reliability, incident response, and performance/cost optimization. Deep hands-on experience with LLM inference and/or fine-tuning of open-weight models ...

Senior Software Engineer, GenAI Platform

Hiring Organisation
Deliveroo
Location
London, UK
Employment Type
Full-time
record of designing and owning production services, APIs, data pipelines, or ML infrastructure at scale. Experience operating systems in production, including observability, debugging, reliability, incident response, and performance/cost optimization. Deep hands-on experience with LLM inference and/or fine-tuning of open-weight models ...

Staff Software Engineer, Observability & Profiling

Location
Greater London, England, United Kingdom
kernel, the network stack, or the hardware Have excellent communication skills and enjoy partnering with internal teams to improve their operational visibility and incident response capabilities Are excited about building foundational infrastructure and comfortable navigating ambiguous, high-impact technical challenges, both independently and with a team Preferred qualifications ...

Senior Engineer (Network)

Hiring Organisation
QinetiQ
Location
Orpington, Greater London, UK
Employment Type
Full-time
asked to work at specific times. As part of a commitment to maintaining high service availability and rapid incident response, this role will be required to participate in an on-call rota. This responsibility reflects the trust placed in the role to act swiftly and decisively when critical ...

Staff Software Engineer, Continuous Integration

Hiring Organisation
Humanloop
Location
London, UK
Employment Type
Full-time
that supports thousands of daily builds across multiple cloud providersDevelop intelligent test selection systems that reduce CI time while maintaining code qualityBuild and improve incident response automation, including cluster load shedding, automatic recovery, and observability toolingImprove test infrastructure reliability through flake detection, quarantine systems, and test state managementYou ...

Director, Networks and AI

Location
Greater London, England, United Kingdom
integration of AI and machine learning into network planning, operations, optimisation and maintenance. Identify, assess and scale AI use cases that improve network performance, incident response, capacity planning and operational effectiveness. Shape a safe and governed approach to AI adoption, ensuring appropriate ownership, auditability, human oversight and production ...

Head of Consumer Operations

Hiring Organisation
Blockchain
Location
London, UK
Employment Type
Full-time
high-performing operations team (reconciliation, asset movement, support ops), including workforce planning for distributed or outsourced functions as volume grows. Own consumer-facing incident response for operational disruptions (failed withdrawals, settlement delays, banking/exchange outages), including customer communication protocols during incidents. Growth, Expansion & Product SupportDrive operational readiness ...

Staff Software Engineer, Observability & Profiling

Hiring Organisation
Humanloop
Location
London, UK
Employment Type
Full-time
into the kernel, the network stack, or the hardwareHave excellent communication skills and enjoy partnering with internal teams to improve their operational visibility and incident response capabilitiesAre excited about building foundational infrastructure and comfortable navigating ambiguous, high-impact technical challenges, both independently and with a teamPreferred qualifications10+ years ...

Senior Full Stack Engineer (Realtime & Voice) Customer Experience Platform

Location
Greater London, England, United Kingdom
safety and compliance plumbing enterprise partners audit, including guardrails, content filtering, and PII redaction integration points Keep revenue-critical deployments healthy through observability, alerting, incident response, and SLA performance Build the platform capabilities forward-deployed engineers configure for partner telephony integrations and go-lives Raise the engineering ...

Vice President, Data Science

Hiring Organisation
S&P Global
Location
London, UK
Employment Type
Full-time
explainable, resilient, and regulator-ready. Own the full model lifecycle from problem framing, data exploration, feature engineering, back-testing, validation, deployment, monitoring, tuning, and incident response when models degrade or fail in production. Partner closely with engineering, data platform, product, and business teams to translate complex analytical problems ...

Product Engineer

Location
Greater London, England, United Kingdom
/CD pipelines and deployment strategies to ensure a smooth and reliable delivery process for the whole team Set up alerting, observability and incident response that catches issues before customers do Support data infrastructure and pipelines to ensure platform reliability and data availability Spot recurring operational burden ...

Senior Cloud Security Operations Engineer

Location
Greater London, England, United Kingdom
teams to design secure architectures, automate controls, and manage access to systems at scale. The role emphasizes hands-on engineering, strong collaboration, and rapid incident response. Cohere supports a remote-friendly, globally distributed team with London offices and flexible work options. #J-18808-Ljbffr ...

Lead Engineer- Syft

Location
Greater London, England, United Kingdom
Node.js and AWS Background in coaching and mentoring engineers Strong communication skills Growth mindset for continuous learning and sharing technical knowledge Ability to lead incident responses and resolve failure patterns Collaborative approach to problem-solving #J-18808-Ljbffr ...

Software Engineer

Location
Greater London, England, United Kingdom
work mixes coding (typically in Python, JavaScript/TypeScript, Java, Go or Ruby), technical design, code review, debugging, system architecture decisions, and on-call incident response. UK engineers operate across many product areas: consumer apps, fintech, infrastructure, AI/ML, gaming, defence and healthcare. The career has a famously ...

Lead Engineer- Syft

Location
Greater London, England, United Kingdom
complex technical ideas and build alignment. Your growth mindset drives a passion for continuous learning, experimentation, and sharing technical knowledge. An ability to lead incident responses, resolve failure patterns, and guide quality testing strategies will be highly valued. You offer a collaborative approach to problem-solving, welcoming diverse perspectives ...

Senior AI Software Engineer, Coding Agent

Location
Greater London, England, United Kingdom
skills and the ability to work on complex, ambiguous projects, delivering results in a fast‐paced environment. Experience with observability, performance optimization, and production incident response. Bachelor's degree in Computer Science, AI, Machine Learning, or a related technical field (or equivalent practical experience). Preferred Qualifications Direct experience ...

MLOps Platform Developer / Full-Stack AI Engineer

Hiring Organisation
BluetownOnline Ltd
Location
London, United Kingdom
Employment Type
Permanent
work on large tables. Serverless or cloud deployment experience (Vercel, AWS Lambda, Cloudflare Workers or similar) and comfort operating what you ship: logging, monitoring, incident response. API integration experience against third-party systems (REST, OAuth2/JWT, webhooks), including at least one messy real-world system of record. Python ...

Data, AI & Security Lawyer

Location
Greater London, England, United Kingdom
requirements of investigatory powers legislation such as the Investigatory Powers Act 2016 Experience advising on the legal aspects of data and security incident response. CIPP/E, CIPM, AIGP certification or equivalent. Existing UK government Developed Vetting clearance. Our Package Tailored benefits make a real difference. That ...

Senior Software Engineer - London

Location
Greater London, England, United Kingdom
Engineers sophisticated backend solutions involving API versioning, caching strategies, and complex data migration plans. Operational Maturity: Leads observability and SRE practices; defines SLOs, manages incident responses, and conducts blameless post-mortems. Security & Risk: Oversees operational security, including secrets hygiene and dependency risk management, to ensure a hardened production environment. ...

Senior Software Engineer - London

Hiring Organisation
Hackajob Ltd
Location
South West London, London, United Kingdom
Employment Type
Permanent
Engineers sophisticated backend solutions involving API versioning, caching strategies, and complex data migration plans. Operational Maturity: Leads observability and SRE practices; defines SLOs, manages incident responses, and conducts blameless post-mortems. Security & Risk: Oversees operational security, including secrets hygiene and dependency risk management, to ensure a hardened production environment. ...

Senior Cloud Engineer

Location
Greater London, England, United Kingdom
enhance observability and monitoring, ensuring meaningful alerting, clear operational dashboards, and rapid diagnosis of incidents. Raise the bar for operational maturity, including incident management, root-cause analysis, and continuous improvement. Security, Compliance & Governance Ensure workloads are designed, deployed, and operated in line with cloud security, compliance, and governance requirements. … code (e.g. Terraform), CI/CD pipelines, and automation tooling. Hands-on expertise in SRE and operational excellence, including monitoring, alerting, reliability improvement, and incident response. Proven ability to lead technical initiatives end-to-end, balancing robustness, efficiency, and developer usability. Experience operating and supporting Kubernetes-based workloads ...