2,151 to 2,175 of 2,428 Incident Response Jobs in the UK

Site Reliability Engineer, Big Data (Remote, International)

Hiring Organisation
PulsePoint
Location
United Kingdom, UK
Employment Type
Full-time
hybrid infrastructure: bare-metal on-prem, cloud, and the integration between them. You own the lifecycle from architecture through deployment to capacity planning and incident response. What you'll work onKafka architecture, topic design, governance, partition strategy, throughput and latency optimization. Ceph operations, pool design, placement optimization, capacity planning. … Operational automation, reduce manual work, faster incident response, preventive systems. SQL Server backup and recovery pipelines, basic cluster support. Data team tooling with self-service capabilities and observability. TechnologyApache Kafka for messaging layerHadoop and Ceph as distributed storage layerSQL Server backup and recoveryTerraform, Ansible, Puppet, ArgoCD for operational ...

Technical Account Manager - London London

Location
Greater London, England, United Kingdom
dates, data is privileged, and clients hold Legora to the same standard of care they apply to their own work. Part technical advisor, part incident lead, part operating model owner, you hold the technical health of Legora's most complex client relationships and define the playbook the whole function … their primary point of escalation and continuity, and read usage and adoption signals to get ahead of risk. Lead in a crisis: Act as incident lead on Sev1 and Sev2 events, owning coordination, customer communication, and the path to resolution. Drive post‐incident reviews as trust‐building artifacts ...

Principal Platform Advisor - London London

Location
Greater London, England, United Kingdom
dates, data is privileged, and clients hold Legora to the same standard of care they apply to their own work. Part technical advisor, part incident lead, part operating model owner, you hold the technical health of Legora's most complex client relationships and define the playbook the whole function … their primary point of escalation and continuity, and read usage and adoption signals to get ahead of risk. Lead in a crisis: Act as incident lead on Sev1 and Sev2 events, owning coordination, customer communication, and the path to resolution. Drive post-incident reviews as trust-building artifacts ...

Software Engineer - Identity & Access Management - Full-Stack

Location
Greater London, England, United Kingdom
for. You understand that \"speed\" doesn't mean \"sloppy\"—it means building reliable software that doesn't break at 2am. You will participate in incident response, learn from postmortems, and help the team continuously raise the bar on code quality and operational health Qualities we look for Ownership … follow through when things go wrong. You're proactive about flagging issues and asking for help when you need it, and you treat each incident as an opportunity to learn Strong Engineering Fundamentals: You have a solid grasp of data structures, algorithms, and clean coding principles. You can take ...

AI Infrastructure Engineer, Sandbox Platform

Hiring Organisation
Scale AI
Location
London, United Kingdom
Salary
£ 80 K
debugging skills and the ability to navigate performance/security tradeoffs in production systemsComfort with ambiguity, and the ability to context-switch between reactive incident work and proactive product developmentNice to haves:Experience as a founder or early engineer at an infrastructure-focused startup, owning a product … Exposure to snapshotting and restore techniques (e.g., CRIU, VM snapshots, overlays)Open-source contributions to systems or developer-tools projectsHistory of on-call/incident response for production systemsPLEASE NOTE: Our policy requires a 90-day waiting period before reconsidering candidates for the same role. This allows ...

Payment Operations Lead

Location
Greater London, England, United Kingdom
exceptions. You'll design the workflows, set the guardrails, and audit what the agents produce Write the playbooks: new processor onboarding, new market launch, incident response, dispute handling Onboard new processors, banking partners and local acquiring as we launch new markets Own payments performance Own approval rates … updater, 3DS strategy and exemption handling Run chargeback operations and fraud strategy. Set thresholds, represent disputes, keep merchants inside scheme limits Own the payment incident path. When authorisations drop in one market at, you're the one who spots it and fixes it Own the economics Build full visibility ...

AI Infrastructure Engineer, Sandbox Platform London, UK Apply →

Location
Greater London, England, United Kingdom
skills and the ability to navigate performance/security tradeoffs in production systems Comfort with ambiguity, and the ability to context-switch between reactive incident work and proactive product development Nice to haves: Experience as a founder or early engineer at an infrastructure-focused startup, owning a product … snapshotting and restore techniques (e.g., CRIU, VM snapshots, overlays) Open-source contributions to systems or developer-tools projects History of on-call/incident response for production systems PLEASE NOTE: Our policy requires a 90-day waiting period before reconsidering candidates for the same role. This allows ...

Senior Software Engineer, Data & Orchestration

Hiring Organisation
wayve
Location
London, United Kingdom
Salary
£ 80 K
have the opportunity to help shape foundational infrastructure used across the company, supporting high-priority data runs, batch inference, evaluations and business-critical incident response.Key responsibilitiesDesign, build and evolve scalable data pipeline and orchestration infrastructure for Wayve’s next-generation autonomous driving data platformContribute to the revamp … prioritisationSupport batch inference and performance optimisation workflows, helping teams run evaluations in a cost-effective and scalable wayImprove platform reliability, quality and usability, including incident response workflows aligned to business-defined SLOsCollaborate with partner teams to understand their pipeline requirements and provide the underlying infrastructure that enables them ...

Chief Engineer

Hiring Organisation
Apex Systems US
Location
Hounslow, London, United Kingdom
Employment Type
Contract
Contract Rate
GBP 36 - 46 Hourly
environment. Key Responsibilities Provide technical leadership across all critical electrical and mechanical systems. Ensure maximum uptime and operational resilience of the facility. Lead incident response, troubleshooting and root cause analysis activities. Review and approve SOPs, MOPs, EOPs and switching schedules. Manage critical maintenance activities and shutdowns. Oversee specialist … Critical Cooling Chillers CRAH/CRAC units Chilled water systems Pumps and pressurisation systems Cooling towers HVAC controls and BMS systems Environmental monitoring Operations Incident management Planned maintenance Reactive maintenance Permit-to-work systems Change management Risk assessments and RAMS Contractor management Desired Qualifications HVAP or AP qualification (strongly ...

Senior Platform Engineer: Infra Leadership & Incident Focus

Location
Milton Keynes, England, United Kingdom
Platform Engineer to lead the design, deployment, and operation of large-scale IT infrastructure across Milton Keynes and Glasgow. You will own platform reliability, incident response, and capacity planning, partnering with stakeholders to align technology with business goals. You will mentor engineers, manage vendor relationships, and drive continuous ...

Senior GTM System Builder

Hiring Organisation
Mews
Location
United Kingdom
Salary
£ 60 K
adoption metrics, error rates, and data quality visibility instrumented before deploymentManage vendor relationships at pod level (Gong, Clay, LeanData, ZoomInfo, depending on assignment): SLAs, incident escalation, business reviews, and periodic build-vs-buy re-decisionsSupport other engineers through code review, pairing, and coaching on architecture trade-offs — and contribute … experienceHands-on experience building against Salesforce or equivalent CRM APIs: workflows, event-driven logic, and data integrations across multiple systemsComfortable owning live production systems: incident response, observability, vendor management, and data integrity — draws no hard line between building and runningCan translate operational pain into technical requirements and explain ...

Chief Cybersecurity Operations & Platform Delivery

Location
United Kingdom
Director, Cybersecurity Operations & Platform Delivery to lead our cybersecurity operations and platform delivery functions from a remote UK base. The role involves guiding incident response, discussing risk with executives, and shaping the DYOGUARD service portfolio. You will build roadmaps, drive accountability, and collaborate with sales, engineering, and operations ...

Infrastructure & Network Manager — Hybrid Role & Growth

Location
Accrington, England, United Kingdom
across the organisation. You will oversee networks, servers, cloud services and related technologies, ensuring reliable, secure, and high-performing IT services while guiding policy, incident response, and continuous improvement initiatives. #J-18808-Ljbffr ...

Senior Security Engineer: Cloud & DevSecOps Leader

Location
United Kingdom
technical authority, designing security architectures, implementing resilient controls, and driving automation across cloud and infrastructure environments. The role covers threat detection, vulnerability management, DevSecOps, incident response, compliance, and security assurance. #J-18808-Ljbffr ...

Software Support & Escalation Lead

Location
West of England, England, United Kingdom
resolution of complex challenges and fostering continuous improvement. You will support OIPT software for clients in EMEAI and external staff globally, lead escalation and incident response, and partner with Sales/Engineering to sustain customer uptime. #J-18808-Ljbffr ...

Senior Cyber Threat Intelligence Lead - Hybrid

Location
Portsmouth, England, United Kingdom
threat data into actionable intelligence to support proactive defence and risk reduction across the organisation. Day-to-day you will work with Security Operations, Incident Response, and stakeholders to strengthen enterprise resilience. #J-18808-Ljbffr ...

Senior Security Engineer, Infrastructure — Hybrid

Location
Greater London, England, United Kingdom
Engineering to implement scalable controls and improve visibility, resilience and security maturity across the organisation. The role covers networking, cloud, edge security, IAM and incident response, with emphasis on designing and maturing security controls, SIEM, WAF, Kubernetes security and governance alignment. #J-18808-Ljbffr ...

Cloud Technology Risk & Controls Leader

Location
Glasgow, Scotland, United Kingdom
coordinate audits and regulatory responses to maintain high risk posture. You will drive risk management initiatives, advance AI governance, and strengthen controls across access, incident response, vulnerability, and data protection in cloud environments. #J-18808-Ljbffr ...

International Production Information Security Manager

Location
Greater London, England, United Kingdom
Production Information Security Program across international productions. You will drive consistent tool usage, workflows, and controls and coordinate with global InfoSec, Risk, Threat Intelligence, Incident Response, Training, and Governance teams to align security with enterprise programs. Responsibilities include delivering dashboards on security trends, ensuring compliance, and maturing production ...

Global Cloud SRE Lead - Google Cloud & DevOps Strategy

Location
Glasgow, Scotland, United Kingdom
Infrastructure Platform - Cloud Foundational Services SRE team, collaborating across regions to maintain high availability and strong SLAs while leveraging enterprise AI tools to optimize incident response and reliability. #J-18808-Ljbffr ...

Cyber Threat Intelligence Analyst

Location
Manchester, England, United Kingdom
government sources, and produce concise reports and IOCs for stakeholders. You will monitor the threat landscape, identify new TTPs, track threat actors, and support incident response across Vanguard’s technology footprint. Strong collaboration with hunt teams and leadership is essential. #J-18808-Ljbffr ...

Senior Systems Engineer: Automation & Reliability

Location
Greater London, England, United Kingdom
will troubleshoot issues, automate tasks, and collaborate with development, support, and vendor teams to deliver reliable technology solutions. The role focuses on system stability, incident response, and ongoing improvements across Linux, UNIX, Windows environments with storage and virtualization components. #J-18808-Ljbffr ...

ServiceNow Platform Operations Lead

Location
Dartford, England, United Kingdom
will coordinate with engineering, IT support, MSPs and business stakeholders to ensure reliable service delivery and a positive user experience. You will lead incident response and problem management, drive continual improvement, and shape a modern, scalable platform as part of our digital transformation across the business. #J ...

SOAR Playbook Engineer - MXDR Automations & Integrations

Location
Manchester, England, United Kingdom
Engineer to design, develop, and optimize security automation using Splunk SOAR. You will focus on playbook engineering, Python automation, and API integrations to drive incident response workflows. You will deploy and continuously improve MXDR solutions for diverse clients, translating customer requirements into scalable automations while collaborating with stakeholders ...

Senior Platform Engineer - Private Cloud, VMware & Kubernetes

Location
Caerphilly, Wales, United Kingdom
high uptime for large public sector clients. The position offers hybrid working and a chance to lead automation initiatives, CI/CD pipelines, and incident response while mentoring engineers and shaping platform decisions. #J-18808-Ljbffr ...