1 to 25 of 36 Incident Response Jobs in Cambridgeshire

Senior Manager, Cybersecurity Incident Response

Hiring Organisation
ARM
Location
Cambridge, Cambridgeshire, UK
Employment Type
Full-time
Overview: Interested in defending a global tech company from the latest cyber threats? Arm is seeking a passionate, experienced Senior Manager of Cybersecurity Incident Response to join our growing Cyber Defence Operations (CDO) team, protecting Arm against current and future cyber-attacks! Situated within Arm's Enterprise Security … function, this role will lead Arm's global incident response team across the US, UK and India, including acting as a senior technical and operational leader for major cyber incidents. CDO enables Arm to be successful, delivering scalable and defendable security services that not only provide ...

Site Reliability Engineer

Location
Cambridge, England, United Kingdom
continuous improvement of the Bango Platform end-to-end — from the infrastructure and pipelines that build and deploy it, to the observability and incident response that keep it running to agreed service levels. You combine three things that have historically sat in separate teams: platform and cloud infrastructure … engineering, automation and delivery pipeline ownership, and proactive/reactive reliability engineering including incident response and customer impact management. You design, build and operate the automation, observability and platform capabilities that other engineering teams rely on, and you are equally comfortable diagnosing a live incident ...

Cyber Security Analyst

Hiring Organisation
Carter Jonas
Location
Peterborough, Cambridgeshire, UK
Employment Type
Full-time
team in Peterborough to help protect Carter Jonas's systems, data, and information assets. The role combines hands-on security operations with vulnerability management, incident response, cyber risk, governance, assurance and continual improvement of the organisation's security posture. The post holder will work with IT, Compliance … events across key platforms including SIEM, EDR/XDR, Microsoft Defender, Microsoft Sentinel, email security, cloud security and network monitoring tools. Support timely incident response, containment, remediation, recovery and post-incident review activity with internal teams and third-party providers. Maintain the vulnerability management programme, including scanning ...

EU Resilience Lead

Location
Huntingdon, England, United Kingdom
recovery capabilities for critical services. You Have: 8+ years of experience leading or advising on enterprise technology, resilience, cybersecurity, infrastructure, disaster recovery, business continuity, incident response, or crisis management disciplines for large, complex European or global organizations Experience with advising executive stakeholders and leading senior-level advising … current capabilities, and developing practical improvement roadmaps that help clients harden infrastructure, improve alignment to NIST CSF categories, strengthen recovery and continuity strategies, test response capabilities, and mature program governance over time Experience in a consulting or corporate advisory role, including with a top-tier consulting company, specialized cybersecurity ...

EU Resilience Lead

Location
Cambridgeshire and Peterborough, England, United Kingdom
capabilities for critical services. You Have: 8+ years of experience leading or advising on enterprise technology, resilience, cybersecurity, infrastructure, disa ster recovery, business continuity, incident response, or crisis management disciplines for large, complex European or global organizations Experience with advising executive stakeholders and leading senior-level advising … capabilities, and develop ing practical improvement roadmaps that help clients harden infrastructure, improve alignment to NIST CSF categories, strengthen recovery and continuity strategies, test response capabilities, and mature program governance over time Experience in a consult ing or corporate advisory role, including with a top-tier consulting company ...

Principal AI Platform Engineer

Location
Cambridge, England, United Kingdom
runtime platforms that make AI services reliable, secure, observable and supportable at Arm scale. You will work across Kubernetes, cloud, identity, secrets, networking, telemetry, incident management and automation to provide the production foundation for Arm's AI platform. Participate in production support and our paid on-call rota … high impact incident response. Production AI runtime platforms: Build/deploy and operate the infrastructure for centrally hosted AI platform services, including MCP server infrastructure, model gateway services and supporting control-plane components. Design runtime patterns for isolation, scalability, secure execution, capacity management and cost-aware operation. Automate provisioning ...

SENIOR INFORMATION ASSURANCE ENGINEER

Hiring Organisation
Leidos Innovations UK Limited
Location
Huntingdon, Cambridgeshire, East Anglia, United Kingdom
Employment Type
Permanent
Salary
£75,000
with experience developing assurance artefacts such as SyOPs, RMADs, Security Impact Assessments, risk assessments, and accreditation evidence. Experience leading assurance initiatives, audits, control assessments, incident response activity, and vulnerability management using tools such as Nessus, Trivy, Tenable, or similar platforms. Ability to brief and influence senior stakeholders, mentor … Security Impact Assessments, Security Management Plans, risk assessments, and governance documentation. Drive maturity of assurance processes, tooling, reporting, and governance across multiple programmes. Lead incident response capability development, including response plans, playbooks, testing, exercising, and Tabletop Exercises (TTXs). Manage vulnerability, risk, audit, control assessment, and compliance ...

Senior Detection and Response Engineer

Location
Cambridge, England, United Kingdom
queries and monitoring logic across cloud, endpoint, network and application environments. Develop tooling and automation to improve security telemetry, alert enrichment, investigation workflows and response times. Conduct threat hunting using adversary behaviours, TTPs and frameworks such as MITRE ATT&CK, incorporating findings into security controls and detections. Create … continuously improve incident runbooks, playbooks and detection processes based on findings from real-world investigations. Work with the external SOC and internal engineering teams to strengthen monitoring coverage, investigate escalations and continuously improve detection and response capability. Participate in an on-call rotation. Requirements Hands-on experience investigating ...

Principal Software Engineer-AI

Hiring Organisation
Appcast
Location
Croydon, Cambridgeshire, UK
agentic runtime, auth, data retrieval, eval tooling)Experience running AI systems in production at scale, including observability, cost and capacity planning, regression detection, and incident response for AI-powered applicationsExperience operating production distributed systems on AWS/Azure, with a strong grasp of reliability, observability, and incident response at scaleDeep knowledge of cloud-native technologies, serverless applications, event-driven architectures, data and inference pipelines, relational, NoSQL, and vector databases, and modern software architecture patternsProven track record of owning multi-year technical strategy and architectural roadmaps, guiding teams from AI prototype through production deployment, and influencing ...

Technical Support Manager

Location
Cambridge, England, United Kingdom
team's time-zone spread to ensure responsive support well beyond any single region. This position owns end-to-end employee support operations including incident response, request fulfillment, onboarding, offboarding, and executive‐level support. The Technical Support Manager will drive continuous improvement of support workflows, knowledge bases … coverage model using EMEA and India working hours, with well‐defined cross‐region handoff processes Own end‐to‐end employee support operations including incident response, request fulfillment, onboarding, offboarding, and white‐glove support for leadership Define and maintain SLAs and quality standards; track and report on ticket volume ...

Technical Support Manager

Hiring Organisation
Roku
Location
Cambridge, Cambridgeshire, UK
Employment Type
Full-time
team's time-zone spread to ensure responsive support well beyond any single region. This position owns end-to-end employee support operations including incident response, request fulfillment, onboarding, offboarding, and executive-level support. The Technical Support Manager will drive continuous improvement of support workflows, knowledge bases … coverage model using EMEA and India working hours, with well-defined cross-region handoff processesOwn end-to-end employee support operations including incident response, request fulfillment, onboarding, offboarding, and white-glove support for leadershipDefine and maintain SLAs and quality standards; track and report on ticket volume, resolution time ...

Site Reliability Engineer

Location
Cambridge, England, United Kingdom
internal developer platform* Collaborate with **DevSecOps** to integrate security, compliance, and resilience practices* Contribute to cross-team initiatives that improve reliability across the stack### **Incident & Operational Excellence*** Play a key role in **incident response**, particularly within your specialism* Contribute to **on-call rotations** and continuous improvement ...

IT Internal Systems Apprentice

Hiring Organisation
NETSUPPORT LTD
Location
NETSUPPORT HOUSE, TOWNGATE EAST, MARKET DEEPING, PETERBOROUGH, England, United Kingdom
Employment Type
Advanced Apprenticeship
Salary
£15,392 a year
location customers, including equipment checks, hardware installations, cabling and site visits Assisting with cyber security operations including vulnerability management, endpoint protection, security monitoring and incident response activities Supporting business projects involving infrastructure upgrades, automation and new technology implementations Working closely with other departments across the business to resolve ...

Senior Platform Engineer

Location
Cambridge, England, United Kingdom
facilitate seamless product feature releases, adhering to established TechOps peer-review processes. Monitor application-level health within the cluster, collaborating with TechOps on incident response and driving post-mortems for engineering-led deployments. Maintain clear documentation, deployment runbooks, and developer self-service guides to streamline the engineering onboarding ...

Platform and Application Engineer

Hiring Organisation
Appcast
Location
Cambridge, Cambridgeshire, UK
responsibility from requirements through to planning, implementation, and delivery and will participate in production support and our paid on-call rota for high impact incident response.At the heart of our approach is a genuine passion for improving the lives of other Engineers at Arm. We maintain a complete focus ...

Senior DevOps Engineer

Location
Cambridge, England, United Kingdom
responsibility from requirements through to planning, implementation, and delivery and will participate in production support and our paid on-call rota for high impact incident response. At the heart of our approach is a genuine passion for improving the lives of other Engineers at Arm. We maintain a complete ...

Staff Private Cloud Engineer

Location
Cambridge, England, United Kingdom
patterns for reliable and repeatable builds. Develop platform services, APIs (FastAPI) and automation using Python. Drive improvements in reliability, scalability and performance. Lead incident response, fixing and root cause analysis. Work across compute, networking and storage layers. Collaborate with engineering teams to improve platform usability. Mentor engineers ...

InfraOps Engineer

Location
Cambridge, England, United Kingdom
Build self-service internal platform tools and robust CI/CD pipelines enabling high-velocity deployments. SRE & Reliability: Apply SRE principles, lead high-severity incident responses, champion root-cause analysis, and ensure strict SLA/SLO delivery. What we’re looking for Experience: 5+ years in Platform Engineering ...

Senior Engineer, Data Infrastructure

Location
Cambridge, England, United Kingdom
Head of Data Infrastructure. The successful candidate will normally work from the campus at least three days per week, participate in operational support and incident escalation, and must have permission to work in the UK. Responsibilities: Lead the architecture and evolution of our AWS-based infrastructure across cloud services … data movement and lifecycle management. Partner with Cloud Engineers and Data Engineers to deliver integrated solutions, providing practical mentoring through technical design, code review, incident response and operational improvement. Work with AI scientists, computational biologists and scientific teams to translate model development, training, evaluation, inference and data requirements ...

Senior Engineer, Data Infrastructure

Location
Cambridge, England, United Kingdom
Head of Data Infrastructure. The successful candidate will normally work from the campus at least three days per week, participate in operational support and incident escalation, and must have permission to work in the UK. Responsibilities: Lead the architecture and evolution of our AWS-based infrastructure across cloud services … data movement and lifecycle management. Partner with Cloud Engineers and Data Engineers to deliver integrated solutions, providing practical mentoring through technical design, code review, incident response and operational improvement. Work with AI scientists, computational biologists and scientific teams to translate model development, training, evaluation, inference and data requirements ...

Infrastructure Team Lead

Hiring Organisation
Speechmatics
Location
Cambridge, Cambridgeshire, UK
Employment Type
Full-time
operational stability across our infrastructure, including hardware monitoring, troubleshooting and disaster recovery testingActing as the escalation point for technical and operational issues, leading major incident response and driving resolutionLeading infrastructure strategy, including capacity and project planningDriving standardisation, automation and optimisation to reduce toilLeading vendor relationships with hardware vendors ...

Staff Full Stack Software Engineer

Hiring Organisation
Appcast
Location
Cambridge, Cambridgeshire, UK
aligned to business outcomes.Please note: There will be the requirement to participate in production support and our paid on-call rota for high impact incident response.Responsibilities:Working hands-on with a variety of technologies, to provide a first-class engineering experience for Arm's hardware and software engineers. ...

Senior Software Engineer, Infrastructure / Efficiency / Productivity

Hiring Organisation
Roku
Location
Cambridge, Cambridgeshire, UK
Employment Type
Full-time
First Automation & Intelligent Tooling Define and lead an AI-first automation roadmap for Engineering Infrastructure and Enterprise Tooling (e.g., reducing waste, accelerating reviews, improving incident response, increasing release confidence).Architect and ship AI/LLM-enabled workflow automation across the SDLC (e.g. risk assessment, automated build, test … change-impact analysis, auto-triage of build/test failures).Establish policies and guardrails for AI usage in internal tools: data handling, prompt/response logging, model/provider selection, evaluation, fallback behaviors, and abuse prevention.Dev Productivity & Enterprise Tooling (Global Scale)Set technical direction and standards for the globally ...

Operations Team Lead (Production & Reliability)

Location
Cambridge, England, United Kingdom
Operational readiness for new releases Safe production access and change coordination Production is a high-discipline environment. You make sure it stays that way. Incident Management You own the full lifecycle: High‐signal alerting and fast detection Structured incident response Clear internal and customer communication Blameless postmortems … Escalations are fast and predictable. Monitoring & Reliability Define SLIs/SLOs for critical systems Improve visibility across availability, latency, errors, and saturation Track MTTR, incident frequency, and escalation trends Drive reliability roadmap initiatives We measure reliability, and improve it continuously. Team Leadership Lead and grow the Operations team ...

Junior Cloud Reliability Engineer

Hiring Organisation
Jagex
Location
Cambridge, Cambridgeshire, UK
Employment Type
Full-time
platform through Infrastructure-as-Code. Contribute to the management of our Linux virtual machine fleet, including monitoring, configuration management, and workload optimisation. Participate in incident response and post-mortems, developing your understanding of how complex systems fail and how to make them more resilient. What we're looking ...