1 to 25 of 496 Remote Root Cause Analysis Jobs

Service Reliability Analyst

Location
Greater London, England, United Kingdom
Sitting at the intersection between technical teams, internal stakeholders, and external clients, your primary purpose is to deliver service excellence by ensuring incident communications, Root Cause Analysis (RCA) reports, and technical workarounds are clear, structured, and business-focused. In this early-career, development-focused position, you will … responsible for translating complex diagnostic details into actionable insights, driving ticket quality, coordinating offshore analysis tasks, and tracking preventive actions to ensure long-term system stability across Finova's financial technology platform. About you: Communication & Technical Translation: Exceptional written and verbal English communication skills, with a proven ability ...

SIAM Problem Analyst

Location
United Kingdom
focused, ensuring that underlying causes of incidents are investigated and addressed through effective supplier governance, collaboration, and service integration. The Problem Analyst will drive root cause analysis, coordinate permanent resolutions, and maintain oversight of problem records while ensuring suppliers meet their problem management obligations and performance commitments. … Through proactive trend analysis, continual improvement, and data-driven insights, the role contributes to improved service stability, reduced incident volumes, enhanced end-user experience, increased operational resilience, and the delivery of measurable business value. Hybrid working: The places that you work from day to day will vary according ...

Continuous Improvement and Organisational Learning Lead

Hiring Organisation
Great British Nuclear
Location
Warrington, Cheshire, United Kingdom
Employment Type
Permanent
Salary
£75900/annum
improvement and operational excellence across GBE-N and its delivery partners. You will be responsible for developing and maintaining robust processes for incident investigation, root cause analysis, and organisational learning. You will work closely with Integrated Project Teams (IPTs), Delivery Partners, Technology Partners, Construction Partner … notifiable events to enforcing authorities, capturing and sharing organisational learning across GBE-N and its partners. Develop and implement a structured investigation process and root cause analysis methodology, leading and supporting investigations into events and incidents, identifying root and contributory causes where required. Design and implement ...

IT Problem Manager

Hiring Organisation
Surrey County Council
Location
Reigate, Surrey, United Kingdom
Employment Type
Permanent
Salary
GBP 55,486 - 60,898 Annual
challenges, driving continual improvement and influencing technical teams to deliver outstanding customer outcomes. You will play a key role in reducing recurring incidents, identifying root causes and developing Problem Management capability across a large and diverse technology estate supporting essential public services. As Surrey County Council continues its digital … identification, investigation and resolution of recurring and high-impact IT issues. Working across technical, service delivery, security and supplier teams, you will lead root cause analysis activities, manage known errors and help implement permanent solutions that improve service reliability and performance. This is a highly collaborative role ...

CPE (Customer Premises Equipment)Technical Investigation Engineer - Home Hub

Hiring Organisation
Flint UK Technology Services
Location
Ipswich, Suffolk, United Kingdom
Employment Type
Contract
Contract Rate
GBP Annual
client's Home Hubs and associated CPE devices. The successful candidate will take ownership of challenging technical issues from initial escalation through to root cause identification, acting as the bridge between In-Life Operations and the firmware development teams. You will be responsible for recreating real-world customer … technical evidence, and driving resolution through supplier and development teams. Above all, they are looking for someone who is technically curious, tenacious in pursuing root cause, and an excellent communicator who can clearly articulate complex technical issues to engineers, suppliers, and senior stakeholders alike. What ...

Lead Security Operations Center Analyst (f/m/d)

Location
Reading, England, United Kingdom
incidents. You will lead investigations into sophisticated threats such as advanced persistent threats (APTs), malware outbreaks, and targeted attacks, whilst performing hands on analysis of security events, forensic evidence collection, and root cause analysis. You will also drive the development and enhancement of detection capabilities across SIEM … proactive threat hunting activities, leveraging threat intelligence, application logs, and infrastructure telemetry to uncover indicators of compromise or stealthy threat activity. Perform in-depth analysis of logs, API configurations and traffic, container environments, network data, application and infrastructure architecture, as well as data center hosting environments to support threat ...

Broadband Test Engineer

Hiring Organisation
Infosec
Location
Ipswich, Suffolk, East Anglia, United Kingdom
Employment Type
Contract, Work From Home
Contract Rate
£300 - £350 per day
resolving complex issues across Home Hubs, Wi-Fi platforms and broadband customer equipment. This is a highly technical role focused on fault investigation, root cause analysis and defect resolution rather than traditional support. Required Experience Home Hub, gateway or broadband CPE troubleshooting TCP/IP networking … technologies and RF fundamentals Packet capture analysis Linux environments Python scripting FTTP and broadband technologies IPv4/IPv6 Core dump and crash analysis Technical log analysis Root cause investigation Technical Environment Home Hubs Broadband CPE Digital Voice TR-069/TR-369 Linux Python ...

Infrastructure Support Engineer

Location
United Kingdom
cloud platforms and helping protect the business against emerging security threats. If you love solving complex technical challenges, thrive on getting to the root cause of issues, and enjoy sharing your expertise with others, we'd love to hear from you. The role The go-to expert … complex technical issues, acting as the highest point of escalation within our IT support team. You'll investigate and resolve critical incidents, drive root cause analysis and work closely with internal teams and external partners to deliver a secure, reliable and high-performing technology environment. Acting ...

Operations Support & Insight Coordinator

Hiring Organisation
Lightfoot
Location
Exeter, Devon, United Kingdom
Employment Type
Full-Time
Salary
Competitive salary
ownership of repeatable reporting, systems, administrative and support activity. The role combines operational insight, case support, Salesforce and Lightfoot Portal administration, issue investigation and Root Cause Analysis, process documentation and systems testing. It will provide flexible capacity across Operations where required, while existing teams retain ownership … straightforward Salesforce and Portal queries, gathering relevant evidence and escalating genuine system or development issues to specialist resource. Support operational issue investigation and Root Cause Analysis, collating case, system and customer information, identifying recurring themes and coordinating follow up with the appropriate owner. Maintain Operations process documentation ...

Alpha Data Content & Configuration Analyst, Assistant Vice President

Location
Greater London, England, United Kingdom
quality reference data across CRD (e.g., securities, issuers, pricing, curves) Investigate and resolve data exceptions, discrepancies, and breaks impacting trading and downstream processes Perform root cause analysis and drive remediation of recurring data issues Support implementation of data quality rules, controls, and governance frameworks Instrument setup … front-to-back investment data flows (Front Office, Middle Office, Back Office) Hands-on experience in: Data quality management, Exception handling and reconciliation, Root cause analysis Proficient in SQL and data analysis techniques Excellent analytical and problem-solving capabilities High level of attention to detail with ...

Continuous Improvement Specialist - Stockless Fulfilment

Location
Yeovil, England, United Kingdom
demand, ensuring alignment to strategic priorities, customer impact and operational performance within capacity constraints and a defined improvement pipeline. Lead end-to-end discovery, analysis and redesign of Stockless Fulfilment processes across order capture, supplier routing, fulfilment, dispatch, tracking, exceptions, returns and refunds. Apply structured problem-solving techniques (e.g. … root cause analysis, value stream mapping, Pareto analysis) to identify inefficiencies, risks, bottlenecks and failure demand. Design and document future-state processes that improve flow, reduce manual effort and increase straight-through processing, ensuring solutions are practical, scalable and implementable. Maintain clear, standardised process documentation, SOPs ...

Site Reliability Engineer / Senior Engineer

Location
Greater London, England, United Kingdom
excellence Establishing and enhancing comprehensive monitoring, observability, logging, and alerting systems to provide deep visibility into system health and performance Conducting in‐depth technical analysis of production platforms to identify, prioritize, and remediate performance, stability, and resilience issues Exploring and partnering with teams to prototype, and implement AI/… solutions to enhance production support efficiency, including predictive incident detection, intelligent alerting, automated root cause analysis, and proactive anomaly detection Leading root cause analysis (RCA) activities and implement preventative measures Building strong partnerships with global stakeholders across development, infrastructure, and production functions to proactively ...

Data Engineer (SC Cleared)

Hiring Organisation
scrumconnect ltd
Location
City, Newcastle Upon Tyne, United Kingdom
Employment Type
Permanent
Salary
GBP 65,000 - 75,000 Annual
services for storage, compute, and analytics, you will help deliver reliable, well-governed data assets to downstream users. You will apply strong data analysis skills to identify root causes of data issues, work with dimensional data models and slowly changing dimensions, and implement infrastructure as code using Terraform. … distributed cloud infrastructure. Workflow orchestration Configure and manage Apache Airflow DAGs for task orchestration, ensuring reliable scheduling, monitoring, and execution of data processing workflows. Root cause analysis Perform data analysis to identify and resolve root causes of pipeline failures and data quality issues - including reviewing ...

Senior Business Systems Analyst

Location
Cambridge, England, United Kingdom
System Integration: Oversee end-to-end supply chain flows and maintain smooth integration between Oracle EBS and downstream platforms (Anaplan, SAP, ASCP).* Escalation & Root Cause Analysis: Act as the primary escalation point for critical ERP issues, conducting deep root cause analysis ...

Senior Business Systems Analyst

Location
Greater London, England, United Kingdom
System Integration: Oversee end-to-end supply chain flows and maintain smooth integration between Oracle EBS and downstream platforms (Anaplan, SAP, ASCP).* Escalation & Root Cause Analysis: Act as the primary escalation point for critical ERP issues, conducting deep root cause analysis ...

Data Engineer

Location
Newcastle upon Tyne, England, United Kingdom
services for storage, compute, and analytics, you will help deliver reliable, well-governed data assets to downstream users. You will apply strong data analysis skills to identify root causes of data issues, work with dimensional data models and slowly changing dimensions, and implement infrastructure as code using Terraform. … distributed cloud infrastructure. Workflow orchestration Configure and manage Apache Airflow DAGs for task orchestration, ensuring reliable scheduling, monitoring, and execution of data processing workflows. Root cause analysis Perform data analysis to identify and resolve root causes of pipeline failures and data quality issues - including reviewing ...

Continual Improvement and Innovation Manager

Hiring Organisation
Tiger Resourcing Group Ltd
Location
Redhill, Surrey, South East, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£80,000
opportunities to improve reliability, efficiency and customer outcomes. Develop structured improvement roadmaps aligned with contractual objectives, service strategy and wider business priorities. Facilitate workshops, root cause analysis and improvement sessions with operational and technical stakeholders. Develop business cases and improvement proposals, clearly defining objectives, costs, risks, dependencies … ITIL Continual Improvement practices and service management principles . Experience working across service management, operational support, engineering and architecture teams . Experience facilitating workshops, root cause analysis and stakeholder engagement activities. Strong analytical and problem-solving capability. Experience analysing service performance metrics, operational data and service reporting ...

Senior Systems Support Analyst Team Leader

Location
Chichester, England, United Kingdom
Level 1 and Level 2 Service Desk colleagues. Support the management of major incidents, ensuring timely resolution and effective communication with stakeholders. Conduct root cause analysis and contribute to problem management activities to prevent recurring issues. Develop, maintain, and promote knowledge management resources to improve service quality … Windows operating systems, Active Directory, and enterprise technologies. Proven experience leading technical investigations and resolving complex IT incidents. Experience supporting major incident response, root cause analysis, and problem management activities. Strong leadership, mentoring, and coaching skills with the ability to support and develop technical colleagues. Experience using ...

Senior Incident Response Consultant 2

Location
Oxford, England, United Kingdom
neutralize cyber threats. Specializing in industry-standard forensic tools and Sophos technologies, the team provides comprehensive investigations, response actions, remediation guidance, and root cause analysis to combat a wide range of cybersecurity incidents. As a Senior Incident Response Consultant on the Sophos DFIR team, you will … have been taken by both your team and the customer to effectively neutralize the threat. Additionally, you will be tasked with conducting a thorough root cause analysis to determine the origin of the incident, including identifying whether any data exfiltration occurred, provided the necessary evidence is available. ...

Senior Specialist Engineer - Networks

Hiring Organisation
UK Health Security Agency
Location
Birmingham, Leeds, Liverpool or London (Canary Wharf), E14 4PU, United Kingdom
Salary
£56185.00 to £70566.00
with security standards and evolving threat landscapes. Act as the senior network authority during Major Incident Response Team (MIRT) activities, leading diagnosis, resolution, communication, Root Cause Analysis (RCA) and Post Incident Reviews (PIR). Develop incident runbooks, continuity plans and mitigation strategies to reduce risk, improve service … major incidents, leading network diagnosis and resolution activities, collaborating within multi-disciplinary Major Incident Response Teams, and managing complex, high-pressure situations. Experience producing Root Cause Analysis (RCA) reports, Post-Incident Reviews (PIRs), and maintaining incident runbooks, playbooks, and escalation procedures while driving continuous service improvement ...

Data Quality Analyst (global role – in a virtual working environment)

Hiring Organisation
Grant Thornton International Ltd
Location
United Kingdom
quality issues Produce regular reporting on data quality of member firms. Support ongoing training for member firms to continually improve data quality Issue Identification, Root Cause Analysis & Resolution Detect data defects and anomalies in member firm GRD data submissions Perform root cause analysis … monitoring and improving data quality Understanding of data quality controls, governance principles and data lifecycle management Strong Excel skills, including Pivot Tables and data analysis Ability to interpret data and communicate findings clearly to both technical and non-technical stakeholders Process improvement mindset with the ability to identify inefficiencies ...

Senior Linux DevOps Engineer

Hiring Organisation
RedTech Recruitment Ltd
Location
City of London, London, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£90,000
Ansible Experience building and managing CI/CD pipelines using GitLab CI/CD, GitHub Actions, Jenkins or similar Experience with incident response, root cause analysis and driving improvements to the reliability and performance of production systems Commercial experience with Microsoft Azure or another major cloud platform … Linux-based production environments Build and improve containerised environments using Docker and Kubernetes Diagnose complex infrastructure, networking and performance issues, supporting incident response and root cause analysis Develop automation, Infrastructure as Code and CI/CD processes to improve engineering efficiency and reliability Implement and enhance monitoring ...

Site Reliability Engineer

Hiring Organisation
REVYBE IT RECRUITMENT LIMITED
Location
City of London, London, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£85,000
improve SLIs, SLOs, and reliability metrics Proactively identify and resolve performance, availability, and reliability issues Lead and contribute to incident response, troubleshooting, and root cause analysis Automate operational processes and eliminate repetitive manual tasks Work closely with software engineers to improve deployment processes, system reliability, and developer … logging, tracing, alerting, and system health Experience troubleshooting complex production environments Understanding of SLIs, SLOs, SLAs, and error budgets Experience with incident management and root cause analysis Good understanding of cloud networking, security, and infrastructure fundamentals Strong scripting/automation skills A strong understanding of reliability, scalability ...

Network & Infrastructure Tooling / Automation Specialist/ Architect - freelance - hybrid, London, UK

Location
Greater London, England, United Kingdom
unified operational and observability platforms. Enable traffic engineering, QoS, and policy enforcement through code-driven workflows. Observability, Operations & Resilience Develop automation for fault detection, root cause analysis, and remediation. Integrate telemetry, logs, and metrics into observability platforms. Support SRE-style practices including error budgets, reliability metrics … OSPF, EIGRP, Hybrid WAN, SD-WAN, MPLS, Traffic engineering, QoS, ExpressRoute, Direct Connect Observability & Operations : Observability platforms, Telemetry, Logs, Metrics, Fault detection automation, Root cause analysis, Remediation automation, SRE practices, Reliability metrics Platform, Cloud & Service Integration: IPAM integration, CMDB integration, Platform engineering, DevOps, Lifecycle management, Cloud networking ...

Software Engineer Lead - Site Reliability

Location
Telford, England, United Kingdom
failover, degradation handling and recovery testing. Embed reliability, security, performance and operational-readiness expectations into engineering delivery. Support major incidents, post-incident reviews and root-cause analysis, ensuring improvement actions are owned and delivered. Use DORA, availability, recovery and operational metrics to identify risks and evidence progress. … expertise in Terraform, GitHub, GitHub Actions, CI/CD pipeline design, deployment automation and release support. Practical knowledge of incident management, problem management, root-cause analysis and operational readiness. Good awareness of Site Reliability Engineering principles, with the ability to apply them pragmatically to DevOps and platform ...