101 to 125 of 377 Problem Management Jobs in the UK

Senior Wintel Engineer

Location
Redhill, England, United Kingdom
technological partner for the core business operations of its clients worldwide. It stands at the forefront of key sectors including Transport, Defence, Air Traffic Management and Space, alongside advanced Information Technology services delivered through Minsait, and cutting-edge capabilities in Sovereign AI, Cybersecurity, and Cyberdefence via IndraMind. The company … global benchmark in innovative transportation and mobility solutions. It is recognised as one of the world’s top three companies in public transportation management systems. Indra’s technology supports the daily journeys of over 78 million people, helping to reduce more than 10 million tonnes of CO2 emissions annually ...

HPC Senior Hardware Engineer

Location
Haywards Heath, England, United Kingdom
across our UK Data Centre.This role is responsible for ensuring the reliability, performance, and scalability of mission-critical HPC platforms, while leading hardware lifecycle management and infrastructure projects. You will work across compute, GPU, storage, networking, and data centre infrastructure, providing technical leadership and supporting the delivery of world … Engineering, Facilities, and technology vendors to deliver reliable, secure, and high-performing infrastructure.**Key Responsibilities****HPC Infrastructure & Hardware*** Lead the installation, maintenance, and lifecycle management of HPC server infrastructure, including CPU and GPU platforms.* Install, upgrade, and replace server components including processors, GPUs, storage, networking, and system hardware.* Perform ...

Technical Lead

Hiring Organisation
London Stock Exchange Group
Location
London, UK
Employment Type
Full-time
Jenkins/GitHub)Configuring and running Code/Binary scans using solutions like SonarQube, Semgrep, Blackbuck, Trivy, GitLeaks Veracode, etc. Configuring and using Secrets management tools like Vault and Cloud native solutionsBroad knowledge of SDLC Tools, specifically Build, Test and Deploy Automation tools, e.g., Maven, Gradle, Selenium, Ansible, etc. … Linux Servers. Cloud services (AWS/Azure/GCP). Managing incidents, change requests, service requests and driving TRT (Technical Recovery Team) calls. Strong problem solving skills on these platforms Minimum knowledge and understanding of financial markets are desirable. Ability to work independently and in a team environment. Ability ...

Senior Consultant - Cloud Operating Model Transformation

Location
Manchester, England, United Kingdom
clients in designing, assessing and evolving cloud operating models that are automation‐first and AI‐ready. You will work hands‐on across technology, service management and organisational change to help clients realise tangible business outcomes from cloud adoption, with a particular focus on how GenAI, agentic workflows and intelligent … through standardisation, automation, improved data flows and readiness for AI‐enabled operations, and translate findings into the design of future‐state operating models, service management processes and organisational structures that are standardisation‐led, automation‐first and AI‐ready. Future State Design & Operating Model Definition: Design target cloud operating models ...

IT Support Engineer - 3rd Line

Hiring Organisation
Netteam tX Ltd
Location
Manchester, North West, United Kingdom
Employment Type
Permanent
Salary
£50,000
escalation point for complex incidents, recurring problems and engineering-level change. You will own advanced investigation through to resolution, drive Root Cause Analysis and Problem Management, and work with vendors and internal teams to deliver permanent fixes, reduce repeat demand and improve service resilience. You will maintain clear … telephony, prioritised to SLA and customer impact Provide clear and timely customer updates throughout complex incidents, Problems and engineering activity Own assigned Problem records, identify recurring incidents, complete Root Cause Analysis, and maintain Known Errors and workarounds Validate escalations, provide technical guidance, and return those that do not meet ...

Service Delivery Manager

Hiring Organisation
Nomios
Location
Basingstoke, Hampshire, UK
Employment Type
Full-time
service-related enquiries. The Service Delivery Manager will help deliver customer value by maximising the IT service quality. You will build knowledge management of our customer's environment by understanding and aligning with three main service streams: people, processes, and technology. The successful candidate will be result-orientated, focused … possess exceptional attention to detail. You should enjoy problem-solving and getting things done in a fast-paced environment, and should be motivated to succeed, with a desire to continuously learn and improve. Responsibilities Whilst no two days will be the same, the types of activities that the Service ...

Technical Authority

Hiring Organisation
MASS Consultants
Location
Corsham, Wiltshire, South West, United Kingdom
Employment Type
Permanent
Salary
£80,000
improve resilience, scalability, or efficiency. Track emerging technologies and recommend their adoption where value is clear. Change & Governance Lead technical elements of change management, ensuring risks are properly assessed and mitigated. Provide expert input to CAB and drive continuous improvement in governance processes. Capture technical lessons learned and embed … them into engineering practice. Problem Management & Resilience Lead root-cause analysis and coordinate long-term remediation for recurring incidents. Work with engineering teams to reduce technical debt and increase platform stability. Drive proactive risk reduction across infrastructure and networks. Stakeholder, Supplier & Customer Engagement Act as the technical authority ...

Print Management Engineer

Hiring Organisation
MSP Talent Bridge Ltd
Location
Reading, Berkshire, United Kingdom
Employment Type
Full-Time
Salary
£33,000 - £35,000 per annum
function. This is a remote/hybrid technical support role focused on managing print environments for a range of customers — resolving incidents, administering print management platforms, and ensuring printing services run smoothly through proactive remote monitoring and troubleshooting. You'll work closely with customers, service desk teams and third … high and disruption to a minimum. What You'll Be Doing Remote Print Support Provide remote support for printers, multifunction devices (MFDs) and print management systems Diagnose and resolve printing, scanning, copying and connectivity issues Monitor print environments and proactively address issues before they impact users Manage and support ...

Application Support Analyst

Location
Bradford, England, United Kingdom
JSON, XML, payloads and log structures Support and delivery tooling – Runbooks, data‐flow mappings and recovery procedures Support and delivery tooling – Defect, release and problem management Essential experience Experience in Application Support, Integration Support or a similar technical role. Hands‐on experience supporting business‐critical production systems … integrated enterprise platforms. Strong troubleshooting skills across applications, APIs, integrations, file transfers, logs and monitoring tools. Experience of incident and problem management, root cause analysis and defect tracking in production environments. Ability to prioritise concurrent issues by business impact and communicate clear diagnostics to engineering teams. Useful ...

Senior ITSM Process & Governance Analyst

Location
Dunmurry, Northern Ireland, United Kingdom
Shearman is seeking a Senior ITSM Process Support Analyst to administer ITSM processes across the Global Service Desk. The role involves coordinating Change & Release Management, Problem Management and Service Design activities, preparing management information, and escalating risks to process owners or governance forums. The successful candidate ...

Senior Platform Engineer

Location
Kettering, England, United Kingdom
Exchange Online. Provide efficient and effective escalation path for junior members of the team Own day-to-relationships with extended technical partners in the management of IT services, working in a collaborative and transparent culture Promote technical implementations into production Ensuring IT assets are maintained in the IT Asset … updated, secure, and plans in place prior to End of Life To lead any root cause analysis and resolution of problems. Own and drive problem management for the relevant technical areas. Knowledge and Expertise Essential: Experience in designing, implementing and supporting Azure Cloud services (IaaS, PaaS, Networks) Experience ...

Lead Product Manager

Location
Bracknell, England, United Kingdom
business, IT, data, security, architecture, quality, and compliance functions. Own product delivery governance, prioritisation cadence, backlog health, release planning, capacity allocation, and management of scope, time, cost, and value trade-offs. Drive operational excellence through reliable product operations, incident and problem management, root-cause analysis, and post … cross-functional teams and delivering measurable outcomes through Agile and Lean product practices. Strong product strategy skills, including vision setting, roadmap development, prioritisation, portfolio management, and value measurement through OKRs/KPIs. Experience managing complex digital product portfolios in enterprise environments. Understanding of digital R&D workflows in laboratory ...

Enterprise Integration Product Manager

Hiring Organisation
Experis
Location
Sheffield, South Yorkshire, United Kingdom
Employment Type
Contract
Contract Rate
£480 - £530/day
Role Description: Principal responsibilities: Define and maintain a multi-quarter product roadmap covering currency, resilience, security improvements, automation and consumer experience. Own platform lifecycle management across the estate: standard versions, upgrade plans, deprecation/exit strategy, and technical debt prioritization.Shape and priorities demand using transparent criteria (risk reduction, service … KPIs (incident trends, MTTR, change failure rate, platform performance indicators). Drive operational excellence through: Major incident review leadership and root cause elimination Problem management, reliability engineering initiatives and runbook maturity Ensure DR/HA expectations are defined, tested, and evidenced. Success measures (example KPIs): Reduction in high ...

HPC Senior Hardware Engineer

Hiring Organisation
CGG
Location
Haywards Heath, West Sussex, UK
Employment Type
Full-time
Data Centre. This role is responsible for ensuring the reliability, performance, and scalability of mission-critical HPC platforms, while leading hardware lifecycle management and infrastructure projects. You will work across compute, GPU, storage, networking, and data centre infrastructure, providing technical leadership and supporting the delivery of world-class … Operations, Platform Engineering, Facilities, and technology vendors to deliver reliable, secure, and high-performing infrastructure. Key ResponsibilitiesHPC Infrastructure & HardwareLead the installation, maintenance, and lifecycle management of HPC server infrastructure, including CPU and GPU platforms. Install, upgrade, and replace server components including processors, GPUs, storage, networking, and system hardware. Perform ...

Principal AI Platform Engineer

Location
Cambridge, England, United Kingdom
platforms that make AI services reliable, secure, observable and supportable at Arm scale. You will work across Kubernetes, cloud, identity, secrets, networking, telemetry, incident management and automation to provide the production foundation for Arm's AI platform. Participate in production support and our paid on-call rota for high … platform services, including MCP server infrastructure, model gateway services and supporting control-plane components. Design runtime patterns for isolation, scalability, secure execution, capacity management and cost-aware operation. Automate provisioning, configuration, upgrades and lifecycle management using infrastructure-as-code and GitOps patterns. Reliability, observability and support: Define ...

Senior IT Platform Engineer

Hiring Organisation
E.surv Limited
Location
Kettering, Northamptonshire, East Midlands, United Kingdom
Employment Type
Permanent
Exchange Online. Provide efficient and effective escalation path for junior members of the team Own day-to-relationships with extended technical partners in the management of IT services, working in a collaborative and transparent culture Promote technical implementations into production Ensuring IT assets are maintained in the IT Asset … updated, secure, and plans in place prior to End of Life To lead any root cause analysis and resolution of problems. Own and drive problem management for the relevant technical areas. Knowledge and Expertise Essential: Experience in designing, implementing and supporting Azure Cloud services (IaaS, PaaS, Networks) Experience ...

Internal Audit, Asset & Wealth Management Technology Audit, Vice President, Birmingham

Location
Birmingham, England, United Kingdom
that Goldman Sachs maintains effective controls by assessing the reliability of financial reports, monitoring the firm’s compliance with laws and regulations, and advising management on developing smart control solutions. Our group has unique insight on the financial industry and its products and operations. We’re looking for detail … defense, Internal Audit’s mission is to independently assess the firm’s internal control structure, including the firm’s governance processes and controls, risk management and capital and anti-financial crime frameworks, raise awareness of control risk and monitor the implementation of management’s control measures. In doing ...

Senior Service Delivery Manager

Location
Greater London, England, United Kingdom
difficult client conversations and clearly articulate remediation plans in a way that is reassuring. Solution focussed and outcome driven. Service operations Implement incident and problem management processes using best practice such as ITIL or Agile Service Management. Coordinate and manage the resolution of major incidents and subsequent root … cause analyses. Champion governance, risk and engagement processes and be responsible for others following the processes. Manage change using robust change management processes that prevent scope creep. Ability to manage workflows with popular ticket management tools such as ServiceNow, Jira Service Desk, Zendesk etc. Create, run and report ...

Lead Linux System Administrator

Location
Greater London, England, United Kingdom
operational efficiency.The team is responsible for the evaluation, certification, integration, and ongoing support of Linux-based infrastructure, including hardware platforms, Red Hat Linux, configuration management, system services such as DNS, DHCP, and NTP, and a variety of internally developed tools.In the Technology division, we leverage innovation to build … that power our Firm, enabling our clients and colleagues to redefine markets and shape the future of our communities.This is a Lead Infrastructure Production Management & Reliability Engineering position at Director level (equivalent to AVP) which is part of the job family responsible for maintaining the stability and reliability ...

Systems Administrator (Associate)

Location
Basingstoke, England, United Kingdom
maintenance activities. You engage in automation activities, perform root cause analysis (RCA), and remediation. Knowledge of production support process including incident/change/problem management, call triaging, and critical issue resolution procedures. ESSENTIAL JOB FUNCTIONS AND RESPONSIBILITIES: Infrastructure Operations and Production Support of container technologies and orchestration … platforms (Docker UCP, Kubernetes, OpenShift, etc.) OpenShift/Docker/Kubernetes deployment, configuration, scaling and management of containerized applications. Contribute to automation and self‐service patterns (templates, scripts, IaC modules) that enable rapid, standardized environment creation for experimentation and production GenAI workloads Openshift Enterprise Support Working with tools surrounding ...

Software Engineer Lead - Site Reliability

Location
City of Edinburgh, Scotland, United Kingdom
starting approach with evidence of identifying issues, shaping solutions and delivering improvements without heavy direction. Strong practical Azure experience, including Azure DevOps, Azure API Management or comparable API platform capabilities. Hands‐on expertise in Terraform, GitHub, GitHub Actions, CI/CD pipeline design, deployment automation and release support. Practical … knowledge of incident management, problem management, root‐cause analysis and operational readiness. Good awareness of Site Reliability Engineering principles, with the ability to apply them pragmatically to DevOps and platform improvement work. Experience improving observability through monitoring, logging, tracing, alerting, service‐health dashboards and actionable telemetry. Experience ...

Software Engineer Lead - Site Reliability

Location
City of Westminster, England, United Kingdom
starting approach with evidence of identifying issues, shaping solutions and delivering improvements without heavy direction. Strong practical Azure experience, including Azure DevOps, Azure API Management or comparable API platform capabilities. Hands-on expertise in Terraform, GitHub, GitHub Actions, CI/CD pipeline design, deployment automation and release support. Practical … knowledge of incident management, problem management, root-cause analysis and operational readiness. Good awareness of Site Reliability Engineering principles, with the ability to apply them pragmatically to DevOps and platform improvement work. Experience improving observability through monitoring, logging, tracing, alerting, service-health dashboards and actionable telemetry. Experience ...

Software Engineer Lead - Site Reliability

Location
Telford, England, United Kingdom
starting approach with evidence of identifying issues, shaping solutions and delivering improvements without heavy direction. Strong practical Azure experience, including Azure DevOps, Azure API Management or comparable API platform capabilities. Hands-on expertise in Terraform, GitHub, GitHub Actions, CI/CD pipeline design, deployment automation and release support. Practical … knowledge of incident management, problem management, root-cause analysis and operational readiness. Good awareness of Site Reliability Engineering principles, with the ability to apply them pragmatically to DevOps and platform improvement work. Experience improving observability through monitoring, logging, tracing, alerting, service-health dashboards and actionable telemetry. Experience ...

Service Assurance Manager

Location
West of England, England, United Kingdom
Service Assurance Manager leads the Service Assurance team and is accountable for the effective governance of IT service management across HL's dual technology estate, spanning heritage systems, cloud services, and integrated third party platforms. The role ensures consistent execution of service management processes, strong major incident leadership … risk appetite. What you'll be doing Lead and manage the Service Assurance team, setting clear priorities, standards, and performance expectations Manage IT service management processes (Incident & Problem, Change & Release, Knowledge, and SACM) across IT Service Operations Participate in on‐call and out‐of‐hours cover, providing business ...

Service Assurance Manager

Location
Bristol, England, United Kingdom
About the role The Service Assurance Manager leads the Service Assurance team and is accountable for the effective governance of IT service management across HL’s dual technology estate, spanning heritage systems, cloud services, and integrated third party platforms. The role ensures consistent execution of service management processes … risk appetite. What you’ll be doing Lead and manage the Service Assurance team, setting clear priorities, standards, and performance expectations Manage IT service management processes (Incident & Problem, Change & Release, Knowledge, and SACM) across IT Service Operations Participate in on‐call and out‐of‐hours cover, providing business ...