176 to 200 of 432 Problem Management Jobs

Systems Operations Team Lead

Location
Plymouth, England, United Kingdom
Systems Operations Lead will be responsible for the overall management and performance of the Systems Operations team.Ensures that the Systems Manager is informed with regards to the operational capacity and functionality of the server estate. This includes improvements, risks, changes, and problems.Manages multiple infrastructure projects, delivery of new initiatives … range of specialist and complex D&I issues, following the relevant technical strategies, policies, standards, and practices.Responsible for ensuring a structured approach for problem, incident, and change management processes. Primarily remote working with some on site duties as and when required. The role is part ...

Service Support Analyst

Location
Greater London, England, United Kingdom
active member of Made Tech’s communities of practice, sharing knowledge and learning from others All team members share responsibility for maintaining our Service Management System. This includes actively contributing to its upkeep, ensuring documentation remains accurate and current, and staying aware of ongoing service activities, improvements, and compliance … digital and/or data services in the production environment Strong customer service background, working in 2nd line support A logical, analytical approach to problem solving Knowledge of network and firewall infrastructure An understanding of the UK public sector Familiar with solutions that use scripting (Python, SQL etc. ...

z/OS Mainframe Systems Programmer

Location
Greater London, England, United Kingdom
system security controls and ensure compliance with organizational standards. Create and maintain technical documentation, procedures, and operational runbooks. Participate in root cause analysis and problem management activities. Provide technical guidance and mentoring to junior team members. Support production implementations and planned maintenance activities, including off-hours and weekend … maintaining multiple IBM and CA/Broadcom mainframe products. Experience working in large enterprise environments, preferably within Banking or Financial Services. Strong troubleshooting and problem-solving capabilities. Ability to independently manage complex technical assignments. Excellent verbal and written communication skills. Experience supporting mission-critical production environments with stringent ...

Service Delivery and Applications Support

Hiring Organisation
Noir
Location
London, South East England, United Kingdom
Employment Type
Full-Time
Salary
£70,000 - £75,000 per annum
Service Delivery and Support Manager - Leading Investment Bank - London - 12-months FTC (Tech stack: IT Service Management, ITIL, Major Incident Management, ServiceNow, Ivanti, Jira, Microsoft Intune, SCCM, Active Directory, Group Policy, Windows 11, Microsoft Teams, Microsoft 365, Endpoint Management) Are you an experienced Service Delivery Manager … financial services environment? This is an opportunity to take ownership of service delivery across a sophisticated investment bank, combining people leadership, ITSM, major incident management and continuous improvement. Our client is an established investment bank with a sophisticated technology environment supporting critical business operations. As part of the Technology ...

Service Reporting Manager

Location
Greater London, England, United Kingdom
insights that enhance service delivery, efficiency, and customer experience. This role analyses ticket trends, performance metrics, and operational behaviours to drive process improvement, proactive problem management, and informed decision-making across the Service Desk and wider IT organisation. Responsibilities Data Analytics Design and deliver advanced analytics to identify … opportunities for service improvement Define and implement best practices for data collection, storage and analysis. Develop predictive models and insights to support proactive problem management. Create dashboards and visuals that translate complex data into actionable insights for technical and non-technical audiences. Stakeholder Collaboration Partner with service desk leadership ...

IT Service Delivery Manager (FTC)

Location
Greater London, England, United Kingdom
year fixed term contract Location: London or Amsterdam (Hybrid) Salary: 55,000 – 70,000 GBP + 10% Bonus Job responsibilities: 1. Service Desk & Incident Management L1 triage & ownership: Act as regional L1 triage layer — validate severity/impact, capture data and route incidents to the correct resolver group, keeping … knowledge and control within future autonomy. Major incident management: Coordinate major incidents end-to-end (e.g. authentication/M365 outages, store-down events), driving service vendor and other resolvers to resolution and clear user communication. Problem management: Turn recurring incidents into root-cause fixes and reduce avoidable ...

Expertise Centre Manager

Location
Windsor, England, United Kingdom
amazing team. Mission — Why the Job Exists Lead and develop the UK Expertise Centre supporting BoxTop, our business‐critical supply chain and freight management software. Your mission is to build a high‐performing, technically capable and resilient support operation that delivers consistently strong service, resolves complex customer issues effectively … KPIs, with customers kept informed through to resolution. Build and develop a high-performing UK Expertise Centre team with clear expectations, effective workload management and strong accountability for individual, team and customer outcomes. Increase the technical capability and self‐sufficiency of the team by identifying knowledge gaps and implementing ...

AWS Iaas Sys Admin

Location
Birmingham, England, United Kingdom
automation solutions using a broad range of DevOps toolsets. Expertise in Infrastructure as Code tools, such as Terraform is preferred. Expertise with code repository management, code merge and quality checks, continuous integration, and automated deployment and management using Ansible, Jenkins, and Git. Knowledge of Docker, Kubernetes, Puppet, Chef … Maven, Ant, Ivy, and UrbanCode would be advantageous. Experience in DevSecOps, including secret management, tools integration to harden the baseline, and privilege management. Associate‐level cloud certification in Azure or AWS. Hands‐on experience in Azure and AWS cloud technologies, including compute, networking, storage, and security services. Expertise ...

SC Cleared Linux Engineer

Hiring Organisation
Experis
Location
Telford, Shropshire, United Kingdom
Employment Type
Contract
Contract Rate
£450 - £500/day
both SUSE Linux Enterprise Server (SLES) and Red Hat Enterprise Linux (RHEL) Deliver operational support across live, development, and test environments. Perform incident resolution, problem management, and root cause analysis. Implement, test, and deploy infrastructure changes in accordance with change management processes. Support infrastructure migration and transformation … infrastructures. Strong understanding of Linux server performance tuning, patching, and system maintenance. Experience working within managed service and operational support environments. Strong analytical and problem-solving skills. Excellent communication and stakeholder management abilities. Technical Knowledge Linux & UNIX SUSE Linux Enterprise Server (SLES) Red Hat Enterprise Linux (RHEL) Linux ...

Endpoint Support Team Lead

Location
Cardiff, Wales, United Kingdom
support function within the University managed service across 5 campuses, providing technical direction, workload coordination and hands‐on escalation support across endpoint management, application deployment, remediation and service improvement activities. Working as part of the Technical Services team, the Endpoint Support Team Lead combines strong technical expertise with people … point of escalation for complex technical issues, service risks and delivery challenges, ensuring incidents and requests are progressed and resolved within agreed SLAs. Endpoint Management: Lead the administration and support of Microsoft Intune-managed Windows devices , including configuration, compliance, security, patching and device lifecycle management. Autopilot & Application Deployment: Oversee ...

Sr. Consultant Employee Central

Location
Greater London, England, United Kingdom
Service Operations (ITIL)Triage, troubleshoot, and resolve incidents and service requests (P1P4) across EC.Execute change requests for minor enhancements and fixes following change control.Drive problem management: root cause analysis, preventive and corrective actions.Maintain a clear knowledge base, runbooks, and support SOPs.Configuration and Data StewardshipMaintain and adjust EC configuration … Foundation Objects, MDF objects, Business Rules, Workflows, Event Reasons, Position Management, Picklists, and Role-Based Permissions.Perform data loads and mass updates using standard import templates; handle data corrections with audit traceability.Ensure data quality through validations, spot checks, and reconciliation.Release and TestingLead release readiness and impact assessments for SuccessFactors releases.Own ...

AWS IaaS Sys Admin

Location
Birmingham, England, United Kingdom
automation solutions using a broad range of DevOps toolsets. Expertise in Infrastructure as Code tools, such as Terraform is preferred. Expertise with code repository management, code merge and quality checks, continuous integration, and automated deployment and management using Ansible, Jenkins, and Git. Knowledge of Docker, Kubernetes, Puppet, Chef … Maven, Ant, Ivy, and UrbanCode would be advantageous. Experience in DevSecOps, including secret management, tools integration to harden the baseline, and privilege management. Associate-level cloud certification in Azure or AWS. Hands-on experience in Azure and AWS cloud technologies, including compute, networking, storage, and security services. Expertise ...

Problem and Change Manager

Hiring Organisation
Hackajob Ltd
Location
York, North Yorkshire, Yorkshire, United Kingdom
Employment Type
Permanent
hackajob is partnering directly with OneAdvanced to hire for this role. Own and manage the Incident, Problem and Change processes for our customers, ensuring all necessary communication is carried out, coordinating and aligning resources to allow timely resolution of incidents, driving out recurring problems at root cause, and enabling … reporting and notifications to key stakeholders Own the major incident lifecycle end to end, from declaration through to satisfactory resolution as confirmed by Service Management Proactively manage major incidents to satisfactory resolution in a timely manner, ensuring minimal business impact, initiating escalation procedures as appropriate, chairing post-incident reviews ...

IT Service Operations Manager 24 / 7

Hiring Organisation
ASDA
Location
Leeds, West Yorkshire, United Kingdom
Employment Type
Full-Time
Salary
Competitive salary
steps. Maintain real-time situational awareness across all critical services, particularly during overnight batch, replenishment, and early-morning trading readiness windows. Incident & Major Incident Management Act as the on-shift Major Incident Manager, working alongside our partner to drive rapid restoration of service. Make judgement calls on escalation — waking … leaders, engaging suppliers, invoking service continuity plans. Own customer, colleague and business impact assessment during incidents, ensuring the right level of response. Monitoring & Event Management Ensure monitoring tools and dashboards are actively used and events acted upon promptly. Identify noise, false positives and monitoring gaps and feed them into ...

Infrastructure Engineer

Location
Potters Bar, England, United Kingdom
resolved in a timely manner, while also taking ownership of the company's Microsoft Server, Desktop, Microsoft 365, Azure cloud, networking, and endpoint management environments. The role works closely with the Developers, Azure Architects, and Network Architects to support and evolve the organisation's infrastructure, and provides exposure … timely manner Diagnosing and resolving technical issues remotely and on site Undertaking small- to medium-sized IT and infrastructure projects as instructed by management Providing desktop and server support Supporting and maintaining Microsoft Server/Desktop operating systems, Microsoft 365, and Azure environments Setting up and configuring new laptops ...

Head of Technology Operations & Service Delivery

Location
City of Westminster, England, United Kingdom
focused on customer, intermediary and colleague outcomes. Forward thinking and driven to embed process automation through AI adoption. Key areas of responsibility: IT service management Own and continuously improve IT service management policies, standards, controls and tooling. Establish effective incident, problem, request, change, release, configuration and knowledge … arrangements and service reporting for all material technology services. Lead major incidents, stakeholder communications, post-incident reviews and remediation tracking. Service ownership & business service management Maintain a comprehensive service catalogue with clear owners, support models and criticality classifications. Map technology services and third-party dependencies supporting Important Business Services ...

Major Incident Manager | Berkshire | Contract £285 per day

Location
Reading, England, United Kingdom
stakeholders, you’ll ensure incidents are managed professionally, communications remain clear and everyone stays focused on achieving the best possible outcome. Alongside major incident management, you’ll contribute to the wider Service Management function by supporting Change, Problem and Continual Service Improvement activities. Your experience and judgement … comfortable coordinating multiple technical teams during periods of high pressure. We’re particularly interested in people who can demonstrate: Previous Major Incident Management experience within an enterprise IT environment. A strong understanding of ITIL Service Management principles. Excellent communication skills with the ability to engage confidently at both ...

Platform Engineer - Enterprise Platforms

Location
Greater London, England, United Kingdom
applications across public cloud and hybrid environments. This is a broad Platform Engineering role covering AWS, Azure, GCP, Windows, Linux, infrastructure automation, Terraform, configuration management and enterprise platform operations. You’ll help design, build, automate and operate the organisation’s enterprise platform estate. Responsibilities include: Provisioning, scaling and optimising … Linux environments Supporting SQL Server workloads and enterprise application hosting Managing AWS services including S3, FSx, EBS and WorkSpaces/AppStream Developing configuration management solutions using tools such as Puppet Maintaining monitoring and observability capabilities Supporting enterprise backup and recovery platforms Building and maintaining CI/CD pipelines using ...

4th Line Cloud Support Engineer

Location
Basingstoke, England, United Kingdom
across automation, resilience design and technical leadership. The Role You will be involved in: Responding to complex escalations from 3rd Line engineers Assisting in Problem Management investigation and root cause analysis Extensive work with VMware products and associated monitoring Creating opportunities to automate manual tasks to drive efficiency … reduce downtime Identifying single points of failure and designing resilience into the service Managing obsolescence across applications and operating systems Collaborative working with the Management Team Daily contact with customers and stakeholders through the IT Service Management toolset Managing high profile senior escalations and requests Key Experience Required ...

Senior Major Incident Manager

Location
Belfast City District, Northern Ireland, United Kingdom
Senior Major Incident Manager to join and lead the team. You'll be responsible for maintaining a robust, efficient and well governed 24x7 incident management capability, ensuring effective coordination when incidents occur and driving continuous improvement across the organisation. You will act as an escalation point for complex … Everyone is Valued. These guide how we work, collaborate, and deliver exceptional results. What You'll Be Doing Leading and supporting the Major Incident Management team. Ensuring all incidents progress through to resolution within SLAs. Managing escalations when required. Delivering executive summaries during P1 incidents. Producing weekly and monthly ...

Systems Administrator

Hiring Organisation
Lucid Support Services Ltd
Location
Guildford, Surrey, United Kingdom
Employment Type
Contract
Contract Rate
GBP Annual
Global IT to deliver global projects in the region, and request new services and solutions as needed. Facilitate operations of the IT environment including problem management, incident management, configuration and change management. Contribute to global IT initiatives with professional knowledge and local insights. Ensure proper operation ...

Director of Platform Operations

Location
United Kingdom
progress without overwhelming staff. And where everyone working in a school is reminded why they got into education every day. Our MIS and school management tools are already making a difference in over 12,000 schools and trusts. Giving time and power back to staff, turning data into clear … product teams, R&D leadership, and the executive, with credible attribution of loss to cause. Engineer out recurrence. Drive systemic reliability improvement through problem management, tracking repeat causes, single points of failure, and architectural weak points to closure with named owners and dates. Security engineering and vulnerability management ...

Senior Platform Reliability Engineer

Hiring Organisation
Ricoh
Location
London, United Kingdom
Employment Type
Permanent
predictable service outcomes. What you will be doing Deliver standards for availability, latency, performance, capacity, and scalability. Take part in root-cause analysis and problem management for major incidents. Champion blameless post-mortem culture and ensure actions are tracked and closed. Drive infrastructure-as-code and automation across … Azure (IaaS, PaaS, networking, identity, storage) and on-prem data centre operations. Hands-on skills with infrastructure-as-code (Terraform, ARM/Bicep), configuration management (Ansible, PowerShell DSC), CI/CD pipelines (Azure DevOps, GitHub Actions). Experience with monitoring/observability tools, alert design, and dashboarding. Commercial awareness ...

Platform Engineer

Location
Crawley, England, United Kingdom
secure for internal users and engineering teams. This is a varied role with a strong operational focus, covering infrastructure support, systems administration, troubleshooting, lifecycle management and service improvement across a mixed technology estate. You will work closely with platform engineers, developers, security specialists and other stakeholders to support reliable … service availability, performance and capacity through monitoring, troubleshooting and proactive maintenance. Support containerised services and the underlying infrastructure they rely on. Maintenance and lifecycle management Carry out routine maintenance activities including patching, upgrades and health checks. Support the ongoing management of hardware and software lifecycles to keep services ...

Site Reliability Engineering Professional

Location
Ipswich, England, United Kingdom
modern Site Reliability Engineering (SRE) and Platform Operations model. As a Site Reliability Engineering Professional, you will play a key role in the operational management of BT International's global core platforms, ensuring they are reliable, secure, scalable, and deliver an exceptional customer experience. Working within our 2nd Line … escalation point for complex operational issues. Deliver against operational KPIs, SLAs, OLAs, and customer service targets. Drive continuous improvement through root cause analysis, problem management, automation, and defect reduction. Work closely with Engineering and Product teams to improve platform reliability, resilience, and operational readiness. Build and maintain high ...