376 to 400 of 644 Incident Management Jobs in the UK excluding London

Platform Engineer III - (Pipelines & Developer Experience)

Location
Leeds, England, United Kingdom
release pipelines. Experience with scripting and programming (.NET preferred; familiarity with Go, Python, PowerShell beneficial). Knowledge of observability tooling, chaos testing, and incident management. Strong analytical and problem‐solving abilities, with the capability to closely collaborate with engineering teams. Highly outcome‐oriented, pragmatic, and capable of balancing quality ...

Senior Infrastructure Engineer (Linux & Cloud Automation)

Hiring Organisation
Adria Solutions
Location
Manchester, North West, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£60,000
automation Manage Azure Key Vault, secrets, and certificate lifecycle processes Implement infrastructure changes and contribute to CAB processes Resolve technical escalations and support major incident response Participate in a shared on-call rota What We're Looking For Strong experience in infrastructure engineering, cloud operations, or managed services Strong … Linux administration skills across RHEL, CentOS, Ubuntu, and/or Oracle Linux Hands-on Ansible experience for patching, automation, and configuration management Good Microsoft Azure administration skills, including VMs, networking, backup, and governance Experience with Infrastructure-as-Code using Bicep and/or ARM templates Azure DevOps ...

Finance Systems Accountant

Location
Reading, England, United Kingdom
transformation. You will play a key role in shaping and supporting group-wide finance system initiatives, with a particular focus on Oracle Enterprise Performance Management (EPM) – Planning/PBCS. Working across the full project lifecycle, you will contribute to the design, development, and delivery of high-impact solutions, while … transition into BAU support. Key Responsibilities Deliver functional enhancements and change requests within Oracle Planning/PBCS, ensuring solutions align with business requirements Provide incident management and system support, resolving both functional and technical issues Support data integration, validation, and reconciliation processes across finance systems Maintain robust controls ...

Monitoring & Observability Engineer

Location
Reading, England, United Kingdom
operational responses that improve reliability and reduce downtime. What You'll Be Working On Monitor live systems, triage operational alerts and provide first-line incident response to maximise system availability. Develop and maintain dashboards that deliver clear visibility into system health, trends and operational performance. Analyse recurring incidents … tooling and operational processes. Define monitoring thresholds, alerting strategies and operational metrics that support reliable live quantum computing services. Produce documentation covering monitoring processes, incident response, escalation procedures and operational handovers. Contribute to automation and continuous improvement initiatives that reduce manual intervention and improve service reliability. What ...

Senior Engineering Manager

Location
Leeds, England, United Kingdom
Headquartered in Leeds and part of the global Simpro Group network, our platform connects jobs, people, and performance across plumbing, HVAC, fire & security, facilities management, electrical contracting, and more.Trusted by over 2,200 high-growth businesses, we've processed billions in invoiced work and scheduled tens of millions … Expertise in bridging legacy monoliths (.NET 4.6.2, WebForms, jQuery) with modern architecture (C# .NET 8, Micro-frontends, Microservices on AWS EKS).Full-Stack Platform Management: Ability to lead diverse development streams across Web (React, WebForms), Mobile (iOS/Swift, Android/Kotlin, Kotlin Multiplatform), and Backend systems.Cloud & DevOps Leadership ...

Senior Ofcom Regulatory Manager

Location
Welwyn Garden City, England, United Kingdom
making a contribution to the success of their operations. (As appointed) Membership and trusted adviser for relevant executive committees of Telefonica and other management fora (e.g. security, incident management) You will need A regulatory professional with proven experience, ideally gained in a regulated company, regulatory authority ...

The Core Engineering - Software Engineer - Associate - Birmingham

Location
Birmingham, England, United Kingdom
engineers who thrive on solving operational problems and improving efficiency and would like to apply their skills to Compliance Engineering SRE. Job Responsibilities: Proactive management of our production services by measuring and monitoring availability, capacity and overall system health. Shaping software before go-live through activities such as system … design consulting, capacity planning and launch reviews. Scaling and evolving systems by pushing for changes that improve capacity and reliability. Practicing sustainable incident management in a blameless postmortem culture. Identifying and building improvements to system behavior, control and monitoring tools. Defining and maintaining Service Level Indicators (SLIs ...

Pega Developer/Pega DevOPs Architect

Hiring Organisation
iXceed Solutions
Location
Telford, Shropshire, United Kingdom
Employment Type
Contract
Contract Rate
GBP Annual
cutover activities for a large-scale Pega transformation programme. The ideal candidate will possess deep expertise in Pega DevOps, CI/CD automation, environment management, release governance, and production deployment assurance. The role will be responsible for establishing deployment guardrails, ensuring operational readiness, defining rollback strategies, and supporting successful … operations. Collaborate with delivery, infrastructure, operations, and business teams to ensure deployment readiness and risk mitigation. Establish best practices for CI/CD pipeline management, automated deployments, and release governance. Required Skills & Experienc Strong experience designing and managing CI/CD pipelines in enterprise environments. Expertise in release strategy ...

Operations Support Engineer

Location
Metropolitan Borough of Solihull, England, United Kingdom
issue but ensure it doesn’t reoccur. Responsibilities The Operations Support Engineer have full responsibility for the service delivered to our clients, including. Relationship Management Co-ordinate with key stakeholders. Executing against the stakeholder management and communication plan, to provide visibility of service performance and incident management. …/or clients understand it. Experience of risk control management. Understanding software development lifecycles. Comfortable with multi-tasking and adapting to changing priorities. Incident Management experience – Managing incidents including business expectations and communication A self-motivated achiever who gains satisfaction from providing excellent customer service Beneficial Financial services ...

Engineering Manager

Hiring Organisation
Randstad Technologies Recruitment
Location
Manchester, United Kingdom
Employment Type
Contract
Contract Rate
£80 - £90/hour
HR. System Ownership & Architecture: Maintain end-to-end service health, lead deployment/operations, reduce system risks via clear documentation, and guide technical architecture. Incident Management & Quality: Oversee live production incident response, conduct postmortems, maintain high code standards, and participate in on-call rotations as needed. Continuous ...

Engineering Manager

Hiring Organisation
Randstad Technologies
Location
Manchester, Lancashire, United Kingdom
Employment Type
Full-Time
Salary
£80.00 - £90.00 per hour
HR. System Ownership & Architecture: Maintain end-to-end service health, lead deployment/operations, reduce system risks via clear documentation, and guide technical architecture. Incident Management & Quality: Oversee live production incident response, conduct postmortems, maintain high code standards, and participate in on-call rotations as needed. Continuous ...

Sr Engineering Manager

Hiring Organisation
Randstad Technologies Recruitment
Location
Manchester, United Kingdom
Employment Type
Contract
Contract Rate
£90 - £95/hour
HR. System Ownership & Architecture: Maintain end-to-end service health, lead deployment/operations, reduce system risks via clear documentation, and guide technical architecture. Incident Management & Quality: Oversee live production incident response, conduct postmortems, maintain high code standards, and participate in on-call rotations as needed. Continuous ...

Sr Engineering Manager

Hiring Organisation
Randstad Technologies
Location
Manchester, Lancashire, United Kingdom
Employment Type
Full-Time
Salary
£90.00 - £95.00 per hour
HR. System Ownership & Architecture: Maintain end-to-end service health, lead deployment/operations, reduce system risks via clear documentation, and guide technical architecture. Incident Management & Quality: Oversee live production incident response, conduct postmortems, maintain high code standards, and participate in on-call rotations as needed. Continuous ...

Product Engineer - M365 Unified Collaboration & Gen AI

Hiring Organisation
Barclays
Location
Knutsford, Cheshire, UK
Employment Type
Full-time
Ensure the reliability, availability, and scalability of the systems, platforms, and technology through the application of software engineering techniques, automation, and best practices in incident response. AccountabilitiesBuild Engineering: Development, delivery, and maintenance of high-quality infrastructure solutions to fulfil business requirements ensuring measurable reliability, performance, availability, and ease … use. Including the identification of the appropriate technologies and solutions to meet business, optimisation, and resourcing requirements. Incident Management: Monitoring of IT infrastructure and system performance to measure, identify, address, and resolve any potential issues, vulnerabilities, or outages. Use of data to drive down mean time to resolution. ...

Service Desk Analyst x4 - Inverness

Hiring Organisation
Adecco
Location
Inverness, Highlands, United Kingdom
Employment Type
Permanent
Salary
£31000 - £32000/annum + shift allowance
days off. We are seeking an experienced Service Desk Analyst to join a fast-paced IT Operations team, providing operational oversight and incident coordination across critical services. You will be responsible for monitoring operational alerts, coordinating technical teams and suppliers, managing escalations, and ensuring service disruptions are effectively prioritised … resolved. You will play a key role in maintaining service availability while providing clear communication to stakeholders throughout the incident life cycle. Key Responsibilities Monitor alerts, events, and incidents, assessing impact and ensuring appropriate action is taken. Coordinate with technical teams, suppliers, and stakeholders to manage incidents and service ...

People Systems Analyst (6 month FTC)

Location
Belfast City District, Northern Ireland, United Kingdom
Benefits or Payroll is desirable. Comfortable analysing and resolving system and process issues in a structured manner. Strong organisational skills, with experience in operating incident management tools. Ability to prioritise and manage multiple requests within agreed service levels. Familiarity with Microsoft Office suite. Experience using ServiceNow and JIRA … work well within a global and sometimes virtual team. Understanding of HR data governance, confidentiality requirements and system administration best practices. Good stakeholder management skills, with the ability to communicate clearly with technical and non-technical audiences. Ability to handle sensitive people data with discretion and in line with ...

Senior Site Reliability Engineer

Location
Nottingham, England, United Kingdom
collaborate with Architecture, Engineering, Security, and Platform teams to ensure reliability is built into systems from day one. While this is not a people‐management you will work closely with global teams and may occasionally be called upon for major incidents or critical issues. This position requires a highly … Design and evolve monitoring and alerting solutions that improve visibility, reduce toil, and strengthen system health. Continuously drive reliability improvements across our environments through incident reduction, performance tuning, and building resilient patterns. Partner with Security teams to ensure our platforms meet compliance, security, and risk‐management expectations. Influence ...

Systems Support Engineer (Linux)

Hiring Organisation
Cubic Corporation
Location
Redhill, Surrey, UK
Employment Type
Full-time
short notice. Essential Job Duties and Responsibilities: Resolve technical queries escalating from field maintenance, projects, internal resolver groups and customers. Proactively manage the 'Incident Management' queue systems ensuring that all tickets are picked up within a timely manner and resolved within KPI.As part of troubleshooting, you will … required to escalate identified bugs and issues to engineering using JIRA.Manage the monthly security and application patching to 'Front End Devices' through Patch Management tooling and processes for Windows and Unix environments whilst looking at continuous improvements to the process (Preferably using Ivanti EPM).Support the Linux patching process ...

The Core Engineering - Site Reliability Engineering - Associate - Birmingham

Location
Birmingham, England, United Kingdom
planning, operational readiness reviews, and launch reviews. Scale and evolve systems by driving changes that improve reliability, capacity, performance, and operational resilience. Practice sustainable incident management through clear escalation, effective remediation, and a blameless postmortem culture. Identify and implement improvements to system behavior, controls, observability, and monitoring tools. … financial markets, technology, and continuous learning. ABOUT GOLDMAN SACHS The Goldman Sachs Group, Inc. is a leading global investment banking, securities and investment management firm that provides a wide range of financial services to a substantial and diversified client base that includes corporations, financial institutions, governments and individuals. Founded ...

The Core Engineering - Site Reliability Engineering - Associate - Birmingham

Location
Birmingham, England, United Kingdom
planning, operational readiness reviews, and launch reviews. Scale and evolve systems by driving changes that improve reliability, capacity, performance, and operational resilience. Practice sustainable incident management through clear escalation, effective remediation, and a blameless postmortem culture. Identify and implement improvements to system behavior, controls, observability, and monitoring tools. … financial markets, technology, and continuous learning. ABOUT GOLDMAN SACHS The Goldman Sachs Group, Inc. is a leading global investment banking, securities and investment management firm that provides a wide range of financial services to a substantial and diversified client base that includes corporations, financial institutions, governments and individuals. Founded ...

Lead Site Reliability Engineer

Hiring Organisation
London Stock Exchange Group
Location
Nottingham, UK
Employment Type
Full-time
collaborate with Architecture, Engineering, Security, and Platform teams to ensure reliability is built into systems from day one. While this is not a people‐management or shift‐based role, you will work closely with global teams and may occasionally be called upon for major incidents or critical issues. This … Design and evolve monitoring and alerting solutions that improve visibility, reduce toil, and strengthen system health. Continuously drive reliability improvements across our environments through incident reduction, performance tuning, and building resilient patterns. Partner with Security teams to ensure our platforms meet compliance, security, and risk‐management expectations. Lead ...

Senior Platform Owner - Customer Engagement

Location
Skipton, England, United Kingdom
Society. As our Senior Platform Owner , you will lead the evolution of the Dynamics 365 and Power Platform ecosystem, spanning CRM, marketing automation, case management, colleague engagement tools, workflow orchestration, and low‐code applications. This platform is a central enabler of Skipton’s purpose and transformation goals, powering journeys … Platform is the engine room of how Skipton understands, supports and communicates with its members. Powering Dynamics 365, Power Platform solutions, marketing automation, case management and workflow tooling, CEP is the central nervous system that ensures colleagues have the insight, context and capability to deliver human, meaningful interactions ...

Platform Engineer

Location
Leeds, England, United Kingdom
pipelines to enable frequent, automated and reliable software delivery.Improve platform reliability through observability, monitoring, automated testing and disaster recovery practices.Support production systems, participate in incident response and contribute to continuous service improvements.Collaborate with engineers, architects and stakeholders to deliver scalable platform solutions that meet business needs.Contribute to platform standards … Terraform.Knowledge of CI/CD technologies, GitLab CI, GitHub Actions, ArgoCD and modern software delivery practices.Familiarity with observability, security and operational excellence, including monitoring, incident management and platform reliability.Experience working in agile environments with a passion for automation, continuous learning and emerging technologies, including AI-powered engineering tools.What ...

Telecoms Specialist Engineer

Hiring Organisation
Flotek
Location
Bridgend, Mid Glamorgan, Wales, United Kingdom
Employment Type
Permanent
Salary
£35,000
cannot offer sponsorship or relocation assistance, so you must already have the right to work in the UK. What you will be doing: Incident management: Be the first point of contact for telecoms incidents, logging, categorising and prioritising faults in line with ITIL best practice. Service level management … customers and stakeholders, managing expectations during incidents. On-call support: Support telecoms services during scheduled weekend and out-of-hours cover, following major incident escalation paths. What we are looking for: Essential Requirements: You will be a logical problem-solver with a customer-first mindset who thrives ...

Head of Platform Engineering

Location
Manchester, England, United Kingdom
improvement Identify capability gaps and support hiring, development, and succession planning Own resource planning, including capacity, skills mix, and on-call structure Operational Resilience & Incident Ownership Establish and run a two-tier resolver model: Ops on-call for standard support and runbook-driven issues, Product Engineering (Buyer/Seller … minimum 3 engineers per domain, or full team where smaller) so resolution doesn't depend on any one individual Drive a genuinely blameless Post-Incident Review culture, building on the existing Incident Management SOP Security & Risk Embed secure coding practice, SAST/DAST scanning, and OWASP-aligned ...