26 to 50 of 54 Incident Management Jobs in the East of England

Site Reliability Engineer (SRE)

Location
Cambridge, England, United Kingdom
large-scale software systems through a blend of software engineering and systems administration. Key responsibilities involve automating operational tasks,improving observability, andcontributing to incident management, while also collaborating with developmentand technologyteams to build more reliable and scalable applications. Join Altium as a Senior Site Reliability Engineer to ensure … Altium Cloud Platforms. Key Responsibilities: Understanding how an Altium Cloud Platform works Pioneer improvements in observability, including logging, monitoring, and application performance management (APM), ensuring system reliability and proactive issue detection. Develop and implement reliability frameworks and patterns that standardize and elevate the resilience of our SaaS products across ...

Information Governance Assurance Officer

Location
Norwich, England, United Kingdom
standards, and guidelines. Risk Assessments: Complete Data Protection Impact Assessments (DPIAs) and identify privacy risks associated with new projects, systems, or policies. Audit & Asset Management: Develop and manage the internal IG audit schedule, participate in spot-check audits, and support data flow mapping and Information Asset Register activities. Reporting … Escalation: Collate, analyze, and present IG risk and assurance data for senior bodies, including the Caldicott and Information Governance Assurance Committee and Hospital Management Board. Incident Management: Process and monitor responses to IG incidents and data breaches using Datix, ensuring actions are completed and potential patient harm ...

Infrastructure Engineer

Location
Potters Bar, England, United Kingdom
resolved in a timely manner, while also taking ownership of the company's Microsoft Server, Desktop, Microsoft 365, Azure cloud, networking, and endpoint management environments. The role works closely with the Developers, Azure Architects, and Network Architects to support and evolve the organisation's infrastructure, and provides exposure … timely manner Diagnosing and resolving technical issues remotely and on site Undertaking small- to medium-sized IT and infrastructure projects as instructed by management Providing desktop and server support Supporting and maintaining Microsoft Server/Desktop operating systems, Microsoft 365, and Azure environments Setting up and configuring new laptops ...

Windows Engineer (DV Security Clearance)

Hiring Organisation
Certain Advantage
Location
Stevenage, Hertfordshire, South East, United Kingdom
Employment Type
Contract
Contract Rate
£74.32 per hour, Benefits Overtime Rate
status: Inside IR35 (Umbrella) Windows Engineer Job Description: Digital Excellence function providing IT Services to the business. Responsibilities: Provide operational BAU support, administration & configuration management across Infrastructure and Workplace Windows Environments in restricted secret & above networks. Responsible for maintaining a good working knowledge of the business IT environment. Planning … line with agreed SLAs, following escalation processes as appropriate. Acting as Tier 2/Tier 3 support for Windows related issues & problem resolution, performing incident response, investigation & remediation within SLA timeframes. Co-ordinate with incident & Technical managers, following incident management process during high-severity incidents. Identify ...

Vice President, Data Center Manager – EMEA

Location
Essex, England, United Kingdom
ways: Provide leadership and oversight of data center operations across EMEA, ensuring safe, secure, and compliant execution in a 24x7 mission-critical environment. Lead incident response, change execution, and escalation coordination across infrastructure, network, facilities, and security teams. Oversee vendor access, onsite activities, and service delivery to maintain operational … security, and compliance standards. Drive execution of infrastructure initiatives, including hardware deployments, lifecycle management, structured cabling, and capacity planning. Maintain strong operational governance through effective risk management, control adherence, and documentation discipline. Partner with cross-functional teams to deliver complex infrastructure activities and major change events with minimal ...

Senior Lead Software Engineer - LLM Ops Platform Reliability

Hiring Organisation
Hackajob Ltd
Location
Milton, Cambridgeshire, UK
Software Engineer at our client within the AI and Machine Learning Platform team, you will build and scale AI infrastructure that modernizes traditional infrastructure management and site reliability engineering through applied AI. You will own the reliability, performance, and cost-efficiency of the large language model inference platform … production at scale, with deep instrumentation and strong operational rigor. You will partner across engineering to deliver secure software, improve stability, and lead incident response and continuous improvement. Job responsibilities Design, develop, troubleshoot, and deliver secure, high-quality production software and services for AI infrastructure Build backend services ...

Lead Software Engineer - LLM Ops Platform Reliability

Hiring Organisation
Hackajob Ltd
Location
Milton, Cambridgeshire, UK
Software Engineer at our client in the AI and Machine Learning Platform team, you will build and scale AI infrastructure that modernizes traditional infrastructure management and site reliability engineering through applied AI. You will own the reliability, performance, and cost-efficiency of the LLM inference platform end to end. … production at scale, with deep instrumentation and strong operational rigor. You will partner across engineering to deliver secure software, improve stability, and lead incident response and continuous improvement. Job Responsibilities Design, develop, troubleshoot, and deliver secure, high-quality production software and services for AI infrastructure Build backend services ...

Performance Manager (6 Months FTC, Shifts)

Hiring Organisation
M Group
Location
St. Ives, Cambridgeshire, East Anglia, United Kingdom
Employment Type
Permanent
best technology, manage assets and refresh systems. With 24/7 national operations, we keep things running smoothly, while operating comprehensive network or service management repair and maintenance to keep everything running operationally. Want to come and be a part of it? We're looking for a proactive … teams, customers, and stakeholders to maintain service excellence. Alongside operational leadership, you'll play a key role in developing your team through coaching, performance management, wellbeing support, and creating an inclusive environment where colleagues feel empowered to succeed. What youll bring Experience leading teams within a Service Desk, National ...

Performance Manager (6 Months FTC, Shifts)

Hiring Organisation
M Group
Location
Cambridge, Cambridgeshire, UK
best technology, manage assets and refresh systems. With 24/7 national operations, we keep things running smoothly, while operating comprehensive network or service management repair and maintenance to keep everything running operationally. Want to come and be a part of it? We're looking for a proactive … teams, customers, and stakeholders to maintain service excellence. Alongside operational leadership, you'll play a key role in developing your team through coaching, performance management, wellbeing support, and creating an inclusive environment where colleagues feel empowered to succeed. What you'll bring Experience leading teams within a Service Desk ...

CISCO Telephony & Videoconferencing Engineer

Hiring Organisation
Sanderson Government and Defence
Location
Stevenage, Hertfordshire, South East, United Kingdom
Employment Type
Permanent
continuous improvement of Cisco telephony and video conferencing platforms , ensuring resilient communications across secure Defence networks. You'll lead project implementations, support major incident resolution, and provide technical expertise across complex Cisco UC environments. Key Responsibilities Deliver and implement Cisco Telephony and Video Conferencing projects. … Cisco Unified Communications technologies. Troubleshoot and resolve complex incidents and support major incident management activities. Conduct root cause analysis and contribute to service improvement initiatives. Produce and maintain HLDs, LLDs, and operational documentation. Perform hardware and configuration testing. Provide technical mentoring and knowledge transfer to engineering teams. Work ...

DevOps & Technical Support Engineer

Location
Hitchin, England, United Kingdom
digital marketing. We value simplicity, avoiding complexity and jargon. This opportunity will see the successful candidate responsible for CDA cloud infrastructure environments and Production incident management, working closely with the Software Development and Operations team. THE SUCCESSFUL CANDIDATE We’re looking for a technical, solution-focused individual, with … portfolio demonstrating creativity and ideas across online and offline Ability to lead scoping sessions with both internal teams and clients Good customer and stakeholder management skills Located within a 1-hour commute of Hitchin, Hertfordshire SALARY Salary dependant on skills and experience Private health and dental cover BENEFITS ...

Senior .Net Developer (Blue Prism)

Location
Hatfield, England, United Kingdom
services using WCF and Web API. Proficiency with Visual Studio (2022 or newer) and TFS (DevOPS 2022). Experience in Change and Incident Management environments with strict SLAs. Ability to coordinate application releases and ensure seamless implementation processes. Familiarity with both Agile and Waterfall project delivery methodologies. Hands … experience with DevOps practices, tools, and continuous integration/continuous deployment (CI/CD pipelines). .NET Core Razor pages. Advanced database management with SQL Server versions 2014/2016/2022. Knowledge of Entity Framework 6.0, LINQ, Lambda Expressions, and Razor syntax. Familiarity with PowerShell and SSIS packages. ...

Senior Ofcom Regulatory Manager

Location
Welwyn Garden City, England, United Kingdom
making a contribution to the success of their operations. (As appointed) Membership and trusted adviser for relevant executive committees of Telefonica and other management fora (e.g. security, incident management) You will need A regulatory professional with proven experience, ideally gained in a regulated company, regulatory authority ...

Windchill Support Consultant (M/F)

Location
Watford, England, United Kingdom
Analyze and resolve complex incidents (code, logs, Oracle/SQL databases, Windchill configurations) Handle recurring issues and implement workarounds Ensure compliance with SLAs and incident management processes Deploy fixes and ensure system installation, maintenance, and updates Install and configure environments (integration, validation, pre-production, production) Ensure daily operations ...

Security Engineering Manager - Workplace Technology

Location
Welwyn Garden City, England, United Kingdom
security domains. Desirable Awareness of core technology landscape and retail systems, and how cyber risk translates into customer and business impact. Understanding of cyber incident management models and escalation frameworks across enterprise environments. Experience with product methodologies and service-oriented delivery models. Exposure to data analytics and insights ...

SC-Cleared Google Workspace Pro (L2) - Stevenage

Location
Stevenage, England, United Kingdom
Workspace Support Specialist for an onsite contract in Stevenage, Hertfordshire. This hands-on Level 2 role focuses on Google Workspace administration and strong ITSM incident management. Active SC clearance is essential, with MOD SC preferred. The position requires collaboration across technical teams to maintain high-quality service delivery. #J ...

Senior Data Centre Operations Manager - Multi-Site

Location
Peterborough, England, United Kingdom
ensuring safe operation of power, cooling and M&E systems that support banking services. The role requires proven leadership in data centres, strong stakeholder management and a hands-on approach to incident management, performance, and risk across multiple #J-18808-Ljbffr ...

Maintenance Engineer

Hiring Organisation
NTT Global Data Centers EMEA UK ltd
Location
Dagenham, Essex, South East, United Kingdom
Employment Type
Permanent
operational performance is always maintained to the highest possible standards. Provide engineering services and guidance in general on property matters affecting the on-going management and development of the Data Centers. The role is full time Monday to Friday based at our LON1 site in Dagenham. What you will … diagnosing root causes Monitoring BMS alarms and maintaining equipment performance to ensure correct operation of the critical services of the Data Centers Assisting with incident management and subsequent technical investigations to establish root cause Supporting the Shift Engineering Teams to ensure correct operation of the critical services ...

IT Service Desk Team Lead — Hands-On Leadership

Location
Cambridge, England, United Kingdom
Desk Team Lead to enhance IT support experiences for staff. This role combines leadership and hands-on technical skills, focusing on service improvement and incident management. Candidates should have experience in IT support leadership and a strong commitment to user-focused service delivery. The position requires mostly on-site ...

Cisco Telephony & Video VC Team Lead (Onsite, SC Cleared)

Location
Stevenage, England, United Kingdom
daily rate of up to £869.40 (Umbrella inside IR35). The role requires advanced Cisco Telephony and video conferencing expertise, team leadership, and strong incident management capabilities. You will oversee capabilities across CUCM, Unity, Jabber, and room systems, coordinating with #J-18808-Ljbffr ...

Director of Site Reliability Engineering

Hiring Organisation
Hackajob Ltd
Location
Milton, Cambridgeshire, UK
client within the Corporate Technology and Enterprise Technology Team, you draw upon your advanced knowledge to identify new opportunities to influence critical incident management and improve the end-to-end lifecycle of software development for the firm. You will have the opportunity to manage, design, and implement infrastructure … Identifies and solves problems of high complexity and drives improvements as outcomes Uses enterprise-authorized AI capabilities within the work environment to accelerate complex incident analysis and reliability decisioning, validating outputs and handling operational data according to sensitivity and security requirements. Works with development teams throughout the software life ...

Senior SRE: Cloud Platform Reliability & Automation

Location
Cambridge, England, United Kingdom
reliability, availability, and performance of our large-scale cloud platforms and SaaS products. Your work will automate operational tasks, improve observability, and contribute to incident management while collaborating with development teams to build more reliable and scalable applications across regions. #J-18808-Ljbffr ...

RTLS Support Engineer – Industrial IoT & Smart Factory

Location
Cambridge, England, United Kingdom
Support Analyst to ensure continuity and quality of service for customer installations of Ubisense RTLS systems. You will deliver front-line technical support, incident management, and maintain proactive monitoring across on-site and cloud environments for major customers. The role requires expertise in IT support, log analysis ...

Production Reliability Team Lead

Location
Hemel Hempstead, England, United Kingdom
just keep it running, but build a system that scales, improves reliability, and delivers predictable performance for our enterprise customers. You’ll lead incident management, define monitoring, drive on-call rotations, and mentor the team to ship steady, observable, and blameless production practices. This hands-on role requires ...

Senior Infrastructure Engineer – UK & Ireland

Location
Cambridge, England, United Kingdom
business stakeholders, ensuring high availability and security across critical platforms. You will mentor colleagues, oversee data centre operations, and drive resilience through governance and incident management while shaping infrastructure strategy for ZEISS Ltd. #J-18808-Ljbffr ...