51 to 56 of 56 Incident Response Jobs in the North West

Vice President, DevOps Production Services

Hiring Organisation
Jobleads-UK
Location
Manchester, England, United Kingdom
Support to manage and support critical enterprise AI based applications in a fast‐paced production environment. The role requires hands‐on expertise in monitoring, incident management, troubleshooting, release support, and ensuring high availability and stability of business‐critical platforms. In this role, you’ll make an impact … analysis (RCA) for recurring issues and drive permanent fixes. Analyze production logs, identify failure patterns, and create actionable dashboards to improve service monitoring and incident response. Coordinate with development, infrastructure, database, network, and business teams for issue resolution. Support application deployments, change requests, weekend releases, and post‐release validations. ...

Lead Cloud Platform Engineer

Hiring Organisation
Jobleads-UK
Location
Manchester, England, United Kingdom
improving, and troubleshooting the reconciliation pipeline across environments. Manage and optimise core managed services: EKS, MSK, and RDS, including scaling, patching, cost governance, and incident response. Maintain and extend infrastructure‐as‐code using Terraform, ensuring consistency, modularity, and alignment with AWS best practices. Support and improve GitHub Actions … layer awareness, ability to support and troubleshoot Java/Kotlin and/or Python services running in containers. Strong operational mindset, with experience owning incident management, on‐call frameworks, and post‐incident review processes. Proven experience leading a platform or infrastructure engineering team, including mentoring, performance support ...

Lead Engineer - DevOps

Hiring Organisation
Jobleads-UK
Location
Knutsford, England, United Kingdom
Ensure the reliability, availability, and scalability of the systems, platforms, and technology through the application of software engineering techniques, automation, and best practices in incident response. To be successful as a Lead Engineer, you should have experience with: Observability frameworks, tools, and infrastructure Experience implementing Open Telemetry at scale … Ensure the reliability, availability, and scalability of the systems, platforms, and technology through the application of software engineering techniques, automation, and best practices in incident response. Accountabilities Build Engineering: Development, delivery, and maintenance of high-quality infrastructure solutions to fulfil business requirements ensuring measurable reliability, performance, availability, and ease ...

Lead Engineer - DevOps

Hiring Organisation
Jobleads-UK
Location
Knutsford, England, United Kingdom
Ensure the reliability, availability, and scalability of the systems, platforms, and technology through the application of software engineering techniques, automation, and best practices in incident response. Responsibilities Build Engineering: Development, delivery, and maintenance of high-quality infrastructure solutions to fulfil business requirements ensuring measurable reliability, performance, availability, and ease … use. Identify the appropriate technologies and solutions to meet business, optimisation, and resourcing requirements. Incident Management: Monitor IT infrastructure and system performance to measure, identify, address, and resolve any potential issues, vulnerabilities, or outages. Use data to drive down mean time to resolution. Automation: Develop and implement automated tasks ...

Lead Engineer - DevOps

Hiring Organisation
Jobleads-UK
Location
Knutsford, England, United Kingdom
Ensure the reliability, availability, and scalability of the systems, platforms, and technology through the application of software engineering techniques, automation, and best practices in incident response. To be successful as a Lead Engineer, you should have experience with: Observability frameworks, tools, and infrastructure Experience implementing Open Telemetry at scale … Ensure the reliability, availability, and scalability of the systems, platforms, and technology through the application of software engineering techniques, automation, and best practices in incident response. Accountabilities Build Engineering: Development, delivery, and maintenance of high-quality infrastructure solutions to fulfil business requirements ensuring measurable reliability, performance, availability, and ease ...

Infrastructure Specialist

Hiring Organisation
Jobleads-UK
Location
Knutsford, England, United Kingdom
security policies and procedures to ensure data protection. Performance of regular audits and reviews of RACF access controls, resolving any security issues. Monitoring and response to security alerts; performing root‐cause analysis and implementing corrective actions. Understanding the implementation of Mainframe Certificates and Mainframe cryptography. Experience of Security Best … Ensure the reliability, availability, and scalability of the systems, platforms, and technology through the application of software engineering techniques, automation, and best practices in incident response. Accountabilities Build Engineering: Development, delivery, and maintenance of high‐quality infrastructure solutions to fulfil business requirements ensuring measurable reliability, performance, availability, and ease ...