26 to 50 of 68 Incident Management Jobs in Scotland

Sr Lead Infrastructure Engineer- Devops/AWS

Location
Auchentibber, Scotland, United Kingdom
learning products. As the team takes end-to-end ownership of the platforms it runs, you will build the CI/CD, observability, and incident-management practices that keep those services stable, secure, and performant across international markets. This is a Vice President-level role and an integral … evolves the team's CI/CD pipelines, release automation, and deployment tooling Establishes reliability practices (SLOs, error budgets, runbooks) and leads production incident response and post-incident review Builds and operates observability across the team's AI/ML services (metrics, logging, tracing, alerting) Automates infrastructure provisioning ...

Chief Information Security Officer

Location
Glasgow, Scotland, United Kingdom
privacy and resilience, and develop mitigating actions consistent with Ofgem’s risk appetite. Managing Cyber assurance, data protection, security operations and response capabilities, including Incident Management on Data, leak investigations and independent investigation activity with Cabinet Office. Acting as a security and privacy professional, championing best practice … will need: Direct experience in leading Cyber Security, including ongoing definition and delivery of a robust security and privacy strategy, which incorporates intelligence management; threat and risk analysis; risk mitigation; business continuity preparations; and incident detection, response and recovery. Substantial experience in a combination of business and risk ...

Senior Lead Software Engineer - LLM Ops Platform Reliability

Location
Glasgow, Scotland, United Kingdom
Lead Software Engineer at JPMorganChase within the AI and Machine Learning Platform team, you will build and scale AI infrastructure that modernizes traditional infrastructure management and site reliability engineering through applied AI. You will own the reliability, performance, and cost-efficiency of the large language model inference platform … production at scale, with deep instrumentation and strong operational rigor. You will partner across engineering to deliver secure software, improve stability, and lead incident response and continuous improvement. Job responsibilities Design, develop, troubleshoot, and deliver secure, high-quality production software and services for AI infrastructure Build backend services ...

Senior Lead Software Engineer - LLM Ops Platform Reliability

Hiring Organisation
Hackajob Ltd
Location
Paisley, Renfrewshire, UK
Software Engineer at our client within the AI and Machine Learning Platform team, you will build and scale AI infrastructure that modernizes traditional infrastructure management and site reliability engineering through applied AI. You will own the reliability, performance, and cost-efficiency of the large language model inference platform … production at scale, with deep instrumentation and strong operational rigor. You will partner across engineering to deliver secure software, improve stability, and lead incident response and continuous improvement. Job responsibilities Design, develop, troubleshoot, and deliver secure, high-quality production software and services for AI infrastructure Build backend services ...

Lead Software Engineer - LLM Ops Platform Reliability

Hiring Organisation
Hackajob Ltd
Location
Glasgow, UK
Software Engineer at our client in the AI and Machine Learning Platform team, you will build and scale AI infrastructure that modernizes traditional infrastructure management and site reliability engineering through applied AI. You will own the reliability, performance, and cost-efficiency of the LLM inference platform end to end. … production at scale, with deep instrumentation and strong operational rigor. You will partner across engineering to deliver secure software, improve stability, and lead incident response and continuous improvement. Job Responsibilities Design, develop, troubleshoot, and deliver secure, high-quality production software and services for AI infrastructure Build backend services ...

Lead Software Engineer - LLM Ops Platform Reliability

Location
Paisley, Scotland, United Kingdom
Software Engineer at JPMorgan Chase in the AI and Machine Learning Platform team, you will build and scale AI infrastructure that modernizes traditional infrastructure management and site reliability engineering through applied AI. You will own the reliability, performance, and cost-efficiency of the LLM inference platform end to end. … production at scale, with deep instrumentation and strong operational rigor. You will partner across engineering to deliver secure software, improve stability, and lead incident response and continuous improvement. Job Responsibilities Design, develop, troubleshoot, and deliver secure, high-quality production software and services for AI infrastructure Build backend services ...

Senior Lead Software Engineer - LLM Ops Platform Reliability

Location
Auchentibber, Scotland, United Kingdom
Lead Software Engineer at JPMorganChase within the AI and Machine Learning Platform team, you will build and scale AI infrastructure that modernizes traditional infrastructure management and site reliability engineering through applied AI. You will own the reliability, performance, and cost-efficiency of the large language model inference platform … production at scale, with deep instrumentation and strong operational rigor. You will partner across engineering to deliver secure software, improve stability, and lead incident response and continuous improvement. Job responsibilities Design, develop, troubleshoot, and deliver secure, high-quality production software and services for AI infrastructure Build backend services ...

OT Technical Support Manager

Location
Perth, Scotland, United Kingdom
leadership role focused on ensuring the availability, performance, security, and continuous improvement of high-availability OT platforms. The successful candidate will be responsible for incident management, fault recovery, service delivery, vulnerability remediation, workforce planning, and the ongoing development of technical support teams. You will work closely with operational … aligned to business and regulatory requirements. What You'll Be Doing Lead the operational support and maintenance of critical OT platforms and services. Manage incident response activities, fault diagnosis, service restoration, and root cause analysis. Ensure operational resilience across high-availability environments, supporting business continuity and recovery activities. Oversee ...

Senior Lead Software Engineer - LLM Ops Platform Reliability

Hiring Organisation
JP Morgan Chase
Location
Glasgow, UK
Employment Type
Full-time
Lead Software Engineer at JPMorganChase within the AI and Machine Learning Platform team, you will build and scale AI infrastructure that modernizes traditional infrastructure management and site reliability engineering through applied AI. You will own the reliability, performance, and cost-efficiency of the large language model inference platform … production at scale, with deep instrumentation and strong operational rigor. You will partner across engineering to deliver secure software, improve stability, and lead incident response and continuous improvement. Job responsibilitiesDesign, develop, troubleshoot, and deliver secure, high-quality production software and services for AI infrastructureBuild backend services and APIs that ...

Security Operations Team Lead

Location
Paisley, Scotland, United Kingdom
Security Operations team to proactively identify and address cybersecurity threats, drive operational excellence, and enhance our security posture. You'll lead strategic initiatives, supervise incident response, manage critical security projects, and foster a collaborative environment, ensuring top-tier system performance and adherence to security policies. Power without pause. Heating … Provide hands-on expertise in the delivery of internal security projects and technical support for operational excellence. Act as a technical escalation point for incident management, ensuring prompt resolution and the consistent implementation of security processes and compliance. Collaborate closely with external suppliers and manage key security service ...

Chief Information Security Officer

Hiring Organisation
Public Sector Resourcing CWS
Location
Aberdeen, Aberdeenshire, Scotland, United Kingdom
Employment Type
Permanent
environments - cloud, collaboration, and staff identity and access - as they stand up, so that new services launch on secure foundations. * Establish incident response for real: plans, roles, on-call arrangements and exercises that prove GBE can handle an incident, not just describe one. * Baseline GBE's security maturity … risk and assurance at executive level, including presenting risk clearly and honestly to boards and audit committees. * Operational depth: accountability for detection, response and incident management, including choosing and directing the right delivery model - in-house, managed service or hybrid. * Depth in security standards and compliance - the NCSC ...

Trainee DevOps Engineer | No experience needed (Ref: 7501)

Hiring Organisation
Qualify Nation Recruitment
Location
Glasgow, Lanarkshire, United Kingdom
Employment Type
Full-Time
Salary
£28,000 - £38,000 per annum
Continuous Deployment (CD) Infrastructure as Code (IaC) Containerisation with Docker Container Orchestration Cloud Computing Fundamentals Cloud Platforms (AWS, Microsoft Azure and Google Cloud) Configuration Management Monitoring and Logging Security Best Practices (DevSecOps) Networking Fundamentals Automation and Scripting Incident Management and Reliability Engineering Practical Experience You will work ...

Trainee DevOps Engineer | No experience needed (Ref: 7501)

Hiring Organisation
Qualify Nation Recruitment
Location
Edinburgh, Midlothian, United Kingdom
Employment Type
Full-Time
Salary
£28,000 - £38,000 per annum
Continuous Deployment (CD) Infrastructure as Code (IaC) Containerisation with Docker Container Orchestration Cloud Computing Fundamentals Cloud Platforms (AWS, Microsoft Azure and Google Cloud) Configuration Management Monitoring and Logging Security Best Practices (DevSecOps) Networking Fundamentals Automation and Scripting Incident Management and Reliability Engineering Practical Experience You will work ...

Salesforce.com Commercial Application Manager

Location
Thurso, Scotland, United Kingdom
prioritization, and delivery of enhancements. Stakeholder Engagement: Serve as liaisonbetween business and IT; lead discovery sessions,requirements gathering and solution design. Project/Enhancement Management and Execution Lead implementation activities through all phases of the project lifecycle. Support project planning, scope definition, timelines, resource coordination, and delivery management. Collaborate … operational excellence. Identifyproject risks, dependencies, and mitigation strategies to ensure successful project outcomes. Testing & QA: Ensure quality through testing, UAT, and release. Change Management: Drive user adoption through training, communication, and support. Leadership& Communication Serve as the primary interface between business stakeholders and IT, acting as the trusted advisor ...

Site Reliability Engineer (SRE) - Glasgow, UK

Location
Glasgow, Scotland, United Kingdom
Skills:**Site Reliability Operational Support* Monitor maintain and support businesscritical applications and cloud infrastructure* Ensure high system availability performance scalability and reliability* Participate in incident management root cause analysis and problem resolution* Implement proactive monitoring alerting and observability solutions* Reduce operational overhead through automation and selfhealing mechanisms* Support … scripts and tools to improve operational efficiency* Partner with development teams to streamline software delivery processes* Drive continuous improvement initiatives across DevOps and release management practicesSoftware Development Engineering* Contribute to application enhancements and operational tooling using C and NET technologies* Develop utilities automation scripts and support tools* Assist development ...

Senior Backend Engineer

Hiring Organisation
Inspire People
Location
Edinburgh, Midlothian, Scotland, United Kingdom
Employment Type
Permanent, Part Time, Work From Home
Salary
£80,000
code solutions to support service delivery. * Embed Site Reliability Engineering principles including SLIs, SLOs, error budgets and continual service improvement. * Support live service operations, incident management, troubleshooting and platform resilience. * Collaborate with engineers, architects, product teams and stakeholders whilst coaching and mentoring colleagues. Essential Skills for the Senior ...

Service Desk Analyst x4 - Inverness

Hiring Organisation
Adecco
Location
Inverness, Highlands, United Kingdom
Employment Type
Permanent
Salary
£31000 - £32000/annum + shift allowance
days off. We are seeking an experienced Service Desk Analyst to join a fast-paced IT Operations team, providing operational oversight and incident coordination across critical services. You will be responsible for monitoring operational alerts, coordinating technical teams and suppliers, managing escalations, and ensuring service disruptions are effectively prioritised … resolved. You will play a key role in maintaining service availability while providing clear communication to stakeholders throughout the incident life cycle. Key Responsibilities Monitor alerts, events, and incidents, assessing impact and ensuring appropriate action is taken. Coordinate with technical teams, suppliers, and stakeholders to manage incidents and service ...

Cyber Security Analyst

Location
Aberdeen City, Scotland, United Kingdom
Apply threat modelling principles to complex system and solution designs to identify security risks and appropriate mitigations Supports, monitors and recommends improvements to cyber incident management process Provides input and support to operational projects related to cyber security Experience of working in an organisation distributed across different geographies … preferred) Excellent analytical, problem solving and execution skills (essential) Strong cyber security-specific experience, support by relevant industry certifications (e.g. CySA+, Security+) and risk management knowledge (essential) Knowledge and experience working across a diverse range of cyber security tools, including SIEM technologies, EDR, NIDS etc. (essential) Self-motivated with ...

Cyber Security Analyst

Location
Bellshill, Scotland, United Kingdom
Supports IS Security achieve regulatory and statutory compliance requirements Supports identification of security risks and appropriate mitigations Supports, monitors and recommends improvements to cyber incident management process Provides input and support to operational projects related to cyber security Knowledge and experience working across a diverse range of cyber … with a solid understanding of current cyber threats (essential) Strong cyber security-specific experience, support by relevant industry certifications (e.g. CySA+, Security+) and risk management knowledge (essential) Understanding of assessing data security and governance requirements and identifying suitable controls. (essential) Excellent analytical, problem solving and execution skills (essential) Excellent ...

24/7 Security Incident Analyst – SIEM & Threat Response

Location
Inverness, Scotland, United Kingdom
Capgemini in the UK is seeking a Security Incident Management Analyst to detect, investigate, and respond to threats across government and commercial clients. You’ll monitor security systems, log incidents, and ensure tool health, collaborating with analysts and support teams to drive continuous improvement. The role ...

Remote Staff Software Engineer - Data Platforms

Hiring Organisation
grabjobs
Location
Penicuik, Midlothian, UK
GitHub. Experience in operationally managing software components/service once live, including: observability best practises, logging best practises, error reporting, debugging and live incident management. Experience using tools such as Grafana, Prometheus, New Relic etc. Experience of working with sensitive personal data. Experience working with a health data (genetic ...

Remote Staff Software Engineer - Data Platforms

Hiring Organisation
grabjobs
Location
Kilmacolm, Inverclyde, UK
GitHub. Experience in operationally managing software components/service once live, including: observability best practises, logging best practises, error reporting, debugging and live incident management. Experience using tools such as Grafana, Prometheus, New Relic etc. Experience of working with sensitive personal data. Experience working with a health data (genetic ...

Remote Staff Software Engineer - Data Platforms

Hiring Organisation
grabjobs
Location
West Calder, West Lothian, UK
GitHub. Experience in operationally managing software components/service once live, including: observability best practises, logging best practises, error reporting, debugging and live incident management. Experience using tools such as Grafana, Prometheus, New Relic etc. Experience of working with sensitive personal data. Experience working with a health data (genetic ...

Remote Staff Software Engineer - Data Platforms

Hiring Organisation
grabjobs
Location
Tranent, East Lothian, UK
GitHub. Experience in operationally managing software components/service once live, including: observability best practises, logging best practises, error reporting, debugging and live incident management. Experience using tools such as Grafana, Prometheus, New Relic etc. Experience of working with sensitive personal data. Experience working with a health data (genetic ...

Remote Staff Software Engineer - Data Platforms

Hiring Organisation
grabjobs
Location
Clydebank, West Dunbartonshire, UK
GitHub. Experience in operationally managing software components/service once live, including: observability best practises, logging best practises, error reporting, debugging and live incident management. Experience using tools such as Grafana, Prometheus, New Relic etc. Experience of working with sensitive personal data. Experience working with a health data (genetic ...