51 to 68 of 68 Incident Management Jobs in Scotland

Remote Staff Software Engineer - Data Platforms

Hiring Organisation
grabjobs
Location
Campbeltown, Argyll & Bute, UK
GitHub. Experience in operationally managing software components/service once live, including: observability best practises, logging best practises, error reporting, debugging and live incident management. Experience using tools such as Grafana, Prometheus, New Relic etc. Experience of working with sensitive personal data. Experience working with a health data (genetic ...

VP DevOps/SRE: AI/ML Infra, CI/CD & Reliability

Location
Glasgow, Scotland, United Kingdom
Engineer - Vice President to own automation, reliability, and production operations for AI/ML services. You will build CI/CD, observability, and incident-management practices across international markets. Leverage Terraform, Kubernetes, and cloud-native tooling to scale release automation and reliability, while mentoring engineers and enforcing security … change-management standards. #J-18808-Ljbffr ...

DevOps Engineer

Location
Glasgow, Scotland, United Kingdom
DevOps Engineering experience. Strong CI/CD and infrastructure automation knowledge. Experience with cloud technologies and container orchestration. Knowledge of SRE/Observability and incident management. Experience with Configuration & Release Management. Strong troubleshooting and stakeholder communication skills. #J-18808-Ljbffr ...

Head of Engineering

Location
City of Edinburgh, Scotland, United Kingdom
pace and pragmatism required in a fast-moving commercial environment. The Role Lead, develop and grow the engineering team through coaching, mentoring and performance management whilst maintaining the highest standards across development, testing and deployment Drive predictable and high-quality software delivery in a rapidly evolving business, ensuring production … modern software delivery practices data modelling, database performance and scalable system design software quality, testing strategies and operational excellence reliability, observability, deployment and incident management processes The ability to balance technical excellence with commercial priorities Strong communication skills and confidence engaging with senior stakeholders and board-level audiences ...

Operations Technician (EC&I)

Location
Blackburn, Scotland, United Kingdom
commissioning and integration of new assets, as well as identifying opportunities for equipment upgrades and improvements. Participating in emergency response activities, standby rotas and incident management arrangements to support business resilience. Monitoring contractor activities and protecting pipeline assets to ensure technical, operational and safety standards are consistently maintained. ...

Senior Database Administrator

Location
City of Edinburgh, Scotland, United Kingdom
primarily PHP at Podfather) to review how the application talks to the database, ensuring migrations are a routine process rather than a risky event. Incident Management: Take a leading role in debugging complex database issues, managing incidents, and collaborating with support teams to share known workarounds. Required Skills ...

Hybrid Application Support Engineer - 12-Month Contract

Location
Glenrothes, Scotland, United Kingdom
work with infrastructure, network and security teams to ensure environment readiness. The role requires ITIL familiarity, experience with Windows/Linux, and strong incident management skills. #J-18808-Ljbffr ...

Senior DevOps & Cloud Systems Engineer (Remote)

Location
Dundee, Scotland, United Kingdom
office. You will own AWS-based production platforms, manage Linux environments, and drive automation and IaC using Terraform and Ansible. The role demands strong incident management, monitoring, and on-call readiness. You will collaborate with development teams to optimise performance, security, and reliability across systems and CI/ ...

VP DevOps/SRE Lead — Cloud Infra, Kubernetes & CI/CD

Location
Glasgow, Scotland, United Kingdom
Engineer - Vice President to own automation, reliability and production operations of AI/ML platforms. You will build CI/CD, observability, and incident-management practices that keep services stable across international markets. As part of the IPB Tech AIML team, you will lead reliability engineering ...

VP DevOps/SRE for AI & ML Platforms

Location
Auchentibber, Scotland, United Kingdom
level. You will own automation, reliability, and production operations of our AI/ML platforms, building CI/CD pipelines, observability, and incident-management practices that keep services stable, secure, and performant across international markets. Reporting to the Head of AI, IPB Tech, you will mentor engineers, implement ...

Support Engineer

Location
City of Edinburgh, Scotland, United Kingdom
another region can continue work without losing momentum. Protect service quality — recognise high-impact incidents quickly, follow the appropriate escalation process and contribute to incident reviews. Build customer knowledge — create and maintain troubleshooting guides, known-issue documentation and help-centre content. Reduce repeat problems — identify patterns in support demand … software. The ability to create automated tests and submit code fixes for engineering review. Experience supporting AI-enabled or data-intensive applications. Familiarity with incident-management and observability tools. How we work We’re an in-person team based in Edinburgh. You should expect to work from ...

Director of Site Reliability Engineering

Hiring Organisation
Hackajob Ltd
Location
Glasgow, UK
client within the Corporate Technology and Enterprise Technology Team, you draw upon your advanced knowledge to identify new opportunities to influence critical incident management and improve the end-to-end lifecycle of software development for the firm. You will have the opportunity to manage, design, and implement infrastructure … Identifies and solves problems of high complexity and drives improvements as outcomes Uses enterprise-authorized AI capabilities within the work environment to accelerate complex incident analysis and reliability decisioning, validating outputs and handling operational data according to sensitivity and security requirements. Works with development teams throughout the software life ...

Director of Site Reliability Engineering

Hiring Organisation
Hackajob Ltd
Location
Paisley, Renfrewshire, UK
client within the Corporate Technology and Enterprise Technology Team, you draw upon your advanced knowledge to identify new opportunities to influence critical incident management and improve the end-to-end lifecycle of software development for the firm. You will have the opportunity to manage, design, and implement infrastructure … Identifies and solves problems of high complexity and drives improvements as outcomes Uses enterprise-authorized AI capabilities within the work environment to accelerate complex incident analysis and reliability decisioning, validating outputs and handling operational data according to sensitivity and security requirements. Works with development teams throughout the software life ...

IT Service & Security Operations Manager

Location
Addiewell, Scotland, United Kingdom
supporting around 300 local customers and contributing to IT services across an estate of approximately 1,500 users. The role covers IT service delivery, incident management, technical support, projects, security and compliance within Justice Services, Education and Healthcare networks. #J-18808-Ljbffr ...

Production Reliability Leader

Location
City of Edinburgh, Scotland, United Kingdom
systems, shaping processes, leading incidents, building the team, and moving from reactive firefighting to proactive reliability engineering. This hands-on role focuses on monitoring, incident management, on-call rotations, and driving observability with SLIs/SLOs while maintaining a blameless culture and strong ownership. #J-18808-Ljbffr ...

OT Systems Support Manager: Lead Operations & Resilience

Location
Perth, Scotland, United Kingdom
Perth, Scotland. This hybrid role requires 1-2 days on site weekly to lead operational technology support within a Utilities environment. You will oversee incident management, fault diagnosis, and service restoration across high-availability OT platforms, drive patching and vulnerability remediation, and mentor technical teams while aligning with ...

Senior Manager of SRE

Location
Glasgow, Lanarkshire, United Kingdom
that support the firm's commercial goals by harnessing artificial intelligence and machine learning technologies to develop new products, improve productivity, and enhance risk management effectively and responsibly. As a Senior Manager - SRE at our client within the AIML Data Platforms and Chief Data and Analytics Team , you will … high-quality production code, and reviews and debugs code written by others. Adds to team culture of diversity, opportunity, and respect. Develops and maintains incident response procedures, including root cause analysis and postmortem documentation. Required Qualifications, Capabilities, and Skills: Experience with formal training or certification on software engineering concepts. ...

Senior Manager of SRE

Location
Auchentibber, Scotland, United Kingdom
solutionsthat support the firm’s commercial goals by harnessing artificial intelligence and machine learning technologies to develop new products, improve productivity, and enhance risk management effectively and responsibly. As a Senior Manager - SRE at JPMorgan Chase within the AIML Data Platforms and Chief Data and Analytics Team , you will … high‐quality production code, and reviews and debugs code written by others. Adds to team culture of diversity, opportunity, and respect. Develops and maintains incident response procedures, including root cause analysis and postmortem documentation. Required Qualifications, Capabilities, and Skills Experience with formal training or certification on software engineering concepts. ...