26 to 28 of 28 Observability Jobs in Lanarkshire

Lead Infrastructure Engineer - AWS Cloud Support Engineering

Hiring Organisation
Hackajob Ltd
Location
Glasgow, Lanarkshire, Scotland, United Kingdom
Employment Type
Permanent
adherence to resiliency and security expectations Familiarity with working in a large distributed system across a range of technologies including compute, databases, messaging, observability, and telemetry Knowledge of incident, change, and problem management processes and the controls that govern them Understanding of data-driven decision making and a drive … working in a follow-the-sun or globally distributed on-call support model Familiarity with large-scale cloud migration or modernization initiatives Exposure to observability and telemetry tooling in complex distributed environments ABOUT US Our client is a global leader in financial services, providing strategic advice and products ...

Lead Software Engineer (Container Platforms)

Hiring Organisation
Hackajob Ltd
Location
Glasgow, Lanarkshire, Scotland, United Kingdom
Employment Type
Permanent
usability, and self-service for engineering consumers Create and maintain delivery workflows that standardize engineering practices and reduce operational toil across the platform Improve observability and operational readiness through monitoring, logging, tracing, alerting, runbook development, and on-call practices Partner with security and risk stakeholders to implement secure-by-default … Familiarity with infrastructure-as-code and automation practices, including tools such as Terraform, Helm, Kustomize, Argo CD, Flux, or continuous integration systems Experience with observability stacks and site reliability engineering practices, including service level indicators and objectives, incident response, and post-incident reviews Exposure to regulated environments and implementing security ...

Lead SRE - AWS Platform

Hiring Organisation
Hackajob Ltd
Location
Glasgow, Lanarkshire, Scotland, United Kingdom
Employment Type
Permanent
your team to identify comprehensive service level indicators and partner with stakeholders to establish reasonable service level objectives and error budgets Design and implement observability frameworks and alerting strategies, including white and black box monitoring, service level objective-based alerting, and telemetry collection to ensure proactive detection and response Serve … resiliency best practices Fluency in at least one programming language such as Python, Java/Spring Boot, or .NET Proficient knowledge and experience in observability, including white and black box monitoring, service level objective alerting, and telemetry collection across large-scale production environments Proficiency with continuous integration and continuous delivery ...