22 of 22 Observability Jobs in Lanarkshire

Lead SRE- Azure & GCP

Hiring Organisation
Hackajob Ltd
Location
Glasgow, Lanarkshire, Scotland, United Kingdom
Employment Type
Permanent
with good understanding of REST APIs Hands-on experience with cloud-based technologies and tools especially in deployment, monitoring and operations, such as Google Observability, Azure Monitor, Data Dog, Prometheus, Splunk, Elasticsearch and Grafana. Demonstrated experience using enterprise-authorized AI capabilities within the work environment to improve SRE workflows (e.g. ...

Senior Manager of SRE

Hiring Organisation
Hackajob Ltd
Location
Glasgow, Lanarkshire, Scotland, United Kingdom
Employment Type
Permanent
drive outcomes-oriented probing of architectural designs, technical credentials, and applicability for use within existing systems and information architecture. Drives continuous improvement in system observability, alerting, and capacity planning. Collaborates with engineering and data teams to optimize infrastructure and deployment processes, focusing on automation and operational excellence. Performs platform design ...

Java Software Engineer - VP

Hiring Organisation
Henderson Scott
Location
Glasgow, Lanarkshire, Scotland, United Kingdom
Employment Type
Permanent
Data: Proficiency with Docker, Kubernetes, and relational databases (MS SQL, Sybase). Delivery Automation & Monitoring: Strong focus on automated testing, automated release pipelines, and observability tools (Grafana, Prometheus). Mindset: Delivery-focused problem solver with a hands-on approach and strong stakeholder communication skills. Desirable Experience Background in Equity Swaps ...

Software Engineer III - Data Engineering- Corporate Know Your Customer

Hiring Organisation
Hackajob Ltd
Location
Glasgow, Lanarkshire, Scotland, United Kingdom
Employment Type
Permanent
depth in disciplines such as cloud, AI/ML, or data engineering Experience in large-scale data processing, microservices, API design, Kafka, Redis, MemCached, observability tools (Dynatrace, Splunk, Grafana), and orchestration frameworks (Airflow, Temporal) Advanced working knowledge of relational and NoSQL databases, vector stores, data lake architectures, and data governance ...

Platform Engineer

Location
Douglas, Lanarkshire, United Kingdom
management (IAM), and encryption techniques Microsoft Azure certifications are a plus Experience working in regulated or security-conscious enterprise environments Experience with monitoring and observability tooling such as Datadog or similar platforms Benefits of working at Canada Life We believe in recognising and rewarding our people, so we offer ...

Lead Software Data Engineer - Corporate Know Your Customer

Hiring Organisation
Hackajob Ltd
Location
Glasgow, Lanarkshire, Scotland, United Kingdom
Employment Type
Permanent
depth in disciplines such as cloud, AI/ML, or data engineering Experience in large-scale data processing, microservices, API design, Kafka, Redis, MemCached, observability tools (Dynatrace, Splunk, Grafana), and orchestration frameworks (Airflow, Temporal) Advanced working knowledge of relational and NoSQL databases, vector stores, data lake architectures, and data governance ...

Google Cloud Platform Architect

Hiring Organisation
Sanderson Recruitment
Location
Glasgow, Lanarkshire, Scotland, United Kingdom
Employment Type
Contract
Contract Rate
£400 - £500 per day
expertise in Go and/or Python Kubernetes platform engineering experience, including CRDs and API extensions. Strong knowledge of Kubernetes internals, multi-tenancy and observability Proven experience with CI/CD, automation, infrastructure as code and deployment pipelines. Strong understanding of SDLC, Agile delivery, resiliency and security Practical ...

Lead Site Reliability Engineer

Hiring Organisation
Hackajob Ltd
Location
Glasgow, Lanarkshire, Scotland, United Kingdom
Employment Type
Permanent
Fluency in Python & deep knowledge of software applications and technical processes with emerging depth in one or more technical disciplines Proficiency and experience in observability such as white and black box monitoring, SLO alerting, and telemetry collection using tools such as Grafana, Dynatrace, Prometheus, Datadog, Splunk, etc. Proficiency in continuous ...

Lead Site Reliability Engineer - Chief Technology Office

Hiring Organisation
Hackajob Ltd
Location
Glasgow, Lanarkshire, Scotland, United Kingdom
Employment Type
Permanent
.NET) Deep knowledge of software applications and technical processes with emerging depth in one or more technical disciplines Proficiency and hands-on experience in observability practices including white and black box monitoring, service level objective alerting, and telemetry collection using tools such as Grafana, Dynatrace, Prometheus, Datadog, or Splunk Proficiency ...

Python Developer

Hiring Organisation
Diana Duggan UK Limited
Location
Glasgow, Lanarkshire, United Kingdom
Employment Type
Full-Time
Salary
£450.00 per day
opportunities for increased automation and reduced manual intervention. Troubleshoot complex issues across application, automation, infrastructure, and deployment layers. Contribute to improvements in deployment reliability, observability, resilience, and operational support. Participate in code reviews and contribute to engineering standards, design decisions, and best practices. Collaborate with developers, QA engineers, DevOps, architects ...

Network Engineer

Hiring Organisation
SUMMER-BROWNING ASSOCIATES LIMITED
Location
Glasgow, Lanarkshire, Scotland, United Kingdom
Employment Type
Contract
Contract Rate
£0.00 - £0.01 per day
configuration solutions using Ansible, Python/PowerShell, Git and REST APIs. Integrate network provisioning and configuration into CI/CD pipelines. Support monitoring, observability, capacity planning, performance optimisation and continuous service improvement. What We're Looking For Strong experience as a Senior Network Engineer within complex enterprise environments. Excellent knowledge ...

Storage Automation/ Python Developer - Investment Bank (hybrid)

Hiring Organisation
Robert Walters
Location
Glasgow, Lanarkshire, United Kingdom
Employment Type
Full-Time
Salary
£350.00 - £480.00 per day
architect and integrate storage technologies and solutions. Support storage virtualisation, software-defined storage and distributed/scale-out file systems. Develop automation and observability capabilities across large-scale infrastructure environments. Troubleshoot complex issues spanning operating systems, networks and storage infrastructure. Work closely with internal customers to understand requirements and translate ...

Sr Lead Infrastructure Engineer- Devops/AWS

Hiring Organisation
Hackajob Ltd
Location
Glasgow, Lanarkshire, Scotland, United Kingdom
Employment Type
Permanent
machine learning products. As the team takes end-to-end ownership of the platforms it runs, you will build the CI/CD, observability, and incident-management practices that keep those services stable, secure, and performant across international markets. This is a Vice President-level role and an integral part … release automation, and deployment tooling Establishes reliability practices (SLOs, error budgets, runbooks) and leads production incident response and post-incident review Builds and operates observability across the team's AI/ML services (metrics, logging, tracing, alerting) Automates infrastructure provisioning and configuration through infrastructure-as-code Implements operational security, secrets ...

Java Full stack Developer

Hiring Organisation
NEEV LIMITED
Location
Glasgow, Lanarkshire, Scotland, United Kingdom
Employment Type
Contract
Contract Rate
£400 per day
enterprise applications using modern Java technologies, cloud-native architectures, and microservices. The ideal candidate will have hands-on experience with Kafka, Kubernetes, API Security, Observability tools, SQL databases, and Spring-based microservices development . Key Responsibilities Design, develop, and maintain scalable Java-based applications using Java 17+, Spring Boot … mechanisms. Build event-driven solutions using Apache Kafka for real-time data processing and messaging. Deploy, manage, and troubleshoot applications on Kubernetes environments. Implement observability solutions using tools such as Splunk, ELK, Grafana, Prometheus, Dynatrace, or AppDynamics. Optimize application performance, scalability, and reliability. Work closely with business stakeholders, architects ...

Sr Lead AI Platform Engineer

Hiring Organisation
Hackajob Ltd
Location
Glasgow, Lanarkshire, Scotland, United Kingdom
Employment Type
Permanent
Owns the design and build of the team's platform: deployment pipelines, model serving, containerisation, orchestration, and environment management Sets the standard for reliability, observability, and operational excellence across the team's production AI/ML services Builds the tooling and paved paths that let AI engineers ship agentic … record of building deployment and release automation Experience serving, scaling, and monitoring ML models or data-intensive services in production (MLOps) Practical experience with observability tooling (metrics, logging, tracing) and production incident response Experience operating ML/LLM workloads in production (LLMOps, inference reliability, cost/performance management) Strong communication ...

Lead Site Reliability Engineer

Hiring Organisation
Hackajob Ltd
Location
Glasgow, Lanarkshire, Scotland, United Kingdom
Employment Type
Permanent
practices within an application or platform Fluency in at least one programming language such as (e.g., Java, Python, Go, etc.) Proficiency and experience in observability such as white and black box monitoring, SLO alerting, and telemetry collection using tools such as Grafana, Dynatrace, Prometheus, Datadog, Splunk, Elasticsearch, etc. Proficiency … high-availability services Deep understanding of distributed system design principles, networking (TCP/IP, DNS, load balancing), Linux internals. Contributions to open-source observability or telemetry projects. Experience working with agent control planes and management protocols. Hands-on knowledge of OpAMP is highly desirable. ABOUT US J.P. Morgan ...

Lead Software Engineer - LLM Ops Platform Reliability

Hiring Organisation
Hackajob Ltd
Location
Glasgow, Lanarkshire, Scotland, United Kingdom
Employment Type
Permanent
strong engineering fundamentals and site reliability practices to cutting-edge AI platforms. You'll work hands-on with cloud and Kubernetes-based deployments, deep observability, and cost-aware performance tuning. If you enjoy solving hard production problems and making platforms measurably better, you'll find meaningful impact and growth here. … Amazon EKS and Amazon SageMaker, as well as on-prem and local GPU clusters, using reproducible infrastructure as code and continuous delivery pipelines Implement observability (logs, metrics, traces) with dashboards and actionable alerting, including Prometheus metrics and Grafana/Alertmanager integration for LLM and GPU workloads Tune GPU and accelerator ...

Corporate KYC Principle Software Engineer - Executive Director

Hiring Organisation
Hackajob Ltd
Location
Glasgow, Lanarkshire, Scotland, United Kingdom
Employment Type
Permanent
regulated financial services environments Establishes engineering standards for LLM-based applications RAG pipelines, embedding workflows, vector store integrations, and model serving ensuring safety, observability, and reproducibility at scale Drives adoption of advanced technical methods and practices aligned with the latest industry standards and product development methodologies Serves as the function … more disciplines (e.g., cloud, AI/ML, data engineering) Experience in large-scale data processing, microservices, API design, Kafka, Redis, MemCached, observability tools (Dynatrace, Splunk, Grafana), and orchestration frameworks (Airflow, Temporal) Advanced working knowledge of relational and NoSQL databases, vector stores, data lake architectures, and data governance Practical cloud-native ...

Corporate KYC Sr Lead Software Engineer

Hiring Organisation
Hackajob Ltd
Location
Glasgow, Lanarkshire, Scotland, United Kingdom
Employment Type
Permanent
scale data processing, microservices, API design, and orchestration frameworks Working knowledge of relational and NoSQL databases, vector stores, and data lake architectures Familiarity with observability tools and frameworks Practical cloud-native experience (AWS, Azure, or GCP) Ability to communicate effectively with senior leaders and executives Commitment to inclusive, collaborative teamwork … catalog services such as Apache Iceberg Experience with LLM orchestration frameworks and model serving infrastructure or managed endpoints Familiarity with AI evaluation and observability practices for LLM workloads Understanding of agentic design patterns and how to constrain agent autonomy in financial workflows Interest in emerging technologies and continuous learning Employer ...

Senior Lead Software Engineer - LLM Ops Platform Reliability

Hiring Organisation
Hackajob Ltd
Location
Glasgow, Lanarkshire, Scotland, United Kingdom
Employment Type
Permanent
strong engineering fundamentals and site reliability practices to cutting-edge AI platforms. You'll work hands-on with cloud and Kubernetes-based deployments, deep observability, and cost-aware performance tuning. If you enjoy solving hard production problems and making platforms measurably better, you'll find meaningful impact and growth here. … large language models on cloud-based container orchestration platforms and on-premises GPU clusters using reproducible infrastructure as code and continuous delivery pipelines Implement observability across logs, metrics, and traces with dashboards and actionable alerting for large language model and GPU workloads Tune GPU and accelerator capacity, autoscaling, and cost ...

Lead Infrastructure Engineer - AWS Cloud Support Engineering

Hiring Organisation
Hackajob Ltd
Location
Glasgow, Lanarkshire, Scotland, United Kingdom
Employment Type
Permanent
adherence to resiliency and security expectations Familiarity with working in a large distributed system across a range of technologies including compute, databases, messaging, observability, and telemetry Knowledge of incident, change, and problem management processes and the controls that govern them Understanding of data-driven decision making and a drive … working in a follow-the-sun or globally distributed on-call support model Familiarity with large-scale cloud migration or modernization initiatives Exposure to observability and telemetry tooling in complex distributed environments ABOUT US J.P. Morgan is a global leader in financial services, providing strategic advice and products ...

Lead SRE - AWS Platform

Hiring Organisation
Hackajob Ltd
Location
Glasgow, Lanarkshire, Scotland, United Kingdom
Employment Type
Permanent
your team to identify comprehensive service level indicators and partner with stakeholders to establish reasonable service level objectives and error budgets Design and implement observability frameworks and alerting strategies, including white and black box monitoring, service level objective-based alerting, and telemetry collection to ensure proactive detection and response Serve … resiliency best practices Fluency in at least one programming language such as Python, Java/Spring Boot, or .NET Proficient knowledge and experience in observability, including white and black box monitoring, service level objective alerting, and telemetry collection across large-scale production environments Proficiency with continuous integration and continuous delivery ...