126 to 150 of 157 Observability Jobs in the North of England

Lead Site Reliability Engineer (SRE Squad Lead)

Hiring Organisation
Inspire People
Location
Manchester, North West, United Kingdom
Employment Type
Permanent, Part Time, Work From Home
Salary
£80,000
diverse engineering community. Design, build and maintain reliable, secure and scalable cloud-based infrastructure using infrastructure-as-code approaches. Enable teams to develop effective observability practices, including monitoring, logging, metrics and alerting that support proactive service management. Work with teams to define and embed Service Level Indicators (SLIs), Service Level … professionals, helping shape platform strategy, improve service reliability and support the delivery of critical digital services across government. The team is actively investing in observability, service-level management, platform automation, developer experience and cloud engineering. You'll join a culture that values collaboration, continuous learning and the freedom to explore ...

Lead Site Reliability Engineer (SRE Squad Lead)

Hiring Organisation
Inspire People
Location
Darlington, County Durham, England, United Kingdom
Employment Type
Full-Time
Salary
£63,824 - £80,158 per annum, Pro-rata, Inc benefits
diverse engineering community. Design, build and maintain reliable, secure and scalable cloud-based infrastructure using infrastructure-as-code approaches. Enable teams to develop effective observability practices, including monitoring, logging, metrics and alerting that support proactive service management. Work with teams to define and embed Service Level Indicators (SLIs), Service Level … professionals, helping shape platform strategy, improve service reliability and support the delivery of critical digital services across government. The team is actively investing in observability, service-level management, platform automation, developer experience and cloud engineering. You'll join a culture that values collaboration, continuous learning and the freedom to explore ...

Lead Site Reliability Engineer (SRE Squad Lead)

Hiring Organisation
Inspire People
Location
Darlington, County Durham, North East, United Kingdom
Employment Type
Permanent, Part Time, Work From Home
Salary
£80,000
diverse engineering community. Design, build and maintain reliable, secure and scalable cloud-based infrastructure using infrastructure-as-code approaches. Enable teams to develop effective observability practices, including monitoring, logging, metrics and alerting that support proactive service management. Work with teams to define and embed Service Level Indicators (SLIs), Service Level … professionals, helping shape platform strategy, improve service reliability and support the delivery of critical digital services across government. The team is actively investing in observability, service-level management, platform automation, developer experience and cloud engineering. You'll join a culture that values collaboration, continuous learning and the freedom to explore ...

Head of Technology Operations

Hiring Organisation
Jobleads-UK
Location
Halifax, England, United Kingdom
adoption of infrastructure as code (IaC), CI/CD pipelines, and automated testing within platform operations. Champion site reliability engineering (SRE) practices, embedding monitoring, observability, and incident response, and continuously improving performance metrics including application load times, throughput, and error rates. Partner with the Head of Development and the Director … technologies. Deep expertise in Kubernetes, cloud networking, CI/CD pipelines, infrastructure as code, and platform security. Experience leading site reliability engineering (SRE), monitoring, observability, and performance optimisation (load times, application speed, latency management). Experience providing senior‐level escalation support for complex infrastructure, firewall, networking, and systems issues. ITIL ...

Devops Engineer

Hiring Organisation
Jackson Hogg Ltd
Location
Newcastle upon Tyne, Tyne & Wear, United Kingdom
Employment Type
Permanent
Salary
£50000 - £60000/annum
We're looking for a DevOps Engineer to join a growing technology team responsible for building, supporting and evolving critical cloud platforms. This role sits at the heart of a modern AWS environment, helping to ...

Lead Cloud Engineer

Hiring Organisation
Jobleads-UK
Location
Manchester, England, United Kingdom
data orchestration toolsets (e.g., dbt, Apache Airflow), ETL/ELT methodologies, real‐time streaming (e.g., AWS Kinesis, Apache Kafka), Vector databases, and RAG architectures. Observability & FinOps: Experience implementing modern observability tooling (OpenTelemetry) alongside automated cost‐control systems (such as Karpenter, Infracost, OpenCost, or Cloud Custodian). Domain & Sector Experience Regulated ...

Lead Cloud Engineer

Hiring Organisation
Jobleads-UK
Location
Leeds, England, United Kingdom
data orchestration toolsets (e.g., dbt, Apache Airflow), ETL/ELT methodologies, real‐time streaming (e.g., AWS Kinesis, Apache Kafka), Vector databases, and RAG architectures. Observability & FinOps: Experience implementing modern observability tooling (OpenTelemetry) alongside automated cost‐control systems (such as Karpenter, Infracost, OpenCost, or Cloud Custodian). Domain & Sector Experience Regulated ...

Lead AI Production Engineer — Agentic Systems

Hiring Organisation
Jobleads-UK
Location
Newcastle upon Tyne, England, United Kingdom
produce reusable patterns and accelerators that scale beyond engagements. You will own the evaluation harness, orchestration frameworks, and multi‐LLM integration standards while maintaining observability, cost governance, and safety #J-18808-Ljbffr ...

ServiceNow CRM Transformation Lead - Managing Consultant

Hiring Organisation
Capgemini
Location
Manchester, United Kingdom
Employment Type
Full Time
realisation Translate business requirements into workflow-enabled operating models Solution Design & Architecture Design and support implementation of: AI Control Tower (AI lifecycle management, governance, observability) Agentic AI workflows enabling autonomous execution Now Assist/GenAI use cases across workflows Define data, integration, and workflow architectures for AI-enabled ServiceNow solutions ...

Senior Cloud Engineer

Hiring Organisation
Anson Mccade
Location
Manchester, North West, United Kingdom
Employment Type
Permanent
Salary
£70,000
Senior Platform Engineer Deliver and support cloud platform engineering solutions across client environments. Design services with a reliability mindset, using SLIs, SLOs, and observability practices. Implement and maintain Infrastructure as Code using Terraform across environments. Support incident management, problem management, and continuous improvement of production platforms. Contribute to observability solutions … Infrastructure as Code across non-production and production environments. Understanding of SRE principles including SLIs, SLOs, error budgets, resilience, and reliability. Experience with observability and monitoring tools such as Dynatrace or similar. Experience supporting production platforms including incident and problem management. Exposure to AIOps practices and automation for proactive issue ...

AI Software Engineering Associate Director

Hiring Organisation
Jobleads-UK
Location
Newcastle upon Tyne, England, United Kingdom
production-grade agentic systems at enterprise scale: multi-agent orchestration across complex environments, RAG pipelines, policy-based routing, memory management, and programme-level lifecycle observability Define RAG pipeline standards across engagements: establish chunking and embedding strategies, set quality benchmarks, and ensure metric-backed tradeoff decisions are documented and transferable … standard design practice across providers including OpenAI, Anthropic, Vertex AI, and open-source models Own LLMOps at programme scale: eval strategy, prompt governance, observability tooling standards, safety monitoring and cost controls across multiple concurrent systems Lead client engineering engagements at senior level — facilitate architecture design sessions, lead proof-of-concept ...

AI Software Engineering Manager

Hiring Organisation
Jobleads-UK
Location
Newcastle upon Tyne, England, United Kingdom
production‐grade agentic systems at enterprise scale: multi‐agent orchestration across complex environments, RAG pipelines, policy‐based routing, memory management, and programme‐level lifecycle observability Define RAG pipeline standards across engagements: establish chunking and embedding strategies, set quality benchmarks, and ensure metric‐backed tradeoff decisions are documented and transferable … standard design practice across providers including OpenAI, Anthropic, Vertex AI, and open‐source models Own LLMOps at programme scale: eval strategy, prompt governance, observability tooling standards, safety monitoring and cost controls across multiple concurrent systems Lead client engineering engagements at senior level — facilitate architecture design sessions, lead proof‐of‐concept ...

Senior Cloud Platform Engineer (GCP | Kubernetes | DevSecOps)

Hiring Organisation
GCS
Location
Sheffield, South Yorkshire, United Kingdom
Employment Type
Contract
Contract Rate
£600 - £620/day
repeatable, automated deployments. Implement and maintain CI/CD pipelines and GitOps deployment workflows. Manage cloud networking, connectivity and platform security. Implement platform observability including logging, monitoring, metrics and distributed tracing. Automate platform provisioning, configuration management and operational tasks. Support deployment and operation of identity platform components and supporting services. … Pipelines GitOps Linux Administration Networking and Load Balancing Service Mesh Technologies (Istio/Envoy) Container Platforms Secrets Management PKI and Certificate Management Workload Identity Observability (Logging, Monitoring and Tracing) Scripting and Automation DevSecOps Site Reliability Engineering (SRE) Performance and Capacity Management Operational Support Global Deployment Strategies Positive can-do attitude ...

Senior / Principal DevOps Engineer

Hiring Organisation
Hays Technology
Location
Ramsbottom, Lancashire, United Kingdom
Employment Type
Contract
Contract Rate
GBP 700,000 - 800,700 Daily
best practices across engineering teams and onboard products onto shared platforms. Build and maintain secure, scalable, and high-performing cloud infrastructure in AWS. Implement observability, monitoring, and operational insights across multiple environments. Improve deployment processes, reduce friction, and enable self-service capabilities for development teams. Support cloud and infrastructure incident … focus on automation. Experience with containerisation and workload orchestration technologies. Scripting and programming experience with tools such as Python and Bash. Strong understanding of observability, reliability, and operational best practices. Knowledge of information security principles and experience embedding security throughout the software delivery lifecycle. If you're interested in this ...

SRE Managing Consultant - Cloud Operating Model

Hiring Organisation
Capgemini
Location
Manchester, United Kingdom
Employment Type
Full Time
Budgets : Establish service measures and targets (SLIs/SLOs) and introduce Error Budgets to enable data-driven trade-offs between reliability and delivery velocity. Observability & Operational Insight: Shape observability approaches (metrics/logs/traces) and operational monitoring models that make reliability risks visible and actionable, improving operational decision-making. … large‐scale delivery contexts; associate‐level certifications are desirable but not mandatory. Design, establish, and evolve SRE‐led centres of excellence (e.g. Reliability, Observability, or Operational Excellence), setting enterprise‐level standards for SLIs/SLOs, incident management, observability, and continuous improvement across cloud and hybrid platforms. Exposure to modern observability ...

Site Reliability Engineer

Hiring Organisation
17918
Location
Manchester, Lancashire, United Kingdom
daily and process more than 1.5 million bets per hour at peak. Job Description As a Site Reliability Engineer, you will enhance system reliability, observability and performance through a strong engineering approach and assist with incident resolution and best practices. You will have strong software engineering skills, approaching system reliability … observability as a software problem protecting, providing for, and progressing the performance and availability of our critical systems. Using your engineering expertise, you will implement solutions that enhance reliability, including service instrumentation with OpenTelemetry and improved logging practices. You will leverage AI tools and LLM platforms in your daily work ...

DevOps Engineer

Hiring Organisation
Opus Recruitment Solutions Ltd
Location
Leeds, West Yorkshire, England, United Kingdom
Employment Type
Contractor
Contract Rate
£400 - £450 per day
InsideIR35 | Hybrid 1 day onsite in Leeds | 6 month Initial contract DevOps Engineer to work on the development of an large scale observability platform, experienced in cloud techologies, infrastracture as code, building and operating distributed systems at scale, proficient in devops engineering, reviewing, writing and testing code, working with source … control mechanisms and and deploying infrastructure. Key experience we are looking for: Previous experience building and supporting large-scale AWS observability and monitoring platforms. Strong Python development background with experience creating automation and engineering tooling. Hands-on Kubernetes (K8s) experience deploying, managing and troubleshooting containerised workloads. Experience using Grafana ...

DevOps Engineer

Hiring Organisation
Oscar Associates (UK) Limited
Location
Manchester, North West, United Kingdom
Employment Type
Permanent
Salary
£70,000
scalable, reliable and cost-efficient as it moves into full production. Working closely with engineering teams, you'll drive automation, improve deployment pipelines, strengthen observability and ensure the platform performs under high-volume, real-time workloads. This is a hands-on position with genuine ownership and plenty of opportunity … enhancing CI/CD pipelines with blue/green deployments and automated rollback Driving platform reliability, resilience and scalability Developing monitoring, alerting and observability across the environment Managing cloud costs and implementing best FinOps practices Participating in a small production on-call rota Technology AWS ECS Fargate Terraform Aurora ...

Vice President, DevOps Production Services

Hiring Organisation
Jobleads-UK
Location
Manchester, England, United Kingdom
enterprise applications and ensure platform stability, resiliency, and availability. Monitor application health, system performance, batch jobs, interfaces, and alerts using enterprise monitoring and observability tools. Investigate, troubleshoot, and resolve production incidents within defined SLAs. Perform root cause analysis (RCA) for recurring issues and drive permanent fixes. Analyze production logs, identify … Cloud experience preferred. Knowledge of automation/scripting using Python, Shell, or PowerShell. Exposure to DevOps/SRE practices, CI/CD pipelines, and observability tooling. Strong communication skills with the ability to provide concise incident and executive status updates. #J-18808-Ljbffr ...

Software Engineer

Hiring Organisation
RWS
Location
Sheffield, United Kingdom
Employment Type
Full Time
security, reliability, and operation of the services you build, with a DevSecOps approach throughout Improving engineering practices including CI/CD pipelines, automated testing, observability, security scanning, and deployment workflows Collaborating closely with product managers, designers, domain experts, and other engineers to deliver meaningful outcomes Mentoring and supporting engineers across … DevSecOps mindset with experience owning the security and operation of services you build Understanding of modern delivery practices: CI/CD, automated testing, observability, and production ownership Ability to work across the full development lifecycle, from early design through to deployment and operations Clear communication skills and a collaborative working ...

Site Reliability Engineer (H/F)

Hiring Organisation
17918
Location
Manchester, Lancashire, United Kingdom
million requests daily and process more than 1.5 million bets per hour at peak. As a Site Reliability Engineer, you will enhance system reliability, observability and performance through a strong engineering approach and assist with incident resolution and best practices. You will have strong software engineering skills, approaching system reliability … observability as a software problem protecting, providing for, and progressing the performance and availability of our critical systems. Using your engineering expertise, you will implement solutions that enhance reliability, including service instrumentation with OpenTelemetry and improved logging practices. You will leverage AI tools and LLM platforms in your daily work ...

DevOps Engineer

Hiring Organisation
Fruition Group
Location
Leeds, West Yorkshire, Yorkshire, United Kingdom
Employment Type
Contract
Contract: Inside IR35 We're seeking an experienced Senior DevOps Engineer to join a small, highly skilled engineering team delivering a large-scale enterprise observability platform as they move away from Splunk This is an opportunity to work on a critical cloud platform supporting the migration of numerous services onto … modern monitoring and logging solution. What you'll be doing * Support and enhance a large-scale observability platform. * Help engineering teams onboard and migrate their services. * Build and maintain dashboards, log pipelines and alerting. * Develop and manage cloud infrastructure using Terraform across Azure and AWS. * Produce technical documentation and operational ...

Digital Senior Full Stack Engineer

Hiring Organisation
Leeds Building Society
Location
Leeds, West Yorkshire, Yorkshire, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£75,000
services. You'll lead complex technical delivery, champion modern engineering practices and help shape high-quality solutions through clean architecture, automation, CI/CD, observability and secure-by-default development. Just as importantly, you'll coach and mentor other engineers, raise standards across the squad and define ways of working. … leading code/design reviews; uplifting test automation and quality gates. Ability to influence stakeholders across Product, Architecture, InfoSec, Risk and Operations; governance experience. Observability experience: metrics, logs, traces; operational ownership of services. Experience of supporting UI/UX Design would be beneficial And in return ...

Senior Vice President, Full-Stack Engineer

Hiring Organisation
Jobleads-UK
Location
Manchester, England, United Kingdom
underlying workflow engine (e.g., Camunda) to enable extensibility, portability, and enterprise-scale orchestration. Drive delivery excellence across workflow and decisioning platforms, embedding observability, resilience, auditability, and performance at scale, while establishing engineering standards across CI/CD, testing, security, and data architecture. Qualifications Bachelor’s or Master’s degree … platform design and optimisation. Proven track record of delivering production-grade platforms, embedding engineering excellence across test automation, CI/CD, observability, resilience, and traceability while driving continuous improvement of SDLC practices at scale. Hands-on technical leader who can actively contribute to solution design and critical builds, while defining ...

DevSecOps Capability Manager

Hiring Organisation
WRK DIGITAL LTD
Location
Skipton, North Yorkshire, Yorkshire, United Kingdom
Employment Type
Permanent
improvement Strategy, Governance & Technical Direction Set DevSecOps strategy across pipelines and security automation Establish governance for CI/CD, IaC, and cloud delivery Define observability standards (SLOs, tracing, dashboards) Embed security into pipelines (SAST, SCA, DAST, secrets, IaC scanning) Govern "Golden Path" templates and adoption Operational Oversight & Risk Management Oversee …/CD, DevSecOps, and security integration Strong cloud, containerisation, and IaC knowledge Proven ability to improve DORA and engineering performance metrics Experience with observability and monitoring frameworks Strong background in security tooling (SAST, SCA, DAST, scanning tools) Solid understanding of cloud security, IAM, and zero-trust principles Experience working ...