176 to 200 of 243 Observability Jobs in the North West

Founding Engineer

Location
Manchester, England, United Kingdom
each customer The systems that keep long-running agents dependable in production, including durable memory, scheduled work, sandbox lifecycle, tool execution, recovery and observability The trust layer between an agent and a founder, so work moves from draft to delivered with the right approvals, clear status and safe ways ...

ServiceNow AI & Enterprise Automation Lead - Managing Consultant

Location
Manchester, England, United Kingdom
value* Translate business requirements into AI-enabled workflow solutions**Solution Design & Architecture*** Design and support implementation of:* AI Control Tower (AI lifecycle management, governance, observability)* Agentic AI workflows enabling autonomous execution* Now Assist/GenAI use cases across workflows* Define data, integration, and workflow architectures for AI-enabled ServiceNow solutions ...

Senior Platform Engineer

Hiring Organisation
Anson Mccade
Location
Manchester, North West, United Kingdom
Employment Type
Permanent
Salary
£65,000
Senior Platform Engineer Deliver and support cloud platform engineering solutions across client environments. Design services with a reliability mindset, using SLIs, SLOs, and observability practices. Implement and maintain Infrastructure as Code using Terraform across environments. Support incident management, problem management, and continuous improvement of production platforms. Contribute to observability solutions … Infrastructure as Code across non-production and production environments. Understanding of SRE principles including SLIs, SLOs, error budgets, resilience, and reliability. Experience with observability and monitoring tools such as Dynatrace or similar. Experience supporting production platforms including incident and problem management. Exposure to AIOps practices and automation for proactive issue ...

Senior Software Engineer- Python

Location
Manchester, England, United Kingdom
specialists (traders, data scientists, platform) without needing to be an expert in their field on day one Raising the bar across software (typing, testing, observability, tooling, and thoughtful refactors), tooling, and ways of working What you'll do Design, build and operate the Python backend services behind Asset Backed Trading … modelling, optimisation. Event driven or serverless architectures at meaningful scale Infrastructure as Code (CDK, CloudFormation, or Terraform) Experience with Pydantic, strict typing, and mypy Observability in production (CloudWatch, Grafana) Exposure to energy, trading, optimisation, or other constraint heavy domains Working in a monorepo with shared libraries and multiple deployable services ...

Vice President, Full-Stack Engineer

Hiring Organisation
Hackajob Ltd
Location
Manchester, North West, United Kingdom
Employment Type
Permanent
engineering teams; set clear objectives, coach talent, and foster succession planning. Own end-to-end delivery for critical software: requirements, architecture, implementation, testing, deployment, observability, and reliability. Raise engineering excellence and resilience: best practices and automation across code, testing, microservices/APIs, performance, and infrastructure; secure-by-design with threat … scalable, observable, testable systems; strong API design. Strong DevOps practices: CI/CD (e.g., GitLab), automated testing (JUnit/Spock), code reviews, telemetry/observability (Splunk, AppDynamics), containers (Docker), and cloud. Hands-on AI development using modern tools and IDEs (e.g., Windsurf) and experience integrating AI into product workflows. Excellent ...

Vice President, Full-Stack Engineer Opportunities

Hiring Organisation
Hackajob Ltd
Location
Manchester, North West, United Kingdom
Employment Type
Permanent
engineering teams; set clear objectives, coach talent, and foster succession planning.. Own end-to-end delivery for critical software: requirements, architecture, implementation, testing, deployment, observability, and reliability. Raise engineering excellence and resilience: best practices and automation across code, testing, microservices/APIs, performance, and infrastructure; secure-by-design with threat … scalable, observable, testable systems; strong API design. Strong DevOps practices: CI/CD (e.g., GitLab), automated testing (JUnit/Spock), code reviews, telemetry/observability (Splunk, AppDynamics), containers (Docker), and cloud Hands-on AI development using modern tools and IDEs (e.g., Windsurf) and experience integrating AI into product workflows Excellent ...

Senior Software Engineer - Backend

Location
Manchester, England, United Kingdom
caching and data-access strategies Building reliable transactional workflows Developing applications and services within AWS Deploying and operating containerised applications Improving monitoring, alerting and observability Exploring AI-assisted engineering and modern development tooling At Senior level, you’ll be expected to understand the wider system rather than only the individual … traffic or data volumes Performance optimisation, caching and latency reduction Designing for resilience and failure Containerisation and orchestration technologies such as Docker and Kubernetes Observability, monitoring and operating production systems Automated testing and modern engineering practices We don’t expect candidates to have worked with every technology in our stack. ...

Senior DevOps / Platform Engineer - Autonomous Vulnerability Research (Harness Engineering)

Location
Greater Manchester, England, United Kingdom
reach sanctioned targets. Build CI/CD pipelines with integrated supply-chain security: SBOMs, image signing, artefact provenance, and automated policy gates. Deliver comprehensive observability - logs, metrics, distributed traces, and cost telemetry - across long-running, non-deterministic agent workloads. Build evidence-capture pipelines: immutable audit trails, artefact retention, and reproducible … egress control, network policy, and segmentation in cloud-native environments. Strong CI/CD engineering skills and experience embedding security controls into delivery pipelines. Observability expertise across logging, tracing, and metrics, including designing for auditability and evidence retention. A security-first mindset with the judgement to balance researcher velocity against ...

Senior DevOps / Platform Engineer - Autonomous Vulnerability Research (Harness Engineering)

Hiring Organisation
Lloyds Banking Group
Location
Manchester, UK
Employment Type
Full-time
reach sanctioned targets. Build CI/CD pipelines with integrated supply-chain security: SBOMs, image signing, artefact provenance, and automated policy gates. Deliver comprehensive observability - logs, metrics, distributed traces, and cost telemetry - across long-running, non-deterministic agent workloads. Build evidence-capture pipelines: immutable audit trails, artefact retention, and reproducible … egress control, network policy, and segmentation in cloud-native environments. Strong CI/CD engineering skills and experience embedding security controls into delivery pipelines. Observability expertise across logging, tracing, and metrics, including designing for auditability and evidence retention. A security-first mindset with the judgement to balance researcher velocity against ...

Platform Reliability Engineer: Cloud & Automation

Location
Knutsford, England, United Kingdom
support queries, and collaborating with engineering teams to sustain smooth operation. You will work with Development, Infrastructure and Platform teams to enhance monitoring, observability, #J-18808-Ljbffr ...

Network Telemetry & Splunk Engineer (Hybrid)

Location
Warrington, England, United Kingdom
Limited is seeking an experienced Network Splunk Developer to join a large enterprise network environment in Chester. The role focuses on telemetry, monitoring and observability across a complex infrastructure, with hybrid work and a £550 daily rate on a 12-month contract. You will collect, normalise and onboard telemetry, develop ...

Pre-Sales Solutions Engineer — Public Sector (UK)

Location
Manchester, England, United Kingdom
facing role engages with business and technical leaders to align Cisco products with public sector needs. You will craft multi-architecture solutions across Networking, Observability, Security, Collaboration, Compute and AI, deliver technical demonstrations, and support BOM and RFP/RFI processes to drive customer outcomes. #J-18808-Ljbffr ...

Full Stack Engineer - Specialist

Hiring Organisation
Hackajob Ltd
Location
Manchester, North West, United Kingdom
Employment Type
Permanent
stack with a strong Java backend emphasis, working within agile teams that own the full product lifecycle from design and build through to deployment, observability, and iteration. We are particularly interested in candidates holding aMasters in Computer Science or Artificial Intelligence, ideally with some industrial placement or internship experience … ways: Design, build, and maintain backend services, batches and APIs, contributing to UI components as needed. Own end-to-end delivery: implementation, testing, deployment, observability, and reliability. Write clean, well-tested code; participate in code reviews and continuous improvement. Collaborate with product, design, and operations to translate business needs into ...

Backend Engineer

Location
Manchester, England, United Kingdom
using containers and modern cloud/platform technologies Implement and maintain CI/CD pipelines for automated testing, deployment and release management Establish strong observability, monitoring, logging and alerting capabilities Contribute to technical design reviews, engineering standards and best practices Work closely with AI Engineers, Data Scientists, Enterprise Architects, Security … OAuth, SAML and SSO API Gateways and enterprise integrations Cloud-native development CI/CD and DevOps practices Distributed systems and high-availability architectures Observability, monitoring and operational tooling Enterprise Integration Platforms AI/Agent Orchestration Platforms Why Join? This is an opportunity to work on a next-generation enterprise ...

Site Reliability Engineer (SRE)

Hiring Organisation
Spencer Rose Ltd
Location
Manchester, Lancashire, United Kingdom
Employment Type
Contract
Contract Rate
GBP Daily
operational excellence of cloud-hosted services on Google Cloud Platform. This is a hands-on engineering role spanning SRE practices, production Kubernetes, infrastructure automation, observability, CI/CD, incident response and continuous service improvement. About the role The Senior Site Reliability Engineer will work with Cloud Platform, Software Engineering, Product … shared services. Define and operate service level indicators, service level objectives and error-budget practices that connect technical health to customer impact. Design actionable observability using Dynatrace, including instrumentation, dashboards, distributed tracing, service health views and SLO-based alerting. Build modular, reusable and maintainable Terraform code for secure cloud infrastructure ...

Senior Specialist, Production Services Application Support Analyst

Hiring Organisation
Hackajob Ltd
Location
Manchester, North West, United Kingdom
Employment Type
Permanent
Lead technical coordination during major incidents, helping drive rapid diagnosis, recovery, stakeholder communication, and root cause remediation. Drive continuous improvement initiatives focused on automation, observability, service reliability, operational efficiency, and reduction of manual processes. Evaluate production risks associated with application releases, infrastructure changes, and platform enhancements to ensure safe … resolve complex technical issues under pressure. Understanding of enterprise application architecture, distributed systems, cloud technologies, middleware, databases, and infrastructure components. Experience with monitoring, observability, automation, and operational tooling used to support highly available production platforms. Analytical and problem-solving skills with the ability to identify root causes and implement sustainable ...

Vice President, DevOps Production Services

Hiring Organisation
Hackajob Ltd
Location
Manchester, North West, United Kingdom
Employment Type
Permanent
enterprise applications and ensure platform stability, resiliency, and availability. Monitor application health, system performance, batch jobs, interfaces, and alerts using enterprise monitoring and observability tools. Investigate, troubleshoot, and resolve production incidents within defined SLAs. Perform root cause analysis (RCA) for recurring issues and drive permanent fixes. Analyze production logs, identify … Cloud experience preferred. Knowledge of automation/scripting using Python, Shell, or PowerShell. Exposure to DevOps/SRE practices, CI/CD pipelines, and observability tooling. Strong communication skills with the ability to provide concise incident and executive status updates. Our culture allows us to run our company better ...

Lead Software Engineer

Location
Manchester, England, United Kingdom
high standards through code reviews, mentoring engineers, and contributing to CI/CD and testing best practices Applying site reliability engineering principles to improve observability, resilience, and incident response Monitoring, profiling, and optimising application performance, including capacity planning to meet SLAs Identifying technical debt and recommending improvements to systems, processes … containerisation Strong debugging, performance optimisation, and problem-solving skills across distributed systems Experience mentoring engineers and contributing to technical leadership within teams Knowledge of observability, monitoring, and site reliability engineering practices Experience with Java and/or telecoms technologies (e.g. VoIP, WebRTC) is highly beneficial What do we offer ...

Agentic Platform Engineer · Manchester, UK ·

Location
Manchester, England, United Kingdom
Protocol (MCP). Build production-grade agent services using Python, cloud-native architectures, event-driven design, automation and Infrastructure as Code. Implement robust evaluation, observability and continuous improvement capabilities, including testing, tracing, telemetry and performance optimisation. Embed security, governance and responsible AI principles through least-privilege access, policy enforcement, auditability … workflow state, retrieval-augmented generation and human-in-the-loop patterns. Experience building secure integrations with enterprise APIs, repositories, cloud services, CI/CD, observability or ITSM platforms. Experience creating evaluation frameworks for agent quality, task completion, safety, reliability, latency and cost. Strong understanding of agent security, including workload identity ...

Engineering Manager (Remote - UK)

Hiring Organisation
Reonomy
Location
Manchester, Greater Manchester, United Kingdom
Salary
£ 70 K
building and evolving AWS-native and cloud-based platforms, alongside legacy systemsSet and uphold strong engineering standards across code quality, testing, CI/CD, observability and documentationStay close to technical decisions through design reviews, architecture discussions and hands-on coachingBalance new feature delivery with technical debt, reliability, security and long … distributed systemsProficiency in at least one modern programming languageStrong grasp of system design and software engineering fundamentalsExperience with Infrastructure as Code, CI/CD, observability and secure production systemsAble to communicate technical ideas clearly to both technical and non-technical stakeholdersUnlock your Altus Experience!If you’re looking to advance ...

Senior Engineering Manager - 9-10 month FTC

Location
Manchester, England, United Kingdom
practices, including test‐driven development and automated testing, helping teams build quality into the development process. Support strong operational ownership through CI/CD, observability, production support and the principle that teams own the systems they build and run. Help teams identify and address technical debt sustainably while maintaining appropriate … outcomes and translating these into clear engineering priorities. Strong understanding of modern software engineering practices, including automated testing and TDD principles, CI/CD, observability and production ownership. Experience facilitating technical decisions, bringing the right people and evidence together and constructively challenging thinking when needed. Comfortable balancing feature delivery ...

Data Architect

Hiring Organisation
PA Consulting
Location
Manchester, Greater Manchester, United Kingdom
Salary
£ 70 K
Company DescriptionWe believe in the power of ingenuity to build a positive human future. We challenge where it matters and own the outcome. As strategies, technologies, and innovation collide, we create opportunity from complexity. Our ...

Database Reliability Engineer

Location
Manchester, England, United Kingdom
Cloud Portability: Use CNPG and cloud-native patterns to ensure our database layer remains provider-agnostic, allowing seamless deployment across AWS and GCP Evolve Observability & Monitoring: Build deep, proactive monitoring and alerting for our global database fleet. You will ensure we have the visibility to detect performance regressions and health … Cloud Portability: Use CNPG and cloud-native patterns to ensure our database layer remains provider-agnostic, allowing seamless deployment across AWS and GCP Evolve Observability & Monitoring: Build deep, proactive monitoring and alerting for our global database fleet. You will ensure we have the visibility to detect performance regressions and health ...

Vice President, Production Services Application Support

Location
Manchester, England, United Kingdom
Lead technical coordination during major incidents, helping drive rapid diagnosis, recovery, stakeholder communication, and root cause remediation. Drive continuous improvement initiatives focused on automation, observability, service reliability, operational efficiency, and reduction of manual processes. Evaluate production risks associated with application releases, infrastructure changes, and platform enhancements to ensure safe … Lead technical coordination during major incidents, helping drive rapid diagnosis, recovery, stakeholder communication, and root cause remediation. Drive continuous improvement initiatives focused on automation, observability, service reliability, operational efficiency, and reduction of manual processes. Evaluate production risks associated with application releases, infrastructure changes, and platform enhancements to ensure safe ...

Vice President, Production Services Application Support

Hiring Organisation
Hackajob Ltd
Location
Manchester, North West, United Kingdom
Employment Type
Permanent
Lead technical coordination during major incidents, helping drive rapid diagnosis, recovery, stakeholder communication, and root cause remediation. Drive continuous improvement initiatives focused on automation, observability, service reliability, operational efficiency, and reduction of manual processes. Evaluate production risks associated with application releases, infrastructure changes, and platform enhancements to ensure safe … resolve complex technical issues under pressure. Deep understanding of enterprise application architecture, distributed systems, cloud technologies, middleware, databases, and infrastructure components. Experience with monitoring, observability, automation, and operational tooling used to support highly available production platforms. Strong analytical and problem-solving skills with the ability to identify root causes ...