251 to 275 of 338 Observability Jobs in the North of England

Senior Software Engineer- Python

Location
Manchester, England, United Kingdom
specialists (traders, data scientists, platform) without needing to be an expert in their field on day one Raising the bar across software (typing, testing, observability, tooling, and thoughtful refactors), tooling, and ways of working What you'll do Design, build and operate the Python backend services behind Asset Backed Trading … modelling, optimisation. Event driven or serverless architectures at meaningful scale Infrastructure as Code (CDK, CloudFormation, or Terraform) Experience with Pydantic, strict typing, and mypy Observability in production (CloudWatch, Grafana) Exposure to energy, trading, optimisation, or other constraint heavy domains Working in a monorepo with shared libraries and multiple deployable services ...

Vice President, Full-Stack Engineer

Hiring Organisation
Hackajob Ltd
Location
Manchester, North West, United Kingdom
Employment Type
Permanent
engineering teams; set clear objectives, coach talent, and foster succession planning. Own end-to-end delivery for critical software: requirements, architecture, implementation, testing, deployment, observability, and reliability. Raise engineering excellence and resilience: best practices and automation across code, testing, microservices/APIs, performance, and infrastructure; secure-by-design with threat … scalable, observable, testable systems; strong API design. Strong DevOps practices: CI/CD (e.g., GitLab), automated testing (JUnit/Spock), code reviews, telemetry/observability (Splunk, AppDynamics), containers (Docker), and cloud. Hands-on AI development using modern tools and IDEs (e.g., Windsurf) and experience integrating AI into product workflows. Excellent ...

Vice President, Full-Stack Engineer Opportunities

Hiring Organisation
Hackajob Ltd
Location
Manchester, North West, United Kingdom
Employment Type
Permanent
engineering teams; set clear objectives, coach talent, and foster succession planning.. Own end-to-end delivery for critical software: requirements, architecture, implementation, testing, deployment, observability, and reliability. Raise engineering excellence and resilience: best practices and automation across code, testing, microservices/APIs, performance, and infrastructure; secure-by-design with threat … scalable, observable, testable systems; strong API design. Strong DevOps practices: CI/CD (e.g., GitLab), automated testing (JUnit/Spock), code reviews, telemetry/observability (Splunk, AppDynamics), containers (Docker), and cloud Hands-on AI development using modern tools and IDEs (e.g., Windsurf) and experience integrating AI into product workflows Excellent ...

Senior Software Engineer - Backend

Location
Manchester, England, United Kingdom
caching and data-access strategies Building reliable transactional workflows Developing applications and services within AWS Deploying and operating containerised applications Improving monitoring, alerting and observability Exploring AI-assisted engineering and modern development tooling At Senior level, you’ll be expected to understand the wider system rather than only the individual … traffic or data volumes Performance optimisation, caching and latency reduction Designing for resilience and failure Containerisation and orchestration technologies such as Docker and Kubernetes Observability, monitoring and operating production systems Automated testing and modern engineering practices We don’t expect candidates to have worked with every technology in our stack. ...

Oracle OCI Lead Engineer

Location
Leeds, England, United Kingdom
with industry standards. They will need to be able to design, implement and maintain OCI infrastructure, specifically platform services like identity and access management, observability, and core infrastructure, including virtual machines, storage solutions and networking components. Responsibilities include technical leadership, architectural reviews, platform support and mentoring junior engineers. Responsibilities include … SAML federation, Cloud Guard, Vault, and KMS. Network expertise – dynamic routing gateways, transit routing, domain name services, IPSec tunnels, remote peering connections, and FastConnect. Observability and monitoring – logging, monitoring, alarms, events, notifications, and external third-party feeds (e.g. Splunk). Compute, database, and storage knowledge – compute instances, patching, hardening, functions ...

Network Splunk Developer

Hiring Organisation
McGregor Boyall Associates Limited
Location
Chester, Cheshire, United Kingdom
Employment Type
Contract
Contract Rate
GBP Annual
days per week We are looking for an experienced Network Splunk Developer to join a large enterprise network environment, supporting telemetry, monitoring and observability across a complex infrastructure estate click apply for full job details ...

Network Telemetry & Splunk Engineer (Hybrid)

Location
Warrington, England, United Kingdom
Limited is seeking an experienced Network Splunk Developer to join a large enterprise network environment in Chester. The role focuses on telemetry, monitoring and observability across a complex infrastructure, with hybrid work and a £550 daily rate on a 12-month contract. You will collect, normalise and onboard telemetry, develop ...

Enterprise Ecommerce Solution Architect

Location
West Yorkshire, England, United Kingdom
senior leadership to drive digital transformation. Responsibilities include translating business needs into robust architectures, governing major initiatives, and reviewing designs for scalability, resilience and observability across complex platforms. #J-18808-Ljbffr ...

Head of Production Reliability & Operations

Location
Sheffield, England, United Kingdom
incidents, and build a scalable team to move from firefighting to proactive reliability engineering. The role demands hands-on leadership, setting standards, and driving observability, incident management, and on-call practices. You will partner with engineering and product to ensure uptime, performance, and predictable releases. #J-18808-Ljbffr ...

Hybrid Enterprise Integration Product Manager

Location
Sheffield, England, United Kingdom
Sheffield is seeking an Enterprise Integration Product Manager for a hybrid, contractor role. You will own a multi-quarter roadmap across middleware, messaging, and observability, driving platform standards and governance, while aligning with regulatory and regional constraints. The role requires deep IBM MQ knowledge, experience with ACE/ ...

Pre-Sales Solutions Engineer — Public Sector (UK)

Location
Manchester, England, United Kingdom
facing role engages with business and technical leaders to align Cisco products with public sector needs. You will craft multi-architecture solutions across Networking, Observability, Security, Collaboration, Compute and AI, deliver technical demonstrations, and support BOM and RFP/RFI processes to drive customer outcomes. #J-18808-Ljbffr ...

Production Reliability Lead

Location
Leeds, England, United Kingdom
lead incidents, build a scalable team, and move us from reactive firefighting to proactive reliability engineering. You’ll own the live systems: ensure stability, observability, and safe releases, while building runbooks, incident playbooks, and a culture of ownership. Collaborate across product, engineering and support to improve MTTR, establish SLIs/ ...

Production Reliability Lead: Scale & Incident Mastery

Location
Newcastle upon Tyne, England, United Kingdom
incidents, build the team, and move us from reactive firefighting to proactive reliability engineering. This hands-on role requires calm communication and a strong observability mindset for reliable systems. #J-18808-Ljbffr ...

Incident Commander - 24/7 Digital Reliability Lead

Location
Leeds, England, United Kingdom
roles, and coordinating cross-functional teams to drive rapid resolution for high-priority incidents. The ideal candidate has 5+ years in IT operations or observability with strong analytics, alerting, and dashboard skills, plus a background in gaming or digital commerce. #J-18808-Ljbffr ...

M365 Copilot Specialist — Enterprise Incident & IAM Expert

Location
Sheffield, England, United Kingdom
across M365 services. You will handle complex escalations in the Admin Centre, Entra, Conditional Access, SharePoint/OneDrive permissions, and Teams, while contributing to observability and monitoring efforts. The role requires 5–8+ years in M365 support, strong Copilot skills, PowerShell and Graph API for troubleshooting, and a proven track ...

Full Stack Engineer - Specialist

Hiring Organisation
Hackajob Ltd
Location
Manchester, North West, United Kingdom
Employment Type
Permanent
stack with a strong Java backend emphasis, working within agile teams that own the full product lifecycle from design and build through to deployment, observability, and iteration. We are particularly interested in candidates holding aMasters in Computer Science or Artificial Intelligence, ideally with some industrial placement or internship experience … ways: Design, build, and maintain backend services, batches and APIs, contributing to UI components as needed. Own end-to-end delivery: implementation, testing, deployment, observability, and reliability. Write clean, well-tested code; participate in code reviews and continuous improvement. Collaborate with product, design, and operations to translate business needs into ...

Backend Engineer

Location
Manchester, England, United Kingdom
using containers and modern cloud/platform technologies Implement and maintain CI/CD pipelines for automated testing, deployment and release management Establish strong observability, monitoring, logging and alerting capabilities Contribute to technical design reviews, engineering standards and best practices Work closely with AI Engineers, Data Scientists, Enterprise Architects, Security … OAuth, SAML and SSO API Gateways and enterprise integrations Cloud-native development CI/CD and DevOps practices Distributed systems and high-availability architectures Observability, monitoring and operational tooling Enterprise Integration Platforms AI/Agent Orchestration Platforms Why Join? This is an opportunity to work on a next-generation enterprise ...

Backend Engineer 1861

Location
Leeds, England, United Kingdom
using containers and modern cloud/platform technologies Implement and maintain CI/CD pipelines for automated testing, deployment and release management Establish strong observability, monitoring, logging and alerting capabilities Contribute to technical design reviews, engineering standards and best practices Work closely with AI Engineers, Data Scientists, Enterprise Architects, Security … OAuth, SAML and SSO API Gateways and enterprise integrations Cloud-native development CI/CD and DevOps practices Distributed systems and high-availability architectures Observability, monitoring and operational tooling Enterprise Integration Platforms AI/Agent Orchestration Platforms Why Join? This is an opportunity to work on a next-generation enterprise ...

Senior Software Developer – Scala

Location
Newcastle upon Tyne, England, United Kingdom
frontend engineers, QA, product owners, solution designers, and other backend developers to deliver high-quality product increments. Support production stability by investigating issues, improving observability, and continuously reducing technical debt. Required Skills and Experience: Extensive professional backend development experience, ideally in enterprise, SaaS, or cloud-based product environments. Strong hands … problems. Nice to Have: Experience with Squeryl, Doobie, or similar Scala data access libraries. Familiarity with Grafana or Kibana dashboards, alerting, logging, and production observability practices. Experience with large distributed systems, horizontal scaling, resilient service design, or high-throughput Play/Pekko applications. Previous experience in Payroll ...

Senior Specialist, Production Services Application Support Analyst

Hiring Organisation
Hackajob Ltd
Location
Manchester, North West, United Kingdom
Employment Type
Permanent
Lead technical coordination during major incidents, helping drive rapid diagnosis, recovery, stakeholder communication, and root cause remediation. Drive continuous improvement initiatives focused on automation, observability, service reliability, operational efficiency, and reduction of manual processes. Evaluate production risks associated with application releases, infrastructure changes, and platform enhancements to ensure safe … resolve complex technical issues under pressure. Understanding of enterprise application architecture, distributed systems, cloud technologies, middleware, databases, and infrastructure components. Experience with monitoring, observability, automation, and operational tooling used to support highly available production platforms. Analytical and problem-solving skills with the ability to identify root causes and implement sustainable ...

Vice President, DevOps Production Services

Hiring Organisation
Hackajob Ltd
Location
Manchester, North West, United Kingdom
Employment Type
Permanent
enterprise applications and ensure platform stability, resiliency, and availability. Monitor application health, system performance, batch jobs, interfaces, and alerts using enterprise monitoring and observability tools. Investigate, troubleshoot, and resolve production incidents within defined SLAs. Perform root cause analysis (RCA) for recurring issues and drive permanent fixes. Analyze production logs, identify … Cloud experience preferred. Knowledge of automation/scripting using Python, Shell, or PowerShell. Exposure to DevOps/SRE practices, CI/CD pipelines, and observability tooling. Strong communication skills with the ability to provide concise incident and executive status updates. At BNY, our culture allows us to run our company ...

Lead Software Engineer

Location
Manchester, England, United Kingdom
high standards through code reviews, mentoring engineers, and contributing to CI/CD and testing best practices Applying site reliability engineering principles to improve observability, resilience, and incident response Monitoring, profiling, and optimising application performance, including capacity planning to meet SLAs Identifying technical debt and recommending improvements to systems, processes … containerisation Strong debugging, performance optimisation, and problem-solving skills across distributed systems Experience mentoring engineers and contributing to technical leadership within teams Knowledge of observability, monitoring, and site reliability engineering practices Experience with Java and/or telecoms technologies (e.g. VoIP, WebRTC) is highly beneficial What do we offer ...

Technical Lead - GCP

Location
Sheffield, England, United Kingdom
operational resilience. Lead solution design for new integrations and platform capabilities, balancing pace with long-term maintainability. Set engineering standards (coding, testing, documentation, observability, security) and ensure consistent adoption across the team. Provide hands‐on contribution where needed (critical ETLs, framework components, design spikes, performance fixes). ETL orchestration & data … Implement end‐to‐end data lineage (source → platform → vendor), including metadata capture and a visual representation that supports auditability and faster incident resolution. Enhance observability across pipelines and services: logging, metrics, alerting, and dashboards; drive measurable improvements in MTTR and change failure rate. Stakeholder management & ways of working Partner with ...

Agentic Platform Engineer · Manchester, UK ·

Location
Manchester, England, United Kingdom
Protocol (MCP). Build production-grade agent services using Python, cloud-native architectures, event-driven design, automation and Infrastructure as Code. Implement robust evaluation, observability and continuous improvement capabilities, including testing, tracing, telemetry and performance optimisation. Embed security, governance and responsible AI principles through least-privilege access, policy enforcement, auditability … workflow state, retrieval-augmented generation and human-in-the-loop patterns. Experience building secure integrations with enterprise APIs, repositories, cloud services, CI/CD, observability or ITSM platforms. Experience creating evaluation frameworks for agent quality, task completion, safety, reliability, latency and cost. Strong understanding of agent security, including workload identity ...

Senior Engineering Manager - 9-10 month FTC

Location
Manchester, England, United Kingdom
practices, including test‐driven development and automated testing, helping teams build quality into the development process. Support strong operational ownership through CI/CD, observability, production support and the principle that teams own the systems they build and run. Help teams identify and address technical debt sustainably while maintaining appropriate … outcomes and translating these into clear engineering priorities. Strong understanding of modern software engineering practices, including automated testing and TDD principles, CI/CD, observability and production ownership. Experience facilitating technical decisions, bringing the right people and evidence together and constructively challenging thinking when needed. Comfortable balancing feature delivery ...