25 of 25 Permanent Observability Jobs in Cheshire

Senior Software Engineer

Hiring Organisation
Hackajob Ltd
Location
Knutsford, Cheshire, North West, United Kingdom
Employment Type
Permanent
Practical experience with microservices and API design, with a clear understanding of service boundaries, integration contracts, and non-functional requirements such as resilience, scalability, observability, and failure handling. Experience working with event-streaming or messaging platforms (e.g. Kafka or equivalent), including concepts such as topics, partitions, schemas, consumer groups ...

Service Reliability Engineer

Hiring Organisation
Barclays
Location
Knutsford, Cheshire, United Kingdom
Salary
£ 70 K
engineering teams to maintain the smooth operation of platform services.You will collaborate with Development, Infrastructure, and Platform teams to troubleshoot issues, improve monitoring and observability, support platform releases, and identify opportunities for automation and operational improvement. You will also contribute to incident and problem management activities, helping to resolve production ...

Service Reliability Engineer

Location
Knutsford, England, United Kingdom
teams to maintain the smooth operation of platform services. You will collaborate with Development, Infrastructure, and Platform teams to troubleshoot issues, improve monitoring and observability, support platform releases, and identify opportunities for automation and operational improvement. You will also contribute to incident and problem management activities, helping to resolve production ...

Lead DevOps Engineer

Hiring Organisation
Pathfinder Business Solutions Ltd
Location
Chester, Cheshire, North West, United Kingdom
Employment Type
Permanent
Salary
£80,000
teams on resilient, cloud-native services. Youll lead improvements across automation, GitOps, infrastructure as code and CI/CD, reducing manual effort while improving observability, SLOs, security, resilience and reliability. Youll also act as a senior escalation point for complex production issues and participate in the shared on-call rota. ...

Lead DevOps Engineer

Hiring Organisation
Shortlist Recruitment
Location
Chester, Cheshire, UK
Employment Type
Full-time
deployment solutions using IaC and GitOps practicesBuild and maintain CI/CD pipelines to enable development teams to deploy applications quickly and reliablyImprove monitoring, observability and incident response to enhance system reliability and resilienceContribute to the ongoing replatforming of services towards AWS and Kubernetes, identifying opportunities to automate, improve performance ...

Software Engineer

Hiring Organisation
Hackajob Ltd
Location
Knutsford, Cheshire, North West, United Kingdom
Employment Type
Permanent
Practical experience with microservices and API design, with a clear understanding of service boundaries, integration contracts, and non-functional requirements such as resilience, scalability, observability, and failure handling. Experience working with event-streaming or messaging platforms (e.g. Kafka or equivalent), including concepts such as topics, partitions, schemas, consumer groups ...

Head of Cyber, Platforms & IT

Location
Chester, England, United Kingdom
enforce secure‐by‐design principles, including cybersecurity standards, cloud architecture guardrails and operational controls Lead DevOps and Site Reliability Engineering (SRE) maturity, embedding monitoring, observability, automated testing and structured incident response Drive adoption of automation and AI‐enabled tooling to improve anomaly detection, incident management, vulnerability management and operational efficiency ...

MongoDB Site Reliability Engineer

Location
Knutsford, England, United Kingdom
paced environment, your role will be essential to ensuring our infrastructure remains resilient, secure, and scalable. You’ll work on automating operations, enhancing system observability, and driving continuous improvements that reduce downtime and improve efficiency. If you’re motivated by solving, multi-layered problems and building systems that perform reliably ...

AI Platform Engineer

Location
Knutsford, England, United Kingdom
agentic workflows across the software delivery lifecycle. OpenShift Experience building or operating platforms using OpenShift or a similar enterprise Kubernetes distribution. Agent security, observability, and evaluation: Knowledge of AI guardrails, evaluation frameworks, identity, access control, monitoring and operational assurance. Knowledge of agent evaluation, guardrails, identity, access control, monitoring and operational ...

AI Security Operations (SecOps) Specialist

Location
Macclesfield, England, United Kingdom
fail-safe behavior; drive secure, supportable integration of agents with enterprise platforms and Cybersecurity services via governed APIs and service identities. AI Security Observability: Build telemetry strategies that reconstruct agent intent and actions—including prompts and instructions where policy permits—tool calls, identities, memory updates, policy decisions, outputs, errors ...

Cloud and Infrastructure Operations Manager (AWS)

Hiring Organisation
Radius Payment Solutions
Location
Chester, Cheshire, United Kingdom
Salary
£ 70 K
architecture, cloud networking and security best practices.Experience managing highly available, business-critical infrastructure services.Strong knowledge of Infrastructure as Code and automation tooling.Experience implementing monitoring, observability and operational tooling.Experience managing major incidents and operational service delivery.Strong understanding of infrastructure security principles and governance.Excellent stakeholder management and communication skills.Highly DesirableExposure to Azure ...

Cloud and Infrastructure Operations Manager (AWS)

Hiring Organisation
Radius Payment Solutions
Location
Chester, Cheshire, UK
Employment Type
Full-time
security best practices. Experience managing highly available, business-critical infrastructure services. Strong knowledge of Infrastructure as Code and automation tooling. Experience implementing monitoring, observability and operational tooling. Experience managing major incidents and operational service delivery. Strong understanding of infrastructure security principles and governance. Excellent stakeholder management and communication skills. Highly ...

Lead Developer

Location
Winsford, England, United Kingdom
technical environment. Clear communicator who can bridge business and engineering. Bonus Points Experience scaling a SaaS platform . Terraform or Infrastructure as Code. Observability tooling (monitoring, logging, tracing). Performance tuning at scale. Experience hiring or growing engineering teams. #J-18808-Ljbffr ...

Engineering Manager - Stream Aligned

Location
Knutsford, England, United Kingdom
that owns delivery end to end, from discovery through to production operation. Strong understanding of modern software delivery practices, including CI/CD, testing, observability, deployability, and operational readiness. Evidence of improving delivery flow, DORA metrics, release confidence, incident response, or sustainable engineering practices. Experience partnering closely with Product Managers ...

Senior SRE - Cloud Reliability & Automation

Location
Knutsford, England, United Kingdom
embedding SRE practices and maturity across diverse stakeholder groups. You will apply advanced programming, automation, and data‐driven approaches to reduce incident impact, improve observability, and accelerate delivery. #J-18808-Ljbffr ...

Network Splunk Developer

Location
Warrington, England, United Kingdom
days per week We are looking for an experienced Network Splunk Developer to join a large enterprise network environment, supporting telemetry, monitoring and observability across a complex infrastructure estate. Key Responsibilities Collect, normalise and onboard network device telemetry, metrics, logs and events into Splunk. Develop and maintain Splunk dashboards, searches ...

Senior Service Reliability Engineer

Location
Knutsford, England, United Kingdom
reliability and customer experience expectations. Working across Engineering, Infrastructure, Security, and Product teams, you will identify opportunities to automate manual processes, improve monitoring and observability, strengthen resilience, and reduce operational risk. You will also contribute to service reviews, support governance and control activities, and provide clear communication to stakeholders during … reliability for critical business services. Cloud, Platform & Engineering Expertise –Strong hands-on knowledge of AWS cloud technologies, microservices, APIs, containerized platforms (OpenShift/Kubernetes), observability tooling, automation, CI/CD, Infrastructure as Code, and modern platform engineering practices. Senior Stakeholder & Operational Leadership–Proven ability to lead critical incidents, drive root ...

Senior Service Reliability Engineer

Hiring Organisation
Barclays
Location
Knutsford, Cheshire, United Kingdom
Salary
£ 70 K
meet reliability and customer experience expectations.Working across Engineering, Infrastructure, Security, and Product teams, you will identify opportunities to automate manual processes, improve monitoring and observability, strengthen resilience, and reduce operational risk. You will also contribute to service reviews, support governance and control activities, and provide clear communication to stakeholders during … service reliability for critical business services.Cloud, Platform & Engineering Expertise –Strong hands-on knowledge of AWS cloud technologies, microservices, APIs, containerized platforms (OpenShift/Kubernetes), observability tooling, automation, CI/CD, Infrastructure as Code, and modern platform engineering practices.Senior Stakeholder & Operational Leadership–Proven ability to lead critical incidents, drive root cause ...

Platform Reliability Engineer: Cloud & Automation

Location
Knutsford, England, United Kingdom
support queries, and collaborating with engineering teams to sustain smooth operation. You will work with Development, Infrastructure and Platform teams to enhance monitoring, observability, #J-18808-Ljbffr ...

Network Telemetry & Splunk Engineer (Hybrid)

Location
Warrington, England, United Kingdom
Limited is seeking an experienced Network Splunk Developer to join a large enterprise network environment in Chester. The role focuses on telemetry, monitoring and observability across a complex infrastructure, with hybrid work and a £550 daily rate on a 12-month contract. You will collect, normalise and onboard telemetry, develop ...

Senior Site Reliability Engineer

Location
Knutsford, England, United Kingdom
ability to articulate these relationships through narrative, diagrams, and documentation. Keen interest in researching, evaluating and directly engaging with technologies to improve predictability, observability and performance, combined with enthusiasm for teaching others and lifelong learning. Some other highly valued skills include: A deep understanding of systems engineering, including operating systems … tools, and infrastructure‐as‐code. Interest and experience in innovative uses for artificial intelligence, solving technology problems faster and at greater scale. Experience with observability tools and techniques for instrumentation, gathering information and extending the boundaries of the known. You may be assessed on the key critical skills relevant ...

Senior QA Engineer

Hiring Organisation
SRG
Location
Warrington, Cheshire, United Kingdom
Employment Type
Full-Time
Salary
£45,000 - £50,000 per annum
both fast and reliable. You'll help move quality earlier into the process (shift-left), while also using real production insights to improve decisions (observability-led quality). There's strong scope to influence how QA operates within the team, from testing strategy through to continuous improvement. What … considered early in design and development Leading exploratory testing to uncover issues real users might experience Using monitoring, metrics, and logs to drive observability-led quality and improve production outcomes Identifying risks early and helping the team make informed decisions Improving QA processes, standards, and ways of working across ...

Network Automation Developer

Location
Chester, Cheshire, United Kingdom
operational and network infrastructure tasks. Work with REST APIs and JSON to integrate systems and exchange data. Assist with monitoring, alerting and network observability solutions. Troubleshoot automation and infrastructure issues. Follow established engineering standards, designs and operational processes. Work alongside senior automation engineers on defined development tasks and enhancements. Support … wireless and basic SD-WAN. Experience with REST APIs, JSON and Git . Basic Linux knowledge and troubleshooting experience. Familiarity with monitoring, telemetry and observability concepts. Good communication and collaboration skills. A willingness to learn and develop further in network automation. Experience with Django/Flask, Infrastructure as Code ...

Network Automation Developer

Hiring Organisation
MCGREGOR BOYALL ASSOCIATES LIMITED
Location
Chester, Cheshire, UK
operational and network infrastructure tasks. Work with REST APIs and JSON to integrate systems and exchange data. Assist with monitoring, alerting and network observability solutions. Troubleshoot automation and infrastructure issues. Follow established engineering standards, designs and operational processes. Work alongside senior automation engineers on defined development tasks and enhancements. Support … wireless and basic SD-WAN. Experience with REST APIs, JSON and Git . Basic Linux knowledge and troubleshooting experience. Familiarity with monitoring, telemetry and observability concepts. Good communication and collaboration skills. A willingness to learn and develop further in network automation. xehkeey Experience with Django/Flask, Infrastructure as Code ...

Service Engineer

Hiring Organisation
Hackajob Ltd
Location
Knutsford, Cheshire, North West, United Kingdom
Employment Type
Permanent
line with ITIL processes, and working with engineering teams to improve application reliability, performance, and operational resilience. You will also leverage monitoring and observability tools to proactively identify issues and minimise service disruption. To be successful as an Application Support Engineer, you should have: Strong AWS knowledge from an application … such as GitLab pipelines, with an understanding of release and deployment practices Experience using monitoring and automation tools such as Kibana, AppDynamics and other observability platforms to support service performance and operational excellence You may be assessed on the key critical skills relevant for success in the role, such ...