101 to 125 of 140 Observability Jobs in the South West

Senior Elixir Engineer (Sovereign Territory Surveillance)

Location
Gloucester, England, United Kingdom
telemetry Developing rich LiveView interfaces Working with large geospatial datasets Designing fault‐tolerant distributed services Helping shape technical strategy across the platform Improving deployment, observability and developer tooling You’ll ideally have experience with: Elixir and OTP Phoenix and LiveView Designing distributed systems Building APIs and backend services Distributed relational ...

Principal Software Architect

Location
Bristol, England, United Kingdom
/software ecosystem. Assess the architectural impact of new technologies. Be aware of the usability, performance, reliability, maintainability, testability, security and observability constraints on the software architecture. Prototyping and validating architectural concepts through proof-of-concept implementations. Contribute to future and/or related product definitions with a forward-looking ...

Lead SRE - Chase UK

Hiring Organisation
Hackajob Ltd
Location
South West London, London, United Kingdom
Employment Type
Permanent
possess an interest in the financial sector and focus on addressing our customer needs. We work in teams focused on improving the reliability, resilience, observability, and operability of customer-facing digital banking services. We build automation, define measurable reliability practices, reduce operational friction, and partner with engineering teams to ensure … knowledge of microservice infrastructure components, including service discovery, ingress, networking, and load balancing. Experience with Kubernetes. Experience with cloud computing services. Familiarity with common observability and reliability toolchains such as Grafana, Prometheus, Elasticsearch, Kibana, or Jaeger. Ability to use AI-assisted engineering tools responsibly, including validating outputs, understanding failure modes ...

Software Engineer

Hiring Organisation
IBM SIXworks Limited
Location
Taunton, Somerset, United Kingdom
Employment Type
Full-Time
Salary
Competitive salary
Work on Technology That Protects What Matters At SiXworks , we build secure digital solutions that support Defence and National Security missions . Our teams work on complex problems where reliability, security, and speed of innovation ...

Infrastructure Engineering Specialist

Hiring Organisation
Randstad Digital
Location
Bristol, Avon, South West, United Kingdom
Employment Type
Contract
Contract Rate
£650 - £700 per hour
Produce technical documentation (LLDs, SOPs) and deliver knowledge transfer to support teams. Key Requirements Strong, hands-on experience designing and deploying enterprise monitoring and observability suites . Practical experience integrating with major network and firewall management consoles . Deep knowledge of SNMP, Syslog, NetFlow, WMI, and REST APIs . Solid ...

Security Consultant

Hiring Organisation
Appcast
Location
Plymouth, Devon, UK
other highly valued skills may includeExperience in AI governance and risk management frameworksExposure to conversational AI, contact centre technologies, or digital channelsAwareness of AI observability, monitoring, and assurance practicesYou may be assessed on the key critical skills relevant for success in role, such as risk and controls, change and transformation ...

Site Reliability Engineer

Hiring Organisation
Anson Mccade
Location
Gloucester, Gloucestershire, South West, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£65,000
reliability and performance. Automating repetitive operational tasks and reducing manual intervention wherever possible. Monitoring and troubleshooting systems across the application and infrastructure stack. Improving observability and instrumentation to identify issues and measure system performance. Working alongside development and product teams to build scalable and resilient services. Responding to production incidents … including Bash or PowerShell. Cloud platforms such as AWS, Azure or OpenStack. Infrastructure automation and configuration management. CI/CD and deployment tooling. Monitoring, observability and troubleshooting of production systems. Docker, containers and/or microservices. Diagnosing issues across different levels of the technology stack. Working within Agile engineering teams. ...

Lead Software Engineer - Application Owner

Location
Bournemouth, England, United Kingdom
after-action reviews, and closure of follow-up actions. Define and continuously improve production readiness standards, including release safety and rollback strategy, dependency awareness, observability requirements, and operational runbooks. Contribute hands‐on to design and delivery, including system design, code reviews, automation, and complex troubleshooting, with secure‐by‐design … with strong operational accountability, including controls, resiliency and recovery, and remediation tracking. Strong system design fundamentals and cloud‐native operational patterns, including scalability, reliability, observability, and dependency management. Hands‐on experience with Go‐based services and modern CI/CD practices. Experience operating workloads on AWS and Kubernetes ...

Lead Site Reliability Engineer

Hiring Organisation
Hackajob Ltd
Location
South West London, London, United Kingdom
Employment Type
Permanent
undergoing a multi year convergence and modernization journey. You will play a pivotal role in shaping our next generation SRE patterns, reliability frameworks, observability strategy, and performance engineering capabilities across globally distributed systems. This role is ideal for an SRE specialist who thrives in fast paced front office environments, enjoys … Deep knowledge of reliability engineering principles: SLIs/SLOs, real-time telemetry, disaster recovery planning, capacity planning, and performance tuning. Experience designing and implementing observability frameworks for mission critical systems. Proven ability to lead incident response and drive long term remediation. Solid programming skills in Python, Java, or Kotlin, with ...

Senior Platform Engineer

Location
Warminster, England, United Kingdom
Support and maintain existing simulation and training systems, as well as existing deployment and virtualisation tools. Apply SRE practices to improve system reliability, including observability (metrics, logs, tracing), incident response, and root cause analysis. What We Are Looking For: This is not a pure cloud or greenfield platform role. … failures Pragmatic and delivery-focused, with a bias toward keeping systems running. Strong collaborator across engineering disciplines Adopts an SRE mindset, focusing on reliability, observability, and continuous improvement of running systems. Key Technical Proficiencies: Expert working knowledge of Kubernetes, Helm, Teraform, Ansible, and Docker. Understanding of Distributed Systems in production. ...

QA Test Infrastructure Engineer (Contract)

Location
Taunton, England, United Kingdom
Work on Technology That Protects What Matters AtSiXworks, we build secure digital solutions that supportDefence and National Security missions. Our teams work on complex problems where reliability, security, and speed of innovation matter. We’re ...

E2 SAP Basis Engineer E2

Location
Swindon, England, United Kingdom
capabilities across implementation and operations. You’ll contribute to the evolution of our SAP Application Lifecycle Management tooling, helping establish modern approaches to observability, engineering effectiveness and change governance across our SAP platforms. At Nationwide we offer hybrid working wherever possible. More rewarding relationships are supported through our hybrid approach ...

DevOps & Infrastructure Engineer

Location
Gloucester, England, United Kingdom
Security customers, spanning both on-premise environments and cloud-based solutions. You’ll lead hands-on DevOps and infrastructure engineering across CI/CD, observability, infrastructure-as-code and platform automation, helping teams build secure, reliable and scalable services in demanding environments. What you’ll be doing: You’ll lead … cloud-based solutions. Develop and maintain CI/CD pipelines, GitOps workflows and automated deployment approaches using tools such as ArgoCD. Implement and improve observability using Prometheus, Grafana, logging and alerting to support resilient platform operations. Use infrastructure-as-code and platform automation with Helm, Go and Terraform to deliver ...

Principal DevSecOps Engineer

Hiring Organisation
83zero Limited
Location
Bristol, Avon, South West, United Kingdom
Employment Type
Permanent, Work From Home
workflows * Establish secure-by-design engineering practices and enforce security and technical standards * Lead Infrastructure as Code (IaC) practices across teams and environments * Drive observability, monitoring, logging and audit controls * Support incident response, patching, compliance reporting and technical debt remediation * Partner with developers and delivery teams to improve engineering quality … Security & compliance - Trivy, vulnerability management, HashiCorp Vault, cert-manager * Containers & cloud - Docker, AWS EKS, AWS IAM, S3 and network policies * Infrastructure as Code - Terraform * Observability - Grafana, Loki * Automation - Python and Bash * Experience delivering within the UK Government Digital Service (GDS) lifecycle on a public sector engagement Why join ...

Senior Associate, Full-Stack Engineer

Hiring Organisation
Hackajob Ltd
Location
South West London, London, United Kingdom
Employment Type
Permanent
ways: Design, build, and maintain backend services, batches and APIs, contributing to UI components as needed. Own end-to-end delivery: implementation, testing, deployment, observability, and reliability. Write clean, well-tested code; participate in code reviews and continuous improvement. Collaborate with product, design, and operations to translate business needs into … microservices Proficiency in Java with Spring. Experience with CI/CD, automated testing (JUnit/Spock), and containers (Docker). Familiarity with microservices, observability/telemetry (e.g., Splunk, AppDynamics), and cloud deployments. Curiosity to understand the business domain and translate product strategy into technical solutions. How we work: Agile (Scrum ...

Platform Site Reliability Engineer

Location
Gloucester, England, United Kingdom
provision of tooling for our support organisation Apply ITSM frameworks: Incident, Major Incident, Change Management, and service improvement. Maintain and enhance Radiant’s observability stack: Prometheus, Grafana, and custom monitoring integrations Operate and support services in 24x7 production environments, including on-call rotation Contribute to Incident postmortem analyses, root cause … DHCP, VLANs, routing, switching Strong experience with API interrogation Strong experience with infrastructure scripting and automation (Bash, Python, Ansible) Deep understanding of observability principles and tools (Prometheus, Grafana preferred) Strong grasp of ITSM and service operation best practices Excellent communication and mentorship skills Comfortable interfacing with internal stakeholders and external ...

Infrastructure Site Reliability Engineer

Location
Gloucester, England, United Kingdom
resolution and provision of tooling for our support organisation Apply ITSM frameworks: Incident, Major Incident, Change Management, and service improvement. Maintain and enhance ’s observability stack: Prometheus, Grafana, and custom monitoring integrations Operate and support services in 24x7 production environments, including on-call rotation Contribute to Incident postmortem analyses, root … Strong networking fundamentals: TCP/IP, DNS, DHCP, VLANs, routing, switching Strong experience with infrastructure scripting and automation (Bash, Python, Ansible) Deep understanding of observability principles and tools (Prometheus, Grafana) Hands-on experience operating orchestration platforms (Kubernetes, MAAS, Tinkerbell) Strong grasp of ITSM and service operation best practices Excellent communication ...

Software Engineering Manager

Location
Bristol, England, United Kingdom
risks early and transparently. Establish engineering guardrails across scope, quality, and non-functional requirements, enabling teams to design optimal solutions within them. Champion observability and operational excellence, ensuring system health, SLOs, and alerting are visible and actively managed. Partner with Tech Leads and Architects on system design and evolution, bringing … architecture, including API design and integration, performance optimisation, security, and microservice or event‐driven patterns. Experience with CI/CD, modern development workflows, and observability practices. Proven track record of leading high‐performing teams in fast‐paced, complex or regulated environments. Passion for mentoring and developing engineers through coaching, feedback ...

QA Test Infrastructure Engineer

Location
Cheltenham, England, United Kingdom
QA Test Infrastructure Engineer - Chelmsford, Onsite - Outside IR35 - Highest Security Clearance As a QA Test Infrastructure Engineer, you'll help design, build, and deliver secure digital solutions in highly secure environments. You'll work alongside ...

Lead Site Reliability Engineer (Dynatrace)

Hiring Organisation
SF Partners
Location
South West England, United Kingdom
Employment Type
Full-Time
Salary
£80,000 - £100,000 per annum
looking for an experienced Site Reliability Engineer/Observability Engineer with deep Dynatrace expertise to join a major technology and platform engineering programme. This is not a role for someone who has simply used Dynatrace dashboards. We're looking for an engineer who has been involved in the implementation, configuration … technical SME within complex production environments. What we're looking for Strong hands-on Dynatrace implementation and administration experience Experience designing and implementing observability/monitoring solutions end-to-end Strong SRE and production engineering background Experience configuring instrumentation, metrics, alerting and monitoring Understanding of technologies such as OneAgent, ActiveGate ...

Senior Software Engineer-AI

Hiring Organisation
Hackajob Ltd
Location
South West London, London, United Kingdom
Employment Type
Permanent
operation with limited supervision Hands-on experience with cloud-native technologies, serverless applications, event-driven architectures, data pipelines, relational and NoSQL databases, vector databases, observability tooling, and automated deployment pipelines Solid understanding of algorithms, data structures, scalability, reliability, performance optimization, security best practices, and engineering trade-offs Experience mentoring engineers … technical designs, participate in design reviews, and identify risks, constraints, trade-offs, and alternative approaches Maintain engineering excellence through automated testing, code reviews, observability, monitoring, alerting, operational readiness, and participation in on-call support Apply machine learning operations practices, including prompt versioning, automated evaluation, deployment pipelines, monitoring, and production issue ...

Senior Associate, Full-Stack Engineer Opportunities

Hiring Organisation
Hackajob Ltd
Location
South West London, London, United Kingdom
Employment Type
Permanent
ways: Design, build, and maintain backend services, batches and APIs, contributing to UI components as needed. Own end-to-end delivery: implementation, testing, deployment, observability, and reliability. Write clean, well-tested code; participate in code reviews and continuous improvement. Collaborate with product, design, and operations to translate business needs into … microservices Proficiency in Java with Spring. Experience with CI/CD, automated testing (JUnit/Spock), and containers (Docker). Familiarity with microservices, observability/telemetry (e.g., Splunk, AppDynamics), and cloud deployments. Curiosity to understand the business domain and translate product strategy into technical solutions. Working knowledge of Groovy with ...

Principal Platform Engineer (12 Month FTC)

Hiring Organisation
Hackajob Ltd
Location
South West London, London, United Kingdom
Employment Type
Permanent
delivery, operational processes, and platform management. Establish and promote platform standards, engineering patterns, and best practices across engineering teams. Improve the reliability, scalability, performance, observability, and operational efficiency of the platform. Identify opportunities to reduce engineering friction, eliminate repetitive work, and accelerate software delivery. Establish engineering principles and guardrails that … delivery. Automation of operational processes and platform lifecycle management. Experience establishing repeatable, standardised engineering workflows. Reliability & Performance Engineering Designing platforms for resilience, fault tolerance, observability, and operational excellence. Applying SRE principles and practices to improve availability and reduce operational risk. Performance analysis, capacity planning, scalability engineering, and proactive reliability improvement. ...

Data Architect

Location
Bristol, England, United Kingdom
We believe in the power of ingenuity to build a positive human future.We challenge where it matters and own the outcome.As strategies, technologies, and innovation collide, we create opportunity from complexity. Our teams of interdisciplinary ...

Data Consultant - DV CLEAR

Hiring Organisation
Hays Specialist Recruitment Limited
Location
Corsham, Wiltshire, United Kingdom
Employment Type
Full-Time
Salary
£518.88 per day
Your new company You will be working for a secure client delivering data engineering capability in a highly regulated environment. The client's identity will remain confidential at this stage. The position requires previous experience ...