826 to 850 of 2,062 Observability Jobs in the UK

VodafoneThree - SRE III

Hiring Organisation
VodafoneThree
Location
Greater London, United Kingdom
Employment Type
Full Time
enable teams to validate scalability, reliability, and operational readiness from the earliest stages of design and development. Your expertise in software engineering, performance testing, observability, and automation will help drive the adoption of engineering best practices across the organisation. You will lead initiatives to integrate performance and resilience validation into … service capabilities, providing tooling, frameworks, standards, and guidance for performance, resilience, and chaos testing. Support in defining and championing engineering best practices, including SLOs, observability, capacity planning, and operational readiness to improve service reliability and customer experience. You will collaborate closely Product, Engineering, Platform, Architecture, and SRE teams to design ...

Senior Site Reliability Engineer

Hiring Organisation
Spectrum It Recruitment Limited
Location
Southampton, Hampshire, South East, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£70,000
Have: Practical experience managing large-scale Kubernetes clusters; certifications in Kubernetes are a strong bonus Hands-on familiarity with the Grafana Observability Suite, including tools like Loki, Mimir, and Tempo Background in administering or developing with popular monitoring and automation tools such as Splunk, Datadog, PagerDuty, or Rundeck Experience using … with tools such as Jenkins, GitLab CI/CD, or CircleCI Strong understanding of containerization (e.g., Docker, Kubernetes) and microservices architecture Skilled in using observability and monitoring tools such as Prometheus, Grafana, ELK stack, or AWS CloudWatch Excellent analytical and troubleshooting abilities, especially within complex distributed systems Proven experience handling ...

Principal Site Reliability Engineer, Infrastructure Observability

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
opportunity to grow and make a difference in ways that matter to you. Role Summary In this role as Principal Site Reliability Engineer, Infrastructure Observability you will help formulate, develop, and implement a team of Site Reliability Engineers (SREs) focused on the observability, sustainability, scalability, measurability and recoverability of T. … Proficiency with understanding and explaining incident situations and their recovery plans to prevent recurrence Knowledge/experience driving dashboard standardization across the ecosystem for observability, APM and infrastructure monitoring, and application‐specific logging Knowledge/experience with observability tools such as New Relic, SolarWinds DPA, Elastic Stack, Prometheus, Grafana, Splunk ...

Senior Software Engineer

Hiring Organisation
XCEEDANCE LIMITED
Location
London, UK
Employment Type
Full-time
developments and bring relevant patterns back to the team. Break down stories into tasks, provide estimates, and surface risks early in sprint planning. Maintain observability standards: structured logging, metrics, and distributed tracing. Support CI/CD pipelines and participate in production readiness reviews. Mentor junior engineers through pair programming, code … sprint planning, story decomposition, backlog grooming, retrospectives. Strong unit and component testing discipline; exposure to BDD or contract testing is a plus. Appreciation for observability: structured logging, distributed tracing, alerting hygiene. Desirable Qualifications and Skills Insurance or Insurtech domain knowledge — Policy Administration, Claims, or Underwriting workflows. Familiarity with Kafka ...

Lead Performance Test Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
distributed systems.* Knowledge of cloud-native, containerised and modern application architectures.* Experience integrating performance testing within CI/CD pipelines.* Experience using monitoring, observability and telemetry tools to identify performance bottlenecks.* Strong root-cause analysis, diagnostics and performance tuning skills.* Experience defining performance requirements, workloads, test data, baselines and acceptance … Level Objectives (SLOs) and operational resilience.* Experience working within regulated, public sector or healthcare environments.* Relevant certifications in Performance Testing, Cloud, DevOps, SRE or Observability disciplines.## **Your security clearance**To be successfully appointed to this role, it is a requirement to obtain Security Check (SC) clearance. To obtain SC clearance ...

Lead Software Developer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
discussions. Participate in release planning, deployment activities, and production support. Take responsibility for the production services developed by the team. Improve system robustness, resilience, observability, performance, and operational stability. Conduct code reviews and provide quality assurance for work in progress. Troubleshoot complex technical issues and drive root cause analysis. Identify … environments. Familiarity with Government Technology Code of Practice and Service Standards. Experience with CI/CD pipelines and automated testing frameworks. Knowledge of monitoring, observability, and operational support practices in production environments. Experience leading distributed or multi-supplier teams. Pension Scheme - contributions matched up to 10% Private medical cover Income ...

Data Engineer - Security & Intelligence

Hiring Organisation
Hackajob Ltd
Location
Gloucester, Gloucestershire, South West, United Kingdom
Employment Type
Permanent
Salary
£85,000
support high-performance analytical environments. Contribute to DevSecOps and cloud-native delivery, including automated deployment, infrastructure provisioning and containerised environments. Support the ongoing operation, observability and resilience of production data platforms. Skills Required Experience in data engineering using modern programming languages such as Python, Java or Scala, and SQL. Hands … Familiarity with Infrastructure as Code and modern cloud-native data architectures. Experience working in DevSecOps environments, including Docker, Kubernetes, CI/CD pipelines, and observability/monitoring tooling. Ability to work directly with stakeholders to understand data requirements and translate them into robust, scalable data engineering solutions. Comfortable working ...

Senior Fullstack Engineer (Python + React.js)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
minimal downtime. Write unit and integration tests to maintain code reliability and ensure high- quality releases. Continuously monitor and optimize backend performance using observability tools such as Datadog, Cloud Watch or similar. Participate in design discussions and decision-making to enhance system robustness and scalability. Maintain technical documentation to ensure … handling asynchronous communication. Experience with Infrastructure as Code (IaC) tools like Terraform or CloudFormation for managing cloud infrastructure. Knowledge of observability and monitoring tools, such as Cloud Watch or Datadog, to track and troubleshoot system performance. Familiarity with serverless architectures (e.g., AWS Lambda) and event‐-driven programming paradigms. Exposure ...

Senior Platform Engineer (Remote UK Only)

Hiring Organisation
Jobleads-UK
Location
Bristol, England, United Kingdom
workflows for Kubernetes‐based deployments (using Helm, Kustomize, ArgoCD, or similar) with automated guardrails to ensure fast, repeatable, and safe code delivery. Implement application observability: Set up application‐level metrics, logging, and alerting within the namespaces, ensuring engineering teams have the visibility they need to monitor workload health. Create developer … Docker Compose in production environments Understanding of networking fundamentals Strong scripting ability: Bash and Python Experience with GitOps tooling: ArgoCD or Flux Experience with observability tooling (Prometheus, Grafana, Loki, Alertmanager or equivalent) Ability to think creatively within constraints and plan pragmatically around them: our stack is real‐world, not greenfield ...

Senior Platform Engineer (Remote UK Only)

Hiring Organisation
Jobleads-UK
Location
Cardiff, Wales, United Kingdom
workflows for Kubernetes‐based deployments (using Helm, Kustomize, ArgoCD, or similar) with automated guardrails to ensure fast, repeatable, and safe code delivery. Implement application observability: Set up application‐level metrics, logging, and alerting within the namespaces, ensuring engineering teams have the visibility they need to monitor workload health. Create developer … Docker Compose in production environments Understanding of networking fundamentals Strong scripting ability: Bash and Python Experience with GitOps tooling: ArgoCD or Flux Experience with observability tooling (Prometheus, Grafana, Loki, Alertmanager or equivalent) Ability to think creatively within constraints and plan pragmatically around them: our stack is real‐world, not greenfield ...

Senior Platform Engineer (Remote UK Only)

Hiring Organisation
Jobleads-UK
Location
Leeds, England, United Kingdom
workflows for Kubernetes‐based deployments (using Helm, Kustomize, ArgoCD, or similar) with automated guardrails to ensure fast, repeatable, and safe code delivery. Implement application observability: Set up application‐level metrics, logging, and alerting within the namespaces, ensuring engineering teams have the visibility they need to monitor workload health. Create developer … Docker Compose in production environments Understanding of networking fundamentals Strong scripting ability: Bash and Python Experience with GitOps tooling: ArgoCD or Flux Experience with observability tooling (Prometheus, Grafana, Loki, Alertmanager or equivalent) Ability to think creatively within constraints and plan pragmatically around them: our stack is real‐world, not greenfield ...

Senior Platform Engineer (Remote UK Only)

Hiring Organisation
Jobleads-UK
Location
Manchester, England, United Kingdom
workflows for Kubernetes‐based deployments (using Helm, Kustomize, ArgoCD, or similar) with automated guardrails to ensure fast, repeatable, and safe code delivery. Implement application observability: Set up application‐level metrics, logging, and alerting within the namespaces, ensuring engineering teams have the visibility they need to monitor workload health. Create developer … Docker Compose in production environments Understanding of networking fundamentals Strong scripting ability: Bash and Python Experience with GitOps tooling: ArgoCD or Flux Experience with observability tooling (Prometheus, Grafana, Loki, Alertmanager or equivalent) Ability to think creatively within constraints and plan pragmatically around them: our stack is real‐world, not greenfield ...

Senior Platform Engineer (Remote UK Only)

Hiring Organisation
Jobleads-UK
Location
City of Edinburgh, Scotland, United Kingdom
workflows for Kubernetes‐based deployments (using Helm, Kustomize, ArgoCD, or similar) with automated guardrails to ensure fast, repeatable, and safe code delivery. Implement application observability: Set up application‐level metrics, logging, and alerting within the namespaces, ensuring engineering teams have the visibility they need to monitor workload health. Create developer … Docker Compose in production environments Understanding of networking fundamentals Strong scripting ability: Bash and Python Experience with GitOps tooling: ArgoCD or Flux Experience with observability tooling (Prometheus, Grafana, Loki, Alertmanager or equivalent) Ability to think creatively within constraints and plan pragmatically around them: our stack is real‐world, not greenfield ...

Senior DevOps Engineer London - UK

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
workflows for Kubernetes‐based deployments (using Helm, Kustomize, ArgoCD, or similar) with automated guardrails to ensure fast, repeatable, and safe code delivery. Implement application observability: Set up application‐level metrics, logging, and alerting within the namespaces, ensuring engineering teams have the visibility they need to monitor workload health. Create developer … Docker Compose in production environments Understanding of networking fundamentals Strong scripting ability: Bash and Python Experience with GitOps tooling: ArgoCD or Flux Experience with observability tooling (Prometheus, Grafana, Loki, Alertmanager or equivalent) Ability to think creatively within constraints and plan pragmatically around them: our stack is real‐world, not greenfield ...

Senior DevOps Engineer

Hiring Organisation
Jobleads-UK
Location
Manchester, England, United Kingdom
workflows for Kubernetes-based deployments (using Helm, Kustomize, ArgoCD, or similar) with automated guardrails to ensure fast, repeatable, and safe code delivery. Implement application observability: Set up application-level metrics, logging, and alerting within the namespaces, ensuring engineering teams have the visibility they need to monitor workload health. Create developer … Docker Compose in production environments Understanding of networking fundamentals Strong scripting ability: Bash and Python Experience with GitOps tooling: ArgoCD or Flux Experience with observability tooling (Prometheus, Grafana, Loki, Alertmanager or equivalent) Ability to think creatively within constraints and plan pragmatically around them: our stack is real-world, not greenfield ...

Senior DevOps Engineer

Hiring Organisation
Jobleads-UK
Location
Cardiff, Wales, United Kingdom
workflows for Kubernetes-based deployments (using Helm, Kustomize, ArgoCD, or similar) with automated guardrails to ensure fast, repeatable, and safe code delivery. Implement application observability: Set up application-level metrics, logging, and alerting within the namespaces, ensuring engineering teams have the visibility they need to monitor workload health. Create developer … Docker Compose in production environments Understanding of networking fundamentals Strong scripting ability: Bash and Python Experience with GitOps tooling: ArgoCD or Flux Experience with observability tooling (Prometheus, Grafana, Loki, Alertmanager or equivalent) Ability to think creatively within constraints and plan pragmatically around them: our stack is real-world, not greenfield ...

Senior DevOps Engineer

Hiring Organisation
Jobleads-UK
Location
West of England, England, United Kingdom
workflows for Kubernetes-based deployments (using Helm, Kustomize, ArgoCD, or similar) with automated guardrails to ensure fast, repeatable, and safe code delivery. Implement application observability: Set up application-level metrics, logging, and alerting within the namespaces, ensuring engineering teams have the visibility they need to monitor workload health. Create developer … Docker Compose in production environments Understanding of networking fundamentals Strong scripting ability: Bash and Python Experience with GitOps tooling: ArgoCD or Flux Experience with observability tooling (Prometheus, Grafana, Loki, Alertmanager or equivalent) Ability to think creatively within constraints and plan pragmatically around them: our stack is real-world, not greenfield ...

Senior DevOps Engineer

Hiring Organisation
Jobleads-UK
Location
City of Edinburgh, Scotland, United Kingdom
workflows for Kubernetes-based deployments (using Helm, Kustomize, ArgoCD, or similar) with automated guardrails to ensure fast, repeatable, and safe code delivery. Implement application observability: Set up application-level metrics, logging, and alerting within the namespaces, ensuring engineering teams have the visibility they need to monitor workload health. Create developer … Docker Compose in production environments Understanding of networking fundamentals Strong scripting ability: Bash and Python Experience with GitOps tooling: ArgoCD or Flux Experience with observability tooling (Prometheus, Grafana, Loki, Alertmanager or equivalent) Ability to think creatively within constraints and plan pragmatically around them: our stack is real-world, not greenfield ...

Senior DevOps Engineer

Hiring Organisation
Jobleads-UK
Location
Oxford, England, United Kingdom
workflows for Kubernetes-based deployments (using Helm, Kustomize, ArgoCD, or similar) with automated guardrails to ensure fast, repeatable, and safe code delivery. Implement application observability: Set up application-level metrics, logging, and alerting within the namespaces, ensuring engineering teams have the visibility they need to monitor workload health. Create developer … Docker Compose in production environments Understanding of networking fundamentals Strong scripting ability: Bash and Python Experience with GitOps tooling: ArgoCD or Flux Experience with observability tooling (Prometheus, Grafana, Loki, Alertmanager or equivalent) Ability to think creatively within constraints and plan pragmatically around them: our stack is real-world, not greenfield ...

Senior Platform Engineer (Remote UK Only)

Hiring Organisation
Jobleads-UK
Location
Newcastle upon Tyne, England, United Kingdom
workflows for Kubernetes-based deployments (using Helm, Kustomize, ArgoCD, or similar) with automated guardrails to ensure fast, repeatable, and safe code delivery. Implement application observability: Set up application-level metrics, logging, and alerting within the namespaces, ensuring engineering teams have the visibility they need to monitor workload health. Create developer … Docker Compose in production environments Understanding of networking fundamentals Strong scripting ability: Bash and Python Experience with GitOps tooling: ArgoCD or Flux Experience with observability tooling (Prometheus, Grafana, Loki, Alertmanager or equivalent) Ability to think creatively within constraints and plan pragmatically around them: our stack is real-world, not greenfield ...

Senior Cloud Engineer

Hiring Organisation
Jobleads-UK
Location
Manchester, England, United Kingdom
Implement the data pipelines, workflow orchestrations, and specialised compute footprints needed to support enterprise AI applications, using technologies like AWS Bedrock and AWS AgentCore. Observability & Reliability: Build robust monitoring and observability pipelines to ensure the health, performance, and security of distributed cloud applications and AI models. FinOps Standards: Embed automated … implementing components of data pipelines, including real-time streaming tools (AWS Kinesis, Kafka), data orchestration (dbt, Airflow), and managing vector databases for RAG architectures. Observability & Cost Management: SRE/Platform experience with the practical application of real-time monitoring and cloud cost optimisation using native CSP tools or utilities like ...

Senior Cloud Engineer

Hiring Organisation
Jobleads-UK
Location
Leeds, England, United Kingdom
Implement the data pipelines, workflow orchestrations, and specialised compute footprints needed to support enterprise AI applications, using technologies like AWS Bedrock and AWS AgentCore. Observability & Reliability: Build robust monitoring and observability pipelines to ensure the health, performance, and security of distributed cloud applications and AI models. FinOps Standards: Embed automated … implementing components of data pipelines, including real-time streaming tools (AWS Kinesis, Kafka), data orchestration (dbt, Airflow), and managing vector databases for RAG architectures. Observability & Cost Management: SRE/Platform experience with the practical application of real-time monitoring and cloud cost optimisation using native CSP tools or utilities like ...

Senior Cloud Engineer

Hiring Organisation
Jobleads-UK
Location
City of Edinburgh, Scotland, United Kingdom
Implement the data pipelines, workflow orchestrations, and specialised compute footprints needed to support enterprise AI applications, using technologies like AWS Bedrock and AWS AgentCore. Observability & Reliability: Build robust monitoring and observability pipelines to ensure the health, performance, and security of distributed cloud applications and AI models. FinOps Standards: Embed automated … implementing components of data pipelines, including real-time streaming tools (AWS Kinesis, Kafka), data orchestration (dbt, Airflow), and managing vector databases for RAG architectures. Observability & Cost Management: SRE/Platform experience with the practical application of real-time monitoring and cloud cost optimisation using native CSP tools or utilities like ...

Senior Cloud Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Implement the data pipelines, workflow orchestrations, and specialised compute footprints needed to support enterprise AI applications, using technologies like AWS Bedrock and AWS AgentCore. Observability & Reliability: Build robust monitoring and observability pipelines to ensure the health, performance, and security of distributed cloud applications and AI models. FinOps Standards: Embed automated … implementing components of data pipelines, including real-time streaming tools (AWS Kinesis, Kafka), data orchestration (dbt, Airflow), and managing vector databases for RAG architectures. Observability & Cost Management: SRE/Platform experience with the practical application of real-time monitoring and cloud cost optimisation using native CSP tools or utilities like ...

Senior Cloud Engineer

Hiring Organisation
Jobleads-UK
Location
Manchester, England, United Kingdom
Implement the data pipelines, workflow orchestrations, and specialised compute footprints needed to support enterprise AI applications, using technologies like AWS Bedrock and AWS AgentCore. Observability & Reliability: Build robust monitoring and observability pipelines to ensure the health, performance, and security of distributed cloud applications and AI models. FinOps Standards: Embed automated … implementing components of data pipelines, including real-time streaming tools (AWS Kinesis, Kafka), data orchestration (dbt, Airflow), and managing Vector Databases for RAG architectures. Observability & Cost Management: SRE/Platform experience with the practical application of real-time monitoring and cloud cost optimisation using native CSP tools or utilities like ...