1,626 to 1,650 of 2,655 Remote/Hybrid Observability Jobs

Remote Head of DevOps

Hiring Organisation
1inch
Location
Kirkwall, Orkney Islands, UK
secrets management and infrastructure provisioning. Service Reliability: Define and implement SLIs, SLOs, and SLAs to ensure the stability and performance of critical services. Observability & Incident Management: Oversee the full monitoring stack and establish formal incident response and post-mortem processes. Security & Compliance: Enforce "security by design" and maintain strict adherence … GitOps practices. Cloud Architecture: Proficiency in managing multi-cloud environments (AWS, Hetzner, GCP). Technical Depth: Strong understanding of service mesh and microservices architecture. Observability: Solid experience with observability tools for metrics, logging, and tracing. Security Mindset: Practical experience implementing security best practices within CI/CD. Communication: Ability ...

Remote Head of DevOps

Hiring Organisation
1inch
Location
Wemyss Bay, Inverclyde, UK
secrets management and infrastructure provisioning. Service Reliability: Define and implement SLIs, SLOs, and SLAs to ensure the stability and performance of critical services. Observability & Incident Management: Oversee the full monitoring stack and establish formal incident response and post-mortem processes. Security & Compliance: Enforce "security by design" and maintain strict adherence … GitOps practices. Cloud Architecture: Proficiency in managing multi-cloud environments (AWS, Hetzner, GCP). Technical Depth: Strong understanding of service mesh and microservices architecture. Observability: Solid experience with observability tools for metrics, logging, and tracing. Security Mindset: Practical experience implementing security best practices within CI/CD. Communication: Ability ...

Remote Head of DevOps

Hiring Organisation
1inch
Location
Airdrie, North Lanarkshire, UK
secrets management and infrastructure provisioning. Service Reliability: Define and implement SLIs, SLOs, and SLAs to ensure the stability and performance of critical services. Observability & Incident Management: Oversee the full monitoring stack and establish formal incident response and post-mortem processes. Security & Compliance: Enforce "security by design" and maintain strict adherence … GitOps practices. Cloud Architecture: Proficiency in managing multi-cloud environments (AWS, Hetzner, GCP). Technical Depth: Strong understanding of service mesh and microservices architecture. Observability: Solid experience with observability tools for metrics, logging, and tracing. Security Mindset: Practical experience implementing security best practices within CI/CD. Communication: Ability ...

Remote Head of DevOps

Hiring Organisation
1inch
Location
Llantwit Major, Vale of Glamorgan, UK
secrets management and infrastructure provisioning. Service Reliability: Define and implement SLIs, SLOs, and SLAs to ensure the stability and performance of critical services. Observability & Incident Management: Oversee the full monitoring stack and establish formal incident response and post-mortem processes. Security & Compliance: Enforce "security by design" and maintain strict adherence … GitOps practices. Cloud Architecture: Proficiency in managing multi-cloud environments (AWS, Hetzner, GCP). Technical Depth: Strong understanding of service mesh and microservices architecture. Observability: Solid experience with observability tools for metrics, logging, and tracing. Security Mindset: Practical experience implementing security best practices within CI/CD. Communication: Ability ...

Remote Head of DevOps

Hiring Organisation
1inch
Location
Remote, UK
secrets management and infrastructure provisioning. Service Reliability: Define and implement SLIs, SLOs, and SLAs to ensure the stability and performance of critical services. Observability & Incident Management: Oversee the full monitoring stack and establish formal incident response and post-mortem processes. Security & Compliance: Enforce "security by design" and maintain strict adherence … GitOps practices. Cloud Architecture: Proficiency in managing multi-cloud environments (AWS, Hetzner, GCP). Technical Depth: Strong understanding of service mesh and microservices architecture. Observability: Solid experience with observability tools for metrics, logging, and tracing. Security Mindset: Practical experience implementing security best practices within CI/CD. Communication: Ability ...

Remote Head of DevOps

Hiring Organisation
1inch
Location
Bath, Somerset, UK
secrets management and infrastructure provisioning. Service Reliability: Define and implement SLIs, SLOs, and SLAs to ensure the stability and performance of critical services. Observability & Incident Management: Oversee the full monitoring stack and establish formal incident response and post-mortem processes. Security & Compliance: Enforce "security by design" and maintain strict adherence … GitOps practices. Cloud Architecture: Proficiency in managing multi-cloud environments (AWS, Hetzner, GCP). Technical Depth: Strong understanding of service mesh and microservices architecture. Observability: Solid experience with observability tools for metrics, logging, and tracing. Security Mindset: Practical experience implementing security best practices within CI/CD. Communication: Ability ...

Remote Head of DevOps

Location
Ferndown, Dorset, United Kingdom
secrets management and infrastructure provisioning. Service Reliability: Define and implement SLIs, SLOs, and SLAs to ensure the stability and performance of critical services. Observability & Incident Management: Oversee the full monitoring stack and establish formal incident response and post-mortem processes. Security & Compliance: Enforce "security by design" and maintain strict adherence … GitOps practices. Cloud Architecture: Proficiency in managing multi-cloud environments (AWS, Hetzner, GCP). Technical Depth: Strong understanding of service mesh and microservices architecture. Observability: Solid experience with observability tools for metrics, logging, and tracing. Security Mindset: Practical experience implementing security best practices within CI/CD. Communication: Ability ...

Remote Head of DevOps

Location
Rochester, Kent, United Kingdom
secrets management and infrastructure provisioning. Service Reliability: Define and implement SLIs, SLOs, and SLAs to ensure the stability and performance of critical services. Observability & Incident Management: Oversee the full monitoring stack and establish formal incident response and post-mortem processes. Security & Compliance: Enforce "security by design" and maintain strict adherence … GitOps practices. Cloud Architecture: Proficiency in managing multi-cloud environments (AWS, Hetzner, GCP). Technical Depth: Strong understanding of service mesh and microservices architecture. Observability: Solid experience with observability tools for metrics, logging, and tracing. Security Mindset: Practical experience implementing security best practices within CI/CD. Communication: Ability ...

Remote Head of DevOps

Location
Bedworth, Warwickshire, United Kingdom
secrets management and infrastructure provisioning. Service Reliability: Define and implement SLIs, SLOs, and SLAs to ensure the stability and performance of critical services. Observability & Incident Management: Oversee the full monitoring stack and establish formal incident response and post-mortem processes. Security & Compliance: Enforce "security by design" and maintain strict adherence … GitOps practices. Cloud Architecture: Proficiency in managing multi-cloud environments (AWS, Hetzner, GCP). Technical Depth: Strong understanding of service mesh and microservices architecture. Observability: Solid experience with observability tools for metrics, logging, and tracing. Security Mindset: Practical experience implementing security best practices within CI/CD. Communication: Ability ...

Site Reliability Engineer- Spacetime UK

Location
Greater London, England, United Kingdom
system for a platform that transforms how networks of satellites, ground stations, and fleets are interconnected and orchestrated. You will be building the core observability stack that ensures the reliability of systems critical to the operation of satellite megaconstellations and missions to deep space. This is a greenfield/brownfield … expert, helping to define and implement the strategy and building the tools that empower our engineers. You will support the roadmap to mature our observability stack, moving from cloud-native tools to a robust, scalable, and insightful platform built on best-in-class technologies (Prometheus, OpenTelemetry, etc.). ...

Hybrid Linux Automation Engineer – Travel Expensed

Location
Milton, Scotland, United Kingdom
automation across large-scale enterprise infrastructure, including Linux, VMware, and F5 environments. You will build automation to accelerate patching and changes, strengthen validation, improve observability with Prometheus, Grafana, and Airflow, and work with Python and Ansible within a collaborative engineering team. #J-18808-Ljbffr ...

GenAI Full-Stack Engineer (Python/TypeScript) – Hybrid

Location
Belfast City District, Northern Ireland, United Kingdom
/Azure cloud architecture, delivering production GenAI systems used across global operations. This hybrid role focuses on feature implementation, system integration, cloud architecture and observability tooling, ensuring secure coding and scalable deployments. #J-18808-Ljbffr ...

Hybrid AI-Driven SRE & Reliability Engineer

Location
United Kingdom
bet365 Group is seeking a Site Reliability Engineer to enhance system reliability, observability, and performance. You will treat reliability as a software problem, protecting uptime and driving improvements across critical systems. Responsibilities include building tools, dashboards, and automation, contributing to live incident resolution and post-mortems, and mentoring teammates ...

Senior Backend Engineer: Chaos & Reliability (Remote)

Location
Greater London, England, United Kingdom
guide product direction, and collaborate with friendly colleagues who live our FAITH values. You’ll design and run chaos experiments, improve load testing and observability, and introduce new tooling to boost reliability. Global teams collaborate on scalable solutions. #J-18808-Ljbffr ...

Remote Data Engineer for AI Data Platform

Location
United Kingdom
with a strong emphasis on privacy, security, and robust data practices. As part of a collaborative team, you’ll design data pipelines, APIs, and observability, applying IaC and container orchestration tools to keep our platform scalable, reliable, and self-serve for internal teams. #J-18808-Ljbffr ...

Value Stream Engineering Lead — Hybrid London, Cloud Native

Location
Greater London, England, United Kingdom
lead cross-functional teams across backend and frontend, guide API designs, event-driven and microservices architectures, and promote best practices in CI/CD, observability and secure, scalable cloud native solutions. #J-18808-Ljbffr ...

Senior ML Engineer - Build ML Core for Payments (Remote)

Location
United Kingdom
turn promising ML ideas into shipped solutions and measurable outcomes. You will lead the development of scalable ML infrastructure, versioning, CI/CD and observability, while mentoring other engineers and clarifying technical direction in a rapidly evolving area. #J-18808-Ljbffr ...

Backend Engineer - Cloud APIs (Remote UK)

Location
West of England, England, United Kingdom
mainly remote setup with occasional visits to Bristol or London. The role emphasizes building robust APIs, microservices, and scalable architectures, with a focus on observability, testing, and production readiness. #J-18808-Ljbffr ...

Senior DevOps Engineer - Remote, FinOps & Compliance

Location
Greater London, England, United Kingdom
contribute to secure multi-account governance, IAM/SOC 2 alignment, and cost management while supporting a distributed engineering team. You will shape reliability, observability, and automation using Terraform, CDK, or Pulumi, in a fully remote setup with a global team and a strong focus on security and operational excellence. ...

Web App Developer - React/Python, Azure Cloud, Hybrid

Location
Greater London, England, United Kingdom
Vite frontends, contributing to reliability, performance, and developer experience. In this hybrid London role, you’ll collaborate with engineering teams on CI/CD, observability, security and scalable cloud deployments, with a strong focus on delivering robust software and reusable engineering patterns. #J-18808-Ljbffr ...

Engineering Lead for Acquisition & AI Growth

Location
Greater London, England, United Kingdom
raise standards while shaping how AI tools are used across the team. You’ll own end-to-end delivery, drive CI/CD and observability improvements, and balance reliability with cost across AWS services. Hybrid work is offered in a dynamic, AI-driven environment. #J-18808-Ljbffr ...

Lead Engineer: Greenfield Systems & Tech Strategy (Hybrid)

Location
Cardiff, Wales, United Kingdom
role focuses on Java (Spring Boot) on the backend, React/TypeScript on the frontend, PostgreSQL multitenancy, and AWS cloud hosting. You’ll drive observability and secure design across the stack. #J-18808-Ljbffr ...

Senior Platform Engineer: Cloud Infra Leader (Remote)

Location
Greater London, England, United Kingdom
shared platform layer across B2B and B2C products in a high-growth fintech environment. You will build and improve scalable, secure cloud infrastructure, enhance observability and CI/CD, and collaborate with product and engineering teams to enable rapid, reliable releases. #J-18808-Ljbffr ...

Senior Custody Apps Support AVP — Global Hybrid

Location
Belfast City District, Northern Ireland, United Kingdom
seeking a Global Custody Production Support professional to ensure stability and reliability of custody and settlement platforms. You will apply SRE principles, automation, and observability to reduce toil and drive continuous service improvements. The role focuses on distributed systems, cloud-native tech, and cross-functional teamwork in a fast-paced ...

Hybrid Technical Lead - Full-Stack Java & Microservices

Location
Lancaster, England, United Kingdom
quality engineering on a modern, scalable software platform. This full-stack leadership role demands hands-on Java expertise, Spring frameworks, microservices, Kafka, and modern observability tooling. You will mentor engineers, shape architectures, and lead agile ceremonies in a fast-paced, collaborative environment. The role supports a hybrid working model with ...