1,276 to 1,300 of 2,392 Remote/Hybrid Observability Jobs

Senior Software Engineer in Test (SET) New London

Location
Greater London, England, United Kingdom
performing engineering teams. In this role, you’ll define and drive quality strategy forplatform and infrastructure-level products — from container orchestration and microservices to observability tooling and CI/CD pipelines.This is a hands‐on engineering position within cross‐functional teams where quality iseveryone’s responsibility but you’ll lead … platform behavesreliably under real-world conditions Be a Technical Leader in Quality Engineering Establish standards and practices for testing distributed, event-driven systems Enable observability-driven debugging by working closely with platform and service teams Automate validation of operational characteristics like availability, latency, throughput and recoverability Contribute to security posture ...

Senior Backend Engineer

Hiring Organisation
Inspire People
Location
South West London, London, United Kingdom
Employment Type
Permanent, Part Time, Work From Home
Salary
£80,000
part of a multidisciplinary agile team, you will design, build and run platform services that underpin critical digital products, helping development teams improve observability, monitoring, CI/CD and service resilience. As a Senior Backend Engineer (Site Reliability), you will: * Design, build and operate reliable, secure and scalable cloud platform … services supporting critical digital products. * Develop and maintain platform tooling, automation, observability, monitoring and CI/CD capabilities. * Build software solutions using Python and modern engineering practices. * Write clean, maintainable code and infrastructure-as-code solutions to support service delivery. * Embed Site Reliability Engineering principles including SLIs, SLOs, error budgets ...

Senior Backend Engineer

Location
United Kingdom
part of a multidisciplinary agile team, you will design, build and run platform services that underpin critical digital products, helping development teams improve observability, monitoring, CI/CD and service resilience. As a Senior Backend Engineer (Site Reliability), you will: Design, build and operate reliable, secure and scalable cloud platform … services supporting critical digital products. Develop and maintain platform tooling, automation, observability, monitoring and CI/CD capabilities. Build software solutions using Python and modern engineering practices. Write clean, maintainable code and infrastructure-as-code solutions to support service delivery. Embed Site Reliability Engineering principles including SLIs, SLOs, error budgets ...

Senior Backend Engineer

Hiring Organisation
Inspire People
Location
Edinburgh, Midlothian, Scotland, United Kingdom
Employment Type
Permanent, Part Time, Work From Home
Salary
£80,000
part of a multidisciplinary agile team, you will design, build and run platform services that underpin critical digital products, helping development teams improve observability, monitoring, CI/CD and service resilience. As a Senior Backend Engineer (Site Reliability), you will: * Design, build and operate reliable, secure and scalable cloud platform … services supporting critical digital products. * Develop and maintain platform tooling, automation, observability, monitoring and CI/CD capabilities. * Build software solutions using Python and modern engineering practices. * Write clean, maintainable code and infrastructure-as-code solutions to support service delivery. * Embed Site Reliability Engineering principles including SLIs, SLOs, error budgets ...

Senior Backend Engineer

Hiring Organisation
Inspire People
Location
Birmingham, West Midlands, United Kingdom
Employment Type
Permanent, Part Time, Work From Home
Salary
£80,000
part of a multidisciplinary agile team, you will design, build and run platform services that underpin critical digital products, helping development teams improve observability, monitoring, CI/CD and service resilience. As a Senior Backend Engineer (Site Reliability), you will: * Design, build and operate reliable, secure and scalable cloud platform … services supporting critical digital products. * Develop and maintain platform tooling, automation, observability, monitoring and CI/CD capabilities. * Build software solutions using Python and modern engineering practices. * Write clean, maintainable code and infrastructure-as-code solutions to support service delivery. * Embed Site Reliability Engineering principles including SLIs, SLOs, error budgets ...

Senior Backend Engineer

Hiring Organisation
Inspire People
Location
Manchester, North West, United Kingdom
Employment Type
Permanent, Part Time, Work From Home
Salary
£80,000
part of a multidisciplinary agile team, you will design, build and run platform services that underpin critical digital products, helping development teams improve observability, monitoring, CI/CD and service resilience. As a Senior Backend Engineer (Site Reliability), you will: * Design, build and operate reliable, secure and scalable cloud platform … services supporting critical digital products. * Develop and maintain platform tooling, automation, observability, monitoring and CI/CD capabilities. * Build software solutions using Python and modern engineering practices. * Write clean, maintainable code and infrastructure-as-code solutions to support service delivery. * Embed Site Reliability Engineering principles including SLIs, SLOs, error budgets ...

Senior Backend Engineer

Hiring Organisation
Inspire People
Location
Cardiff, South Glamorgan, Wales, United Kingdom
Employment Type
Permanent, Part Time, Work From Home
Salary
£80,000
part of a multidisciplinary agile team, you will design, build and run platform services that underpin critical digital products, helping development teams improve observability, monitoring, CI/CD and service resilience. As a Senior Backend Engineer (Site Reliability), you will: * Design, build and operate reliable, secure and scalable cloud platform … services supporting critical digital products. * Develop and maintain platform tooling, automation, observability, monitoring and CI/CD capabilities. * Build software solutions using Python and modern engineering practices. * Write clean, maintainable code and infrastructure-as-code solutions to support service delivery. * Embed Site Reliability Engineering principles including SLIs, SLOs, error budgets ...

Senior Backend Engineer

Hiring Organisation
Inspire People
Location
Belfast, County Antrim, Northern Ireland, United Kingdom
Employment Type
Permanent, Part Time, Work From Home
Salary
£80,000
part of a multidisciplinary agile team, you will design, build and run platform services that underpin critical digital products, helping development teams improve observability, monitoring, CI/CD and service resilience. As a Senior Backend Engineer (Site Reliability), you will: * Design, build and operate reliable, secure and scalable cloud platform … services supporting critical digital products. * Develop and maintain platform tooling, automation, observability, monitoring and CI/CD capabilities. * Build software solutions using Python and modern engineering practices. * Write clean, maintainable code and infrastructure-as-code solutions to support service delivery. * Embed Site Reliability Engineering principles including SLIs, SLOs, error budgets ...

Senior Backend Engineer

Hiring Organisation
Inspire People
Location
Darlington, County Durham, North East, United Kingdom
Employment Type
Permanent, Part Time, Work From Home
Salary
£80,000
part of a multidisciplinary agile team, you will design, build and run platform services that underpin critical digital products, helping development teams improve observability, monitoring, CI/CD and service resilience. As a Senior Backend Engineer (Site Reliability), you will: * Design, build and operate reliable, secure and scalable cloud platform … services supporting critical digital products. * Develop and maintain platform tooling, automation, observability, monitoring and CI/CD capabilities. * Build software solutions using Python and modern engineering practices. * Write clean, maintainable code and infrastructure-as-code solutions to support service delivery. * Embed Site Reliability Engineering principles including SLIs, SLOs, error budgets ...

Senior Backend Developer

Hiring Organisation
Inspire People
Location
Darlington, County Durham, North East, United Kingdom
Employment Type
Permanent, Part Time, Work From Home
Salary
£80,000
part of a multidisciplinary agile team, you will design, build and run platform services that underpin critical digital products, helping development teams improve observability, monitoring, CI/CD and service resilience. As a Senior Backend Engineer (Site Reliability), you will: * Design, build and operate reliable, secure and scalable cloud platform … services supporting critical digital products. * Develop and maintain platform tooling, automation, observability, monitoring and CI/CD capabilities. * Build software solutions using Python and modern engineering practices. * Write clean, maintainable code and infrastructure-as-code solutions to support service delivery. * Embed Site Reliability Engineering principles including SLIs, SLOs, error budgets ...

Software Engineering Manager

Hiring Organisation
Halian Technology Limited
Location
Central London, London, United Kingdom
Employment Type
Permanent
making sound architectural and design decisions. Champion software quality, security, performance, and operational excellence. Encourage modern engineering practices, including CI/CD, automated testing, observability, and cloud-native development. Stakeholder Management Build strong relationships with business and technology stakeholders. Communicate progress, risks, and dependencies effectively. Align engineering activities with organisational … would be beneficial: .NET, Java, Python, or Node.js, React Microservices architecture RESTful APIs Kubernetes and Docker AWS, Azure, or GCP CI/CD tooling Observability and monitoring platforms Modern data platforms and event-driven architectures There is a 2 - 3 stage interview process, with interview slots now available with ...

Senior Cloud Platform DevOps Engineer

Location
City Of London, England, United Kingdom
robust enterprise cloud security practices, ensuring comprehensive data protection, secure network topologies, and compliance as the platform scales. Implement and enhance comprehensive platform telemetry, observability, logging, and monitoring systems to guarantee high system reliability and performance. What we're looking for: Strong proficiency in Python and modern web frameworks (FastAPI … cloud providers (AWS, GCP, Azure) and local physical servers. Experience integrating or orchestrating generative AI workflows (ComfyUI, LoRAs, LLMs, diffusion models). Knowledge of observability, LLM evaluation, and prompt tracing frameworks (Langfuse, Langgraph). Experience working within an Agile environment. Understanding of GPU workload scaling and compute optimisation ...

Senior Data Platform Engineer - Data Enablement

Location
Greater London, England, United Kingdom
engineering.depop.com/What You’ll Do Pave a path for data as product: Champion data as a first‐class citizen by introducing robust data observability and governance tooling into the platform Software engineering: Develop microservices, libraries, data pipelines Technical design: implement and evolve platform services that enable teams to work … automation‐first mindset Experience delivering data compliance & privacy solutions, ensuring that we uphold data subject rights for our customers Experience introducing a data governance & observability stack enabling rich data lineage, data contracts, SLA/SLO, tagging and data quality monitoring capabilities both on our own platform but also for data ...

Principal AI Engineer - Hybrid

Hiring Organisation
Genesis10
Location
Columbus, Ohio, United States
Employment Type
Permanent
Salary
USD Hourly
adoption, support, and continuous improvement Establish and promote responsible AI, privacy, access control, governance, cost management, and model-risk practices Lead the implementation of observability capabilities for logging, metrics, token utilization, tracing, and operational health on GCP-hosted services Partner with product management, security, architecture, and lines of business … Experience designing and integrating secure REST APIs and OAuth 2.0/OpenID Connect authorization flows Strong understanding of software development lifecycle, CI/CD, observability, reliability, security, privacy, and responsible AI practices Desired Skills: Hands-on experience using Claude Code and GitHub Copilot in an engineering workflow Experience building ...

Lead Platform Operations Engineer

Location
Greater London, England, United Kingdom
maintain platform standards, patterns, and best practices Own Platform Reliability, Security & Performance Lead incident response, root cause analysis, and platform improvements Implement robust monitoring, observability, and alerting strategies Drive security improvements aligned to ISO27001, SOC2, and modern SDLC practices Ensure strong governance across infrastructure, applications, and data Deliver Scalable & Secure … Experience implementing security tooling (SAST, DAST, container scanning, WAF) Strong knowledge of cloud security, encryption, TLS/SSL, certificates, and access control Experience with observability, monitoring, and alerting tools Security & Compliance Practical experience implementing ISO27001 and SOC2 controls Knowledge of OWASP methodologies and secure development lifecycle practices Experience with vulnerability ...

DevOps Engineer

Location
Greater London, England, United Kingdom
operations. Participate in on‐call rotation, incident response, and post‐incident reviews, driving toward blameless root‐cause analysis and durable fixes. Implement and maintain observability across infrastructure and applications (metrics, logs, traces, dashboards, and alerting). Apply DevSecOps practices by embedding security checks into the development lifecycle (SAST/DAST … Vault, SOPS) Understanding of compliance frameworks relevant to cloud environments and experience with policy‐as‐code frameworks for automated compliance and guardrails. Exposure to observability platforms such as Datadog, Prometheus/Grafana, or the OpenTelemetry ecosystem. Experience with container image hardening and scanning (Trivy, Grype, or similar). Experience with ...

Lead Java Developer — Real-Time Risk & Cloud (Hybrid)

Location
Greater London, England, United Kingdom
full lifecycle from design to production support, integrating new analytics and data sets across global teams. The role emphasizes scalable microservices, streaming data, and observability with ELK, Prometheus and Grafana. Hybrid work model and competitive benefits are offered. #J-18808-Ljbffr ...

Lead DevOps Engineer for Low-Latency Trading Platform (Hybrid)

Location
Greater London, England, United Kingdom
infrastructure across cloud and on-prem environments. The role emphasizes reliability, security, and fast delivery, with hands-on leadership across DevOps, platform engineering, and observability initiatives. You will build CI/CD pipelines, automate provisioning and deployment, and collaborate with engineering, security, and operations teams to improve platform readiness ...

Senior Platform Engineer – Remote UK (Cloud/SRE)

Location
Greater London, England, United Kingdom
Hudl in London, United Kingdom, is seeking a Senior Engineer to join our Platform Engineering team. You’ll work on site reliability, cloud infrastructure, observability and production operations to keep Hudl’s platform highly available, scalable and secure. You’ll lead with technical excellence, mentor engineers and drive innovation using ...

Remote NOC Engineer: Cloud Reliability & Automation (UK)

Location
West of England, England, United Kingdom
Ideal candidates will have Linux administration, AWS experience, and hands-on Terraform/Docker work, plus scripting in Python, Bash or Go and strong observability tooling knowledge. #J-18808-Ljbffr ...

Remote Cloud Reliability Architect (Java/C#)

Location
United Kingdom
Bristol or London, aligning with SRE and backend engineering standards. You'll partner with Tech Operations and broader engineering teams to improve deployment safety, observability, and capacity planning, while mentoring staff and delivering maintainable code in Java or C#. #J-18808-Ljbffr ...

Cloud Platform Operations Lead: Azure, Kubernetes & SRE

Location
Greater London, England, United Kingdom
will lead hands-on engineering and inspire a team of Platform Operations engineers. You'll shape platform strategy, manage incident response, and advance observability, automation, and governance across infrastructure, applications, and data. Remote/hybrid working options available with global teams. #J-18808-Ljbffr ...

Remote AWS SRE: Build Resilient Cloud Platforms

Location
England, United Kingdom
join a globally operating AI-driven cloud platform team. This fully remote UK role involves maintaining production systems on AWS, implementing automation and observability, and partnering with software, platform, cloud and security engineers to improve reliability. You will handle 24/7 incidents, build resilient cloud services, and drive continuous ...

Senior SRE: Hybrid Cloud Reliability Engineer

Location
Horsell, England, United Kingdom
call rotations and on-site client engagements. You will strengthen SRE capabilities through training, mentoring, and hands-on practice with OpenShift, Kubernetes, and observability stacks. #J-18808-Ljbffr ...

Azure Data Platform DevOps Engineer – SC Eligible

Location
Leeds, England, United Kingdom
SFIA Level 4) to support Azure-based data platforms for a UK public sector client. You’ll contribute to IaC, CI/CD, and observability while collaborating with senior engineers and data teams. The role emphasizes automation, platform reliability, and adherence to security and change governance within a hybrid work ...