226 to 250 of 393 Remote Observability Jobs

Site Reliability Engineer- Spacetime UK

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
system for a platform that transforms how networks of satellites, ground stations, and fleets are interconnected and orchestrated. You will be building the core observability stack that ensures the reliability of systems critical to the operation of satellite megaconstellations and missions to deep space. This is a greenfield/brownfield … expert, helping to define and implement the strategy and building the tools that empower our engineers. You will support the roadmap to mature our observability stack, moving from cloud-native tools to a robust, scalable, and insightful platform built on best-in-class technologies (Prometheus, OpenTelemetry, etc.). ...

Developer Enablement, Technical Architect – Release on Demand (SVP)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
graceful degradation, and zero-downtime deployments. Define SLOs, own the SRE practice for the platform, and be accountable when things need fixing.* **Define the Observability Strategy.** Establish a comprehensive observability framework — distributed tracing, structured logging, metrics, dashboards, and alerting. The platform must be fully understood at all times. You will … NoSQL databases: PostgreSQL, MongoDB or Couchbase* Demonstrated SRE or platform engineering experience — SLOs, incident management, reliability engineering at scale* Experience defining and implementing observability strategies: distributed tracing, structured logging, metrics and alerting* Proven experience leading technical projects and mentoring engineers**Highly Desirable skills*** Experience with SDLC tooling, release management platforms ...

Senior SRE: Cloud Reliability & Automation (Remote)

Hiring Organisation
Jobleads-UK
Location
Manchester, England, United Kingdom
continually improve a highly available cloud platform supporting mission-critical production services. The role spans cloud infrastructure, software engineering, and operations, driving automation, observability, and best practices across the SDLC to strengthen platform reliability and incident response. #J-18808-Ljbffr ...

Senior SRE: Remote, Impactful Cloud Reliability

Hiring Organisation
Jobleads-UK
Location
Bristol, England, United Kingdom
continuously improve a highly available cloud platform for mission-critical services. The role spans cloud infrastructure, software engineering, and operations to drive automation, observability, and reliability across the development lifecycle. You will strengthen platform reliability, improve incident response, and embed SRE practices. #J-18808-Ljbffr ...

Senior Platform Engineer – Developer Experience (Hybrid)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
will build and maintain internal tools, templates, and paved roads that enable self-service for product teams, improve CI/CD, and enhance observability with logs, metrics, and traces. #J-18808-Ljbffr ...

Hybrid Principal SDE — Cloud-Native Platform Leader

Hiring Organisation
Jobleads-UK
Location
Manchester, England, United Kingdom
engineering standards and mentor teams across AWS and Azure, delivering secure, scalable cloud-native solutions in a regulated environment. You will drive DevSecOps disciplines, observability, and end-to-end ownership while collaborating with product and commercial teams to align technical outcomes with business goals. #J-18808-Ljbffr ...

Hybrid Principal SDE — Cloud-Native Platform Leader

Hiring Organisation
Jobleads-UK
Location
Reigate, England, United Kingdom
engineering standards and mentor teams across AWS and Azure, delivering secure, scalable cloud-native solutions in a regulated environment. You will drive DevSecOps disciplines, observability, and end-to-end ownership while collaborating with product and commercial teams to align technical outcomes with business goals. #J-18808-Ljbffr ...

Hybrid Principal SDE — Cloud-Native Platform Leader

Hiring Organisation
Jobleads-UK
Location
Glasgow, Scotland, United Kingdom
engineering standards and mentor teams across AWS and Azure, delivering secure, scalable cloud-native solutions in a regulated environment. You will drive DevSecOps disciplines, observability, and end-to-end ownership while collaborating with product and commercial teams to align technical outcomes with business goals. #J-18808-Ljbffr ...

Lead Network Engineer

Hiring Organisation
DGH Recruitment
Location
City of London, London, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£95,000
Microsoft Azure networking including virtual networks, subnets, NSGs, route tables and hybrid connectivity. - Experience supporting network resilience, high availability, disaster recovery testing, monitoring and observability across enterprise LAN, WAN, wireless, data centre and cloud services. - Experience producing and maintaining HLDs, LLDs, standards, implementation plans and operational documentation. Lead Network Engineer ...

Site Reliability & Network Systems Administrator

Hiring Organisation
Franklin Bates Limited
Location
Leamington Spa, Warwickshire, West Midlands, United Kingdom
Employment Type
Permanent
Salary
£55,000
setup, configuration and ongoing optimisation Managing networking, routing, firewalls, VPNs and connectivity Infrastructure monitoring, alerting and performance optimisation Implementing and managing web monitoring and observability platforms Incident response, root cause analysis and continual service improvement Capacity planning, disaster recovery and business continuity Supporting security best practice, patch management and infrastructure ...

Identity & Access Lead - BPL

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Respect, Integrity, Service, Excellence and Stewardship, and the Barclays Mindset of Empower, Challenge and Drive. Qualifications Extensive experience of Identity & Access Management (IAM). Observability Pipeline experience partnering with peers and CISO. Experience of Joiner‐Mover‐Leaver pipeline creation and innovation. Tailscale – ability to input on and architect a zero ...

Platform & Workplace Engineering Director

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
responsible for leading both our Cloud Platform and Enterprise Service orgs: Leading and developing two engineering teams — Cloud Platform (cloud infrastructure, observability, SRE, incident response) and Enterprise Services (IT support, IAM, Okta, Slack and core SaaS) — setting technical direction, hiring and performance standards across both. Delivering the reliability, performance ...

Enterprise Account Executive, EMEA Sales EMEA (Remote)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
powerful ways to monitor, troubleshoot, and optimize their AI systems. That’s where we come in. Arize AI is the leading AI & Agent Engineering observability and evaluation platform, empowering AI engineers to ship high-performing, reliable agents and applications. From first prototype to production scale, Arize AX unifies build, test ...

Engineering Manager, Payments

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
team and drive positive change.Bonus points if:You are familiar with any one of the following technologies : Scala, Java or Typescript.You are familiar with observability, tracking and data pipeline tools and methodologies.You have previously worked in an App first business.Additional InformationHealth + Mental WellbeingPMI and cash plan healthcare access with ...

Backend Engineer, Platform

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
platform and product development from discovery to implementation. Contribute to high‐level architectural decisions for the core application and associated services. Improve scalability, reliability, observability, and operational safety of critical backend systems. Work on foundational platform capabilities: backend systems, service‐to‐service communication, authentication, infrastructure architecture, messaging, database scalability, deployment … Experience with PostgreSQL, MySQL, or DynamoDB. Experience building scalable web applications or backend systems handling significant traffic and data volume. Experience improving reliability, scalability, observability, and operational maturity in production systems. Experience in a DevOps culture: CI/CD pipelines, observability, incident response, and production ownership. Experience with platform capabilities ...

Senior Backend Engineer | AI Platform

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
high degree of autonomy and ownership, as you'll be responsible for designing scalable solutions that empower multiple engineering teams while ensuring reliability, observability, and cost efficiency. What are we looking for: 5+ years of experience in Software Engineering, Backend Engineering, or Platform Engineering. Strong experience building and maintaining backend … LangChain, LangGraph, CrewAI, or similar. Experience working with cloud platforms such as Google Cloud Platform (preferred), AWS, or Azure. Strong understanding of system reliability, observability, monitoring, and incident management. Experience with Infrastructure as Code and cloud-native architectures. Previous experience working within a Platform Engineering team is a strong plus. ...

DevOps Engineer

Hiring Organisation
WeDo Technology Solutions Limited
Location
Croydon, Surrey, England, United Kingdom
Employment Type
Full-Time
Salary
£75,000 - £85,000 per annum
pipelines across multiple engineering teams• Automate infrastructure and deployments using Infrastructure as Code• Support Azure cloud infrastructure and AKS environments• Improve platform monitoring, observability, and operational efficiency• Troubleshoot production and deployment issues to maintain platform reliability• Collaborate closely with Software Engineers, Platform Engineers, and Security teams to improve delivery …/Kubernetes• Strong Terraform or Infrastructure as Code experience• Experience building and maintaining CI/CD pipelines• Good understanding of monitoring, logging, and observability tools• Strong troubleshooting and problem-solving skills• Experience working within Agile engineering teams Why Apply? You'll be joining an engineering organisation operating at genuine enterprise ...

Senior Platform Engineer

Hiring Organisation
SF Partners Admin
Location
United Kingdom
Employment Type
Permanent, Work From Home
Developing Infrastructure as Code using Terraform - Creating CI/CD pipelines and automation - Building Internal Developer Platforms (IDPs) and self-service tooling - Improving observability, reliability and platform performance - Working closely with architects and engineering teams on large-scale transformation programmes To be suitable, you must have experience with most … following: - AWS - Kubernetes - Terraform - Linux - Git - CI/CD - Infrastructure as Code - Platform Engineering or DevOps Desirable but not essential - Experience with GitOps, Observability (Grafana, Prometheus, OpenTelemetry), SRE or DevSecOps. What's on offer? - £75,000 - £110,000 depending on experience and location - Hybrid working (upto 30% onsite, 70% working ...

Senior Platform Engineer

Hiring Organisation
Anson Mccade
Location
Manchester, North West, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£70,000
Code using Terraform across production and non-production environments Driving DevSecOps best practice and improving engineering standards across delivery teams Implementing SRE principles including observability, monitoring, SLIs/SLOs and platform reliability Supporting production environments, incident management and continuous service improvements Mentoring engineers and acting as a technical leader within … DevSecOps engineering experience within enterprise environments Proven Terraform experience building Infrastructure as Code Good understanding of Site Reliability Engineering (SRE) principles Experience with observability and monitoring tools such as Dynatrace, Grafana, Prometheus or similar Knowledge of CI/CD pipelines and modern cloud-native engineering practices Experience supporting live production ...

Senior DevOps Engineer

Hiring Organisation
Halian Technology Limited
Location
Reading, Berkshire, South East, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£95,000
reliability, and availability Implement self-service tooling to empower development teams Drive DevOps best practices across the digital product lifecycle Develop and enhance monitoring, observability, and incident response processes Support global engineering teams delivering high-traffic platforms Key Requirements Proven experience supporting digital product delivery in a DevOps or platform … with Infrastructure as Code (Terraform, Ansible, Puppet or similar) Hands-on experience with Kubernetes, Docker, and cloud platforms (AWS preferred) Experience with monitoring/observability tools (Prometheus, Grafana, ELK, APM tools) Solid understanding of system performance, scalability, and resilience Strong collaboration and communication skills within cross-functional product teams Desirable ...

Site Reliability Engineer (AWS)

Hiring Organisation
Spectrum It Recruitment Limited
Location
Southampton, Hampshire, South East, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£60,000
issues and restoring services quickly and effectively Developing automation to reduce manual operational tasks and improve platform resilience Building and improving monitoring, alerting and observability across cloud environments Working alongside Software, Platform, Cloud and Security Engineers to improve reliability and operational excellence Contributing to post-incident reviews and driving continuous … with exposure to: Linux systems administration AWS cloud infrastructure Kubernetes and Docker Production support and incident management Python, Bash or Go scripting Monitoring and observability platforms such as Grafana, Prometheus, Datadog, Splunk or CloudWatch Networking fundamentals including DNS, TCP/IP and load balancing A passion for automation, continuous improvement ...

Site Reliability Engineer (AWS)

Hiring Organisation
Spectrum IT Recruitment
Location
Birmingham, West Midlands, West Midlands (County), United Kingdom
Employment Type
Permanent
issues and restoring services quickly and effectively Developing automation to reduce manual operational tasks and improve platform resilience Building and improving monitoring, alerting and observability across cloud environments Working alongside Software, Platform, Cloud and Security Engineers to improve reliability and operational excellence Contributing to post-incident reviews and driving continuous … with exposure to: Linux systems administration AWS cloud infrastructure Kubernetes and Docker Production support and incident management Python, Bash or Go scripting Monitoring and observability platforms such as Grafana, Prometheus, Datadog, Splunk or CloudWatch Networking fundamentals including DNS, TCP/IP and load balancing A passion for automation, continuous improvement ...

Principal Software Development Engineer

Hiring Organisation
Jobleads-UK
Location
Manchester, England, United Kingdom
/CD pipelines, infrastructure as code, automation frameworks, and database-as-code practices using Redgate Flyway. Own critical customer systems, ensuring operational resilience, observability, performance optimisation, and rapid incident response. Collaborate with Product, Delivery, Operations, and Commercial teams to shape technical solutions, delivery plans, and strategic outcomes. Promote secure … Connect or Genesys Cloud. Proven ability to design and deliver secure, scalable, and resilient cloud-native solutions within complex enterprise environments. Strong understanding of observability, operational support, reliability engineering, and end-to-end ownership practices. Knowledge of regulated financial services environments, including UK GDPR and FCA Consumer Duty requirements. Excellent ...

Principal Software Development Engineer

Hiring Organisation
Jobleads-UK
Location
Reigate, England, United Kingdom
pipelines, Infrastructure as Code, automation frameworks, and database-as-code practices using Redgate Flyway. Take ownership of critical customer systems, ensuring operational resilience, observability, performance optimisation, and rapid incident response. Collaborate closely with Product, Delivery, Operations, and Commercial teams to shape technical solutions, delivery plans, and strategic outcomes. Promote secure … Connect or Genesys Cloud. Proven ability to design and deliver secure, scalable, and resilient cloud-native solutions within complex enterprise environments. Strong understanding of observability, operational support, reliability engineering, and end-to-end ownership practices. Knowledge of regulated financial services environments, including UK GDPR and FCA Consumer Duty requirements. Excellent ...

Staff Engineer - Data

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
create unnecessary complexity, risk, duplicated capability or long‐term support burden. Raise the quality bar for data products through clear ownership, robust testing, reconciliation, observability, lineage, documentation, performance and supportability. Collaborate with cross‐functional teams to address security, GDPR, PII handling, role‐based access, auditability and data governance are designed … services across batch, streaming and event‐driven patterns. Deep understanding of engineering practice: clean design, testing strategy, CI/CD, infrastructure as code, observability, performance, security, incident response and DevSecOps. Experience with cloud data services and modern data stacks. Relevant technologies may include Snowflake, Azure/AWS/GCP data ...