301 to 325 of 502 Remote/Hybrid Observability Jobs

Sr. Software Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
while architecting systems that handle high-throughput data pipelines. Participate in building robust export pipelines, streaming architectures, webhook integrations and MCP servers. Maintain high observability and reliability standards using tools like Coralogix, CloudWatch, and Grafana. Participate in on-call rotation and incident response for owned services. What You'll Bring … static site generators). Familiarity with authentication, API gateways, and rate limiting strategies. Experience in compliance standards for APIs and data handling. Experience with observability tools and practices. Languages: Golang (primary) with some TypeScript Monitoring: Coralogix, Grafana, CloudWatch CI/CD & IaC: GitHub Actions, Terraform What We Offer Generous paid ...

Site Reliability Engineer- Spacetime UK

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
system for a platform that transforms how networks of satellites, ground stations, and fleets are interconnected and orchestrated. You will be building the core observability stack that ensures the reliability of systems critical to the operation of satellite megaconstellations and missions to deep space. This is a greenfield/brownfield … expert, helping to define and implement the strategy and building the tools that empower our engineers. You will support the roadmap to mature our observability stack, moving from cloud-native tools to a robust, scalable, and insightful platform built on best-in-class technologies (Prometheus, OpenTelemetry, etc.). ...

Developer Enablement, Technical Architect – Release on Demand (SVP)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
graceful degradation, and zero-downtime deployments. Define SLOs, own the SRE practice for the platform, and be accountable when things need fixing.* **Define the Observability Strategy.** Establish a comprehensive observability framework — distributed tracing, structured logging, metrics, dashboards, and alerting. The platform must be fully understood at all times. You will … NoSQL databases: PostgreSQL, MongoDB or Couchbase* Demonstrated SRE or platform engineering experience — SLOs, incident management, reliability engineering at scale* Experience defining and implementing observability strategies: distributed tracing, structured logging, metrics and alerting* Proven experience leading technical projects and mentoring engineers**Highly Desirable skills*** Experience with SDLC tooling, release management platforms ...

Senior SRE: Cloud Reliability & Automation (Remote)

Hiring Organisation
Jobleads-UK
Location
Manchester, England, United Kingdom
continually improve a highly available cloud platform supporting mission-critical production services. The role spans cloud infrastructure, software engineering, and operations, driving automation, observability, and best practices across the SDLC to strengthen platform reliability and incident response. #J-18808-Ljbffr ...

Remote ML Platform Engineer - Scale AI Infrastructure

Hiring Organisation
Jobleads-UK
Location
United Kingdom
operate complex models, merging software engineering, cloud infrastructure, and ML to advance platform capabilities. Collaborate with researchers and product engineers, focusing on automation, observability, and developer experience. #J-18808-Ljbffr ...

Remote Senior Full-Stack Engineer — AI‐Powered Platform

Hiring Organisation
Jobleads-UK
Location
United Kingdom
development in a fully remote environment. You will build scalable services in Java/Node.js, craft frontend with React and TypeScript, and drive reliability, observability, and CI/CD improvements. #J-18808-Ljbffr ...

Senior Platform Engineer – Developer Experience (Hybrid)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
will build and maintain internal tools, templates, and paved roads that enable self-service for product teams, improve CI/CD, and enhance observability with logs, metrics, and traces. #J-18808-Ljbffr ...

Hybrid Principal SDE — Cloud-Native Platform Leader

Hiring Organisation
Jobleads-UK
Location
Reigate, England, United Kingdom
engineering standards and mentor teams across AWS and Azure, delivering secure, scalable cloud-native solutions in a regulated environment. You will drive DevSecOps disciplines, observability, and end-to-end ownership while collaborating with product and commercial teams to align technical outcomes with business goals. #J-18808-Ljbffr ...

Hybrid Principal SDE — Cloud-Native Platform Leader

Hiring Organisation
Jobleads-UK
Location
Glasgow, Scotland, United Kingdom
engineering standards and mentor teams across AWS and Azure, delivering secure, scalable cloud-native solutions in a regulated environment. You will drive DevSecOps disciplines, observability, and end-to-end ownership while collaborating with product and commercial teams to align technical outcomes with business goals. #J-18808-Ljbffr ...

Hybrid Principal SDE — Cloud-Native Platform Leader

Hiring Organisation
Jobleads-UK
Location
Manchester, England, United Kingdom
engineering standards and mentor teams across AWS and Azure, delivering secure, scalable cloud-native solutions in a regulated environment. You will drive DevSecOps disciplines, observability, and end-to-end ownership while collaborating with product and commercial teams to align technical outcomes with business goals. #J-18808-Ljbffr ...

Lead Network Engineer

Hiring Organisation
DGH Recruitment
Location
City of London, London, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£95,000
Microsoft Azure networking including virtual networks, subnets, NSGs, route tables and hybrid connectivity. - Experience supporting network resilience, high availability, disaster recovery testing, monitoring and observability across enterprise LAN, WAN, wireless, data centre and cloud services. - Experience producing and maintaining HLDs, LLDs, standards, implementation plans and operational documentation. Lead Network Engineer ...

Site Reliability & Network Systems Administrator

Hiring Organisation
Franklin Bates Limited
Location
Leamington Spa, Warwickshire, West Midlands, United Kingdom
Employment Type
Permanent
Salary
£55,000
setup, configuration and ongoing optimisation Managing networking, routing, firewalls, VPNs and connectivity Infrastructure monitoring, alerting and performance optimisation Implementing and managing web monitoring and observability platforms Incident response, root cause analysis and continual service improvement Capacity planning, disaster recovery and business continuity Supporting security best practice, patch management and infrastructure ...

Gen AI - Solutions Architect - Home based

Hiring Organisation
Square One Resources
Location
London, United Kingdom
Employment Type
Contract
Contract Rate
£500 - £517/day
support Design secure, scalable, and governed AI solutions aligned with enterprise architecture standards. Ensure compliance with security, privacy, governance, and Responsible AI requirements. Support observability, interoperability, monitoring, and operational excellence across AI platforms. Embed controls for identity, access management, auditability, transparency, and accountability. Evaluate AI technologies, vendors, and platforms ...

Identity & Access Lead - BPL

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Respect, Integrity, Service, Excellence and Stewardship, and the Barclays Mindset of Empower, Challenge and Drive. Qualifications Extensive experience of Identity & Access Management (IAM). Observability Pipeline experience partnering with peers and CISO. Experience of Joiner‐Mover‐Leaver pipeline creation and innovation. Tailscale – ability to input on and architect a zero ...

Platform & Workplace Engineering Director

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
responsible for leading both our Cloud Platform and Enterprise Service orgs: Leading and developing two engineering teams — Cloud Platform (cloud infrastructure, observability, SRE, incident response) and Enterprise Services (IT support, IAM, Okta, Slack and core SaaS) — setting technical direction, hiring and performance standards across both. Delivering the reliability, performance ...

Enterprise Hybrid Cloud Platform Architect (Advisory) -Manager - National Security

Hiring Organisation
KPMG UK
Location
England, United Kingdom
lowest common denominator capability set. You must understand the importance of both enterprise MI, for long term decision making, and (near) real time observability for operations. You must have a strong knowledge of traditional platform delivery approaches, technologies and op models, and a thorough appreciation of the capabilities ...

Enterprise Account Executive, EMEA Sales EMEA (Remote)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
powerful ways to monitor, troubleshoot, and optimize their AI systems. That’s where we come in. Arize AI is the leading AI & Agent Engineering observability and evaluation platform, empowering AI engineers to ship high-performing, reliable agents and applications. From first prototype to production scale, Arize AX unifies build, test ...

Engineering Manager, Payments

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
team and drive positive change.Bonus points if:You are familiar with any one of the following technologies : Scala, Java or Typescript.You are familiar with observability, tracking and data pipeline tools and methodologies.You have previously worked in an App first business.Additional InformationHealth + Mental WellbeingPMI and cash plan healthcare access with ...

Backend Engineer, Platform

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
platform and product development from discovery to implementation. Contribute to high‐level architectural decisions for the core application and associated services. Improve scalability, reliability, observability, and operational safety of critical backend systems. Work on foundational platform capabilities: backend systems, service‐to‐service communication, authentication, infrastructure architecture, messaging, database scalability, deployment … Experience with PostgreSQL, MySQL, or DynamoDB. Experience building scalable web applications or backend systems handling significant traffic and data volume. Experience improving reliability, scalability, observability, and operational maturity in production systems. Experience in a DevOps culture: CI/CD pipelines, observability, incident response, and production ownership. Experience with platform capabilities ...

Network Engineer

Hiring Organisation
HCLTech
Location
City of London, London, United Kingdom
Modern Ops and AI-first operating model. The role focuses on Network Infrastructure as Code (NetIaC), CI/CD pipelines, AI-driven operations (AIOps), observability integration, and SRE-led reliability engineering. Key Responsibilities Develop and manage Network Infrastructure as Code (NetIaC) using Python, Ansible, and Terraform for provisioning and lifecycle … ITSM workflows. Drive AI/ML use cases such as WAN capacity forecasting, anomaly detection, predictive analytics, and self-healing networks. Integrate and manage observability platforms (SolarWinds Orion, Elastic, Grafana, ZDX) for proactive monitoring and insights. Provide engineering and support for MCP (Model Context Protocol) and AI agent integrations. Ensure ...

Network Automation Consultant

Hiring Organisation
HCLTech
Location
London Area, United Kingdom
Modern Ops and AI-first operating model. The role focuses on Network Infrastructure as Code (NetIaC), CI/CD pipelines, AI-driven operations (AIOps), observability integration, and SRE-led reliability engineering. Key Responsibilities Develop and manage Network Infrastructure as Code (NetIaC) using Python, Ansible, and Terraform for provisioning and lifecycle … ITSM workflows. Drive AI/ML use cases such as WAN capacity forecasting, anomaly detection, predictive analytics, and self-healing networks. Integrate and manage observability platforms (SolarWinds Orion, Elastic, Grafana, ZDX) for proactive monitoring and insights. Provide engineering and support for MCP (Model Context Protocol) and AI agent integrations. Ensure ...

Senior Backend Engineer | AI Platform

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
high degree of autonomy and ownership, as you'll be responsible for designing scalable solutions that empower multiple engineering teams while ensuring reliability, observability, and cost efficiency. What are we looking for: 5+ years of experience in Software Engineering, Backend Engineering, or Platform Engineering. Strong experience building and maintaining backend … LangChain, LangGraph, CrewAI, or similar. Experience working with cloud platforms such as Google Cloud Platform (preferred), AWS, or Azure. Strong understanding of system reliability, observability, monitoring, and incident management. Experience with Infrastructure as Code and cloud-native architectures. Previous experience working within a Platform Engineering team is a strong plus. ...

DevOps Engineer

Hiring Organisation
WeDo Technology Solutions Limited
Location
Croydon, Surrey, England, United Kingdom
Employment Type
Full-Time
Salary
£75,000 - £85,000 per annum
pipelines across multiple engineering teams• Automate infrastructure and deployments using Infrastructure as Code• Support Azure cloud infrastructure and AKS environments• Improve platform monitoring, observability, and operational efficiency• Troubleshoot production and deployment issues to maintain platform reliability• Collaborate closely with Software Engineers, Platform Engineers, and Security teams to improve delivery …/Kubernetes• Strong Terraform or Infrastructure as Code experience• Experience building and maintaining CI/CD pipelines• Good understanding of monitoring, logging, and observability tools• Strong troubleshooting and problem-solving skills• Experience working within Agile engineering teams Why Apply? You'll be joining an engineering organisation operating at genuine enterprise ...

Senior Platform Engineer

Hiring Organisation
Anson Mccade
Location
Manchester, North West, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£70,000
Code using Terraform across production and non-production environments Driving DevSecOps best practice and improving engineering standards across delivery teams Implementing SRE principles including observability, monitoring, SLIs/SLOs and platform reliability Supporting production environments, incident management and continuous service improvements Mentoring engineers and acting as a technical leader within … DevSecOps engineering experience within enterprise environments Proven Terraform experience building Infrastructure as Code Good understanding of Site Reliability Engineering (SRE) principles Experience with observability and monitoring tools such as Dynatrace, Grafana, Prometheus or similar Knowledge of CI/CD pipelines and modern cloud-native engineering practices Experience supporting live production ...

Senior DevOps Engineer

Hiring Organisation
Halian Technology Limited
Location
Reading, Berkshire, South East, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£85,000
reliability, and availability Implement self-service tooling to empower development teams Drive DevOps best practices across the digital product lifecycle Develop and enhance monitoring, observability, and incident response processes Support global engineering teams delivering high-traffic platforms Key Requirements Proven experience supporting digital product delivery in a DevOps or platform … with Infrastructure as Code (Terraform, Ansible, Puppet or similar) Hands-on experience with Kubernetes, Docker, and cloud platforms (AWS preferred) Experience with monitoring/observability tools (Prometheus, Grafana, ELK, APM tools) Solid understanding of system performance, scalability, and resilience Strong collaboration and communication skills within cross-functional product teams Desirable ...