1,376 to 1,400 of 1,781 Permanent Observability Jobs

EMEA Regional Sales Director — AI SaaS Growth Leader

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
leading AI observability platform is seeking a Sales Director for the EMEA region. You will lead and scale a high-performing sales team to drive revenue and new customer acquisition. The ideal candidate has proven sales leadership in high-growth SaaS or AI companies, with a track record of meeting ...

London-Based French-Speaking UK Mid-Market Sales Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
drive pre-sales success across the London market. You will provide technical direction to the sales team and demonstrate Rubrik’s data resilience, observability, and remediation tools to new and existing customers. You will work with Mid-Market AEs to generate pipeline, run customer demos and POCs, and engage ...

AI-First Engineering Manager, Manage Squad Lead

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Reporting to the CPO, you will mentor four engineers, partner with Product and Design, own hiring, ensure quality and platform health, and drive security, observability and governance across the domain. #J-18808-Ljbffr ...

Public Sector AI Sales Exec - Education

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
full sales cycle—from cold outreach to close—driving new logos and expanding existing health customers by showcasing Elastic’s capabilities in search, observability, and security. You will engage with executives, navigate complex procurement, and leverage MEDDPICC to forecast accurately while collaborating with cross-functional teams to maximize deal velocity ...

Public Sector AI Account Executive

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Public Sector Account Executive focused on education. You will own the full sales cycle—from prospecting to close—driving adoption of AI-powered search, Observability and Security across new mid-market accounts and existing health customers. You will articulate Elastic's value to CIOs and procurement teams, negotiate high-stakes ...

Data Services Sales Leader – Remote-First, EMEA/APJ

Hiring Organisation
Jobleads-UK
Location
Windsor, England, United Kingdom
Manager for our Data Services portfolio to lead a high-performing team of sales specialists across EMEA and APJ, driving the ARR target for Observability and Cyber Resilience solutions in large enterprise accounts. You will develop and execute a comprehensive sales strategy, oversee pipeline reviews and revenue forecasting, coach ...

HPC Engineer — NVLink/NVSwitch & GPU Infra

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
support NVLink/NVSwitch platforms in a large data center environment. This role involves troubleshooting Linux and networking issues, while enhancing automation and observability in operations. Ideal candidates will possess strong system administration and hands-on debugging skills. We offer a competitive salary range ...

Engineering Manager: Real-Time Quant Frameworks

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
shape platforms that support high-impact research. This role includes overseeing systems that simplify workloads across large-scale compute environments, guiding the development of observability tools, and driving the strategy of critical scheduling platforms. The ideal candidate will have experience in building scalable production systems and leading technical teams. ...

Lead Customer Success Manager - Technical Data/product dashboards

Hiring Organisation
MLR Associates
Location
London, United Kingdom
Employment Type
Permanent
Salary
GBP 60,000 - 90,000 Annual
Exposure to AI-powered products, including understanding how to set realistic expectations for clients. Familiarity with tools such as Linear, Notion, Slack, HubSpot, or observability ...

Engineering Lead - APAC Expansion (Korea)

Hiring Organisation
Wise
Location
United Kingdom
Employment Type
Full Time
user experience and make data-driven decisions to fix customer pain points Engineering Best Practices: You ensure your team follows defined standards for observability, security, and infrastructure usage. What you can expect as an Engineering Lead at Wise: Scope: You will be responsible for defining and building the Korea engineering ...

Senior Observability Architect & Tech Lead

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
leading global music company seeks a Senior Observability Engineer to drive technical excellence and develop observability strategies. The role involves leading the architectural design and implementation of observability solutions using tools like Dynatrace and Grafana. Candidates should have substantial experience in SRE, DevOps, or Observability, possess strong programming skills ...

Observability Engineer

Hiring Organisation
Peel Ports Group
Location
Great Crosby, Merseyside, UK
Employment Type
Full-time
Observability Engineer Are you an experienced Observability Engineer with a passion for improving service performance, resilience and operational excellence? We are seeking a skilled professional to lead our observability capability, delivering effective monitoring, telemetry, dashboards and insights that enable proactive service management and continuous improvement across our technology services... LFWQ1 ...

Observability Engineer

Hiring Organisation
17918
Location
Liverpool, Merseyside, United Kingdom
Observability Engineer Are you an experienced Observability Engineer with a passion for improving service performance, resilience and operational excellence? We are seeking a skilled professional to lead our observability capability, delivering effective monitoring, telemetry, dashboards and insights that enable proactive service management and continuous improvement across our technology services... WKCL1 ...

Site Reliability Engineer II

Hiring Organisation
Mastercard
Location
O Fallon, Missouri, United States
Employment Type
Permanent
Salary
USD Annual
complex situations, requiring occasional guidance typically only in unfamiliar or highly complex scenarios. They will demonstrate growing consistency and reliability in applying the skills. Observability - Ability to use scripting and tooling to implement observability solutions, enabling the collection, analysis, and visualization of metrics, logs, and traces to support incident detection … anticipate issues, identify risks, and drive preventative improvements that enhance application performance and availability. Specific Tool/Systems: Strong knowledge of ITSM practices, observability, and monitoring using tools such as Splunk and Dynatrace Experience operating and supporting applications on PCF and AWS platforms Proven ability to implement CI/ ...

Principal Engineer - Edge Delivery & Observability

Hiring Organisation
Financial Times
Location
Greater London, United Kingdom
Employment Type
Full Time
unique opportunities to support every step of your career. The FT is looking for a Principal Engineer (Individual Contributor) to lead our Edge Delivery & Observability work. About the teams There are two teams in this area. The Edge Delivery & Observability (EDO) team looks after the cloud edge capabilities like CDNs … along with the monitoring and observability infrastructure at the FT. Examples of the kind of work this team tackles are: Managing and improving our central solution for observability tools like Graphite, Grafana, Splunk, Prometheus and Cloudflare. Providing self service APIs and tools that enable other delivery teams to utilise ...

Lead DevOps Engineer - Real Time Platform

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
global scale. This is an individual contributor leadership role , where you’ll define infrastructure architecture, raise operational standards, and ensure resilience, security, and observability across a mission‐critical platform. AI-First Engineering This team operates with an AI-first approach. We expect hands‐on experience with AI development tooling: terminal … augmented IDEs, and automated workflows. You are ultimately accountable for production quality, security, and correctness. This means owning infrastructure review, security validation, system observability, and operational guardrails. WHAT YOU'LL DO Design and operate cloud infrastructure on AWS to support low‐latency, always‐on real‐time workloads. Own infrastructure ...

site reliability engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
robust incident management frameworks and lead major incident response activities for critical systems Implement blameless postmortems and deliver systemic improvements across production environments Establish observability strategies with standardized tooling for metrics, logs, and tracing to support distributed systems Adopt and enforce SRE practices, including SLIs, SLOs, SLAs, and error budgets … operational tooling to reduce manual processes Требования Strong background in Site Reliability Engineering, DevOps, or platform operations in complex, distributed environments Expertise in observability platforms, troubleshooting distributed systems, and telemetry-driven insights Hands‐on experience with automation, Infrastructure as Code (Terraform or CloudFormation), and CI/CD practices Deep understanding ...

Senior SRE / Platform Engineer- Global Prime Brokerage & Financing Platform

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
optimize CI/CD pipelines to improve software delivery for large-scale distributed systems using Amazon CodeBuild, GitHub Actions & Terraform Enterprise Implement and maintain observability solutions to establish real-time monitoring and proactive incident response using Datadog and AWS CloudWatch Ensure high availability and performance of relational and time-series … with event streaming (i.e. Kafka/Kinesis) Deep understanding of distributed systems architecture, fault tolerance, disaster recovery, and performance tuning Hands-on experience with observability and monitoring tools (e.g. Datadog, ELK, CloudWatch) Proven track record managing infrastructure with Terraform or AWS CDK Strong background in incident management, system reliability ...

Lead Product Manager AIOPs

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
responsible for S&P Global's enterprise AIOps platform and strategy, driving the modernization of IT Operations and Site Reliability Engineering (SRE) through intelligent observability, event intelligence, automation, and AI-driven insights. DTS Platform & Tools – Service Enablement: We serve as thought leaders in AIOps, partnering across IT Operations, SRE, engineering … solving, prioritization, and decision‐making skills. What We’re Looking For: Basic Required Qualifications: 10+ years of experience in product management, IT operations, SRE, observability, platform engineering, or related enterprise technology roles. Strong understanding of AIOps concepts, including event correlation, anomaly detection, root cause analysis, noise reduction, predictive analytics ...

DevOps Engineer

Hiring Organisation
DGH Recruitment Ltd
Location
City of London, London, United Kingdom
Employment Type
Permanent
Salary
£80000 - £100000/annum
platform reliability. Key Responsibilities - Design, deploy, and manage AI platforms and agent infrastructure - Build and maintain CI/CD pipelines and DevOps workflows - Implement observability, monitoring, and logging solutions - Optimise performance, scalability, and cost efficiency - Support AI teams with infrastructure, deployment, and integration - Ensure platform security, compliance, and high availability …/CD, automation, and DevOps best practices - Experience with Kubernetes/containerisation technologies - Strong programming skills (e.g. Python, Go, Node.js) - Experience with observability tools (e.g. OpenTelemetry, Datadog) - Understanding of security, performance optimisation, and scalability Desirable Skills - Experience working on AI/ML platforms or deployments - Exposure to large-scale distributed ...

Site Reliability Engineering Manager

Hiring Organisation
Jobleads-UK
Location
City of Westminster, England, United Kingdom
Reliability Engineers. Shape and deliver our Site Reliability Engineering roadmap alongside the Head of Platform. Champion modern engineering practices including SLIs, SLOs, error budgets, observability and automation. Improve the reliability, scalability and performance of our cloud platforms and digital services. Partner with Engineering, Security, Data and Product teams to embed … operational excellence from design through to production. Drive the adoption of our observability platform, helping teams gain deeper insight into the health and performance of their services. Lead incident learning, continuous improvement and automation initiatives that reduce operational toil. Provide technical leadership across AWS, Kubernetes, Infrastructure as Code, CI/ ...

Junior Azure Engineer

Hiring Organisation
COMPUTACENTER (UK) LIMITED
Location
South East London, London, United Kingdom
Employment Type
Permanent
containers Implement security and governance controls (RBAC, Azure Policy, Management Groups) Build and support landing zones and foundational cloud environments Manage monitoring and observability using Azure Monitor, Log Analytics, and alerts Support CI/CD pipelines using Azure DevOps or GitHub Actions Automate operational tasks using PowerShell, Azure … experience with Infrastructure-as-Code (Bicep, ARM, or Terraform) Solid understanding of Azure networking and cloud architecture principles Experience with monitoring, logging, and observability tools Ability to troubleshoot and resolve complex cloud issues Experience with automation and scripting (PowerShell, Azure CLI) Strong collaboration and communication skills Desirable Experience with Azure ...

Senior Cloud Platform Engineer (GCP | Kubernetes | DevSecOps)

Hiring Organisation
Jobleads-UK
Location
Bolsterstone, England, United Kingdom
repeatable, automated deployments. Implement and maintain CI/CD pipelines and GitOps deployment workflows. Manage cloud networking, connectivity and platform security. Implement platform observability including logging, monitoring, metrics and distributed tracing. Automate platform provisioning, configuration management and operational tasks. Support deployment and operation of identity platform components and supporting services. … Pipelines GitOps Linux Administration Networking and Load Balancing Service Mesh Technologies (Istio/Envoy) Container Platforms Secrets Management PKI and Certificate Management Workload Identity Observability (Logging, Monitoring and Tracing) Scripting and Automation DevSecOps Site Reliability Engineering (SRE) Performance and Capacity Management Operational Support Global Deployment Strategies Positive can-do attitude ...

Lead Product Manager AIOPs

Hiring Organisation
S&P Global
Location
Greater London, United Kingdom
Employment Type
Full Time
responsible for S&P Global's enterprise AIOps platform and strategy, driving the modernization of IT Operations and Site Reliability Engineering (SRE) through intelligent observability, event intelligence, automation, and AI-driven insights. DTS Platform & Tools - Service Enablement: We serve as thought leaders in AIOps, partnering across IT Operations, SRE, engineering … solving, prioritization, and decision-making skills. What We're Looking For: Basic Required Qualifications: 10+ years of experience in product management, IT operations, SRE, observability, platform engineering, or related enterprise technology roles. Strong understanding of AIOps concepts, including event correlation, anomaly detection, root cause analysis, noise reduction, predictive analytics ...

Azure Platform Engineering Consultant

Hiring Organisation
Morgan McKinley
Location
Newbury, Berkshire, England, United Kingdom
Employment Type
Full-Time
Salary
£75,000 - £85,000 per annum
platform templates and landing zone patterns. CI/CD & Automation: Build and refine automated deployment pipelines, environment management, and release practices. Platform Quality: Embed observability (monitoring, logging, alerting), resilience, security, and FinOps principles directly into platform assets. Co-Delivery & Knowledge Transfer: Work closely alongside client engineering teams to pair, document … Core compute, networking, storage, identity, security, and platform services. Infrastructure as Code: Strong proficiency with Terraform AND Terragrunt using modular, reusable implementation patterns. DevOps & Observability: Strong experience with CI/CD tools (Azure DevOps/GitHub Actions) and monitoring stacks (Prometheus, Grafana, Azure Monitor, etc.). FinOps: Practical knowledge ...