1,751 to 1,775 of 2,430 Observability Jobs in London

Production AI Engineer - Industrial Field Systems

Location
Greater London, England, United Kingdom
sites, solving real problems in field service and regulation-compliant workflows. The role emphasizes design, reliability, and end-to-end ownership—from prompts to observability and on-site deployment—within a small, fast-paced team that values practical impact. #J-18808-Ljbffr ...

Executive Director Trading Tech and AI Engineering

Location
Greater London, England, United Kingdom
leadership, Quants and Functional leads to drive the technology roadmap and production readiness. You will build and mentor a high-performing team, fostering ownership, observability, security and a production-first culture to deliver fast, scalable and reliable desk workflows across #J-18808-Ljbffr ...

Public Sector AI Sales Exec - Education

Location
Greater London, England, United Kingdom
full sales cycle—from cold outreach to close—driving new logos and expanding existing health customers by showcasing Elastic’s capabilities in search, observability, and security. You will engage with executives, navigate complex procurement, and leverage MEDDPICC to forecast accurately while collaborating with cross-functional teams to maximize deal velocity ...

Public Sector AI Account Executive

Location
Greater London, England, United Kingdom
Public Sector Account Executive focused on education. You will own the full sales cycle—from prospecting to close—driving adoption of AI-powered search, Observability and Security across new mid-market accounts and existing health customers. You will articulate Elastic's value to CIOs and procurement teams, negotiate high-stakes ...

Engineering Manager: Lead, Ship & Scale Impact

Location
Greater London, England, United Kingdom
ensuring high-quality software and fast delivery. You will work closely with Product, Design and Operations to make tradeoffs explicit, raise standards across velocity, observability, and product engagement, and grow the team through hiring and feedback. #J-18808-Ljbffr ...

Infra Engineer – Secure, Scalable Infrastructure (London)

Location
Greater London, England, United Kingdom
access controls, PII protection, and compliance in a data-rich environment with real-world stakes. You’ll standardize the stack, improve observability, and drive reliability, latency, and cost efficiency as we scale, while enabling teams to answer their own questions without needing engineering tickets. #J-18808-Ljbffr ...

Principal Engineer I, Prepurchase Platform (Remote)

Location
City Of London, England, United Kingdom
demand on-sales. You will write production code daily, influence technical direction, and collaborate across multiple teams within the Prepurchase domain. You will drive observability, resilience patterns, and AI-assisted enhancements while embedding across services or working horizontally. #J-18808-Ljbffr ...

Systems Integration Engineer

Location
Greater London, England, United Kingdom
APIs, and data pipelines into resilient end-to-end business workflows. You will design and implement integrations between critical systems, with focus on reliability, observability, and long-term maintainability in complex client environments. 02 - RESPONSIBILITIES What You Will Own Core Responsibilities. Design integration architecture for API, event, and batch workflows. ...

Network Engineer - F5 / DDI (BlueCat)

Hiring Organisation
Oscar Associates (UK) Limited
Location
London, United Kingdom
Employment Type
Contract
Contract Rate
£60 - £64 per day
technical escalation point for business-impacting network incidents, working the problem until it's resolved Proactive reviews - DR testing, vulnerability checks, identifying gaps in observability Keeping runbooks, wikis, and documentation current so triage is fast and repeatable Secondary support on adjacent security infrastructure - firewalls, proxy, content inspection Enforcing production governance ...

Product Manager, Secure and Exchange London

Location
Greater London, England, United Kingdom
resolution, request construction, encryption, token injection, and response handling Expand beyond fallback use cases (fraud tools, 3DS, loyalty, partner migrations) Ensure reliability, latency, and observability on the critical payment path Build debugging and visibility tools for developers Developer Experience Own the public API surface across Vault and Forward Drive SDKs ...

Engineering Manager

Location
Greater London, England, United Kingdom
Operations to make tradeoffs explicit and keep the team focused on what matters most. You continuously raise standards across developer velocity, code quality, observability, incident response, user analytics, and product engagement. When things break, you’re involved. When execution slows, you diagnose the root cause. When the bar needs ...

API Engineer

Location
Greater London, England, United Kingdom
languages, we will give you plenty of room to learn and grow with your team. Maintain the reliability of our systems by building for observability and participating in incident response as needed. You can find more about how we work on our engineering page. Salary range for Senior Engineer ...

Station Control Engineer

Location
Greater London, England, United Kingdom
engineer and work with us to close the gap to a sustainable future. Your new role comprises of a broad set of skills across observability, protection and control allowing you to both contribute to the specific technical solutions within each of those areas as well as having a holistic understanding ...

Staff Product Manager (Creator Platform Intelligence)

Location
Greater London, England, United Kingdom
that proactively surface issues before they impact creators. Identify opportunities to automate operational processes and reduce manual intervention. Partner with engineering teams to improve observability across podcast publishing and distribution. Cross-functional Leadership Translate complex technical concepts into customer and business value. Partner with Engineering on solution design without prescribing ...

Senior Observability Architect & Tech Lead

Location
Greater London, England, United Kingdom
leading global music company seeks a Senior Observability Engineer to drive technical excellence and develop observability strategies. The role involves leading the architectural design and implementation of observability solutions using tools like Dynatrace and Grafana. Candidates should have substantial experience in SRE, DevOps, or Observability, possess strong programming skills ...

Lead DevOps Engineer - Real Time Platform

Location
Greater London, England, United Kingdom
global scale. This is an individual contributor leadership role , where you’ll define infrastructure architecture, raise operational standards, and ensure resilience, security, and observability across a mission‐critical platform. AI-First Engineering This team operates with an AI-first approach. We expect hands‐on experience with AI development tooling: terminal … augmented IDEs, and automated workflows. You are ultimately accountable for production quality, security, and correctness. This means owning infrastructure review, security validation, system observability, and operational guardrails. WHAT YOU'LL DO Design and operate cloud infrastructure on AWS to support low‐latency, always‐on real‐time workloads. Own infrastructure ...

Site Reliability Engineer - Banking & Finance

Location
Greater London, England, United Kingdom
software engineering and infrastructure. You'll develop internal platforms, tooling, and automation across Linux, distributed systems, and cloud-native technologies, helping improve reliability, observability, and operational efficiency across a global production environment. Responsibilities: Design and develop internal infrastructure tooling and automation. Build and maintain monitoring, observability and configuration management platforms. …/Must Have: Strong experience programming with Python, Go and/or C++ Strong Linux knowledge and understanding of distributed systems. Experience with monitoring, observability or SRE practices. Experience with CI/CD pipelines, Git and infrastructure automation. Familiarity with Kubernetes and containerised workloads. Strong analytical and troubleshooting skills. Benefits ...

Senior Mobile Engineering Manager

Hiring Organisation
Innova Solutions
Location
London, United Kingdom
Employment Type
Contract
Contract Rate
£48 - £58/hour
Boot, Gradle build automation Strong understanding of: Microservices architecture RESTful APIs Event-driven architectures Secure software development practices Scalable distributed systems Deep experience implementing observability strategies, including: Logging, monitoring, alerting, Performance Analysis and Production Diagnostics. Experience with industry-standard observability and monitoring platforms such as Sentry or Datadog, New Relic ...

Network Automation & OSS Designer

Location
Greater London, England, United Kingdom
designs that span multiple network domains. Key Responsibilities/Job Description Design end-to-end network automation solution architectures spanning orchestration, inventory, compliance, and observability layers. Develop reference architectures and detailed solution blueprints for multi-domain network automation (IP/MPLS, optical, mobile RAN/Core, SD-WAN, cloud). … Terraform, Ansible, Helm) for network resource provisioning. Design Kafka-based event streaming and messaging architectures for real‐time network telemetry and automation triggers. Define observability strategies covering metrics, logs, traces, and network telemetry pipelines. Architect AIOps capabilities including closed‐loop automation, anomaly detection, and predictive analytics for network operations. Integrate ...

MLOps Engineer

Location
City Of London, England, United Kingdom
platform reliability. Key Responsibilities Design, deploy, and manage AI platforms and agent infrastructure Build and maintain CI/CD pipelines and DevOps workflows Implement observability, monitoring, and logging solutions Optimise performance, scalability, and cost efficiency Support AI teams with infrastructure, deployment, and integration Ensure platform security, compliance, and high availability …/CD, automation, and DevOps best practices Experience with Kubernetes/containerisation technologies Strong programming skills (e.g. Python, Go, Node.js) Experience with observability tools (e.g. OpenTelemetry, Datadog) Understanding of security, performance optimisation, and scalability Desirable Skills Experience working on AI/ML platforms or deployments Exposure to large-scale distributed ...

Lead Product Manager AIOPs

Location
Greater London, England, United Kingdom
responsible for S&P Global's enterprise AIOps platform and strategy, driving the modernization of IT Operations and Site Reliability Engineering (SRE) through intelligent observability, event intelligence, automation, and AI-driven insights. DTS Platform & Tools – Service Enablement: We serve as thought leaders in AIOps, partnering across IT Operations, SRE, engineering … solving, prioritization, and decision‐making skills. What We’re Looking For: Basic Required Qualifications: 10+ years of experience in product management, IT operations, SRE, observability, platform engineering, or related enterprise technology roles. Strong understanding of AIOps concepts, including event correlation, anomaly detection, root cause analysis, noise reduction, predictive analytics ...

MLOps Engineer

Hiring Organisation
DGH Recruitment
Location
City of London, London, United Kingdom
Employment Type
Permanent
platform reliability. Key Responsibilities - Design, deploy, and manage AI platforms and agent infrastructure - Build and maintain CI/CD pipelines and DevOps workflows - Implement observability, monitoring, and logging solutions - Optimise performance, scalability, and cost efficiency - Support AI teams with infrastructure, deployment, and integration - Ensure platform security, compliance, and high availability …/CD, automation, and DevOps best practices - Experience with Kubernetes/containerisation technologies - Strong programming skills (e.g. Python, Go, Node.js) - Experience with observability tools (e.g. OpenTelemetry, Datadog) - Understanding of security, performance optimisation, and scalability Desirable Skills - Experience working on AI/ML platforms or deployments - Exposure to large-scale distributed ...

Oracle OSS Stack Lead

Hiring Organisation
Vodafone
Location
London, UK
Employment Type
Full-time
monitoring, and operational processes using DevOps and SRE practices. Support the adoption of CI/CD capabilities using Jenkins, GitHub, and Azure DevOps. Define observability standards using monitoring tools such as Dynatrace, Splunk, and enterprise monitoring platforms. Act as the senior technical escalation point for business stakeholders, support teams … Experience with Unix/Linux administration and Oracle SQL performance troubleshooting. Knowledge of cloud technologies, containerisation, and Kubernetes environments is advantageous. Familiarity with monitoring, observability, and enterprise operations tooling. Understanding of telecommunications network integrations and Oracle Fusion Middleware integration concepts is beneficial. Strong knowledge of Incident, Problem, Change, and Release ...

Senior Software Engineer, GoLang

Location
Greater London, England, United Kingdom
continuous integration and delivery (CI/CD) Make data-guided decisions affecting core business metrics and processes Apply platform and reliability engineering practices, including observability, performance optimisation, analytics, and security best practices Facilitate collaboration between teams and promote continuous improvement Mentor junior engineers on engineering practices, coding standards, and troubleshooting … DevOps Continuous Integration and Delivery (CI/CD) Infrastructure-as-Code Hard Skills Application Development Performance Optimisation Automation Development Containerisation Development Methodologies Design Patterns Observability Analytics Security Best Practices Troubleshooting Soft Skills Mentoring Collaboration Continuous Improvement Industry Keywords Public-Facing Systems Internal Insurance Systems Tools & Technologies AWS Lambda DynamoDB Azure ...