1,776 to 1,800 of 2,452 Observability Jobs in London

Systems Integration Engineer

Location
Greater London, England, United Kingdom
APIs, and data pipelines into resilient end-to-end business workflows. You will design and implement integrations between critical systems, with focus on reliability, observability, and long-term maintainability in complex client environments. 02 - RESPONSIBILITIES What You Will Own Core Responsibilities. Design integration architecture for API, event, and batch workflows. ...

Network Engineer - F5 / DDI (BlueCat)

Hiring Organisation
Oscar Associates (UK) Limited
Location
London, United Kingdom
Employment Type
Contract
Contract Rate
£60 - £64 per day
technical escalation point for business-impacting network incidents, working the problem until it's resolved Proactive reviews - DR testing, vulnerability checks, identifying gaps in observability Keeping runbooks, wikis, and documentation current so triage is fast and repeatable Secondary support on adjacent security infrastructure - firewalls, proxy, content inspection Enforcing production governance ...

Product Manager, Secure and Exchange London

Location
Greater London, England, United Kingdom
resolution, request construction, encryption, token injection, and response handling Expand beyond fallback use cases (fraud tools, 3DS, loyalty, partner migrations) Ensure reliability, latency, and observability on the critical payment path Build debugging and visibility tools for developers Developer Experience Own the public API surface across Vault and Forward Drive SDKs ...

React Native Engineer

Hiring Organisation
Manual
Location
London, United Kingdom
Salary
£ 60 K
using React Native or React, meeting high engineering standardsInfluence mobile architecture to support product growth and a scaling engineering teamUse analytics, feature flags, and observability to measure success and iterate based on real user dataBalance delivery speed with long-term robustness, making informed technical trade-offsCollaborate cross-functionally to shape ...

Engineering Manager

Location
Greater London, England, United Kingdom
Operations to make tradeoffs explicit and keep the team focused on what matters most. You continuously raise standards across developer velocity, code quality, observability, incident response, user analytics, and product engagement. When things break, you’re involved. When execution slows, you diagnose the root cause. When the bar needs ...

Station Control Engineer

Hiring Organisation
Ramboll Group
Location
London, United Kingdom
Salary
£ 70 K
Control engineer and work with us to close the gap to a sustainable future.Your new role comprises of a broad set of skills across observability, protection and control allowing you to both contribute to the specific technical solutions within each of those areas as well as having a holistic understanding ...

API Engineer

Location
Greater London, England, United Kingdom
languages, we will give you plenty of room to learn and grow with your team. Maintain the reliability of our systems by building for observability and participating in incident response as needed. You can find more about how we work on our engineering page. Salary range for Senior Engineer ...

Station Control Engineer

Location
Greater London, England, United Kingdom
engineer and work with us to close the gap to a sustainable future. Your new role comprises of a broad set of skills across observability, protection and control allowing you to both contribute to the specific technical solutions within each of those areas as well as having a holistic understanding ...

Staff Product Manager (Creator Platform Intelligence)

Location
Greater London, England, United Kingdom
that proactively surface issues before they impact creators. Identify opportunities to automate operational processes and reduce manual intervention. Partner with engineering teams to improve observability across podcast publishing and distribution. Cross-functional Leadership Translate complex technical concepts into customer and business value. Partner with Engineering on solution design without prescribing ...

Senior Observability Architect & Tech Lead

Location
Greater London, England, United Kingdom
leading global music company seeks a Senior Observability Engineer to drive technical excellence and develop observability strategies. The role involves leading the architectural design and implementation of observability solutions using tools like Dynatrace and Grafana. Candidates should have substantial experience in SRE, DevOps, or Observability, possess strong programming skills ...

Lead DevOps Engineer - Real Time Platform

Location
Greater London, England, United Kingdom
global scale. This is an individual contributor leadership role , where you’ll define infrastructure architecture, raise operational standards, and ensure resilience, security, and observability across a mission‐critical platform. AI-First Engineering This team operates with an AI-first approach. We expect hands‐on experience with AI development tooling: terminal … augmented IDEs, and automated workflows. You are ultimately accountable for production quality, security, and correctness. This means owning infrastructure review, security validation, system observability, and operational guardrails. WHAT YOU'LL DO Design and operate cloud infrastructure on AWS to support low‐latency, always‐on real‐time workloads. Own infrastructure ...

Site Reliability Engineer - Banking & Finance

Location
Greater London, England, United Kingdom
software engineering and infrastructure. You'll develop internal platforms, tooling, and automation across Linux, distributed systems, and cloud-native technologies, helping improve reliability, observability, and operational efficiency across a global production environment. Responsibilities: Design and develop internal infrastructure tooling and automation. Build and maintain monitoring, observability and configuration management platforms. …/Must Have: Strong experience programming with Python, Go and/or C++ Strong Linux knowledge and understanding of distributed systems. Experience with monitoring, observability or SRE practices. Experience with CI/CD pipelines, Git and infrastructure automation. Familiarity with Kubernetes and containerised workloads. Strong analytical and troubleshooting skills. Benefits ...

Senior Mobile Engineering Manager

Hiring Organisation
Innova Solutions
Location
London, United Kingdom
Employment Type
Contract
Contract Rate
£48 - £58/hour
Boot, Gradle build automation Strong understanding of: Microservices architecture RESTful APIs Event-driven architectures Secure software development practices Scalable distributed systems Deep experience implementing observability strategies, including: Logging, monitoring, alerting, Performance Analysis and Production Diagnostics. Experience with industry-standard observability and monitoring platforms such as Sentry or Datadog, New Relic ...

Network Automation & OSS Designer

Location
Greater London, England, United Kingdom
designs that span multiple network domains. Key Responsibilities/Job Description Design end-to-end network automation solution architectures spanning orchestration, inventory, compliance, and observability layers. Develop reference architectures and detailed solution blueprints for multi-domain network automation (IP/MPLS, optical, mobile RAN/Core, SD-WAN, cloud). … Terraform, Ansible, Helm) for network resource provisioning. Design Kafka-based event streaming and messaging architectures for real‐time network telemetry and automation triggers. Define observability strategies covering metrics, logs, traces, and network telemetry pipelines. Architect AIOps capabilities including closed‐loop automation, anomaly detection, and predictive analytics for network operations. Integrate ...

MLOps Engineer

Location
City Of London, England, United Kingdom
platform reliability. Key Responsibilities Design, deploy, and manage AI platforms and agent infrastructure Build and maintain CI/CD pipelines and DevOps workflows Implement observability, monitoring, and logging solutions Optimise performance, scalability, and cost efficiency Support AI teams with infrastructure, deployment, and integration Ensure platform security, compliance, and high availability …/CD, automation, and DevOps best practices Experience with Kubernetes/containerisation technologies Strong programming skills (e.g. Python, Go, Node.js) Experience with observability tools (e.g. OpenTelemetry, Datadog) Understanding of security, performance optimisation, and scalability Desirable Skills Experience working on AI/ML platforms or deployments Exposure to large-scale distributed ...

Lead Product Manager AIOPs

Location
Greater London, England, United Kingdom
responsible for S&P Global's enterprise AIOps platform and strategy, driving the modernization of IT Operations and Site Reliability Engineering (SRE) through intelligent observability, event intelligence, automation, and AI-driven insights. DTS Platform & Tools – Service Enablement: We serve as thought leaders in AIOps, partnering across IT Operations, SRE, engineering … solving, prioritization, and decision‐making skills. What We’re Looking For: Basic Required Qualifications: 10+ years of experience in product management, IT operations, SRE, observability, platform engineering, or related enterprise technology roles. Strong understanding of AIOps concepts, including event correlation, anomaly detection, root cause analysis, noise reduction, predictive analytics ...

MLOps Engineer

Hiring Organisation
DGH Recruitment
Location
City of London, London, United Kingdom
Employment Type
Permanent
platform reliability. Key Responsibilities - Design, deploy, and manage AI platforms and agent infrastructure - Build and maintain CI/CD pipelines and DevOps workflows - Implement observability, monitoring, and logging solutions - Optimise performance, scalability, and cost efficiency - Support AI teams with infrastructure, deployment, and integration - Ensure platform security, compliance, and high availability …/CD, automation, and DevOps best practices - Experience with Kubernetes/containerisation technologies - Strong programming skills (e.g. Python, Go, Node.js) - Experience with observability tools (e.g. OpenTelemetry, Datadog) - Understanding of security, performance optimisation, and scalability Desirable Skills - Experience working on AI/ML platforms or deployments - Exposure to large-scale distributed ...

Oracle OSS Stack Lead

Hiring Organisation
Vodafone
Location
London, UK
Employment Type
Full-time
monitoring, and operational processes using DevOps and SRE practices. Support the adoption of CI/CD capabilities using Jenkins, GitHub, and Azure DevOps. Define observability standards using monitoring tools such as Dynatrace, Splunk, and enterprise monitoring platforms. Act as the senior technical escalation point for business stakeholders, support teams … Experience with Unix/Linux administration and Oracle SQL performance troubleshooting. Knowledge of cloud technologies, containerisation, and Kubernetes environments is advantageous. Familiarity with monitoring, observability, and enterprise operations tooling. Understanding of telecommunications network integrations and Oracle Fusion Middleware integration concepts is beneficial. Strong knowledge of Incident, Problem, Change, and Release ...

Senior Software Engineer, GoLang

Location
Greater London, England, United Kingdom
continuous integration and delivery (CI/CD) Make data-guided decisions affecting core business metrics and processes Apply platform and reliability engineering practices, including observability, performance optimisation, analytics, and security best practices Facilitate collaboration between teams and promote continuous improvement Mentor junior engineers on engineering practices, coding standards, and troubleshooting … DevOps Continuous Integration and Delivery (CI/CD) Infrastructure-as-Code Hard Skills Application Development Performance Optimisation Automation Development Containerisation Development Methodologies Design Patterns Observability Analytics Security Best Practices Troubleshooting Soft Skills Mentoring Collaboration Continuous Improvement Industry Keywords Public-Facing Systems Internal Insurance Systems Tools & Technologies AWS Lambda DynamoDB Azure ...

Oracle OSS Stack Lead

Location
Greater London, England, United Kingdom
monitoring, and operational processes using DevOps and SRE practices. Support the adoption of CI/CD capabilities using Jenkins, GitHub, and Azure DevOps. Define observability standards using monitoring tools such as Dynatrace, Splunk, and enterprise monitoring platforms. Act as the senior technical escalation point for business stakeholders, support teams … Experience with Unix/Linux administration and Oracle SQL performance troubleshooting. Knowledge of cloud technologies, containerisation, and Kubernetes environments is advantageous. Familiarity with monitoring, observability, and enterprise operations tooling. Understanding of telecommunications network integrations and Oracle Fusion Middleware integration concepts is beneficial. Strong knowledge of Incident, Problem, Change, and Release ...

Senior Database Platform Engineer

Location
Greater London, England, United Kingdom
modern lakehouse architectures. This is a hands‐on engineering role. You'll be troubleshooting performance issues, validating recovery strategies, automating operational processes, improving observability and helping shape the future of our database estate. You will act as the team's database SME, working closely with Platform Engineers, Data Engineers … recovery capabilities Own restore testing and recovery readiness across critical platforms Support high availability solutions and service resilience initiatives Implement proactive monitoring, alerting and observability for database services Participate in incident response, problem management and post-incident reviews Drive continual improvement through automation, root cause analysis and operational learning Reduce ...

Principal Engineer - Electronic Pricing - Financial Services - TWE4745

Hiring Organisation
Twenty Recruitment Group
Location
City of London, London, United Kingdom
Employment Type
Permanent, Work From Home
Code, GitOps and deployment automation. Work closely with engineering, product, architecture, security and operational teams. Investigate complex performance, resilience and production issues. Improve observability, maintainability and engineering practices across the platform. Key Skills Significant commercial software engineering experience within complex production environments. Deep hands-on Java development experience. Strong understanding … automated software delivery. Experience with secure software development, including authentication and authorisation. Comfortable owning technical initiatives and making architecture and design decisions. Experience with observability and production monitoring tooling. Able to work independently while influencing technical and non-technical stakeholders. Experience within financial services, trading or another high-throughput environment ...

Software Engineer - Cloud Compute Platform

Location
Greater London, England, United Kingdom
contribute to include: Workload orchestration across Kubernetes clusters Platform APIs and Kubernetes operators Cloud platform integrations Multi-tenant workload isolation and security Observability, health checks, and operational tooling Workload scheduling, disruption management, and rolling updates What You Will Do: Design and implement services, APIs, and Kubernetes controllers using Go. Contribute … engineers to understand requirements and deliver reliable solutions. Participate in technical design discussions and help evaluate implementation trade-offs. Improve the reliability, scalability, observability, and maintainability of existing systems. Write automated tests, documentation, and operational runbooks. Participate in code reviews and provide constructive feedback to teammates. Help investigate and resolve ...

Lead Product Manager AIOPs

Hiring Organisation
S&P Global
Location
London, UK
Employment Type
Full-time
responsible for S&P Global's enterprise AIOps platform and strategy, driving the modernization of IT Operations and Site Reliability Engineering (SRE) through intelligent observability, event intelligence, automation, and AI-driven insights. DTS Platform & Tools – Service Enablement: We serve as thought leaders in AIOps, partnering across IT Operations, SRE, engineering … solving, prioritization, and decision-making skills. What We're Looking For: Basic Required Qualifications:10+ years of experience in product management, IT operations, SRE, observability, platform engineering, or related enterprise technology roles. Strong understanding of AIOps concepts, including event correlation, anomaly detection, root cause analysis, noise reduction, predictive analytics ...