1,901 to 1,925 of 2,213 Observability Jobs in London

Lead Observability Engineer - Cloud-Native FinTech

Location
Greater London, England, United Kingdom
JPMorgan Chase & Co. in the United Kingdom is leading the buildout of a cutting-edge observability capability for cloud-native microservices. You will guide a team, shape the technical roadmap, and deliver cost-efficient telemetry tooling for metrics, logs, and traces across the software lifecycle. The Senior Lead Software Engineer ...

Senior Java Engineer — Platform & Observability (Hybrid)

Location
Greater London, England, United Kingdom
London is seeking a Senior Java Software Engineer to join the Platform Team. You will help build the core platform that ingests and transforms observability data for multiple applications, working on a hybrid schedule from our London office. You will participate in all phases of the product lifecycle, mentor peers ...

SRE & Reliability Lead — AI-Ops & Observability

Location
Carshalton, England, United Kingdom
services used by internal and external customers. You will drive reliability improvements, advance automation and AI-Ops capabilities, and lead a team focused on observability, incident response, operational excellence, and continuous improvement. Responsibilities include translating priorities into clear plans, line managing team leaders, ensuring RCAs and post-mortems are completed ...

Senior Software Engineer II — Reliability & Observability

Location
Greater London, England, United Kingdom
self-healing systems at scale, delivering platform tooling that engineers across the company adopt for their services. You will own incident management tooling, evolve observability infrastructure with SLOs and real-time signals, and contribute to AI-driven automation that reduces toil and speeds delivery. #J-18808-Ljbffr ...

Staff Data Engineer – Data Quality & Governance

Location
Greater London, England, United Kingdom
adjustments@depop.com. For any other non-disability related questions, please reach out to our Talent Partners. Role We’re building a Data Quality, Observability & Governance Team to improve the reliability, trust, and compliance of Depop’s data ecosystem. As a Staff Data Engineer in this team, you’ll lead … reduce the mean time to detection and resolution of data incidents, by establishing data contracts between producers and consumers, developing robust data observability systems, and embedding governance and GDPR compliance principles across the data lifecycle. You’ll collaborate with product engineering, data platform, analytics, and legal teams to build confidence ...

Software Engineer, Observability

Location
Greater London, England, United Kingdom
shaping our story, you’ll help define what comes next. About the Role: We are looking for a Software Engineer to join our Observability team. Vercel users rely on Observability to monitor and understand their applications’ health and behavior. In this role, you will design, implement, and maintain Observability products … large-scale data ingestion, storage, and processing from distributed systems. Develop cutting-edge visualization tools to provide insights into application behavior and performance. Integrate observability features with popular frontend tools, frameworks, and build systems to enhance developer experience. Write clean, efficient, and well-documented code, ensuring platform reliability through thorough ...

Senior Infrastructure Software Engineer

Location
Greater London, England, United Kingdom
combines developer-first software with cost-efficient, large-scale compute. Teams get the tools they need for experimentation, training, and production inference, with security, observability, and control built in. We serve solo researchers, startups, and large enterprises. Lightning AI operates globally with offices in New York City, San Francisco, Seattle … compute systems Integrate software and automation with hardware management and provisioning systems Improve tooling and workflows for managing infrastructure throughout its production lifecycle Reliability & Observability Build telemetry, logging, and observability capabilities that provide visibility into the health of our infrastructure Develop tools and automation to identify, diagnose, and respond ...

Cloud Advisory Architecture Consultant

Location
Greater London, England, United Kingdom
where GenAI and Agentic play a role. Champion system performance, resilience, and efficiency: Proactively identifying and addressing consumption and scalability challenges. Champion full stack observability using modern full stack observability, SRE and AIOps. Ensure Robustness & Security: Own the design of enterprise-wide applications that are highly available, fault-tolerant ...

Cloud Advisory Architecture Consultant

Location
City Of London, England, United Kingdom
where GenAI and Agentic play a role. Champion system performance, resilience, and efficiency: Proactively identifying and addressing consumption and scalability challenges. Champion full stack observability using modern full stack observability, SRE and AIOps. Ensure Robustness & Security: Own the design of enterprise‐wide applications that are highly available, fault‐tolerant ...

Cloud Advisory Architecture Consultant

Hiring Organisation
Accenture
Location
London, UK
Employment Type
Full-time
where GenAI and Agentic play a role. Champion system performance, resilience, and efficiency: Proactively identifying and addressing consumption and scalability challenges. Champion full stack observability using modern full stack observability, SRE and AIOps. Ensure Robustness & Security: Own the design of enterprise-wide applications that are highly available, fault-tolerant ...

Technical Architect 8

Location
Greater London, England, United Kingdom
with D360 and core Salesforce platform services.* Advanced Integration - Experience integrating Salesforce with external agents via APIs and open standards (MCP, A2A).* Governance & Observability - Familiarity with prompt governance, observability, and monitoring frameworks.* Cross-Platform Background - Background in cross-platform integrations (e.g., Hyperscaler SDKs to Salesforce Flows).* Multimodal Pipelines ...

Technical Architect 8

Hiring Organisation
Salesforce
Location
London, UK
Employment Type
Full-time
fluency with D360 and core Salesforce platform services. Advanced Integration - Experience integrating Salesforce with external agents via APIs and open standards (MCP, A2A).Governance & Observability - Familiarity with prompt governance, observability, and monitoring frameworks. Cross-Platform Background - Background in cross-platform integrations (e.g., Hyperscaler SDKs to Salesforce Flows).Multimodal Pipelines - Prior ...

Technical Account Manager (Observability)

Hiring Organisation
LinuxRecruit
Location
London, UK
Employment Type
Full-time
data volumes surge and observability costs climb, this is a rare opportunity to join a high-growth SaaS company at the forefront of modern observability, helping companies tackle these challenges with smarter, more scalable solutions. As a Technical Account Manager, you'll be the trusted advisor to some … executive reviews, you'll drive tangible outcomes and long-term success. In this role, you'll work hands on with modern cloud infrastructure and observability stacks, making a meaningful impact from day one. Whether it's guiding integrations, troubleshooting technical issues, or shaping long-term customer strategies, your work will ...

Network Engineer

Hiring Organisation
Balyasny Asset Management
Location
London, UK
Employment Type
Full-time
Level: Experience ProfessionalsContact: Maddelyn MohrJob ID: REQ8574Overview: BAM is seeking an innovative, driven self-starter for a Network Engineer position with a focus on observability and automation. This technical role requires advanced engineering experience in implementing, automating and monitoring a highly available global network infrastructure. The hire will have … complex problems. A strong and efficient communicator and collaborator. To be considered a good technical fit, you must have: Extensive experience in network observability and automation. Proficient with modern observability tools and protocols: Prometheus, Grafana, Telegraf, Cisco Telemetry, gNMI, YANG. Solid understanding of traditional network monitoring systems: Zabbix, LibreNMS, SNMP. ...

Senior Busess Development Representative - Corporates Observability London, UK

Location
Greater London, England, United Kingdom
Senior Business Development Representative - Corporates Observability ITRS Is looking for an organised, and driven Corporate Business Development Representative (BDR) to join our EMEA Marketing organisation. Serving as the bridge between Marketing and Sales, you will create qualified pipeline through a blend of outbound enterprise prospecting and inbound lead engagement. … will support the ITRS Observability Platform value stream while also promoting awareness and opportunity creation for ITRS Uptrends Digital Experience Monitoring value stream across the EMEA. Working with Marketing and partnering daily with the Enterprise Sales team, you'll engage with prospects by responding to trial registrations and demo requests ...

Staff Cloud Native Software Engineer

Location
Greater London, England, United Kingdom
that connect AI applications and networking components at scale. In this software engineering role, you'll work on shared Kubernetes-based platforms, deployment patterns, observability foundations, infrastructure architecture, and operational tooling that help internal teams run services safely and efficiently on GPU-backed infrastructure. You'll partner closely with platform … runtime components. Develop operational tooling and automation that make Kubernetes-native services easier for internal teams to deploy, run, and support. Infrastructure Architecture, Reliability & Observability Drive infrastructure architecture decisions around how AI applications and networking components integrate across the platform, weighing trade-offs at a cross-team level. Build observability ...

DDI Lead Architect Engineer

Hiring Organisation
EOS IT Solutions
Location
London, UK
Employment Type
Full-time
design, build, integrate, and take end-to-end ownership of DDI initiatives across a diverse ecosystem. The focus is on driving modernization, automation, and observability while ensuring scalability and security. A key responsibility is also to lead a comprehensive resilience assessment and provide informed recommendations to the bank based … Anycast deploymentsWork hands-on with multiple DDI platforms (Infoblox, Cygna Labs/QIP, EfficientIP) and integrate them into a cohesive ecosystemBuild and enhance DDI observability and monitoring (DNS Tap, telemetry pipelines, dashboards)Define and improve governance, compliance, and control frameworks for DNS/IP servicesEnsure secure DNS and IP management ...

Vice President, Production Services Application Support

Hiring Organisation
The Bank of New York Mellon
Location
London, UK
Employment Type
Full-time
risk while improving platform stability. Drive initiatives through to completion with strong ownership, accountability, urgency, and quality. Champion an automation-first mindset, leveraging AI, observability, and tooling to reduce manual effort and improve service quality. Identify systemic issues and drive sustainable remediation through process simplification, platform improvements, and close partnership … incident, problem, and change management with measurable improvements in stability and service recovery. Demonstrated automation-first and AI-enabled mindset, with experience driving tooling, observability, and process automation. Strong ownership mentality and execution focus, with the ability to take initiatives from concept through delivery and embed sustainable outcomes. Deep technical ...

Backend Java Developer – Data Fabric / Platform Engineering

Location
City Of London, England, United Kingdom
platforms/query engines (e.g., Starburst or similar) Own API contracts with living documentation in CI/CD Build production-grade, testable pipelines Drive observability, reliability, and performance Contribute to architecture decisions (modularity, DI, extensibility) What You Bring (Must-Have) Strong hands-on experience in Java (17/21) + … backend performance Production-grade testing using JUnit 5, Mockito Experience with clean architecture, DI, modular design Comfortable owning CI/CD, code quality, observability Familiarity with Docker, Maven, Jenkins ⭐ Nice to Have Apache Calcite Starburst or federated query engines JVM performance tuning High-throughput service interfaces (REST/gRPC) Data ...

Backend Java Developer - Birmingham • Hybrid • £50–65k • Not London

Location
Greater London, England, United Kingdom
platforms/query engines (e.g., Starburst or similar) Own API contracts with living documentation in CI/CD Build production-grade, testable pipelines Drive observability, reliability, and performance Contribute to architecture decisions (modularity, DI, extensibility) What You Bring (Must-Have) Strong hands-on experience in Java (17/21) + … backend performance Production-grade testing using JUnit 5, Mockito Experience with clean architecture, DI, modular design Comfortable owning CI/CD, code quality, observability Familiarity with Docker, Maven, Jenkins ⭐ Nice to Have Apache Calcite Starburst or federated query engines JVM performance tuning High-throughput service interfaces (REST/gRPC) Data ...

Senior Software Engineer II

Hiring Organisation
Stepstone UK
Location
South East London, London, United Kingdom
Employment Type
Permanent
engineering standards, best practices and reusable patterns while partnering with Enterprise Architecture and influencing technical direction Drive engineering excellence by improving code quality, testing, observability, reliability and operational practices Support end-to-end delivery by guiding teams through complex technical challenges, improving decision-making, and contributing to planning and risk … data lakes/lakehouse architectures, Iceberg or similar table formats, as well as batch and streaming processing Knowledge of data quality, governance, cataloguing and observability tools (e.g. Datadog), with DBT or AI-assisted engineering practices as a plus ...

Sr Director, Enterprise AI

Location
Greater London, England, United Kingdom
vendors, platform integrations such as ServiceNow, Salesforce, and SAP, and custom-build trade-offs; bringing well-reasoned recommendations to executive leadership. Use Dynatrace's observability platform as a strategic advantage to monitor, measure, and continuously improve the reliability, performance, and business impact of internally deployed AI solutions. AI Governance …/partner decisions at the portfolio level, including vendor evaluation, contract negotiation support, and post-implementation performance tracking. Bonus: hands-on experience with observability or monitoring platforms - Dynatrace or similar - applied to AI system performance and reliability. Why you will love being a Dynatracer Dynatrace is a leader in unified ...

Senior Sales Engineer - Public Sector (UK)

Location
Greater London, England, United Kingdom
communicate with customers and Datadog business/technical teams regarding product feedback and competitive landscape Who You Are: Passionate about educating customers on observability risks that are meaningful to their business, and able to build and execute an evaluation plan with a customer Someone with strong written and oral communication … Growth listed above may vary based on the country of your employment and the nature of your employment with Datadog. Datadog is the leading observability and security platform for the AI era, providing businesses with unified visibility across the technology stack to manage complexity at scale. It brings applications, infrastructure ...

Senior Director Technology - Cloud Engineering

Hiring Organisation
Travelport
Location
London, UK
Employment Type
Full-time
onboards development teams and products to this platform. You'll be responsible for driving the move to infrastructure as code, account management, automated pipelines, observability, resilience, operating coverage and cost control. You'll also help define how the platform supports AI, data engineering and high-scale product demand, including where … standards for infrastructure as code, CI/CD, AWS account management, platform guardrails and developer enablement. Improve the operational model for the platform, including observability, incident response, reliability and 24/7 support coverage. Partner with Product and Commercial teams to get ahead of major demand changes, customer commitments ...

Linux Platform Engineer

Location
Greater London, England, United Kingdom
RHEL and Ubuntu) that underpin our compute environment. You'll work across the full Linux stack, from OS builds and provisioning systems through to observability, security and custom tooling. Mathematics and science are at our core; AI pushes them further to unlock bigger thinking and bigger impact. We expect every … automate deployments Leading vulnerability response, including CVE triage and kernel and package remediation Helping design agentic frameworks for safe, autonomous detection and remediation Improving observability and telemetry to track OS health, provisioning and compliance Who are we looking for? Strong hands‐on experience with Linux systems, particularly RHEL and Ubuntu ...