2,051 to 2,075 of 2,372 Observability Jobs in London

Cloud Advisory Architecture Consultant

Hiring Organisation
Accenture
Location
London, UK
Employment Type
Full-time
where GenAI and Agentic play a role. Champion system performance, resilience, and efficiency: Proactively identifying and addressing consumption and scalability challenges. Champion full stack observability using modern full stack observability, SRE and AIOps. Ensure Robustness & Security: Own the design of enterprise-wide applications that are highly available, fault-tolerant ...

Technical Architect 8

Location
Greater London, England, United Kingdom
with D360 and core Salesforce platform services.* Advanced Integration - Experience integrating Salesforce with external agents via APIs and open standards (MCP, A2A).* Governance & Observability - Familiarity with prompt governance, observability, and monitoring frameworks.* Cross-Platform Background - Background in cross-platform integrations (e.g., Hyperscaler SDKs to Salesforce Flows).* Multimodal Pipelines ...

Technical Architect 8

Hiring Organisation
Salesforce
Location
London, UK
Employment Type
Full-time
fluency with D360 and core Salesforce platform services. Advanced Integration - Experience integrating Salesforce with external agents via APIs and open standards (MCP, A2A).Governance & Observability - Familiarity with prompt governance, observability, and monitoring frameworks. Cross-Platform Background - Background in cross-platform integrations (e.g., Hyperscaler SDKs to Salesforce Flows).Multimodal Pipelines - Prior ...

Technical Account Manager (Observability)

Hiring Organisation
LinuxRecruit
Location
London, UK
Employment Type
Full-time
data volumes surge and observability costs climb, this is a rare opportunity to join a high-growth SaaS company at the forefront of modern observability, helping companies tackle these challenges with smarter, more scalable solutions. As a Technical Account Manager, you'll be the trusted advisor to some … executive reviews, you'll drive tangible outcomes and long-term success. In this role, you'll work hands on with modern cloud infrastructure and observability stacks, making a meaningful impact from day one. Whether it's guiding integrations, troubleshooting technical issues, or shaping long-term customer strategies, your work will ...

Network Engineer

Hiring Organisation
Balyasny Asset Management
Location
London, UK
Employment Type
Full-time
Level: Experience ProfessionalsContact: Maddelyn MohrJob ID: REQ8574Overview: BAM is seeking an innovative, driven self-starter for a Network Engineer position with a focus on observability and automation. This technical role requires advanced engineering experience in implementing, automating and monitoring a highly available global network infrastructure. The hire will have … complex problems. A strong and efficient communicator and collaborator. To be considered a good technical fit, you must have: Extensive experience in network observability and automation. Proficient with modern observability tools and protocols: Prometheus, Grafana, Telegraf, Cisco Telemetry, gNMI, YANG. Solid understanding of traditional network monitoring systems: Zabbix, LibreNMS, SNMP. ...

Staff Cloud Native Software Engineer

Location
Greater London, England, United Kingdom
that connect AI applications and networking components at scale. In this software engineering role, you'll work on shared Kubernetes-based platforms, deployment patterns, observability foundations, infrastructure architecture, and operational tooling that help internal teams run services safely and efficiently on GPU-backed infrastructure. You'll partner closely with platform … runtime components. Develop operational tooling and automation that make Kubernetes-native services easier for internal teams to deploy, run, and support. Infrastructure Architecture, Reliability & Observability Drive infrastructure architecture decisions around how AI applications and networking components integrate across the platform, weighing trade-offs at a cross-team level. Build observability ...

DDI Lead Architect Engineer

Hiring Organisation
EOS IT Solutions
Location
London, UK
Employment Type
Full-time
design, build, integrate, and take end-to-end ownership of DDI initiatives across a diverse ecosystem. The focus is on driving modernization, automation, and observability while ensuring scalability and security. A key responsibility is also to lead a comprehensive resilience assessment and provide informed recommendations to the bank based … Anycast deploymentsWork hands-on with multiple DDI platforms (Infoblox, Cygna Labs/QIP, EfficientIP) and integrate them into a cohesive ecosystemBuild and enhance DDI observability and monitoring (DNS Tap, telemetry pipelines, dashboards)Define and improve governance, compliance, and control frameworks for DNS/IP servicesEnsure secure DNS and IP management ...

Vice President, Production Services Application Support

Hiring Organisation
The Bank of New York Mellon
Location
London, UK
Employment Type
Full-time
risk while improving platform stability. Drive initiatives through to completion with strong ownership, accountability, urgency, and quality. Champion an automation-first mindset, leveraging AI, observability, and tooling to reduce manual effort and improve service quality. Identify systemic issues and drive sustainable remediation through process simplification, platform improvements, and close partnership … incident, problem, and change management with measurable improvements in stability and service recovery. Demonstrated automation-first and AI-enabled mindset, with experience driving tooling, observability, and process automation. Strong ownership mentality and execution focus, with the ability to take initiatives from concept through delivery and embed sustainable outcomes. Deep technical ...

Backend Java Developer – Data Fabric / Platform Engineering

Location
Greater London, England, United Kingdom
platforms/query engines (e.g., Starburst or similar) Own API contracts with living documentation in CI/CD Build production-grade, testable pipelines Drive observability, reliability, and performance Contribute to architecture decisions (modularity, DI, extensibility) What You Bring (Must-Have) Strong hands-on experience in Java (17/21) + … backend performance Production-grade testing using JUnit 5, Mockito Experience with clean architecture, DI, modular design Comfortable owning CI/CD, code quality, observability Familiarity with Docker, Maven, Jenkins ⭐ Nice to Have Apache Calcite Starburst or federated query engines JVM performance tuning High-throughput service interfaces (REST/gRPC) Data ...

Backend Java Developer - Birmingham • Hybrid • £50–65k • Not London

Location
City Of London, England, United Kingdom
platforms/query engines (e.g., Starburst or similar) Own API contracts with living documentation in CI/CD Build production-grade, testable pipelines Drive observability, reliability, and performance Contribute to architecture decisions (modularity, DI, extensibility) What You Bring (Must-Have) Strong hands-on experience in Java (17/21) + … backend performance Production-grade testing using JUnit 5, Mockito Experience with clean architecture, DI, modular design Comfortable owning CI/CD, code quality, observability Familiarity with Docker, Maven, Jenkins ⭐ Nice to Have Apache Calcite Starburst or federated query engines JVM performance tuning High-throughput service interfaces (REST/gRPC) Data ...

Senior Software Engineer II

Hiring Organisation
Stepstone UK
Location
South East London, London, United Kingdom
Employment Type
Permanent
engineering standards, best practices and reusable patterns while partnering with Enterprise Architecture and influencing technical direction Drive engineering excellence by improving code quality, testing, observability, reliability and operational practices Support end-to-end delivery by guiding teams through complex technical challenges, improving decision-making, and contributing to planning and risk … data lakes/lakehouse architectures, Iceberg or similar table formats, as well as batch and streaming processing Knowledge of data quality, governance, cataloguing and observability tools (e.g. Datadog), with DBT or AI-assisted engineering practices as a plus ...

Sr Director, Enterprise AI

Location
Greater London, England, United Kingdom
vendors, platform integrations such as ServiceNow, Salesforce, and SAP, and custom-build trade-offs; bringing well-reasoned recommendations to executive leadership. Use Dynatrace's observability platform as a strategic advantage to monitor, measure, and continuously improve the reliability, performance, and business impact of internally deployed AI solutions. AI Governance …/partner decisions at the portfolio level, including vendor evaluation, contract negotiation support, and post-implementation performance tracking. Bonus: hands-on experience with observability or monitoring platforms - Dynatrace or similar - applied to AI system performance and reliability. Why you will love being a Dynatracer Dynatrace is a leader in unified ...

Senior Sales Engineer - Public Sector (UK)

Location
Greater London, England, United Kingdom
communicate with customers and Datadog business/technical teams regarding product feedback and competitive landscape Who You Are: Passionate about educating customers on observability risks that are meaningful to their business, and able to build and execute an evaluation plan with a customer Someone with strong written and oral communication … Growth listed above may vary based on the country of your employment and the nature of your employment with Datadog. Datadog is the leading observability and security platform for the AI era, providing businesses with unified visibility across the technology stack to manage complexity at scale. It brings applications, infrastructure ...

Senior Director Technology - Cloud Engineering

Hiring Organisation
Travelport
Location
London, UK
Employment Type
Full-time
onboards development teams and products to this platform. You'll be responsible for driving the move to infrastructure as code, account management, automated pipelines, observability, resilience, operating coverage and cost control. You'll also help define how the platform supports AI, data engineering and high-scale product demand, including where … standards for infrastructure as code, CI/CD, AWS account management, platform guardrails and developer enablement. Improve the operational model for the platform, including observability, incident response, reliability and 24/7 support coverage. Partner with Product and Commercial teams to get ahead of major demand changes, customer commitments ...

Linux Platform Engineer

Location
Greater London, England, United Kingdom
RHEL and Ubuntu) that underpin our compute environment. You'll work across the full Linux stack, from OS builds and provisioning systems through to observability, security and custom tooling. Mathematics and science are at our core; AI pushes them further to unlock bigger thinking and bigger impact. We expect every … automate deployments Leading vulnerability response, including CVE triage and kernel and package remediation Helping design agentic frameworks for safe, autonomous detection and remediation Improving observability and telemetry to track OS health, provisioning and compliance Who are we looking for? Strong hands‐on experience with Linux systems, particularly RHEL and Ubuntu ...

Backend Software Engineer – Infrastructure, Foundations

Location
Greater London, England, United Kingdom
compute workloads and efficiently schedule hundreds of thousands of containers hourly Design architecture and opinionated APIs that guide application developers Implement tracing and performance observability in high-scale distributed microservice architectures Build reliable, performant, and scalable systems for storage, authentication, and asset serving Automate deployment, management, and operations of distributed … keywords Distributed Systems Development Java Programming C++ Programming Python Programming Cloud Infrastructure Management ATS Optimization Keywords Hard Skills Software Engineering Data Processing Systems Performance Observability Microservice Architecture API Design Container Scheduling Open-Source Contribution High-Scale Systems Automation Prototyping Soft Skills Strong Communication Skills Team Collaboration Feedback Incorporation Quality Maintenance ...

Data Engineer

Location
Greater London, England, United Kingdom
/CD and infrastructure as code. Create reusable components and maintain clear technical documentation. Quality & Governance (10%) : Implement robust data validation, testing, lineage and observability to ensure high-quality, trusted datasets. Support governance and privacy-conscious data handling. Collaboration & Enablement (10%) : Partner with Data Science, MLOps, Product and commercial teams … cloud environments (preferably AWS) Engineering Best Practice: Knowledge of CI/CD, testing, version control and infrastructure as code Data Quality & Governance: Understanding of observability, validation and maintaining reliable data systems Collaboration & Communication: Ability to translate business and data science needs into scalable solutions and communicate clearly with stakeholders Mindset ...

Integration Architect - SAP SAAS Products

Location
Greater London, England, United Kingdom
4HANA Cloud, SuccessFactors, Ariba, Concur, Datasphere SAC) and SAP BTP. Define patterns, govern APIs/events, ensure secure, resilient data flows, and drive standardization, observability, and compliance. Core Responsibilities Architecture & Standards • Define canonical integration patterns (API-led, event-driven, batch/EDI) and reference architectures on SAP BTP. • Establish guidelines … SuccessFactors, Ariba, Concur, Datasphere, SAC, and S/4HANA Cloud integrations. • Strong security fundamentals (OAuth2, SAML, JWT, SCIM) and compliance awareness. • Hands-on with observability (Cloud ALM), performance tuning, and reliability engineering. • Experience with agile delivery, CI/CD (Git-based pipelines), and test automation for integrations. • Preferred Qualifications ...

Lead Cloud Network Engineer AWS - Hedge Fund

Location
Greater London, England, United Kingdom
network security controls, monitoring and architecture. You'll also design and maintain secure hybrid connectivity between AWS, Azure and on‐premises datacentres, with strong observability, monitoring and capacity planning across the wider network estate. Infrastructure as Code and automation will be central to the role. You'll build reusable Terraform …/CD experience, provisioning, pipelines, network tooling You have a good knowledge of Routing and Switching (BGP), VPNs (IPSec, SSL), Network Monitoring and Observability You're collaborative and pragmatic, able to push back where necessary What's in it for you: Competitive salary, to £140k Pension Private medical care ...

Staff Software Engineer - Customer Data Platform

Location
City of Westminster, England, United Kingdom
outcomes. Work with adjacent teams when needed to align on shared components and dependencies. Actively participate in on-call support, contributing to operational stability, observability, and performance. Coach and support other engineers through pairing, reviews, and mentoring, helping raise the team’s overall capability. Contribute to team OKRs and actively … that balance speed, maintainability, and long-term scalability. You actively identify technical debt, risks, or inefficiencies and take action to address them. You use observability and metrics to validate behavior, debug issues, and improve system health. You regularly support and unblock teammates, helping them deliver more effectively. You contribute ...

Platform Engineer, AI Enablement

Hiring Organisation
wayve
Location
London, UK
Employment Type
Full-time
language models and AI agents. You'll help create a governed model-access layer and a secure production environment for agentic workflows, with safety, observability, and operational excellence built in from the start. Your work will span the full lifecycle—from defining problems and designing systems to implementation, deployment … cost-effective access to AI models and tools. Build the runtime, services, and developer tooling used to run agentic workflows in production. Create observability across cost, performance, reliability, and usage through metrics, tracing, and logging. Implement governance and safety controls that make AI use secure, compliant, and auditable. Develop reusable ...

Software Engineer

Location
Greater London, England, United Kingdom
system. AI harnesses in production. Deploying agentic AI into real-world operational settings — acting on real money, tenancies and legal exposure, with the guardrails, observability and correctness that demands. Non-deterministic LLM working within compliant, secure deterministic software. Our stack We build on NestJS + TypeScript on GCP/… build Small, well-factored services. TDD and DDD as defaults. Trunk-based CI/CD — you ship to production and own it, with tests, observability and clean rollbacks. Lean frameworks, readable code, and we move fast because the tests and boundaries let us. What we're looking for #J ...

Platform Engineer, AI Enablement London, United Kingdom

Location
Greater London, England, United Kingdom
language models and AI agents. You’ll help create a governed model-access layer and a secure production environment for agentic workflows, with safety, observability, and operational excellence built in from the start. Your work will span the full lifecycle—from defining problems and designing systems to implementation, deployment … cost-effective access to AI models and tools. Build the runtime, services, and developer tooling used to run agentic workflows in production. Create observability across cost, performance, reliability, and usage through metrics, tracing, and logging. Implement governance and safety controls that make AI use secure, compliant, and auditable. Develop reusable ...

Lead AI Engineer

Hiring Organisation
Capco
Location
London, UK
Employment Type
Full-time
experience deploying LLMs and multi-modal models at scaleStrong engineering background in Python with proven backend and API development skillsSolid understanding of scalable MLOps, observability, and cloud-native AI deploymentExcellent communication, problem-solving, and project management skills in agile environmentsBonus Points ForExperience with agentic frameworks (e.g., LangChain, LlamaIndex)Experience … deep learning frameworks and front-end developmentFamiliarity with Langfuse, Langsmith, or other LLM observability toolsUnderstanding of Model Context Protocol and bias/hallucination mitigation techniquesPrevious success in integrating GenAI solutions into enterprise-scale systemsWhy Join CapcoDeliver high-impact technology solutions for Tier 1 financial institutionsWork in a collaborative, flat ...

Senior Software Engineer II

Location
Greater London, England, United Kingdom
engineering standards, best practices and reusable patterns while partnering with Enterprise Architecture and influencing technical direction Drive engineering excellence by improving code quality, testing, observability, reliability and operational practices Support end-to-end delivery by guiding teams through complex technical challenges, improving decision-making, and contributing to planning and risk … data lakes/lakehouse architectures, Iceberg or similar table formats, as well as batch and streaming processing Knowledge of data quality, governance, cataloguing and observability tools (e.g. Datadog), with DBT or AI-assisted engineering practices as a plus Additional Information Your benefits We're a community here that cares ...