1,926 to 1,950 of 2,452 Observability Jobs in London

Staff Network Engineer

Location
Greater London, England, United Kingdom
root-cause analysis for performance and stability issues, and systematically reducing reactive toil through runbooks, automation, and measurable SLOs. Set the direction for network observability, telemetry, monitoring, and alerting to provide clear visibility into fabric health, performance, and traffic patterns. Ensure the accuracy and reliability of source-of-truth network … Juniper SRX and/or Palo Alto, including security policy architecture, high-availability design, and multi‐tenant segmentation. Experience designing network telemetry and observability for high-throughput, performance-sensitive environments. Proven ability to lead complex technical decisions and incidents across networking, systems, storage, and HPC/AI workload teams, with ...

Software Engineer - Software Delivery

Hiring Organisation
Neo4J
Location
London, UK
Employment Type
Full-time
internal engineers. This involves things likeWorking closely with internal engineers to identify pain pointsMaking sure the product experience is as good as possibleSetting up observability around how the platform is performing but also how users are interacting with the platformExperience creating abstractions to simplify the local developer workflow via tooling … related build tooling like kustomize or helm is also meriting. Experience integrating software with Google Cloud Platform, AWS and AzureSome experience in common software observability practices such as tracing, logging and metrics exporting.#LI-HybridWhy Join Neo4j Neo4j is, without question, the most popular graph intelligence platform in the world. ...

principal engineer- international technology & Starbucks digital solutions

Hiring Organisation
Starbucks Corporation
Location
London, UK
Employment Type
Full-time
ways of working across EMEA teams. Lead architecture and engineering reviews, providing guidance on solution design and implementation. Champion best practices in testing, observability, reliability, security, automation and operational excellence. Help teams reduce technical debt and improve platform health. Platform Strategy & Solution Ownership Partner with Product teams to deliver business … patterns, architecture and system design. Strong understanding of event-driven architectures, API-first design and integration patterns. Experience with DevOps, CI/CD, automation, observability and operational excellence practices. Strong understanding of security, resilience, scalability and performance engineering. Ability to challenge and assess vendor-supplied architectures and technical solutions. Experience ...

Senior Data Engineer

Location
Greater London, England, United Kingdom
data retrieval layers (pgVector, Pinecone) and ensure efficient embedding pipelines for AI contexts. Partner with AI teams to monitor data latency, cost efficiency, and observability metrics. Collaboration & Governance Partner with Platform Operations and Security to enforce privacy, compliance, and access control frameworks (GDPR, SOC2). Work cross‐functionally with analysts … platform data (Google Ads, Meta, TikTok, DV360, Amazon Ads). Demonstrated expertise in code versioning (GitHub), CI/CD integration, and data observability practices. Ability to write clean, modular, testable code and review peers’ contributions for maintainability and performance. Additional Information Publicis Groupe has fantastic benefits on offer ...

Senior II Product Engineer

Hiring Organisation
9fin
Location
London, UK
Employment Type
Full-time
Engineering, and our editorial and legal domain experts to scope work and ship the right thing. Improve developer experience by investing in tooling, testing, observability, and the paved road so the whole team moves faster. Ramp on legacy areas of the system, find the highest leverage cleanup, and execute … across a product domain. Experience contributing to the design of distributed systems in production, including the operational realities such as failure modes, observability, data consistency, and graceful degradation. A track record of solving scaling problems, whether database scaling, throughput, latency, or cost. You can talk through a real example ...

Agentic Commerce Architect

Hiring Organisation
Accenture
Location
London, UK
Employment Type
Full-time
intelligence — including embedding pipelines, vector retrieval, semantic search, and structured product data — ensuring agent outputs are accurate, grounded, and commercially reliable Embedding AI governance, observability, and responsible AI principles into architecture design — including audit logging, human-in-the-loop escalation points, and performance monitoring — as first-class concerns rather than … deployment, including familiarity with cloud-native services and CI/CD-based delivery practices Understanding of AI governance and responsible AI principles — including observability, tracing, auditability, and how to build guardrails into production AI systems Preferred Experience Experience contributing to technical proposals, reference architectures, or delivery accelerators in a consulting ...

Backend Software Engineer

Location
Greater London, England, United Kingdom
engineering, performance, and reliability problems that sit outside of their scope. Contribute to the shared foundations other engineers depend on: deployment pipelines, service templates, observability, and the environments their work runs in. You should apply if You have built and worked on production backend systems. You will spend your time … whether that means learning a new skill, reaching out to stakeholders to clarify requirements, or suggesting an alternate approach. You treat quality, security, and observability as engineering fundamentals rather than optional extras, and you flag technical debt rather than letting it accumulate silently. You actively seek feedback ...

Agentic Commerce Architect

Location
City Of London, England, United Kingdom
intelligence — including embedding pipelines, vector retrieval, semantic search, and structured product data — ensuring agent outputs are accurate, grounded, and commercially reliable Embedding AI governance, observability, and responsible AI principles into architecture design — including audit logging, human-in-the-loop escalation points, and performance monitoring — as first-class concerns rather than … deployment, including familiarity with cloud-native services and CI/CD-based delivery practices Understanding of AI governance and responsible AI principles — including observability, tracing, auditability, and how to build guardrails into production AI systems Preferred Experience Experience contributing to technical proposals, reference architectures, or delivery accelerators in a consulting ...

Principal Forward Deployed Engineer

Location
Greater London, England, United Kingdom
architecture and design decisions for the team’s solutions, keeping them secure, scalable, and aligned to the enterprise core stack. Take solutions through evaluation, observability, and our AI governance checkpoints, and keep them healthy afterwards. Own the roadmap and the technology choices Form part of the AI Technical Leadership Team … concepts such as fair value, relative value, signal generation, or portfolio construction. Snowflake, Microsoft Fabric/OneLake, Azure AI Foundry, model gateways, or AI observability in production. Vendor selection, commercial negotiation, or licensing at enterprise scale, or contributing to architecture and technology governance forums. Supervisory responsibilities Yes. This role line ...

Application Support Engineer (Third Party Systems), London

Hiring Organisation
Isomorphic Labs
Location
London, UK
Employment Type
Full-time
Isomorphic Labs is applying frontier AI to help unlock deeper scientific insights, faster breakthroughs, and life-changing medicines with an ambition to solve all disease. The future is coming. A future enabled and enriched by ...

Engineering Environments Lead

Hiring Organisation
Adecco
Location
London, United Kingdom
Salary
£ 70 K
Job ID: BROADBEAN_934211790027589Location: London, Greater LondonContract: ContractIndustry: ITRecruiter: Swati MylavarapuE-Mail: swati.mylavarapu.93421.9115@adeccops.aplitrak.comEngineering Environments Lead - (Platform & Environment Strategy)Location: London, Birmingham, Bristol, Manchester 12 months - Inside IR35Our client, a leading global technology company is ...

Data Platform Engineer

Hiring Organisation
ed Resourcing Ltd
Location
London, United Kingdom
Employment Type
Permanent
Salary
GBP 70,000 - 80,000 Annual
Data Platform Engineer £70,000 - £75,000 + 10% bonus + excellent benefits We're looking for a hands-on Data Platform Engineer who enjoys getting stuck into infrastructure, automation and platform engineering. This isn ...

Azure Platform Engineer - DevOps, IaC & Observability

Location
Greater London, England, United Kingdom
technical delivery. You will be designing, deploying and supporting secure, scalable Azure environments, building CI/CD pipelines, IaC and automation, and ensuring observability, resilience and secure governance across Azure. #J-18808-Ljbffr ...

Senior SRE Technical Lead — Reliability & Observability

Location
Greater London, England, United Kingdom
seeking a Technical Lead SRE in Greater London. In this role, you will enhance the reliability engineering capabilities, collaborating with various teams to establish observability standards and ensure operational excellence. The ideal candidate will have over 10 years of experience in SRE or related fields, strong AWS and Kubernetes skills ...

Senior Platform SRE: Reliable, Scalable Observability

Location
Greater London, England, United Kingdom
team to own availability and performance of mission-critical services, and to lead on-call and incident response. You will build tooling, improve observability, and drive platform scalability in collaboration with product teams, while lowering costs and improving developer productivity. The ideal candidate has around 8 years of distributed systems ...

Kubernetes SRE: GitOps, Canary Deployments & Observability

Location
Greater London, England, United Kingdom
secure deployments across dev, staging, and production. Embedded in a hybrid team, you will apply AI-assisted tooling to accelerate delivery and collaborate on observability, policy, and data services for #J-18808-Ljbffr ...

Observability Engineer (Dynatrace) — Telemetry & Performance

Location
Greater London, England, United Kingdom
Computacenter is seeking a Monitoring & Observability Engineer (Dynatrace) to design, implement and manage observability across customer IT estates in the UK. You will collect telemetry, diagnose issues and drive proactive improvements across teams. The role requires strong Dynatrace/Grafana/Splunk experience, scripting skills, cloud familiarity (Azure/ ...

Cloud Advisory Architecture Associate Manager

Hiring Organisation
Accenture
Location
London, UK
Employment Type
Full-time
where GenAI and Agentic play a role. Champion system performance, resilience, and efficiency: Proactively identifying and addressing consumption and scalability challenges. Champion full stack observability using modern full stack observability, SRE and AIOps. Manage & Mentor: Lead teams of architects and engineers, providing technical coaching, career counselling, performance management, and coaching ...

Service Reliability Support Manager

Location
Greater London, England, United Kingdom
batch monitoring and major incident command. It works in partnership with business-aligned Application Support teams and the Technology Centre of Excellence to embed observability engineering, telemetry, and automation — including AI-assisted triage and self-healing — across the estate. Role Summary The Service Reliability Lead owns Marex's central, cross … reactive, manual, ticket-driven model into an engineering-first Service Reliability capability, in which AI-assisted triage and automation absorb routine work and observability is owned as an engineering discipline connected to the Technology Centre of Excellence. The role provides oversight of the existing team, ensuring all current responsibilities continue ...

Head of Digital Platform Enablement

Hiring Organisation
S Merrick LTD
Location
Central London, London, United Kingdom
Employment Type
Permanent
large global enterprise. The role is effectively the Product Owner for Digital Platform Enablement and spans three core pillars: ServiceNow/service management platforms, observability/monitoring, and developer/engineering platforms. The client is moving from a traditional project-led, design-build-run model towards a product-centric, agile … Skills/Experience for the Head of Digital Platform Enablement role: Meaningful ownership of ServiceNow/service management platforms, workflows, automation and self-service Observability strategy across user experience, applications, cloud, infrastructure and networks Developer platforms and tooling including Azure DevOps and/or GitHub CI/CD, Infrastructure ...

Principal Software Engineer-AI

Location
Greater London, England, United Kingdom
production or in platforming (LLM or MCP gateway, agentic runtime, auth, data retrieval, eval tooling) Experience running AI systems in production at scale, including observability, cost and capacity planning, regression detection, and incident response for AI-powered applications Experience operating production distributed systems on AWS/Azure, with a strong … grasp of reliability, observability, and incident response at scale Deep knowledge of cloud-native technologies, serverless applications, event-driven architectures, data and inference pipelines, relational, NoSQL, and vector databases, and modern software architecture patterns Proven track record of owning multi-year technical strategy and architectural roadmaps, guiding teams from ...

Lead Network Operations Engineer

Hiring Organisation
G Research
Location
London, UK
Employment Type
Full-time
organisation's network and security infrastructure across datacentre and office environments. This is a hands-on technical leadership role focused on operational excellence, automation, observability, and incident response — ensuring high availability, resilience, and a strong security posture. Key responsibilities: Own the day-to-day performance, stability, availability, and security … campus environmentsLead response to critical incidents, driving rapid diagnosis, containment, and long-term remediationChampion an automation-first approach, reducing operational toil through tooling, observability, and emerging technologies (including agentic AI)Develop and refine operational tooling, runbooks, and incident response frameworks to ensure consistent, high-quality service deliverySupport an event-driven ...

Front End developer

Location
Greater London, England, United Kingdom
applications. You’ll work across Python services and React/Vite frontends , collaborating closely with engineering teams to improve application reliability, CI/CD, observability, security and developer tooling. Key experience: Docker and CI/CD pipelines Azure cloud experience is essential; AWS is a bonus Good knowledge of Linux … networking, security and observability Experience supporting scalable application platforms and cloud deployments Able to work independently and establish reusable engineering patterns and best practices A great opportunity for someone who enjoys working across development, cloud and platform engineering. Additional Information #TalanUK #J-18808-Ljbffr ...

Data Reliability Engineer

Hiring Organisation
Ashdown Group
Location
London, UK
Employment Type
Full-time
work from home 2 days per week. This is a high-impact role focused on improving data quality, reducing incidents, and building scalable observability across a modern enterprise data platform. You'll help ensure data across the organisation is accurate, reliable, and trusted for critical business decision-making. … style roles, with strong SQL and Python skills and experience working in modern cloud-based data environments. Hands-on experience with data observability tools such as Grafana, Monte Carlo, or Acceldata, and data governance/quality platforms like Informatica, Collibra or Microsoft Purview is highly desirable. Experience within the Azure ...

Senior Software Engineer – Agentic Development Enablement

Location
Greater London, England, United Kingdom
such as Claude Code and GitHub Copilot Design and implement guardrails, controls, and engineering patterns for AI-assisted development Contribute to endpoint and platform observability, telemetry, and policy enforcement Define how controls work consistently across local development environments and CI/CD pipelines Explore changes to development environments, including containerised … production Requirements Strong hands-on background as a software engineer Broad technical understanding across developer tooling, cloud platforms, operating systems, desktop environments, security controls, observability, telemetry, and CI/CD Experience working on developer workflows and engineering ways of working, not only end-user application delivery Ability to work ...