3,276 to 3,300 of 3,971 Observability Jobs

AI-Driven QA Automation Engineer (C#/TypeScript)

Location
Greater London, England, United Kingdom
integrate quality gates into CI/CD, and diagnose production issues. You will migrate an existing C# Selenium framework to Playwright with TypeScript, enhance observability, and contribute to test strategy. #J-18808-Ljbffr ...

AWS Solutions Architect - Microservices & Event-Driven

Location
Cambridge, England, United Kingdom
translate complex requirements into production-grade solutions. You will define end-to-end architectures, lead domain-driven design, and set patterns for resilience, observability, and security. #J-18808-Ljbffr ...

Senior Cloud Infrastructure & Automation Engineer

Location
Bracknell, England, United Kingdom
platform reliability, automation, and scalable systems that power mission-critical CX and CCaaS services. You’ll design and implement IaC, strengthen security and observability, provide 3rd line support, coordinate zero-downtime deployments, and collaborate across disciplines to push forward a culture of continuous improvement. #J-18808-Ljbffr ...

Senior Java Infra Engineer for Elasticsearch Core Platform

Location
United Kingdom
easy to use, performant, maintainable, and secure. You will contribute as a strong individual contributor and lead cross-team projects, designing scalable features, improving observability, and collaborating with open source projects. This role emphasizes autonomy, design excellence, and the ability to communicate technical concepts across teams in a distributed company. ...

Platform Engineer – Analytics Platform & CI/CD

Location
Greater London, England, United Kingdom
seeking a Platform Engineer to join the Analytics Product Engineering team. You will design, build and operate platform capabilities, drive IaC, CI/CD, observability and security, and collaborate with software engineers, data scientists and IT #J-18808-Ljbffr ...

Senior AI Infra Engineer – LLM Ops & Reliability

Location
Auchentibber, Scotland, United Kingdom
serving stacks, optimize performance and costs, and lead incident response across cloud and on-prem GPU clusters. The role emphasizes secure software delivery, observability, and strong engineering fundamentals. You will collaborate with engineering to deliver scalable AI platforms, drive reliability, and build reusable patterns for robust production deployments, with ...

Backend Engineer - Cloud APIs (Remote UK)

Location
West of England, England, United Kingdom
mainly remote setup with occasional visits to Bristol or London. The role emphasizes building robust APIs, microservices, and scalable architectures, with a focus on observability, testing, and production readiness. #J-18808-Ljbffr ...

Lead AI Engineer: Architect Agentic & Scale-Ready ML Systems

Location
Greater London, England, United Kingdom
trust, and performance in production systems. You will collaborate with clients and product teams, apply LangGraph, Kubernetes, and cloud-native tooling, driving governance and observability to reduce hallucinations and latency. #J-18808-Ljbffr ...

Senior ML Engineer - Build ML Core for Payments (Remote)

Location
United Kingdom
turn promising ML ideas into shipped solutions and measurable outcomes. You will lead the development of scalable ML infrastructure, versioning, CI/CD and observability, while mentoring other engineers and clarifying technical direction in a rapidly evolving area. #J-18808-Ljbffr ...

Senior DevOps Engineer - Remote, FinOps & Compliance

Location
Greater London, England, United Kingdom
contribute to secure multi-account governance, IAM/SOC 2 alignment, and cost management while supporting a distributed engineering team. You will shape reliability, observability, and automation using Terraform, CDK, or Pulumi, in a fully remote setup with a global team and a strong focus on security and operational excellence. ...

Web App Developer - React/Python, Azure Cloud, Hybrid

Location
Greater London, England, United Kingdom
Vite frontends, contributing to reliability, performance, and developer experience. In this hybrid London role, you’ll collaborate with engineering teams on CI/CD, observability, security and scalable cloud deployments, with a strong focus on delivering robust software and reusable engineering patterns. #J-18808-Ljbffr ...

Senior Applied AI Engineer — Enterprise AI Platform

Location
Greater London, England, United Kingdom
lead the delivery of enterprise AI products and platform services. You will design, build, and productionise AI systems, govern them through testing and observability, and mentor fellow engineers to raise engineering standards. You will work with Azure AI Foundry, Accio, Nexus, and other components, shaping how the team operates ...

Enterprise AI Copilot Engineer (RAG & Azure)

Location
Leeds, England, United Kingdom
secure AI practices within a modern cloud-native stack. The role focuses on RAG architectures, vector databases, embeddings, and retrieval optimization, with emphasis on observability, governance, and reusable code assets. Collaborative, cross-functional work is expected. #J-18808-Ljbffr ...

Platform Engineer: Cloud-Native Networking Leader

Location
Greater London, England, United Kingdom
from the office five days a week in London. You will design and operate cloud-based networks, build Kubernetes operators in Golang, and improve observability and security across production systems. A strong background in cloud platforms and networking is essential. #J-18808-Ljbffr ...

Product Reliability Engineer – APM

Location
Greater London, England, United Kingdom
## Product Reliability Engineer - APMLondon, UK · Full-time#### About The PositionCoralogix is a modern, full-stack observability platform transforming how businesses process and understand their data. Our unique architecture powers in-stream analytics without reliance on expensive indexing or hot storage. We specialize in comprehensive monitoring of logs, metrics, traces … security events with features such as APM, RUM, SIEM, Kubernetes monitoring, and more, enhancing operational efficiency and reducing observability spending by up to 70%.We seek a **Product Reliability Engineer** who ensures that the Coralogix APM Product and Process exceed the quality and reliability standards, establish a competitive edge ...

Engineering Lead for Acquisition & AI Growth

Location
Greater London, England, United Kingdom
raise standards while shaping how AI tools are used across the team. You’ll own end-to-end delivery, drive CI/CD and observability improvements, and balance reliability with cost across AWS services. Hybrid work is offered in a dynamic, AI-driven environment. #J-18808-Ljbffr ...

AWS Platform Engineer - Build the Next-Gen Cloud Platform

Location
Langley Mill, England, United Kingdom
contract. You will design reusable AWS capabilities and automation that empower analytics, AI and digital services across the organisation. You’ll establish security controls, observability and scalable deployment patterns, collaborating with Architecture, Security and Data Engineering teams to deliver secure, reliable infrastructure at scale. #J-18808-Ljbffr ...

Senior HPC Network Engineer — RDMA, Leaf-Spine, Equity

Location
Greater London, England, United Kingdom
architecture through day-2 operations, including RDMA fabrics, leaf-spine design, and per-tenant isolation. You’ll automate provisioning with Python and Ansible, build observability dashboards, and mentor colleagues while expanding the data-centre and office networks. #J-18808-Ljbffr ...

Associate Site Reliability Engineer, SRE Platforms

Location
Greater London, England, United Kingdom
help build, run and continuously improve the Consolidated Trade Ledger platform in London. This role combines software and systems engineering to boost reliability, observability and incident response for a cloud-native service. The candidate will implement automation, maintain canary and blue/green deployment approaches, and contribute to scalable design ...

Senior Python Backend Engineer – Fintech API Systems

Location
Greater London, England, United Kingdom
Python, FastAPI, SQLAlchemy, and related technologies. Collaborate with cross-functional teams, own complex problems, and deliver production-ready code with strong emphasis on maintainability, observability, and reliability. #J-18808-Ljbffr ...

Lead Engineer: Greenfield Systems & Tech Strategy (Hybrid)

Location
Cardiff, Wales, United Kingdom
role focuses on Java (Spring Boot) on the backend, React/TypeScript on the frontend, PostgreSQL multitenancy, and AWS cloud hosting. You’ll drive observability and secure design across the stack. #J-18808-Ljbffr ...

Senior Data Platform Engineer - Lakehouse & Catalog

Location
Greater London, England, United Kingdom
Engineer to join the Data Lake and Catalog team in London. You will design, build, and operate scalable data lake services, focusing on reliability, observability, and end-to-end platform quality. You will mentor engineers, partner with cross-functional teams, and help evolve the lakehouse architecture using Hive, Spark ...

Senior Platform Engineer: Cloud Infra Leader (Remote)

Location
Greater London, England, United Kingdom
shared platform layer across B2B and B2C products in a high-growth fintech environment. You will build and improve scalable, secure cloud infrastructure, enhance observability and CI/CD, and collaborate with product and engineering teams to enable rapid, reliable releases. #J-18808-Ljbffr ...

Forward Deployed Engineer

Hiring Organisation
Noir
Location
London, South East England, United Kingdom
Employment Type
Full-Time
Salary
£100,000 - £140,000 per annum
Services - London/Hybrid (Tech stack: Forward Deployed Engineer, AI, Agentic AI, LLMs, LLM Integration, AI Agents, Retrieval, Tool Use, Context Orchestration, Evaluations, Guardrails, Observability, Human-in-the-Loop, Automation, Cloud, Data Management, System Integration, Forward Deployed Engineer) Are you a hands-on AI Engineer, Forward Deployed Engineer or technology … environments, combined with a strong understanding of modern AI systems and agentic workflows. Experience across LLM integration, retrieval, tool use, context orchestration, evaluations, guardrails, observability and human-in-the-loop workflows will be highly valuable. Experience within financial services, wealth management, banking, investment, risk, compliance, KYC or operations will ...