1,826 to 1,850 of 2,250 Observability Jobs in London

Senior Developer Experience Security Engineer

Location
Greater London, England, United Kingdom
experience across Motorway.. We have recently built and rolled out a new container platform on top of AWS Fargate, and are currently enhancing our observability, reliability, and developer-focused tooling. We will continue to build and evolve secure, standardised platform capabilities that reduce cognitive load and help teams ship faster … experience across Motorway.. We have recently built and rolled out a new container platform on top of AWS Fargate, and are currently enhancing our observability, reliability, and developer-focused tooling. We will continue to build and evolve secure, standardised platform capabilities that reduce cognitive load and help teams ship faster ...

Software Engineer III - Fullstack (Java, React, Python and AI) Engineer

Hiring Organisation
JP Morgan Chase
Location
London, UK
Employment Type
Full-time
automationCollaborate with product, marketing, and partners to translate requirements into well-designed technical solutionsImprove engineering excellence through code reviews, test automation, CI/CD, observability, performance tuning, and operational best practicesEnsure solutions meet security, privacy, and compliance expectations including consent management, data minimization, and access controlsRequired Qualifications, Capabilities, and Skills … Skills: Experience with campaign management and attribution systemsBackground in AI-enabled content generation, experimentation, or workflow automationSkills in test automation, CI/CD, and observability toolsExperience optimizing performance and scalabilityFamiliarity with consent management and data minimization practicesAbility to drive innovation in personalization and measurementExperience working in cross-functional teamsJ.P. Morgan ...

Senior Data Engineer

Location
Greater London, England, United Kingdom
Insight teams to maintain reliable, accurate and meaningful data models as our products evolve Establish and promote good engineering practices across data pipelines, modelling, observability, performance and reliability Ensure data quality and integrity throughout our pipelines, identifying and addressing issues before they impact downstream users Provide technical guidance to engineers … data transformation skills, particularly using technologies such as Python, Spark and SQL Experience working with cloud-based data platforms Experience designing systems for reliability, observability, performance and data quality Strong understanding of data modelling and the principles behind well-structured analytical data Experience making technical design decisions and taking ownership ...

Data Architect

Location
Greater London, England, United Kingdom
duties, isolation, and alignment to standards such as ISO 27001. Provide architectural direction for cloud data platforms, infrastructure as code, CI/CD, and observability, ensuring designs are cost‐aware and operable. Define the data lineage and controls that support responsible production AI — monitoring, confidence scoring, and human … approaches, and how each changes data architecture requirements, including context window economics and prompt caching. Familiarity with infrastructure as code, CI/CD, and observability for data platforms. Awareness of MCP and tool‐orchestration concepts and their implications for composable, data‐driven AI systems. Fluent use of AI‐assisted development ...

Principal Engineer

Location
Greater London, England, United Kingdom
role moves between deep technical detail and a 10,000 foot view, often on the same day. What You’ll Be Working On Reliability, observability and operational correctness The priority is not simply whether the platform is up, but whether it is doing its job: client feeds arrive completely, surveillance … proportion of engineering time spent on elective work rather than keeping the lights on. That means: Client onboarding becomes increasingly automated end to end. Observability lets teams answer common operational questions without elevated production access. Change is frequent, small and reversible, and incidents do not scale with deployment frequency. Product ...

Senior Network Engineer

Location
Greater London, England, United Kingdom
access patterns. Support hybrid connectivity models (site‐to‐site VPN, client VPN, ExpressRoute, Direct Connect, SD‐WAN). Monitor network performance and reliability using observability and telemetry tools; proactively address capacity and performance issues. Troubleshoot complex network and cross‐domain infrastructure issues spanning network, compute, and cloud layers. Develop … constructs (VPC/VNet design, routing, security groups, load balancers). Experience with SD‐WAN architectures and implementations. Familiarity with network monitoring, logging, and observability tools (e.g., SNMP, NetFlow, Syslog, modern NPM tools). Working knowledge of compute platforms and operating systems (Windows, Linux, virtualization such as VMware/Hyper ...

Principal Software Engineer - Squad Lead Engineer

Location
Greater London, England, United Kingdom
complete complex bug fixes and performance improvements* Define and uphold Definition of Ready/Done including code quality, automated test coverage, security checks, and observability* Establish/maintain CI/CD pipelines, quality gates, and sensible branching/release strategies* Drive a pragmatic quality strategy: test pyramid balance, contract tests … Windows* Experience with relational and non-relational data stores, performance tuning, and data modelling* Knowledge of CI/CD platforms, containers, cloud technologies, observability, and monitoring practices* Understanding of secure coding, performance optimisation, reliability engineering, and incident response **Work in a Way That Works for You**We promote a healthy ...

AI Platform Engineer

Location
Greater London, England, United Kingdom
patterns that support the safe and scalable adoption of AI-assisted development.Improve developer experience through streamlined workflows, tooling integration and self-service capabilities. Develop observability and measurement capabilities to provide insights into engineering productivity, quality and platform adoption. Collaborate with Technical Leads and AI Software Engineers to identify recurring engineering … cost optimisation, including token monitoring, caching strategies and model selection considerations. Knowledge of approaches for managing and reducing token consumption costs.Deep understanding of observability, automated testing and software delivery tooling. Knowledge of platform security, governance and operational controls. Strong programming, automation and problem-solving capabilities. Passionate about improving developer experience ...

Systems & Datacenter Administrator (Ops) — London

Location
City Of London, England, United Kingdom
VyOS. Applications & Data: MySQL, Elasticsearch, Kafka, Java, Apache HTTPD, ... Automation & IaC: Git/GitLab, Chef, Terraform; scripting with Bash/Python. Monitoring/Observability: Centreon, Observium (plus logs pipelines). What You’ll Do Operate and improve Linux fleets (Ubuntu) in production. Manage LXD/LXC container platforms … Elasticsearch, Kafka, Java services, Apache; ability to collaborate with app teams on infra‐adjacent issues. Experience with Centreon and Dynatrace (or equivalent monitoring/observability stacks). Config management/IaC depth (Ansible, Puppet, Terraform modules, Secret management), and CI pipelines in GitLab. Deeper networking (EVPN/VXLAN, BGP, multicast ...

Senior Lead Software Engineer - Mobile Engineering

Hiring Organisation
JP Morgan Chase
Location
London, UK
Employment Type
Full-time
drive measurable improvements in stability and release confidence. Own mobile build/release and operational maturity: CI/CD pipelines, distribution, feature flags, observability, crash/performance monitoring, and incident response. Mentor and coach engineers; support team growth through feedback, technical guidance, and strong engineering culture. Communicate clearly with senior … patterns, accessibility standards, component libraries).Experience with CI/CD for mobile (e.g., build automation, signing, distribution, feature flags, release trains).Experience with observability and production support practices: crash analytics, performance monitoring, logging, alerting, and operational readiness. Experience leading multiple engineers/teams (people leadership or strong matrix leadership), including ...

Principal Engineer - Integration Services

Location
Greater London, England, United Kingdom
duplication and improve interoperability. Providing technical leadership on significant integration decisions spanning multiple platforms, teams and domains. Ensuring integration approaches consider security, resilience, scalability, observability, operability and long‐term maintainability. Working with Architecture and senior engineering leaders to align integration direction with wider FT technology strategy. Establishing a clear view … progress, risks, trade‐offs and technical constraints. Operational Excellence & Reliability Owning the operational performance of shared integration services. Establishing appropriate approaches to monitoring, observability, incident management, problem management and operational readiness. Ensuring critical integrations have clear ownership, appropriate resilience and effective recovery mechanisms. Leading the response to significant incidents affecting ...

Principal Software Engineer - Customer Platforms

Location
Greater London, England, United Kingdom
customers to a desired outcome, without prescribing it Authoritative skills at cloud computing (network, security, serverless, Kubernetes etc) and automation Experience with implementation of Observability and Reliability using market technologies (e.g.: New Relic) Good experience with Performance Engineering (load testing, derivations, tuning, core web vitals, page speed etc.) Expertise … organisation(s) Tech Stack M&S uses a variety of technologies including; Java, Spring, SpringBOOT, Micronaut React, Next.js, Typescript, Angular Azure Cloud, Kubernetes, Dynatrace (observability) SQL Server, MongoDB Ignite, Redis What’s In It For You Working at M&S means being part of something bigger - helping to deliver quality ...

Applied AI Engineer

Hiring Organisation
Nufin
Location
London, UK
Employment Type
Full-time
business impact. Create representative test datasets, evaluation criteria, regression tests, and human-review processes. Measure and improve accuracy, latency, cost, and user experience. Establish observability and feedback loops that make agent behavior understandable and continuously improvable. Develop the AI Application ArchitectureApply context-engineering techniques such as RAG, MCP and knowledge … LlamaIndex, or comparable frameworksContext engineering: RAG, MCP, knowledge graphs, tool use, memory, and retrieval systemsAI evaluation: offline and online evaluations, test datasets, regression testing, observability, and human reviewLanguage models: Gemini, OpenAI, Anthropic, Llama, Mistral, or similarBackend engineering: Python or Java, REST APIs, Kafka, microservices, and distributed systemsData systems: SQL, PostgreSQL ...

Principal Engineer, AI Engineering

Hiring Organisation
Wells Fargo
Location
London, UK
Employment Type
Full-time
chatbots, RAG & Knowledge platforms, Skill based agents, workflow automation, and AI-enabled developer experiences Define engineering standards for prompt engineering, skill engineering, model evaluation, observability, guardrails, red teaming, hallucination detection, and responsible AI adoption. Partner with engineering leads, product owners, cybersecurity, governance, risk, platform, and UI/UX teams … target-state architecture, reusable design patterns, and technical standards for AI systems. Guide architecture decisions across model integration, retrieval design, orchestration, workflow automation, security, observability, and application experience layers. Evaluate emerging AI technologies and translate them into practical, governed, production-ready engineering patterns. Engineering Excellence Lead proof-of-concepts, design ...

Forward Deployed Engineer - Lead AI & Agentic Engineer

Location
City Of London, England, United Kingdom
infrastructure yourself Define AI governance and responsible-use guardrails: data privacy boundaries, LLM governance, and policy enforcement (OPA/Rego) Instrument AI applications for observability (OpenTelemetry, Prometheus) so behaviour and cost stay visible in production Translate enterprise requirements into AI solution roadmaps for senior stakeholders Capture field learnings, codify reusable … Degree in Computer Science, Data Science, Informatics, Engineering, Physics, Mathematics, or a related discipline — or equivalent professional experience Preferred Skills and Experience Familiarity with observability instrumentation for AI systems (OpenTelemetry, Prometheus) Experience with AI/ML frameworks (TensorFlow, PyTorch) and the Hugging Face/open‐source AI ecosystem Experience with ...

Staff Software Engineer

Location
Greater London, England, United Kingdom
from problem framing and design review through to production rollout, migration and decommissioning of what came before Drive operational excellence by improving observability, alerting, capacity planning, load and chaos testing, and incident response, and by closing the loop on the root causes you find Raise the engineering bar across teams … development (we use AWS & Azure), containers, orchestration and infrastructure as code Experience owning production systems on call, with a strong sense of ownership for observability, monitoring and incident response Experience leading complex migrations or re-architectures of live, business-critical systems with no loss of availability or data integrity Excellent ...

Lead Mobile Engineer

Hiring Organisation
Appcast
Location
London, UK
release, and deployment processes, including CI/CD pipelines and app store submissions for Google Play and Apple App Store.Champion automated testing, app stability, observability, and the responsible use of AI to improve developer productivity and software quality.Knowledge, skills and experience required7+ years of professional software engineering experience, with … quality, with experience using mobile testing frameworks and automated testing approaches.Experience owning mobile CI/CD pipelines, build automation, and app store release processes.An observability mindset, with experience using crash reporting, performance monitoring, and analytics tools to improve app reliability.Comfortable working with Git, Jira, Confluence, and modern agile engineering workflows.Proven ...

Staff Software Engineer

Hiring Organisation
Checkout.com
Location
London, UK
Employment Type
Full-time
from problem framing and design review through to production rollout, migration and decommissioning of what came beforeDrive operational excellence by improving observability, alerting, capacity planning, load and chaos testing, and incident response, and by closing the loop on the root causes you findRaise the engineering bar across teams through design … application development (we use AWS & Azure), containers, orchestration and infrastructure as codeExperience owning production systems on call, with a strong sense of ownership for observability, monitoring and incident responseExperience leading complex migrations or re-architectures of live, business-critical systems with no loss of availability or data integrityExcellent communication skills ...

Principal Software Engineer - Squad Lead Engineer

Location
Greater London, England, United Kingdom
complete complex bug fixes and performance improvements Define and uphold Definition of Ready/Done including code quality, automated test coverage, security checks, and observability Establish/maintain CI/CD pipelines, quality gates, and sensible branching/release strategies Drive a pragmatic quality strategy: test pyramid balance, contract tests … Windows Experience with relational and non‐relational data stores, performance tuning, and data modelling Knowledge of CI/CD platforms, containers, cloud technologies, observability, and monitoring practices Understanding of secure coding, performance optimisation, reliability engineering, and incident response Work in a Way That Works for You We promote a healthy ...

Software Engineer - Affiliate Operations

Location
Greater London, England, United Kingdom
remain highly reliable while evolving to support new markets and acquisition channels. Reliability is fundamental to everything we build. We invest heavily in automation, observability, and operational excellence to reduce manual effort. We are also exploring how AI can transform our engineering productivity and marketing platforms. We operate … improve platform effectiveness, launch new capabilities, and reduce manual operational effort across the business. Improve You'll continuously improve our systems through better observability, automation, and thoughtful refactoring. You'll help evolve our architecture and engineering practices to ensure our platforms remain resilient as they scale. Own You'll take ...

Technical Lead, Lending & Savings

Hiring Organisation
Blockchain
Location
London, UK
Employment Type
Full-time
performance, security, and maintainability. Remain hands-on, contributing production-quality code and reviewing critical changes. Drive engineering best practices across system design, testing, deployment, observability, and operational excellence. Mentor engineers through code reviews, technical coaching, and day-to-day leadership. Engineering DeliveryOwn the delivery of technical initiatives from design through … design. Experience working with Redis or other NoSQL technologies. Deep understanding of microservices architecture, APIs, distributed systems, and cloud-native applications. Experience with monitoring, observability, incident response, and production operations. Strong debugging and performance optimisation skills. Demonstrated experience shipping reliable production systems that process financial transactions. LeadershipExperience leading engineering teams ...

Senior Full Stack Engineer (Realtime & Voice) Customer Experience Platform

Location
Greater London, England, United Kingdom
Build the safety and compliance plumbing enterprise partners audit, including guardrails, content filtering, and PII redaction integration points Keep revenue-critical deployments healthy through observability, alerting, incident response, and SLA performance Build the platform capabilities forward-deployed engineers configure for partner telephony integrations and go-lives Raise the engineering … standard part of their workflow, with the judgment to review, correct, and own everything that ships An operable-systems mindset, covering SLAs, observability, on-call rotations, and rollback plans The ability to break down complex problems, make pragmatic tradeoffs, and ship iteratively, backed by strong communication across product, design ...

Senior Applied AI Engineer (Defence Contractor)

Location
Greater London, England, United Kingdom
data ingestion through to inference, owning the whole path rather than a slice of it. Make confidence earned, not asserted. You build the evaluation, observability and guardrails that show how a system actually behaves, its agent behaviour, model performance and failure modes. Set the technical bar. … reasoning Experience with edge or offline AI deployments Familiarity with Kubernetes (EKS/OpenShift) for managing deployed applications MLOps experience: model evaluation, monitoring, reproducibility Observability tooling for agentic systems (model drift, agent behaviour, performance monitoring) Experience with agent orchestration patterns and inter‐agent communication protocols (e.g. A2A) Familiarity with ...

Senior Network Engineer

Location
Greater London, England, United Kingdom
Using BGP, EVPN-VXLAN, JunOS (Juniper QFX/MX), OSPF, Spine-Leaf/IP Fabric, VLANs, VRFs, Linux networking, Ansible, Terraform, Prometheus/Grafana, Observability tooling The adventures that await you after becoming Senior Network Engineer at Hack The Box: Design and implement spine-leaf network architectures across multiple data … series) running JunOS Develop and maintain network automation using Ansible, Terraform or similar infrastructure‐as‐code tooling Establish and improve network monitoring, alerting and observability (Prometheus, Grafana, SNMP, streaming telemetry) Plan and execute network capacity upgrades, site bring‐ups and hardware refresh cycles Collaborate with the platform/systems engineering ...

Endpoint Engineer

Location
Greater London, England, United Kingdom
platform engineering across Windows, Microsoft Intune, and related technologies. This role delivers secure, reliable, and frictionless user experiences through modern device management practices, automation, observability, and Zero Trust principles. How You’ll Make an Impact: Endpoint Platform Engineering (Windows CSP/Intune) Own the migration of legacy Group Policy configurations … e.g., Patch My PC). Strong PowerShell automation expertise and familiarity with Git‐based workflows. Experience with vulnerability management tools and endpoint telemetry/observability platforms. Experience supporting local AI/LLM developer tooling (e.g., Claude Code, Ollama, LM Studio) and GPU‐accelerated workstations. Microsoft certifications related to Endpoint, Azure ...