1,051 to 1,075 of 2,373 Observability Jobs in London

Platform Compute Specialist

Hiring Organisation
SQUAREPOINT CAPITAL
Location
London, UK
Employment Type
Full-time
platforms that support our high-performance, data-intensive workloads. You'll work closely with global engineering teams to design and implement infrastructure-as-code, observability pipelines, and self-healing systems. This is a hands-on engineering role with a strong emphasis on automation, performance tuning, and developer enablement. Main Duties … APIs to abstract infrastructure complexity and improve productivity. Take part in bulk server provisioning which may include using remote hands to complete tasks. Drive observability initiatives by building and integrating telemetry pipelines (metrics, logs, traces).Collaborate with software engineering teams to ensure infrastructure supports application scalability, reliability, and security. Mentor ...

Senior Site Reliability Engineer

Hiring Organisation
Brevan Howard
Location
London, UK
Employment Type
Full-time
operational efficiency. Infrastructure-as-Code (IaC): Own and evolve our declarative infrastructure using Terraform for cloud resources and Helm for Kubernetes application deployment. Monitoring & Observability: Implement and manage robust monitoring, alerting, and logging solutions to ensure clear system visibility and proactive issue identification. Reliability & Performance: Define, measure, and enforce Service … major programming language, preferably Python, for automation and tool development. Tooling & ConceptsCI/CD: Experience setting up and maintaining modern CI/CD pipelines. Observability: Practical experience implementing and managing monitoring and logging tools. Networking: Solid understanding of TCP/IP, load balancing, DNS, and cloud-native networking within Kubernetes. ...

Senior Cloud Platform DevOps Engineer

Location
City Of London, England, United Kingdom
robust enterprise cloud security practices, ensuring comprehensive data protection, secure network topologies, and compliance as the platform scales. Implement and enhance comprehensive platform telemetry, observability, logging, and monitoring systems to guarantee high system reliability and performance. What we're looking for: Strong proficiency in Python and modern web frameworks (FastAPI … cloud providers (AWS, GCP, Azure) and local physical servers. Experience integrating or orchestrating generative AI workflows (ComfyUI, LoRAs, LLMs, diffusion models). Knowledge of observability, LLM evaluation, and prompt tracing frameworks (Langfuse, Langgraph). Experience working within an Agile environment. Understanding of GPU workload scaling and compute optimisation ...

Senior Data Platform Engineer - Data Enablement

Location
Greater London, England, United Kingdom
engineering.depop.com/What You’ll Do Pave a path for data as product: Champion data as a first‐class citizen by introducing robust data observability and governance tooling into the platform Software engineering: Develop microservices, libraries, data pipelines Technical design: implement and evolve platform services that enable teams to work … automation‐first mindset Experience delivering data compliance & privacy solutions, ensuring that we uphold data subject rights for our customers Experience introducing a data governance & observability stack enabling rich data lineage, data contracts, SLA/SLO, tagging and data quality monitoring capabilities both on our own platform but also for data ...

Engineering Manager

Location
Greater London, England, United Kingdom
features and manage the dates, by communicating and delivering on time. Conduct interviews on the hiring processes. Drive CI/CD, test automation, and observability practices across the team. Collaborate with product managers to align priorities and solutions. Observe and enforce the standards set by the Architects. Provide Level … Engineering, or related field. Hands-on expertise with AWS architecture, serverless services, EKS computing, and event-driven design. Experience with CI/CD systems, observability, and infrastructure-as-code (e.g., Terraform). Fluency in one or more backend languages (Java, Python, Node.js). Fluency in SQL (MySQL, Oracle, PostgreSQL ...

Senior Cloud Platform Engineer

Hiring Organisation
Liberis
Location
London, UK
Employment Type
Full-time
on. Build the foundations for AI agents: integrations with Claude, the tools and services agents call, and the orchestration that ties them together. Wire observability, guardrails, and cost controls into the AI agents and tooling we run, so they stay safe and predictable in production. Work hands-on with … Docker) and deployment on Kubernetes or Cloud Run. Infrastructure-as-code with Terraform. Proficient in at least one of .NET (C#), Node.js, or Python. Observability experience with Datadog or cloud-native logging and monitoring (GCP Cloud Operations or Azure Monitor).A track record of driving adoption — building things engineers ...

Lead Platform Operations Engineer

Location
Greater London, England, United Kingdom
maintain platform standards, patterns, and best practices Own Platform Reliability, Security & Performance Lead incident response, root cause analysis, and platform improvements Implement robust monitoring, observability, and alerting strategies Drive security improvements aligned to ISO27001, SOC2, and modern SDLC practices Ensure strong governance across infrastructure, applications, and data Deliver Scalable & Secure … Experience implementing security tooling (SAST, DAST, container scanning, WAF) Strong knowledge of cloud security, encryption, TLS/SSL, certificates, and access control Experience with observability, monitoring, and alerting tools Security & Compliance Practical experience implementing ISO27001 and SOC2 controls Knowledge of OWASP methodologies and secure development lifecycle practices Experience with vulnerability ...

Sr. Technical Lead / Architect

Location
Greater London, England, United Kingdom
integration patterns, messaging, and event-driven architectures; Produce high-level and low-level design documents; Ensure applications meet non-functional requirements including availability, reliability, observability, and security; Drive API-first architecture and reusable component design. Software Development Lead full-stack application development using: Python, SQL, AWS, React JS; Contribute … implement cloud-native architectures; Collaborate with DevOps teams to automate deployments and infrastructure; Define CI/CD pipelines and release strategies; Improve application observability using logging and monitoring tools; Ensure infrastructure scalability and operational excellence. Quality & Security Promote secure coding practices; Ensure applications comply with security and compliance standards; Drive ...

Forward Deployed Engineer

Location
Greater London, England, United Kingdom
plus.* Hands-on experience with cloud platforms (AWS), Docker and Kubernetes; CI/CD in GitLab (or equivalent); infrastructure-as-code (e.g. Terraform); observability and monitoring stacks.* Solid understanding of database systems, SQL, data modelling, ETL pipelines, REST/gRPC APIs and microservices architecture; identity and access management with OIDC …/SAML and Azure Entra ID.* Experience with LLMs, RAG systems, prompt engineering and AI evaluation frameworks; familiarity with MLOps, model deployment and AI observability/guardrails; working use of AI-assisted development tools (e.g. Claude Code).* Familiarity with one or more of: trading, ERP or treasury platforms; workflow ...

Senior Site Reliability Engineer

Location
Greater London, England, United Kingdom
operational efficiency.* Infrastructure-as-Code (IaC): Own and evolve our declarative infrastructure using Terraform for cloud resources and Helm for Kubernetes application deployment.* Monitoring & Observability: Implement and manage robust monitoring, alerting, and logging solutions to ensure clear system visibility and proactive issue identification.* Reliability & Performance: Define, measure, and enforce Service … programming language, preferably Python, for automation and tool development.**Tooling & Concepts*** CI/CD: Experience setting up and maintaining modern CI/CD pipelines.* Observability: Practical experience implementing and managing monitoring and logging tools.* Networking: Solid understanding of TCP/IP, load balancing, DNS, and cloud-native networking within Kubernetes. ...

Snr Lead Software Engineer - Developer Experience Engineering

Hiring Organisation
JP Morgan Chase
Location
London, UK
Employment Type
Full-time
Senior Lead Software Engineer at JPMorganChase within the International Consumer Bank, you will lead a team building and operating a first-class observability capability across our cloud-native microservices. Collaborating closely with product, platform, and development teams, you will influence the technical roadmap, architect, standardize, and build resilient, cost-efficient … SlackHands-on experience designing and implementing cloud multi-region architectures in production environmentsProficiency in one or more additional programming languages beyond primary experienceFamiliarity with observability platforms and telemetry tooling for metrics, logs, and distributed tracing at scale#ICBCareersJ.P. Morgan is a global leader in financial services, providing strategic advice and products ...

DevOps Engineer

Location
Greater London, England, United Kingdom
operations. Participate in on‐call rotation, incident response, and post‐incident reviews, driving toward blameless root‐cause analysis and durable fixes. Implement and maintain observability across infrastructure and applications (metrics, logs, traces, dashboards, and alerting). Apply DevSecOps practices by embedding security checks into the development lifecycle (SAST/DAST … Vault, SOPS) Understanding of compliance frameworks relevant to cloud environments and experience with policy‐as‐code frameworks for automated compliance and guardrails. Exposure to observability platforms such as Datadog, Prometheus/Grafana, or the OpenTelemetry ecosystem. Experience with container image hardening and scanning (Trivy, Grype, or similar). Experience with ...

Senior Software Engineer (Node.js, TypeScript)

Hiring Organisation
Diligent
Location
London, UK
Employment Type
Full-time
reliable solutions. Use AI tools to accelerate coding, debugging, testing, research, and documentation, while validating outputs carefully and applying sound judgment. Strengthen service reliability, observability, and engineering quality by improving monitoring, incident response, testing, and development practices. These are the essentials you'll need to get an interview … these too, but we'll support you if you don't: Familiarity with infrastructure as code, for example CDK or Terraform. Exposure to observability tooling, incident response, or production monitoring practices. Experience working with React or Angular when contributing to end-to-end product delivery. #LIHybridAbout UsDiligent ...

Principal Engineer I — Prepurchase Platform

Location
City Of London, England, United Kingdom
decisions and evolve engineering standards across the Prepurchase domain. Partner with product, security, and SRE teams to align technical decisions with business priorities. Drive observability improvements - ensuring services are instrumented for monitoring, alerting, and rapid incident response. Identifyandeliminatesingle points of failure, improving systemreliabilityand reducing on-call burden. Apply … Proficiencywith AI/ML tools and techniques - using LLMs, AI-assisted development, and automation to accelerate engineering workflows and improve system intelligence. Familiarity with observability tooling: Grafana, Splunk, Prometheus,OpenTracing, or equivalent. Strong understanding of security best practices - OAuth/OIDC, input validation,secretsmanagement. Track recordof leading technical initiatives across ...

Senior Java Developer

Hiring Organisation
London Stock Exchange Group
Location
London, UK
Employment Type
Full-time
Role ProfileThe Senior lead Scala Engineer will play a key role in the engineering and delivery of a critical market‐infrastructure service within the Equities platform, with a strong focus on AWS cloud‐native development. ...

Software Engineer — Observability Instrumentation

Hiring Organisation
G Research
Location
London, UK
Employment Type
Full-time
high-impact research - designing systems that scale, accelerate discovery and support innovation across the firm. Take the next step in your career. The roleThe Observability Engineering Team manages access to G-Research's telemetry platforms, ensuring our engineering teams can effectively produce and consume telemetry for their services. … looking for a technically strong, customer-focused Software Engineer to help make observability easier to adopt across the organisation. This role focuses on the producer side: instrumentation patterns, OpenTelemetry SDKs and the collector configurations that help teams emit consistent, high-quality telemetry. This role is suited to someone who enjoys ...

Terraform Platform Engineer for Multi Cloud Database Modules

Location
Greater London, England, United Kingdom
code at scale, focusing on cloud database platform patterns and module interfaces. You will work with partner teams to ensure guardrails, testing, and observability, and help expand patterns into GCP while improving developer experience and adoption. #J-18808-Ljbffr ...

Senior SRE (Go) - Remote, Unlimited PTO

Location
Greater London, England, United Kingdom
delivery of scalable infrastructure for a fast-growing FinTech payments platform. The role focuses on building resilient Go services, managing authentication, security, networking and observability, and working with Terraform and Kubernetes. Flexible remote work with monthly in-person meets in London. #J-18808-Ljbffr ...

Senior Software Engineer – Scalable, Low-Latency Trading

Location
Greater London, England, United Kingdom
systems. You will work with software, QA and DevOps engineers to drive systemic improvements and support large-scale services. Responsibilities include implementing features, improving observability, mentoring engineers, and coordinating across exchanges and trading venues. #J-18808-Ljbffr ...

Lead Java Developer — Real-Time Risk & Cloud (Hybrid)

Location
Greater London, England, United Kingdom
full lifecycle from design to production support, integrating new analytics and data sets across global teams. The role emphasizes scalable microservices, streaming data, and observability with ELK, Prometheus and Grafana. Hybrid work model and competitive benefits are offered. #J-18808-Ljbffr ...

Lead DevOps Engineer for Low-Latency Trading Platform (Hybrid)

Location
Greater London, England, United Kingdom
infrastructure across cloud and on-prem environments. The role emphasizes reliability, security, and fast delivery, with hands-on leadership across DevOps, platform engineering, and observability initiatives. You will build CI/CD pipelines, automate provisioning and deployment, and collaborate with engineering, security, and operations teams to improve platform readiness ...

Network Automation Engineer - Low-Latency Finance Platform

Location
Greater London, England, United Kingdom
that enables rapid provisioning and reliable operations. This role focuses on network automation at scale, using Python, Ansible, Terraform, and CI/CD, plus observability, on-call duties, and collaboration with security and platform teams. #J-18808-Ljbffr ...

Senior Platform Engineer – Remote UK (Cloud/SRE)

Location
Greater London, England, United Kingdom
Hudl in London, United Kingdom, is seeking a Senior Engineer to join our Platform Engineering team. You’ll work on site reliability, cloud infrastructure, observability and production operations to keep Hudl’s platform highly available, scalable and secure. You’ll lead with technical excellence, mentor engineers and drive innovation using ...

Azure Platform Architect: Multi-Tenant AKS

Location
Greater London, England, United Kingdom
upskilling engineers. You will lead platform operating models, SPI communications, and best-practice cloud-native patterns. You will shape the architecture, governance, and observability stack for scalable, multi-tenant workloads in production, with a strong emphasis on IaC, GitOps, and secure, compliant design. #J-18808-Ljbffr ...

Senior Ruby on Rails Engineer | React & Python Leader

Location
Greater London, England, United Kingdom
ensure deliverables are simple, maintainable, and scalable. The role emphasizes owning code quality, API contracts, and system documentation, with responsibility for performance monitoring and observability to maintain reliability across services. #J-18808-Ljbffr ...