1,701 to 1,725 of 2,250 Observability Jobs in London

Senior Software Engineer / Senior AI Engineer

Location
Greater London, England, United Kingdom
within a defined problem, building and testing tool use, retrieval pipelines, and agent workflows, integrating AI capabilities into enterprise systems, and contributing to evaluation, observability, and guardrails. You will hold a high bar on code quality, flag risks and blockers early, and work alongside host-function stakeholders to make sure … agentic AI solutions to production standards within a defined technical approach. Implement and test tool use, retrieval pipelines, and agent workflows. Contribute to evaluation, observability, and guardrails for agentic systems. Integrate AI capabilities into existing enterprise workflows and systems. Maintain high code quality and documentation so patterns can be reused. ...

Application Architect

Hiring Organisation
London Stock Exchange Group
Location
London, UK
Employment Type
Full-time
communication rather than formal authorityGuide the design of cloud‐native, AWS‐hosted applications, considering resilience, security, operability, and cost. Ensure operational concerns such as observability, supportability, and deployment are considered early in design. You will promote continuous improvements, by contributing to architectural roadmaps and the ongoing evolution of the platform … clearly to both technical and non‐technical audiencesAbility to influence technical decisions and collaborate effectively across teams and culturesExperience in designing solutions with good observability practices embeddedUnderstanding of performance, scalability, and resilience engineeringDeeper exposure to event‐driven architectures and asynchronous messaging at scaleExperience optimising performance, resilience, and cost in production ...

Senior Software Engineer / Senior AI Engineer

Hiring Organisation
Elsevier
Location
London, UK
Employment Type
Full-time
within a defined problem, building and testing tool use, retrieval pipelines, and agent workflows, integrating AI capabilities into enterprise systems, and contributing to evaluation, observability, and guardrails. You will hold a high bar on code quality, flag risks and blockers early, and work alongside host-function stakeholders to make sure … agentic AI solutions to production standards within a defined technical approach. Implement and test tool use, retrieval pipelines, and agent workflows. Contribute to evaluation, observability, and guardrails for agentic systems. Integrate AI capabilities into existing enterprise workflows and systems. Maintain high code quality and documentation so patterns can be reused. ...

Technical Leader

Location
Greater London, England, United Kingdom
manage technical debt, prioritizing improvements that provide meaningful value to the team and the product. Promote engineering best practices around testing, CI/CD, observability, documentation, and operational excellence. Stay current with emerging technologies and industry practices, evaluating when new technologies can provide meaningful improvements to our products and engineering … computer science fundamentals, including data structures, algorithms, concurrency, and system design. Strong understanding of modern software development practices, including CI/CD, TDD, DevOps, observability, and automated quality gates. Confident communication and collaboration skills, with the ability to articulate technical decisions, challenge assumptions, and explain complex technical concepts to both ...

Cloud Platforms Engineer

Location
City Of London, England, United Kingdom
runs the internal platform that PEI’s engineering and data teams build on: the cloud accounts, the reusable infrastructure code, the delivery pipelines, the observability, and the security and cost guardrails that wrap around them. We treat that platform as a product with internal customers – the measure of our work … Desirable: experience with data platform infrastructure such as Databricks, or similar – prior Databricks experience is not required. Desirable: experience with Datadog, or another mature observability platform. Desirable: experience of high‐traffic, international, content‐heavy web platforms. Technical Skills Confident with Linux and containers, and able to debug from the command ...

Engineering Manager (Remote - UK)

Hiring Organisation
Reonomy
Location
London, UK
Employment Type
Full-time
building and evolving AWS-native and cloud-based platforms, alongside legacy systemsSet and uphold strong engineering standards across code quality, testing, CI/CD, observability and documentationStay close to technical decisions through design reviews, architecture discussions and hands-on coachingBalance new feature delivery with technical debt, reliability, security and long … distributed systemsProficiency in at least one modern programming languageStrong grasp of system design and software engineering fundamentalsExperience with Infrastructure as Code, CI/CD, observability and secure production systemsAble to communicate technical ideas clearly to both technical and non-technical stakeholdersUnlock your Altus Experience! If you're looking to advance ...

Product Engineering Manager

Hiring Organisation
Reonomy
Location
London, UK
Employment Type
Full-time
building and evolving AWS-native and cloud-based platforms, alongside legacy systemsSet and uphold strong engineering standards across code quality, testing, CI/CD, observability and documentationStay close to technical decisions through design reviews, architecture discussions and hands-on coachingBalance new feature delivery with technical debt, reliability, security and long … distributed systemsProficiency in at least one modern programming languageStrong grasp of system design and software engineering fundamentalsExperience with Infrastructure as Code, CI/CD, observability and secure production systemsAble to communicate technical ideas clearly to both technical and non-technical stakeholdersThis role is initially offered as a contract position with ...

Senior Backend Engineer

Location
Greater London, England, United Kingdom
engineers to turn agreed product behaviour into implementation-ready technical solution designs covering assumptions, alternatives, service interactions, data and schema changes, failure modes, security, observability, rollout, rollback, and tests. Use agentic coding tools such as Codex, Claude Code, Cursor, or equivalent to generate, refactor, test, review, and document code within … with product and engineering colleagues through concise written analysis and appropriate architecture, sequence, data-flow, and entity-relationship diagrams. Sound judgement around automated testing, observability, production reliability, security, and data integrity. Strong working fluency with agentic coding tools and spec-driven development workflows, including specification writing, code generation, test generation ...

Platform Engineer (10x Openings)

Location
Greater London, England, United Kingdom
into scalable platform features. Integrate security guidance from the security engineering team into platform‐level controls, and remediate findings at the platform layer. Treat observability as a platform concern: instrument services, define meaningful metrics, and build tooling that gives the team visibility into platform health. Own the services you build … storage systems (Ceph or similar) at an engineering level. Background building Kubernetes operators using frameworks such as Kopf, controller‐runtime, or similar. Experience with observability tooling: Prometheus, Grafana, OpenTelemetry, or structured logging in distributed systems. Experience building SaaS or PaaS layers on top of an IaaS platform. Exposure to serverless ...

Senior Backend Software Engineer

Location
Greater London, England, United Kingdom
engineering, performance, and reliability problems that sit outside of their scope. Improve the shared foundations other engineers depend on: deployment pipelines, service templates, observability, and the environments their work runs in. You should apply if You have built and operated production backend systems. You will spend your time working … until you have measured it. Reaching for the profiler, the trace, or the metric is instinct rather than afterthought. You treat quality, security, and observability as engineering fundamentals rather than optional extras, and you make the case for addressing technical debt rather than living with it. You raise ...

Senior Platform Engineer

Location
Greater London, England, United Kingdom
/TLS certificate lifecycle management (ideally with Venafi) and a solid grasp of certificate chains, SNI, and secure delivery requirements* **Troubleshooting and delivery observability** – Edge Diagnostics, debug headers, mPulse (RUM), and traffic reporting APIs, backed by strong network fundamentals (TCP/IP, DNS, HTTP/HTTPS, routing) and a clear … compute (EdgeWorkers/EdgeKV), API Gateway, or service mesh for application-layer traffic management* Akamai DataStream or similar for streaming delivery telemetry into wider observability tooling* Compliance frameworks such as PCI DSS as they apply to edge and CDN configuration # **What makes you stand out** You troubleshoot systematically, tracing ...

Platform & Engineering Senior Manager

Location
Greater London, England, United Kingdom
agent services, with monitoring, support, change control, and lifecycle management.· Set and enforce technical standards for data modelling, ETL or ELT, APIs, testing, deployment, observability, lineage, access, retention, auditability, resilience, and production support, as relevant to the assigned portfolio.· Provide architecture assurance and engineering oversight to internal teams and delivery … least one of the following areas:Data engineering and pipelines, including cloud data platforms, ETL or ELT, APIs, orchestration, data modelling, automated testing, deployment, observability, incident management, and operational reliability.Data governance and agent infrastructure, including data classification, identity and access, lineage, retention, metadata, sensitive-data controls, telemetry, evaluation infrastructure ...

Staff Quality Engineer - Mobile

Hiring Organisation
Lendable
Location
London, UK
Employment Type
Full-time
quality: join specs early, push back on ambiguous acceptance criteria, and surface risk before code is writtenClose the loop on production issues using our observability stack (Datadog, Sentry, Grafana) - tying test coverage back to real customer impactEnsure teams have Service Level Objectives set up and are achieving themRun targeted exploratory … TypeScript, React Native, Expo, EAS, GraphQL, Relay, Jest, React Testing Library, Maestro. Backend: Kotlin, PHP 8 (Symfony), AWS, Postgres, RabbitMQ, Docker, Kubernetes. Tooling and observability: GitHub, GitHub Actions, Jira, Confluence, Datadog, Sentry, Grafana. Why join? See your work matter: Our products are used by millions of customers - the quality ...

Senior Engineering Manager - 9-10 month FTC

Location
Greater London, England, United Kingdom
practices, including test‐driven development and automated testing, helping teams build quality into the development process. Support strong operational ownership through CI/CD, observability, production support and the principle that teams own the systems they build and run. Help teams identify and address technical debt sustainably while maintaining appropriate … outcomes and translating these into clear engineering priorities. Strong understanding of modern software engineering practices, including automated testing and TDD principles, CI/CD, observability and production ownership. Experience facilitating technical decisions, bringing the right people and evidence together and constructively challenging thinking when needed. Comfortable balancing feature delivery ...

EMEA Stress Testing Technology Engineering & Delivery Lead - D

Location
Greater London, England, United Kingdom
lineage, transformation, calculation, aggregation and integration across source platforms, risk engines, model platforms and reporting tools. Drive secure development, automated testing, CI/CD, observability, resilience, performance and operational supportability. Lead platform modernisation, simplification, automation and technical debt reduction, balancing scalability, resilience, cost and regulatory timelines. Key Accountabilities and Responsibilities … data models, code and engineering artefacts and guide teams through complex implementation and production issues. Strong knowledge of automated testing, CI/CD, DevOps, observability, cloud services, containerisation and secure software development. Experience delivering stress testing, risk, capital, liquidity, finance or regulatory reporting technology capabilities. Working knowledge across Credit Risk ...

Forward Deployed Engineer

Hiring Organisation
Janus Henderson
Location
London, UK
Employment Type
Full-time
architecture and design decisions for your solutions, keeping them secure, scalable, and aligned to the enterprise core stack. Take solutions through evaluation, observability, and our AI governance checkpoints, and keep them healthy afterwards. Build investment capability on NexusBuild advanced investment capability on Nexus and make it available across the firm. … portfolio construction. Snowflake, Microsoft Fabric/OneLake, or comparable enterprise data platforms. Azure AI Foundry, model gateways, inference routing, or AI evaluation and observability in production. TypeScript or a second production language, and front-end experience for user-facing AI applications. Supervisory responsibilitiesNo. This is an individual-contributor role. Forward ...

Principal IT Systems Engineer

Location
Greater London, England, United Kingdom
At Pension Insurance Corporation (PIC), we're looking for a Principal IT Systems Engineer to join our IT Operations and Engineering team. This is a senior, hands-on technical leadership role for an experienced engineer ...

Solutions Engineer

Location
Greater London, England, United Kingdom
About Ultralytics: At Ultralytics, we commit to relentless innovation in the AI space and seek team members who resonate with our ambition to produce the world's best YOLO AI models. If you're obsessed ...

AI Engineer

Location
Greater London, England, United Kingdom
Department/Division : Business Technology Duration : Permanent Reports to :IT Development Manger Type of Role : Hybrid Budget Responsibilities : No Reference number : 10740 The Role Reporting to the IT Development Manager, with a dotted line to ...

Data Architect

Hiring Organisation
PA Consulting
Location
London, UK
Employment Type
Full-time
Company DescriptionWe believe in the power of ingenuity to build a positive human future. We challenge where it matters and own the outcome. As strategies, technologies, and innovation collide, we create opportunity from complexity. Our ...

Software Engineer, Real-Time

Hiring Organisation
London Stock Exchange Group
Location
London, UK
Employment Type
Full-time
that show how reliably and quickly market data reaches customers. You will work with experienced engineers to develop production software, measure performance and enable observability of a globally distributed platform. Prior market-data or observability experience is not required; we are looking for strong engineering fundamentals, curiosity and a willingness … hands-on development role for an engineer at an early stage of their career who wants to build experience in mission-critical distributed systems, observability, cloud-native engineering and real-time financial technology. WHAT YOU'LL BE DOINGBuild and improve software components that aggregate, correlate and present data from over ...

AWS EKS DevOps Engineer — CI/CD, IaC & Observability

Location
Greater London, England, United Kingdom
/CD pipelines using GitHub Actions or GitLab CI/CD and ArgoCD. You will collaborate with developers and SREs to improve reliability, observability, and cost efficiency, while applying best practices in IAM, security, and incident response. #J-18808-Ljbffr ...

Senior Software Engineer (GO)

Hiring Organisation
Source Group International
Location
London, UK
Employment Type
Full-time
services using Golang within distributed and microservices-based architectures. Build enterprise-grade AI platform capabilities supporting LLMOps, model deployment, inference services, prompt management, model observability, and governance. Engineer reusable backend blueprints and reference architectures that serve as standardized patterns across the organization. Develop globally scalable systems capable of handling high … highly available data services leveraging MongoDB and other enterprise data platforms. Optimize data flows, event processing, and backend performance for large-scale distributed systems. Observability & Reliability Implement comprehensive observability solutions using: PrometheusOpenTelemetryGrafanaEstablish monitoring, tracing, logging, alerting, and SRE best practices. Drive operational excellence through performance tuning, reliability engineering, and proactive ...

Lead Observability Engineer

Hiring Organisation
Tria
Location
London, United Kingdom
Employment Type
Contract
Location: London, onsite 3 days per week (Sheffield as an alternative) Rate: £tbd/day inside IR35 Duration: 6 months+ Are you a Senior Observability Engineer/SRE Lead, with demonstrable experience of assessing and defining observability and monitoring roadmaps within enterprise scale environments? If so, apply now for this … contract opportunity. The Lead Observability Engineer/SRE Lead will be required to assess a complex hybrid estate, understand how services, platforms, infrastructure and networks should be monitored, and work across multiple internal teams, partners and suppliers to build a consolidated view of existing telemetry, monitoring and alerting capabilities. ...

Staff Software Engineer, Observability & Profiling

Hiring Organisation
Humanloop
Location
London, UK
Employment Type
Full-time
researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems. About the roleAnthropic is seeking Software Engineers to join our Observability team within the Infrastructure organization. The Observability team owns the monitoring and telemetry infrastructure that every engineer and researcher at Anthropic depends on—from metrics … growing by orders of magnitude—and an increasing share of the hardest problems live below the application layer. We're building next-generation observability systems—high-throughput telemetry pipelines, fleet-wide continuous profiling, eBPF-based tracing and network visibility, and agentic diagnostic tools—so engineers can detect, diagnose, and resolve ...