2,726 to 2,750 of 5,618 Observability Jobs

Senior Data Platform Engineer: Cloud Data Pipelines

Location
Greater London, England, United Kingdom
that make datasets economical to store, and efficient to use. Operating batch workloads on AWS, with responsibility for their reliability, resource use, retry behaviour, observability and cost. Improving the shared tooling and infrastructure used to test, deploy, monitor and operate our data services. Making common data workflows easier and safer ...

Software Specialist

Hiring Organisation
Reactive Markets
Location
Addingham, West Yorkshire, UK
open source Cloud-native infrastructure on AWS with colocation in key data centres Market-data and reference-data pipelines across multiple asset classes Observability and monitoring — dashboards, alerting, and metrics that make system behaviour visible Strong C++ with Linux systems programming (3+ years) Experience with performance-sensitive systems (profiling, optimisation ...

Software Specialist

Hiring Organisation
Reactive Markets
Location
Darvel, East Ayrshire, UK
open source Cloud-native infrastructure on AWS with colocation in key data centres Market-data and reference-data pipelines across multiple asset classes Observability and monitoring — dashboards, alerting, and metrics that make system behaviour visible Strong C++ with Linux systems programming (3+ years) Experience with performance-sensitive systems (profiling, optimisation ...

Software Specialist

Hiring Organisation
Reactive Markets
Location
Rhoose, Vale of Glamorgan, UK
open source Cloud-native infrastructure on AWS with colocation in key data centres Market-data and reference-data pipelines across multiple asset classes Observability and monitoring — dashboards, alerting, and metrics that make system behaviour visible Strong C++ with Linux systems programming (3+ years) Experience with performance-sensitive systems (profiling, optimisation ...

Developer Software Development

Hiring Organisation
Reactive Markets
Location
Rhoose, Vale of Glamorgan, UK
open source Cloud-native infrastructure on AWS with colocation in key data centres Market-data and reference-data pipelines across multiple asset classes Observability and monitoring — dashboards, alerting, and metrics that make system behaviour visible Strong C++ with Linux systems programming (3+ years) Experience with performance-sensitive systems (profiling, optimisation ...

Developer Software Development

Hiring Organisation
Reactive Markets
Location
Addingham, West Yorkshire, UK
open source Cloud-native infrastructure on AWS with colocation in key data centres Market-data and reference-data pipelines across multiple asset classes Observability and monitoring — dashboards, alerting, and metrics that make system behaviour visible Strong C++ with Linux systems programming (3+ years) Experience with performance-sensitive systems (profiling, optimisation ...

Cloud DevOps Engineer

Location
Greater London, England, United Kingdom
/CD pipelines through Azure DevOps and GitHub Actions. Your work will include designing and testing backup and disaster recovery arrangements, maintaining monitoring and observability tools, and responding thoughtfully to incidents. You will collaborate with engineering, security and operations colleagues to identify root causes, improve services and embed DevSecOps practices. ...

Software Specialist

Hiring Organisation
Reactive Markets
Location
Milton Keynes, Buckinghamshire, UK
open source Cloud-native infrastructure on AWS with colocation in key data centres Market-data and reference-data pipelines across multiple asset classes Observability and monitoring — dashboards, alerting, and metrics that make system behaviour visible Strong C++ with Linux systems programming (3+ years) Experience with performance-sensitive systems (profiling, optimisation ...

Developer Software Development

Hiring Organisation
Reactive Markets
Location
Grantham, Lincolnshire, UK
open source Cloud-native infrastructure on AWS with colocation in key data centres Market-data and reference-data pipelines across multiple asset classes Observability and monitoring — dashboards, alerting, and metrics that make system behaviour visible Strong C++ with Linux systems programming (3+ years) Experience with performance-sensitive systems (profiling, optimisation ...

Developer Software Development

Hiring Organisation
Reactive Markets
Location
Budleigh Salterton, Devon, UK
open source Cloud-native infrastructure on AWS with colocation in key data centres Market-data and reference-data pipelines across multiple asset classes Observability and monitoring — dashboards, alerting, and metrics that make system behaviour visible Strong C++ with Linux systems programming (3+ years) Experience with performance-sensitive systems (profiling, optimisation ...

Solutions Engineer

Location
Greater London, England, United Kingdom
over architectural purity, and you're explicit about that being a deliberate choice. Take the prototypes that prove out into early production — enough robustness, observability, and compliance posture to handle live flows — in close partnership with the production engineers who will own them. Own the handoff. Instrument it, document ...

Director, Full-Stack Engineer

Hiring Organisation
Hackajob Ltd
Location
Manchester, North West, United Kingdom
Employment Type
Permanent
peer review, automated testing, release governance, production validation, and support readiness. Champion engineering best practices in full stack architecture, reusable design patterns, API strategy, observability, and secure coding. Strengthen platform stability and operational resilience by reducing manual touchpoints, improving exception handling, and enhancing recovery and support processes. Support release planning ...

Site Reliability Engineer II - AI & Corporate Risk Tech

Location
Glasgow, Scotland, United Kingdom
practices within an application or platform Proficiency in at least one programming language such as Python, Java/Spring Boot, or .NET Experience in observability practices such as white and black box monitoring, service level objective alerting, and telemetry collection Proficient knowledge of software applications and technical processes within ...

Site Reliability Engineer

Hiring Organisation
Hackajob Ltd
Location
Glasgow, Lanarkshire, Scotland, United Kingdom
Employment Type
Permanent
practices within an application or platform Proficiency in at least one programming language such as Python, Java/Spring Boot, or .NET Experience in observability practices such as white and black box monitoring, service level objective alerting, and telemetry collection Proficient knowledge of software applications and technical processes within ...

Associate Software Engineer - AI Agents (Satori)

Location
Belfast City District, Northern Ireland, United Kingdom
Agents SDK, MCP, LangChain, or similar) and keep up with the field. Exposure to agentic or LLM‐based systems, and curiosity about the eval, observability, and rollback story that takes them beyond prototypes. Demonstrated learning velocity and curiosity; driven to understand systems beyond surface‐level usage and eager to learn ...

Non-Functional Test Specialist

Location
United Kingdom
Hands-on experience in planning, preparing & executing Resilience/Disaster Recovery/Continuity/Failover testing/Performance Exposure to tools supporting resilience & operational observability (Splunk, Prometheus, Kafka, Chaos tooling etc) Understanding of high-availability architectures, infrastructure redundancy & backup/restore strategies Are familiar with working within large scale & complex ...

Tech Lead

Location
Manchester, England, United Kingdom
engineering culture Improve delivery flow, remove blockers and help the team ship value quickly Champion engineering best practices across testing, CI/CD, observability and operational excellence Work closely with Product, Design and other business stakeholders to turn priorities into successful outcomes Encourage pragmatic adoption of AI tools to improve ...

Sr. Engineer - Platform

Location
Greater London, England, United Kingdom
Your Role We’re hiring a Senior Engineer to join our Platform Engineering team, working with squads dedicated to site reliability, cloud infrastructure, observability and production operations. You’ll help drive operational consistency and ensure that Hudl engineers can build on a highly available, scalable and secure platform. ...

AI Principal Architect

Hiring Organisation
NTT DATA
Location
London, UK
Employment Type
Full-time
agents, orchestration frameworks, or enterprise AI platforms. Deep understanding of AI architecture considerations, including data readiness, model selection, security, responsible AI, governance, integration, scalability, observability, and operations. Experience leading architecture for complex client pursuits, large transformation programs, or multi-tower technology solutions. Bachelor's degree or equivalent work experience. Preferred ...

Software Engineer - Front End - Studio (Core)

Location
Greater London, England, United Kingdom
back-end leads to shape roadmaps and turn ambiguous problems into clear technical plans. Own the operational excellence of your domains, ensuring reliability, observability, and a high SLA. Clearly explain the nuances of system design and front-end paradigms to engineers and stakeholders alike. The Team You will join ...

Customs AI Integration & Automation Engineer (UK)

Location
United Kingdom
allowed to do. Agent orchestration frameworks and standards (LangGraph, MCP, Temporal or equivalent), and experience with long-running or multi-step agent workflows. Observability and cost control for inference at volume. What success looks like in your first year By month three: You understand our customs processes well enough ...

Remote Principal Software Engineer

Location
United Kingdom
live and breathe this approach ourselves: we release new versions of Gearset multiple times a day and we continually invest in improving our own observability and infrastructure tools. This means we can identify and react to issues quickly and delight our users by getting improvements to them as fast ...

Customs AI Integration & Automation Engineer (UK)

Location
Felixstowe, England, United Kingdom
allowed to do. Agent orchestration frameworks and standards (LangGraph, MCP, Temporal or equivalent), and experience with long-running or multi-step agent workflows. Observability and cost control for inference at volume. What success looks like in your first year By month three: You understand our customs processes well enough ...

Remote Principal Software Engineer

Location
Rickmansworth, Hertfordshire, United Kingdom
live and breathe this approach ourselves: we release new versions of Gearset multiple times a day and we continually invest in improving our own observability and infrastructure tools. This means we can identify and react to issues quickly and delight our users by getting improvements to them as fast ...

Remote Principal Software Engineer

Location
Cwmbran, Monmouthshire, United Kingdom
live and breathe this approach ourselves: we release new versions of Gearset multiple times a day and we continually invest in improving our own observability and infrastructure tools. This means we can identify and react to issues quickly and delight our users by getting improvements to them as fast ...