3,651 to 3,675 of 4,093 Permanent Observability Jobs

Engineering Manager: Platform & Growth Lead (Hybrid)

Location
Bristol, England, United Kingdom
client journeys on React and React Native. You'll coach engineers, own delivery, guide architectural decisions, partner with product, and drive observability and high-velocity delivery in a regulated financial services environment. #J-18808-Ljbffr ...

Partner Manager, EMEA

Location
Greater London, England, United Kingdom
About Monte CarloMonte Carlo is the agent trust platform that unifies data and agent observability to monitor, troubleshoot, and improve production AI systems. As enterprises prepare to deploy thousands of agents across business-critical use cases, Monte Carlo provides the reliability infrastructure to support them along this AI transformation, from … cultivating strategic relationships with cloud providers, system integrators, and data consulting partners, enabling them to successfully position, sell, and deliver Monte Carlo’s Data Observability platform. As the connective tissue between Monte Carlo’s field teams and the broader ecosystem, the Partner Manager will drive partner-sourced pipeline, expand ...

Data Engineer

Hiring Organisation
Fort Recruitment
Location
CH41, Birkenhead, Metropolitan Borough of Wirral, Merseyside, United Kingdom
Employment Type
Permanent
Salary
£50000 - £55000/annum + excellent benefits
Engineer - The Role This is a full-time permanent position working for an outstanding Company who will provide full training The role operates the observability and data infrastructure behind an industrial energy-telemetry platform. You will collect and process large volumes of operational data and use that data to build … failures affecting ingestion, transformation, storage, or delivery Investigate performance degradation in queries, storage systems, and data-processing jobs Build dashboards, alerts, notifications, and other observability tools used in daily operations Convert operational data into analysis, internal tools, customer-facing functionality, or new products Integrate data platforms with custom software, industrial ...

Cloud Data Platform Engineer | Platform as Code

Location
Fenny Stratford, England, United Kingdom
mature processes and procedures for secure, scalable cloud data platforms. You will participate in agile platform development, deliver Platform as Code, and ensure observability and cost monitoring across cloud resources, #J-18808-Ljbffr ...

Senior UI Platform Engineer – React/TypeScript Foundations

Location
City Of London, England, United Kingdom
packages, components, and tooling. You will create starter templates, define standards for structure, testing, deployment, and supportability, and partner with other teams on authentication, observability, and frontend/backend integration patterns. This is a platform engineering role spanning teams and regions. #J-18808-Ljbffr ...

Lead SRE: AWS Platform & Reliability Leader

Location
Glasgow, Scotland, United Kingdom
Co. in Glasgow seeks a Lead Site Reliability Engineer to define reliability strategy and drive robust, scalable platforms. You will lead incident response, shape observability, and guide AI-assisted reliability workflows across the SDLC. You will mentor peers, conduct resiliency reviews, and partner with product teams to establish SLOs, error ...

Senior Backend Engineer: Own architecture & resilient systems

Location
Bath, England, United Kingdom
production systems, aiming for scalable, resilient services and robust incident handling. Collaborating with Engineering, Product and Operations, you will drive CI/CD, observability and reliability improvements across a suite of business-critical platforms. #J-18808-Ljbffr ...

AI Platform Engineer: Kubernetes, Cloud & AI Agents

Location
Knutsford, England, United Kingdom
Manchester. You will build scalable Kubernetes platforms, automate cloud infrastructure, and deploy AI agents with governance and monitoring. The role emphasizes secure coding, observability, cost-aware cloud usage, and self-service tooling to enable rapid delivery across teams and business units. #J-18808-Ljbffr ...

Lead Software Engineer - Hybrid, Share Options, £80k+

Location
Wigan, England, United Kingdom
production. You will lead a small cross-functional team, delivering features in a fast-paced SaaS environment. You will help modernise microservices, improve observability, and enhance testability using modern architectural approaches. Hybrid working with two days in the office is available. #J-18808-Ljbffr ...

Senior AI-Powered Backend Engineer

Location
United Kingdom
workflows and decision paths. The Agent Systems team designs inference pipelines and back-end services that orchestrate AI across internal APIs, maintaining reliability and observability as the platform scales. You will prototype quickly, validate agent behaviors, and productionize what proves valuable, contributing to a 0→1 agent platform with cross ...

Senior Value Engineer – Remote: Growth & ROI Storytelling

Location
Pathhead, Scotland, United Kingdom
persuasive value stories. The role emphasizes storytelling with data, enabling GTM teams, and partnerships to articulate Grafana's value propositions, with a focus on observability, open source roots, and a global remote #J-18808-Ljbffr ...

Service Transition & Reliability Consultant

Location
Nottingham, England, United Kingdom
changed services land in production reliably and with robust operational foundations. You will partner with project teams during design, influence outcomes on reliability and observability, and own service readiness across the lifecycle. The role emphasizes hands-on engagement, continuous improvement, and collaboration with SRE, Product Engineering, and Service Operations ...

Senior Mobile Engineering Leader (iOS & Android)

Location
United Kingdom
collaborate with product, design, architecture, security, quality, and operations to deliver reliable, scalable mobile experiences and roadmaps. You will shape architecture, improve release quality, observability, and AI-enabled development, while building a high-performing organization through hiring, mentoring, #J-18808-Ljbffr ...

Senior Platform Engineer - Apollo Core & DevEx

Location
England, United Kingdom
Lead Analytics Engineer and reporting to the Head of Data & Engineering. You’ll lead platform and application engineering, drive architectural decisions, and guide security, observability and deployment pipelines. #J-18808-Ljbffr ...

Senior Quality Engineer - Champion Quality in Cloud Apps

Location
Greater London, England, United Kingdom
perform exploratory testing, and collaborate with Product and Development to build the right software for customers, embracing shift-left quality practices. You will enhance observability, apply risk-based testing, and contribute to automation while growing a culture of quality across teams. #J-18808-Ljbffr ...

Senior Data Platform SRE — Hybrid, Obs & Reliability Leader

Location
Greater London, England, United Kingdom
Senior Site Reliability Engineer to embed within the Data Engineering team. You will own reliability, performance, and operability of the data platform, building observability from the ground up and leading incident response for data outages. You'll balance feature velocity with system stability, applying SRE practices and capacity planning ...

Senior AWS Data Engineer – Secure Public Sector Data

Location
Greater London, England, United Kingdom
shaping engineering practices. You will design data platforms, pipelines and data models, work with architects and security specialists, and ensure governance, security and observability across solutions. #J-18808-Ljbffr ...

Cloud Operations Engineer - Reliability & Automation

Location
England, United Kingdom
premises and AWS cloud services. You will work with MySQL, PostgreSQL and SQL Server, ensure security, availability and disaster recovery while driving automation and observability in a collaborative team. You will collaborate with Software Engineers and Technology teams to deliver reliable platforms and support product delivery, with training opportunities ...

Senior SET: Platform Quality Engineer (Hybrid, London)

Location
Greater London, England, United Kingdom
architecture level, automate validation across environments, and manage delivery risk in distributed systems. You will define standards for testing distributed systems, enable observability-driven debugging, automate validation of availability and latency, and contribute to security posture. #J-18808-Ljbffr ...

Senior Data Platform Engineer | ELT & Analytics

Location
Greater London, England, United Kingdom
senior member, you will influence platform architecture, data products, and engineering standards while collaborating with analytics teams to deliver reliable production solutions and observability across key datasets. #J-18808-Ljbffr ...

Senior SRE: Cloud Platform Reliability & Automation

Location
Cambridge, England, United Kingdom
Engineer to ensure the reliability, availability, and performance of our large-scale cloud platforms and SaaS products. Your work will automate operational tasks, improve observability, and contribute to incident management while collaborating with development teams to build more reliable and scalable applications across regions. #J-18808-Ljbffr ...

Service Transition Lead: Design, Reliability & Readiness

Location
United Kingdom
changed services land in production reliably and with strong operational foundations. You will partner with project teams from design through deployment, focusing on reliability, observability and governance. The role emphasizes hands-on readiness, collaboration with SRE, Product Engineering and Service Operations, and continual improvement of service transition processes within ...

AI-Ops SRE & Operations Leader

Location
Greater London, England, United Kingdom
experienced SRE Manager to lead the reliability function for production services used by internal and external customers. You will drive AI-Ops adoption, automation, observability and incident response. You will own end-to-end incident and problem management, coach team leads, ensure RCAs and post-mortems are completed, and balance ...

AI Platform & SRE Leader — Scale & Govern AI

Location
Greater London, England, United Kingdom
Platform & Site Reliability Engineering Managing Consultant to help clients design, build and scale secure, reliable AI platforms. You will lead platform engineering, SRE, observability and guardrails, moving AI from experimentation to enterprise-scale delivery. You will partner with CIO/CTO stakeholders, shape platform strategies, and guide multidisciplinary delivery teams ...