3,851 to 3,875 of 4,063 Permanent Observability Jobs

Quantitative Developer — Build Fast, Distributed Trading

Location
Greater London, England, United Kingdom
Quadrature is building an automated trading system designed to ingest data, form views, execute in the market, and learn to improve. Quant developers and software engineers join to work across research infrastructure, low-latency processing ...

Remote Network Automation Engineer – Python & Nautobot

Location
United Kingdom
Xalient is seeking an experienced Network Automation Engineer in the UK who will own and expand our automated network platforms, including Nautobot SSoT, IP Fabric operations, and AI-enabled observability. The role combines platform governance ...

Microsoft Specialist - Copilot

Location
Sheffield, England, United Kingdom
errors, unexpected results). Root cause analysis across M365 Admin Centre, Entra, Conditional Access, SharePoint/OneDrive permissions, Teams, Group Policy Objects, Aternity etc Observability and Monitoring Advanced understanding of monitoring and observability tools such as Thousand Eyes and Downdetector Required Qualifications 5–8+ years in Microsoft 365 support ...

Microsoft Specialist - Copilot

Location
Birmingham, England, United Kingdom
errors, unexpected results). Root cause analysis across M365 Admin Centre, Entra, Conditional Access, SharePoint/OneDrive permissions, Teams, Group Policy Objects, Aternity etc Observability and Monitoring Advanced understanding of monitoring and observability tools such as Thousand Eyes and Downdetector Required Qualifications 5–8+ years in Microsoft 365 support ...

Senior Workday Integrations Architect

Location
Hook, England, United Kingdom
operate and continuously improve the integration ecosystem supporting our global workforce. This hands‐on role focuses on Workday integration development, architecture, automation, reliability, security, observability, and operational excellence, with collaboration across Product, HR, and global engineering teams. #J-18808-Ljbffr ...

Backend Engineer (Node.js)

Location
United Kingdom
core: APIs, matching behaviour, list ingestion, and production reliability when regulated teams depend on every response. What you'll do Sanctions feed ingestion and observability What you bring Strong Node.js and Postgres experience Comfort with latency, quotas, and production debugging Interest in sanctions/AML systems or high-trust APIs … data. What you'll do Uptime, backups, and incident readiness What you bring CI/CD, containers, and cloud ops experience Incident response and observability habits Security-minded defaults for regulated SaaS You'll translate AML and sanctions practice into product decisions—thresholds, review language, and workflows analysts can defend ...

Enterprise AI Deployment Architect

Location
Greater London, England, United Kingdom
commercial execution across product, engineering, and sales teams. You will translate business needs into deployment playbooks, drive measurable outcomes, and ensure governance, security, and observability across customer environments. #J-18808-Ljbffr ...

Operations Team Lead: Scale Production Reliability

Location
Nottingham, England, United Kingdom
Complexio, the intelligence layer for enterprise AI, is seeking an Operations Team Lead to own production. You will lead live systems, ensure reliability, observability, and continuous improvement across the platform. This hands‐on role requires building scalable processes, leading incidents, and growing the Ops team while moving from firefighting ...

SRE Engineer – Hybrid Cloud Reliability & Automation

Location
Greater London, England, United Kingdom
ensure ISO 27001 security alignment. This hands‐on role has significant influence across engineering, security and operations teams, with a focus on automation, observability and end‐to‐end service reliability. #J-18808-Ljbffr ...

Site Reliability Engineer — Cloud & Live Ops (Hybrid)

Location
Uxbridge, England, United Kingdom
resilient, secure, and highly available platforms underpinning live, broadcast‐adjacent services. Based at Stockley Park in Uxbridge with hybrid options, the role involves improving observability, incident response, automation, disaster recovery, and collaborating with engineering, operations and project stakeholders. #J-18808-Ljbffr ...

Senior ML Platform & Ops Engineer - Hybrid (London)

Location
Greater London, England, United Kingdom
Preply is hiring a Senior ML Platform/Ops Engineer in London. You will help productionize ML systems with reliability, performance, and observability, working at the intersection of ML, data engineering, and cloud infrastructure. You’ll collaborate with ML Scientists, Backend and Data Engineers to shape the ML lifecycle foundations. ...

Cloud Application Owner & Lead Engineer

Location
Glasgow, Scotland, United Kingdom
design, review, and automation. You will partner with engineering, platform, risk & control, and operations to ensure secure and stable operation, with emphasis on resiliency, observability, #J-18808-Ljbffr ...

Lead Software Engineer, AI Safety & Evals

Location
Greater London, England, United Kingdom
evaluating models, safety guardrails, and governance across internal platforms, ensuring secure and deterministic AI behavior. You will lead continuous verification, red teaming, and observability initiatives, partnering with AI Foundations and engineering teams to scale safe AI across Kraken’s internal tools. #J-18808-Ljbffr ...

Senior ML Engineer, Real-Time Ranking & Personalization

Location
Greater London, England, United Kingdom
personalised rankings across search results and recommendations. You will collaborate with ML Scientists, Backend Engineers and MLOps to build scalable ranking models, improve latency, observability and real-time serving, and help harden the ML infrastructure for ranking systems. #J-18808-Ljbffr ...

ML Platform Engineer: Architect Reliable Production ML

Location
Greater London, England, United Kingdom
Engineer to build and operate platform capabilities that move machine-learning models from experimentation into reliable production services. You will own automation, deployment, observability and operational controls around the ML lifecycle, collaborating with research, software, platform and product teams to ensure reproducibility, scalability, security and dependability of ML systems. #J ...

Technical Lead, Data Pipelines & AI Training

Location
Greater London, England, United Kingdom
pipelines, unify sources for AI training, and mentor engineers across robotics, ML and data governance teams. You'll define roadmaps, shape architecture, improve latency, observability and reliability, and help hire and onboard new engineers in a collaborative, high-impact environment. #J-18808-Ljbffr ...

Real-Time Software Engineer: Low-Latency Data Systems

Location
Nottingham, England, United Kingdom
monitoring services to show how reliably market data reaches customers. You will collaborate with experienced engineers to develop production software, measure performance and enable observability across a globally distributed system. #J-18808-Ljbffr ...

Production Reliability Team Lead

Location
Normanton on Trent, England, United Kingdom
Lead to own production, building a scalable system for reliability. You will lead operational excellence across all live customer-facing systems to ensure reliability, observability, and continuous improvement. This hands-on leadership role requires shaping processes, leading incidents, growing the team, and transitioning from reactive firefighting to proactive reliability engineering. ...

ML Data & Platform Engineer — Hybrid ML Ops & Pipelines

Location
Cambridge, England, United Kingdom
infrastructure to production ML—owning problems end-to-end to accelerate model delivery. You’ll collaborate with the ML team to improve data quality, observability, and MLOps practices, while scaling infrastructure for faster iteration and reliability. #J-18808-Ljbffr ...

Azure Technical Architect - Cloud Security (Remote)

Location
United Kingdom
with delivery teams to drive secure, scalable solutions. You will design and evolve enterprise-scale foundations, ensure best practices in identity, networking, security, and observability, and maintain ADRs and platform documentation to support rapid, compliant #J-18808-Ljbffr ...

Senior IT Operations Analyst & Incident Lead

Location
City of Edinburgh, Scotland, United Kingdom
environments. Based in Edinburgh, the role offers exposure to a global technology-driven transformation within a major banking brand, with a focus on service observability and risk-aware operations. #J-18808-Ljbffr ...

Staff Engineer - Embedded Accountancy Platform Lead

Location
Norwich, England, United Kingdom
collaborate with external partners to deliver scalable financial tooling. As a platform-focused leader, you will ensure robust integration patterns and high standards for observability, performance, and security across the product ecosystem. #J-18808-Ljbffr ...

Mid-Market AI Solutions Account Executive

Location
United Kingdom
assigned territory. You will own the full sales cycle—outreach to close—advocating a value-based approach for our AI-powered search, observability, and security platform. You’ll engage stakeholders, articulate the business value, and navigate complex deals. #J-18808-Ljbffr ...

Linux Automation Engineer — Hybrid (Glasgow)

Location
Paisley, Scotland, United Kingdom
London, to support cross-site initiatives. You will develop automation for patching, upgrades and changes across Linux, VMware and F5 environments, and contribute to observability, validation, and continuous improvement of the platform. #J-18808-Ljbffr ...