1,776 to 1,800 of 2,348 Remote/Hybrid Observability Jobs

Remote Lead Engineer, AI Quality Platform

Location
Greater London, England, United Kingdom
fully remote company with ~270 teammates, seeking a Lead Engineer for the AI Quality engineering team. You’ll design evaluation frameworks, observability tooling, and datasets, while leading a growing team and collaborating with AI Core engineering. You will own the evaluation infrastructure, diagnose quality issues, and drive cost and latency ...

Hybrid Enterprise Integration Product Manager

Location
Sheffield, England, United Kingdom
Sheffield is seeking an Enterprise Integration Product Manager for a hybrid, contractor role. You will own a multi-quarter roadmap across middleware, messaging, and observability, driving platform standards and governance, while aligning with regulatory and regional constraints. The role requires deep IBM MQ knowledge, experience with ACE/ ...

Lead AI Quality Engineer — Remote, Equity

Location
United Kingdom
powered, all-in-one platform for creators and businesses. We’re seeking a Lead Engineer for AI Quality to own evaluation frameworks, observability tooling, and diagnostic infrastructure, while leading the AI Quality engineering team. You’ll drive experiments across prompts and models, improve data pipelines, and collaborate with AI Core ...

Engineering Manager — Hybrid London, Lead High-Impact Teams

Location
Greater London, England, United Kingdom
values. You’ll collaborate with Product Managers, Tech Leads, and Architects to plan work, remove blockers, and scale systems with an emphasis on quality, observability, and inclusivity. #J-18808-Ljbffr ...

Monte Carlo Implementation SME - Hybrid London

Location
Greater London, England, United Kingdom
broader Aladdin/Snowflake data environment, providing hands-on configuration, setup and SME guidance to delivery teams. The role requires deep knowledge of data observability, data quality and data platforms, with strong stakeholder #J-18808-Ljbffr ...

Principal Engineer I, Prepurchase Platform (Remote)

Location
City Of London, England, United Kingdom
demand on-sales. You will write production code daily, influence technical direction, and collaborate across multiple teams within the Prepurchase domain. You will drive observability, resilience patterns, and AI-assisted enhancements while embedding across services or working horizontally. #J-18808-Ljbffr ...

E2 SAP Basis Engineer E2

Location
Swindon, England, United Kingdom
capabilities across implementation and operations. You’ll contribute to the evolution of our SAP Application Lifecycle Management tooling, helping establish modern approaches to observability, engineering effectiveness and change governance across our SAP platforms. At Nationwide we offer hybrid working wherever possible. More rewarding relationships are supported through our hybrid approach ...

Partner Account Executive - London

Hiring Organisation
CISCO Systems
Location
London, UK
Employment Type
Full-time
market engine, and our team acts as strategic business advisors helping them build profitable, high-growth practices around cloud networking, cybersecurity, observability, and AI-ready infrastructure. You will join a supportive, close-knit team of partner account executives who share insights openly, celebrate collective wins, and champion inclusive collaboration. What ...

Partner Account Executive - London

Location
Greater London, England, United Kingdom
market engine, and our team acts as strategic business advisors helping them build profitable, high-growth practices around cloud networking, cybersecurity, observability, and AI-ready infrastructure. You will join a supportive, close-knit team of partner account executives who share insights openly, celebrate collective wins, and champion inclusive collaboration. What ...

API Engineer

Location
Greater London, England, United Kingdom
languages, we will give you plenty of room to learn and grow with your team. Maintain the reliability of our systems by building for observability and participating in incident response as needed. You can find more about how we work on our engineering page. Salary range for Senior Engineer ...

Sr. Snowflake Architect

Location
City of Edinburgh, Scotland, United Kingdom
level policies, Tri-Secret/external key management, network policies, Private Link, SSO/OAuth/MFA, audit trails, and access history.Embed data quality, observability, and lineage using Snowflake Horizon Catalog and native observability features; define SLAs/SLOs, reconciliation frameworks, and incident/RCAs for data reliability.Architect cross-region ...

Observability & AIOps Engineer (all genders)

Hiring Organisation
Lam Research
Location
Villach, Kärnten, Austria
Employment Type
Permanent
Salary
EUR Annual
innovative information system solutions and services. Together, we support users globally with data, information, and systems to achieve their business objectives. Aufgaben As an Observability & AIOps Engineer, you will build the intelligence layer that enables enterprise systems and AI agents to understand operational health in real time. You will transform … autonomous operations by enabling AI agents to perceive, reason about, and respond to complex system behavior. What you'll do Design and implement enterprise observability strategies across cloud, infrastructure, applications, and platforms. Build telemetry pipelines that ingest logs, metrics, traces, events, and topology information. Deploy and tune observability platforms ...

Senior Mobile Engineering Manager

Hiring Organisation
Innova Solutions
Location
London, United Kingdom
Employment Type
Contract
Contract Rate
£48 - £58/hour
Boot, Gradle build automation Strong understanding of: Microservices architecture RESTful APIs Event-driven architectures Secure software development practices Scalable distributed systems Deep experience implementing observability strategies, including: Logging, monitoring, alerting, Performance Analysis and Production Diagnostics. Experience with industry-standard observability and monitoring platforms such as Sentry or Datadog, New Relic ...

Senior / Principal Applied AI Engineer (UK / Europe, Remote)

Location
United Kingdom
production-grade AI agents, agentic workflows, and LLM integrations for internal automation and in-product features for a web-native trading platform. Own implementation, observability, and security while partnering with Product and Platform teams to deliver measurable AI systems in production. Job Description Role Senior/Principal Applied AI Engineer … integrate LLMs into internal systems and the trading product. This is a hands-on engineering role: you will write production code, put evaluation and observability on everything shipped, and operate autonomously in a lean team. Key Responsibilities Build agent loops, tool-calling, structured outputs, planning/state management, retries, guardrails ...

Network Engineer

Hiring Organisation
Laser Digital
Location
Cardiff, United Kingdom
Monitor and analyse performance, recommend improvements, and ensure systems meet demanding SLAs. Support colocation near exchanges and manage high-throughput, low-latency routes. Monitoring & Observability Build and manage comprehensive monitoring and logging systems for network performance, latency, and availability. Implement observability frameworks using modern tools to provide real-time insight ...

Senior AI Product Engineer

Location
Greater London, England, United Kingdom
more junior engineers through pair programming, code review, and design feedback. Raise the engineering bar across the team by promoting good practices in testing, observability, and AI system reliability. Influence cross-team decisions on how AI capabilities integrate with the rest of the Elliptic platform. What you will achieve … technical direction of an AI workstream, including architecture, evaluation, and rollout. Established or improved at least one team practice for building AI systems (evals, observability patterns, prompt management, rollout safety). Mentored junior engineers on AI engineering practices and contributed to their growth. Built strong working relationships across product ...

Platform Engineer

Location
Milton Keynes, England, United Kingdom
Terraform, CloudFormation, or CDK), CI/CD pipelines, API Gateway, Lambda, and Aurora PostgreSQL — and confident applying fundamentals such as high availability, fault tolerance, observability, and cost control in a live environment. Payments or fintech exposure is an advantage but not required; what matters more is a practical, first-principles … rollback processes Support deployment of AI generated applications and tooling Support change control and release management alongside Engineering and the Information Security Officer Reliability, Observability & Data Infrastructure Design and maintain systems for high availability, fault tolerance, and resilience by default Implement logging, monitoring, and alerting (for example CloudWatch) across services ...

Senior DevOps Engineer

Hiring Organisation
Experis
Location
Derby, Derbyshire, United Kingdom
Employment Type
Contract
/CD pipelines and deployment orchestration. Support Kubernetes and OpenShift platform troubleshooting and optimisation. Deliver secure and compliant infrastructure solutions. Implement monitoring, logging, and observability tooling across environments. Collaborate with engineering, architecture, and delivery teams to improve deployment efficiency and platform reliability. Champion automation-first approaches to infrastructure and application … designing and maintaining enterprise-scale CI/CD pipelines . Strong understanding of cloud security and secure delivery practices. Experience implementing monitoring, logging, and observability solutions. Ability to define technical standards, governance, and reusable deployment frameworks. Experience working within large-scale enterprise transformation programmes. Desirable Skills Experience within highly regulated ...

DevOps Engineer (AWS & Cloud Security)

Hiring Organisation
Ernest Gordon Recruitment Limited
Location
London, United Kingdom
Employment Type
Full-Time
Salary
£65,000 - £70,000 per annum
automate deployments using Terraform and Ansible, and build CI/CD pipelines using GitHub Actions. You'll also work across cloud security, networking and observability, while having the opportunity to develop your technical expertise through training and professional certifications. This role would suit an experienced DevOps Engineer looking to work … private cloud environments Automate infrastructure using Terraform and Ansible Build and maintain CI/CD pipelines using GitHub Actions Implement monitoring and observability using Grafana, Prometheus and CloudWatch Manage hybrid networking, IAM, firewalls and VPNs Improve infrastructure security, reliability and performance Support Kubernetes environments, including AWS EKS Join ...

DevOps Engineer (AWS & Cloud Security)

Hiring Organisation
Ernest Gordon Recruitment Limited
Location
Camden, London, Camden Town, United Kingdom
Employment Type
Permanent
Salary
£65000 - £70000/annum + Remote + Progression
automate deployments using Terraform and Ansible, and build CI/CD pipelines using GitHub Actions. You'll also work across cloud security, networking and observability, while having the opportunity to develop your technical expertise through training and professional certifications. This role would suit an experienced DevOps Engineer looking to work … private cloud environments Automate infrastructure using Terraform and Ansible Build and maintain CI/CD pipelines using GitHub Actions Implement monitoring and observability using Grafana, Prometheus and CloudWatch Manage hybrid networking, IAM, firewalls and VPNs Improve infrastructure security, reliability and performance Support Kubernetes environments, including AWS EKS Join ...

Lead Software Engineer - Platform

Location
Greater London, England, United Kingdom
need to scale accordingly, and our platform foundations need to support faster, more reliable delivery. You’ll be central to making that happen – from observability and developer experience to the infrastructure patterns that underpin everything we ship. What you’ll do Alongside the other Lead Engineers you’ll support … infrastructure. You’ll be responsible for the reliability, scalability, and operability of our systems. That means CI/CD pipelines, infrastructure‐as‐code, observability, incident response, and the day‐to‐day health of production. You’ll make sure we can ship with confidence and sleep at night. Shape technical direction. ...

Principal DevSecOps Engineer

Hiring Organisation
83zero Limited
Location
Bristol, Avon, South West, United Kingdom
Employment Type
Permanent, Work From Home
workflows * Establish secure-by-design engineering practices and enforce security and technical standards * Lead Infrastructure as Code (IaC) practices across teams and environments * Drive observability, monitoring, logging and audit controls * Support incident response, patching, compliance reporting and technical debt remediation * Partner with developers and delivery teams to improve engineering quality … Security & compliance - Trivy, vulnerability management, HashiCorp Vault, cert-manager * Containers & cloud - Docker, AWS EKS, AWS IAM, S3 and network policies * Infrastructure as Code - Terraform * Observability - Grafana, Loki * Automation - Python and Bash * Experience delivering within the UK Government Digital Service (GDS) lifecycle on a public sector engagement Why join ...

Site Reliability Engineer (SRE)

Hiring Organisation
Spencer Rose Ltd
Location
Manchester, Lancashire, United Kingdom
Employment Type
Contract
Contract Rate
GBP Daily
operational excellence of cloud-hosted services on Google Cloud Platform. This is a hands-on engineering role spanning SRE practices, production Kubernetes, infrastructure automation, observability, CI/CD, incident response and continuous service improvement. About the role The Senior Site Reliability Engineer will work with Cloud Platform, Software Engineering, Product … shared services. Define and operate service level indicators, service level objectives and error-budget practices that connect technical health to customer impact. Design actionable observability using Dynatrace, including instrumentation, dashboards, distributed tracing, service health views and SLO-based alerting. Build modular, reusable and maintainable Terraform code for secure cloud infrastructure ...

Senior / Principal Applied AI Engineer (UK / Europe, Remote)

Location
United Kingdom
Core Platform Engineering. This is a hands-on engineering seat, not an advisory one. You write and own production code, you put evaluation and observability on everything you ship, and you run autonomously in a lean team. You report to the COO and partner closely with the CTO, the Head … commodity internal needs; build or replace it internally when it’s differentiating or becomes cost-prohibitive. Ship outcomes, not architecture debates. Put evaluation, observability, and cost guardrails on everything - golden datasets, eval harnesses, tracing, fallback chains, latency and spend controls. Nothing ships as an unmeasured demo. Operate inside our security ...

SRE | Permanent | London, Hybrid, AWS

Hiring Organisation
Source Group International
Location
London, UK
Employment Type
Full-time
scalability. Key responsibilities Partner with engineering teams to define, measure, and manage SLOs/SLIs, using error budgets to guide delivery decisions. Enhance observability across services (metrics, logs, traces) to detect and resolve issues proactively. Lead cost optimisation: monitor spend, right-size workloads, tune autoscaling, and improve infrastructure efficiency. Improve … Kubernetes operational experience (on-prem and AWS EKS).Hands-on experience defining and operating SLOs/SLIs, alerting, and incident workflows. Deep understanding of observability and telemetry (monitoring, logging, tracing).Infrastructure as Code with Terraform; experience with GitOps workflows and CI/CD.Scripting proficiency in Python, Bash, or Go. Proven ...