901 to 925 of 3,664 Remote/Hybrid Observability Jobs

DevOps Engineer (AWS & Cloud Security)

Hiring Organisation
Ernest Gordon Recruitment Limited
Location
Camden, London, Camden Town, United Kingdom
Employment Type
Permanent
Salary
£65000 - £70000/annum + Remote + Progression
automate deployments using Terraform and Ansible, and build CI/CD pipelines using GitHub Actions. You'll also work across cloud security, networking and observability, while having the opportunity to develop your technical expertise through training and professional certifications. This role would suit an experienced DevOps Engineer looking to work … private cloud environments Automate infrastructure using Terraform and Ansible Build and maintain CI/CD pipelines using GitHub Actions Implement monitoring and observability using Grafana, Prometheus and CloudWatch Manage hybrid networking, IAM, firewalls and VPNs Improve infrastructure security, reliability and performance Support Kubernetes environments, including AWS EKS Join ...

Principal DevSecOps Engineer

Hiring Organisation
83zero Limited
Location
Bristol, Avon, South West, United Kingdom
Employment Type
Permanent, Work From Home
workflows * Establish secure-by-design engineering practices and enforce security and technical standards * Lead Infrastructure as Code (IaC) practices across teams and environments * Drive observability, monitoring, logging and audit controls * Support incident response, patching, compliance reporting and technical debt remediation * Partner with developers and delivery teams to improve engineering quality … Security & compliance - Trivy, vulnerability management, HashiCorp Vault, cert-manager * Containers & cloud - Docker, AWS EKS, AWS IAM, S3 and network policies * Infrastructure as Code - Terraform * Observability - Grafana, Loki * Automation - Python and Bash * Experience delivering within the UK Government Digital Service (GDS) lifecycle on a public sector engagement Why join ...

Senior Mobile Engineering Manager

Hiring Organisation
Innova Solutions
Location
London, United Kingdom
Employment Type
Contract
Contract Rate
£48 - £58/hour
Boot, Gradle build automation Strong understanding of: Microservices architecture RESTful APIs Event-driven architectures Secure software development practices Scalable distributed systems Deep experience implementing observability strategies, including: Logging, monitoring, alerting, Performance Analysis and Production Diagnostics. Experience with industry-standard observability and monitoring platforms such as Sentry or Datadog, New Relic ...

Site Reliability Engineer (SRE)

Hiring Organisation
Spencer Rose Ltd
Location
Manchester, Lancashire, United Kingdom
Employment Type
Contract
Contract Rate
GBP Daily
operational excellence of cloud-hosted services on Google Cloud Platform. This is a hands-on engineering role spanning SRE practices, production Kubernetes, infrastructure automation, observability, CI/CD, incident response and continuous service improvement. About the role The Senior Site Reliability Engineer will work with Cloud Platform, Software Engineering, Product … shared services. Define and operate service level indicators, service level objectives and error-budget practices that connect technical health to customer impact. Design actionable observability using Dynatrace, including instrumentation, dashboards, distributed tracing, service health views and SLO-based alerting. Build modular, reusable and maintainable Terraform code for secure cloud infrastructure ...

Lead Software Engineer - Platform

Location
Greater London, England, United Kingdom
need to scale accordingly, and our platform foundations need to support faster, more reliable delivery. You’ll be central to making that happen – from observability and developer experience to the infrastructure patterns that underpin everything we ship. What you’ll do Alongside the other Lead Engineers you’ll support … infrastructure. You’ll be responsible for the reliability, scalability, and operability of our systems. That means CI/CD pipelines, infrastructure‐as‐code, observability, incident response, and the day‐to‐day health of production. You’ll make sure we can ship with confidence and sleep at night. Shape technical direction. ...

Senior AI Product Engineer

Location
Greater London, England, United Kingdom
more junior engineers through pair programming, code review, and design feedback. Raise the engineering bar across the team by promoting good practices in testing, observability, and AI system reliability. Influence cross-team decisions on how AI capabilities integrate with the rest of the Elliptic platform. What you will achieve … technical direction of an AI workstream, including architecture, evaluation, and rollout. Established or improved at least one team practice for building AI systems (evals, observability patterns, prompt management, rollout safety). Mentored junior engineers on AI engineering practices and contributed to their growth. Built strong working relationships across product ...

SRE | Permanent | London, Hybrid, AWS

Hiring Organisation
Source Group International
Location
London, UK
Employment Type
Full-time
scalability. Key responsibilities Partner with engineering teams to define, measure, and manage SLOs/SLIs, using error budgets to guide delivery decisions. Enhance observability across services (metrics, logs, traces) to detect and resolve issues proactively. Lead cost optimisation: monitor spend, right-size workloads, tune autoscaling, and improve infrastructure efficiency. Improve … Kubernetes operational experience (on-prem and AWS EKS).Hands-on experience defining and operating SLOs/SLIs, alerting, and incident workflows. Deep understanding of observability and telemetry (monitoring, logging, tracing).Infrastructure as Code with Terraform; experience with GitOps workflows and CI/CD.Scripting proficiency in Python, Bash, or Go. Proven ...

Senior / Principal Applied AI Engineer (UK / Europe, Remote)

Location
United Kingdom
Core Platform Engineering. This is a hands-on engineering seat, not an advisory one. You write and own production code, you put evaluation and observability on everything you ship, and you run autonomously in a lean team. You report to the COO and partner closely with the CTO, the Head … commodity internal needs; build or replace it internally when it’s differentiating or becomes cost-prohibitive. Ship outcomes, not architecture debates. Put evaluation, observability, and cost guardrails on everything - golden datasets, eval harnesses, tracing, fallback chains, latency and spend controls. Nothing ships as an unmeasured demo. Operate inside our security ...

Engineering Manager - Grafana Application Security | EMEA | (Remote, UK)

Hiring Organisation
Grafana Labs
Location
United Kingdom
Salary
£ 70 K
Grafana Labs is the company behind Grafana Cloud, the fully managed observability platform trusted by more than 10,000 organizations to ensure reliability, resolve incidents faster, and optimize telemetry at scale. Built on open source and open standards and designed for interoperability across any stack, Grafana Cloud brings … observability and observability to AI, giving teams (and their agents) unified visibility so they can see, understand, and act on all their disparate data, wherever it lives, and move at the speed of their ambitions. Customers, including Anthropic, Bloomberg, NVIDIA, Microsoft, and Salesforce, rely on Grafana Labs. ...

Software Engineer (Next.js Playwright)

Location
Greater London, England, United Kingdom
maintain applications within Azure. Contribute to CI/CD pipelines using GitHub Actions and Azure DevOps. Monitor application performance and reliability using modern observability tools. Collaborate Across the Business Work closely with clinicians, product managers and fellow engineers to understand user needs. Contribute ideas that improve both the product … Experience with Playwright or other automated testing frameworks. Experience working within Azure. Experience in healthcare, NHS, or regulated environments. Familiarity with application monitoring and observability tools. Experience working in a SaaS or scale-up environment. Mindset Product-minded and user-focused. Pragmatic, collaborative and delivery-oriented. Takes ownership and sees ...

Remote Senior Director, Engineering- X-Ops Platform

Hiring Organisation
Sophos
Location
Essex, United Kingdom
vision, strategy, and operating model for the X-Ops Platform organisation (e.g., Delivery Roadmap, AI First Developer Experience, CI/CD, Observability, SRE, Cloud communicate progress, risks, and trade-offs clearly to executive stakeholders What You Will Bring Cyber Security background and acumen Significant engineering leadership experience, including leading leaders … record of building reliable, secure platforms and improving developer productivity through pragmatic, outcome-driven investment Experience establishing operational excellence practices (incident response, on-call, observability, SLOs, post-incident learning) and driving continuous improvement Ability to set strategy and translate it into execution via roadmaps, prioritisation, and clear measures of success ...

Remote Senior Director, Engineering- X-Ops Platform

Hiring Organisation
Sophos
Location
Dorset, United Kingdom
vision, strategy, and operating model for the X-Ops Platform organisation (e.g., Delivery Roadmap, AI First Developer Experience, CI/CD, Observability, SRE, Cloud communicate progress, risks, and trade-offs clearly to executive stakeholders What You Will Bring Cyber Security background and acumen Significant engineering leadership experience, including leading leaders … record of building reliable, secure platforms and improving developer productivity through pragmatic, outcome-driven investment Experience establishing operational excellence practices (incident response, on-call, observability, SLOs, post-incident learning) and driving continuous improvement Ability to set strategy and translate it into execution via roadmaps, prioritisation, and clear measures of success ...

Remote Senior Director, Engineering- X-Ops Platform

Hiring Organisation
Sophos
Location
Warwickshire, United Kingdom
vision, strategy, and operating model for the X-Ops Platform organisation (e.g., Delivery Roadmap, AI First Developer Experience, CI/CD, Observability, SRE, Cloud communicate progress, risks, and trade-offs clearly to executive stakeholders What You Will Bring Cyber Security background and acumen Significant engineering leadership experience, including leading leaders … record of building reliable, secure platforms and improving developer productivity through pragmatic, outcome-driven investment Experience establishing operational excellence practices (incident response, on-call, observability, SLOs, post-incident learning) and driving continuous improvement Ability to set strategy and translate it into execution via roadmaps, prioritisation, and clear measures of success ...

Remote Senior Director, Engineering- X-Ops Platform

Hiring Organisation
Sophos
Location
Leicestershire, United Kingdom
vision, strategy, and operating model for the X-Ops Platform organisation (e.g., Delivery Roadmap, AI First Developer Experience, CI/CD, Observability, SRE, Cloud communicate progress, risks, and trade-offs clearly to executive stakeholders What You Will Bring Cyber Security background and acumen Significant engineering leadership experience, including leading leaders … record of building reliable, secure platforms and improving developer productivity through pragmatic, outcome-driven investment Experience establishing operational excellence practices (incident response, on-call, observability, SLOs, post-incident learning) and driving continuous improvement Ability to set strategy and translate it into execution via roadmaps, prioritisation, and clear measures of success ...

Remote Senior Director, Engineering- X-Ops Platform

Hiring Organisation
Sophos
Location
Tyne and wear, United Kingdom
vision, strategy, and operating model for the X-Ops Platform organisation (e.g., Delivery Roadmap, AI First Developer Experience, CI/CD, Observability, SRE, Cloud communicate progress, risks, and trade-offs clearly to executive stakeholders What You Will Bring Cyber Security background and acumen Significant engineering leadership experience, including leading leaders … record of building reliable, secure platforms and improving developer productivity through pragmatic, outcome-driven investment Experience establishing operational excellence practices (incident response, on-call, observability, SLOs, post-incident learning) and driving continuous improvement Ability to set strategy and translate it into execution via roadmaps, prioritisation, and clear measures of success ...

Remote Senior Director, Engineering- X-Ops Platform

Hiring Organisation
Sophos
Location
Isle of wight, United Kingdom
vision, strategy, and operating model for the X-Ops Platform organisation (e.g., Delivery Roadmap, AI First Developer Experience, CI/CD, Observability, SRE, Cloud communicate progress, risks, and trade-offs clearly to executive stakeholders What You Will Bring Cyber Security background and acumen Significant engineering leadership experience, including leading leaders … record of building reliable, secure platforms and improving developer productivity through pragmatic, outcome-driven investment Experience establishing operational excellence practices (incident response, on-call, observability, SLOs, post-incident learning) and driving continuous improvement Ability to set strategy and translate it into execution via roadmaps, prioritisation, and clear measures of success ...

Remote Senior Director, Engineering- X-Ops Platform

Hiring Organisation
Sophos
Location
Mid and east antrim, United Kingdom
vision, strategy, and operating model for the X-Ops Platform organisation (e.g., Delivery Roadmap, AI First Developer Experience, CI/CD, Observability, SRE, Cloud communicate progress, risks, and trade-offs clearly to executive stakeholders What You Will Bring Cyber Security background and acumen Significant engineering leadership experience, including leading leaders … record of building reliable, secure platforms and improving developer productivity through pragmatic, outcome-driven investment Experience establishing operational excellence practices (incident response, on-call, observability, SLOs, post-incident learning) and driving continuous improvement Ability to set strategy and translate it into execution via roadmaps, prioritisation, and clear measures of success ...

Forward Deployed AI Engineer

Hiring Organisation
Willis Towers Watson
Location
London, UK
Employment Type
Full-time
enabled systems. You'll bring deep expertise across modern full-stack technologies (.NET, Azure, SQL, React/Angular), along with experience in distributed systems, observability, and AI tooling such as LLMs, retrieval pipelines, agentic workflows, and platforms such as Anthropic Claude. Experience designing and deploying AI agents, leveraging Model Context … orchestration, evaluation loops, and human-in-the-loop controls. Enterprise integration: Integrate AI solutions with enterprise systems, APIs, data platforms, document repositories, workflow tools, observability platforms, and identity and access management services. Production engineering: Ensure AI solutions meet enterprise standards for reliability, scalability, latency, maintainability, cost control, logging, monitoring ...

Remote Senior Director, Engineering- X-Ops Platform

Location
United Kingdom
vision, strategy, and operating model for the X-Ops Platform organisation (e.g., Delivery Roadmap, AI First Developer Experience, CI/CD, Observability, SRE, Cloud & Infrastructure Enablement) Lead, coach, and develop engineering leaders (directors, managers and senior ICs), building high-performing teams with clear ownership and strong engineering culture Own platform … record of building reliable, secure platforms and improving developer productivity through pragmatic, outcome-driven investment Experience establishing operational excellence practices (incident response, on-call, observability, SLOs, post-incident learning) and driving continuous improvement Ability to set strategy and translate it into execution via roadmaps, prioritisation, and clear measures of success ...

Remote Senior Director, Engineering- X-Ops Platform

Location
Wrexham, Denbighshire, United Kingdom
vision, strategy, and operating model for the X-Ops Platform organisation (e.g., Delivery Roadmap, AI First Developer Experience, CI/CD, Observability, SRE, Cloud & Infrastructure Enablement) Lead, coach, and develop engineering leaders (directors, managers and senior ICs), building high-performing teams with clear ownership and strong engineering culture Own platform … record of building reliable, secure platforms and improving developer productivity through pragmatic, outcome-driven investment Experience establishing operational excellence practices (incident response, on-call, observability, SLOs, post-incident learning) and driving continuous improvement Ability to set strategy and translate it into execution via roadmaps, prioritisation, and clear measures of success ...

Remote Senior Director, Engineering- X-Ops Platform

Location
Burntisland, Fife, United Kingdom
vision, strategy, and operating model for the X-Ops Platform organisation (e.g., Delivery Roadmap, AI First Developer Experience, CI/CD, Observability, SRE, Cloud & Infrastructure Enablement) Lead, coach, and develop engineering leaders (directors, managers and senior ICs), building high-performing teams with clear ownership and strong engineering culture Own platform … record of building reliable, secure platforms and improving developer productivity through pragmatic, outcome-driven investment Experience establishing operational excellence practices (incident response, on-call, observability, SLOs, post-incident learning) and driving continuous improvement Ability to set strategy and translate it into execution via roadmaps, prioritisation, and clear measures of success ...

Remote Senior Director, Engineering- X-Ops Platform

Location
Carmarthen, Carmarthenshire, United Kingdom
vision, strategy, and operating model for the X-Ops Platform organisation (e.g., Delivery Roadmap, AI First Developer Experience, CI/CD, Observability, SRE, Cloud & Infrastructure Enablement) Lead, coach, and develop engineering leaders (directors, managers and senior ICs), building high-performing teams with clear ownership and strong engineering culture Own platform … record of building reliable, secure platforms and improving developer productivity through pragmatic, outcome-driven investment Experience establishing operational excellence practices (incident response, on-call, observability, SLOs, post-incident learning) and driving continuous improvement Ability to set strategy and translate it into execution via roadmaps, prioritisation, and clear measures of success ...

Remote Senior Director, Engineering- X-Ops Platform

Location
Cardigan, Cardiganshire, United Kingdom
vision, strategy, and operating model for the X-Ops Platform organisation (e.g., Delivery Roadmap, AI First Developer Experience, CI/CD, Observability, SRE, Cloud & Infrastructure Enablement) Lead, coach, and develop engineering leaders (directors, managers and senior ICs), building high-performing teams with clear ownership and strong engineering culture Own platform … record of building reliable, secure platforms and improving developer productivity through pragmatic, outcome-driven investment Experience establishing operational excellence practices (incident response, on-call, observability, SLOs, post-incident learning) and driving continuous improvement Ability to set strategy and translate it into execution via roadmaps, prioritisation, and clear measures of success ...

Remote Senior Director, Engineering- X-Ops Platform

Location
Banchory, Aberdeenshire, United Kingdom
vision, strategy, and operating model for the X-Ops Platform organisation (e.g., Delivery Roadmap, AI First Developer Experience, CI/CD, Observability, SRE, Cloud & Infrastructure Enablement) Lead, coach, and develop engineering leaders (directors, managers and senior ICs), building high-performing teams with clear ownership and strong engineering culture Own platform … record of building reliable, secure platforms and improving developer productivity through pragmatic, outcome-driven investment Experience establishing operational excellence practices (incident response, on-call, observability, SLOs, post-incident learning) and driving continuous improvement Ability to set strategy and translate it into execution via roadmaps, prioritisation, and clear measures of success ...

Remote Senior Director, Engineering- X-Ops Platform

Location
Caldicot, Monmouthshire, United Kingdom
vision, strategy, and operating model for the X-Ops Platform organisation (e.g., Delivery Roadmap, AI First Developer Experience, CI/CD, Observability, SRE, Cloud & Infrastructure Enablement) Lead, coach, and develop engineering leaders (directors, managers and senior ICs), building high-performing teams with clear ownership and strong engineering culture Own platform … record of building reliable, secure platforms and improving developer productivity through pragmatic, outcome-driven investment Experience establishing operational excellence practices (incident response, on-call, observability, SLOs, post-incident learning) and driving continuous improvement Ability to set strategy and translate it into execution via roadmaps, prioritisation, and clear measures of success ...