376 to 400 of 502 Remote/Hybrid Observability Jobs

Lead Platform Engineer - Cloud & Security (Remote)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Platform Engineer to own the infrastructure and security foundation that our engineering teams rely on. You’ll manage cloud infrastructure end-to-end, establish observability practices, and drive security posture across the platform in a hybrid environment with remote options. Based in London or remote in Europe, you’ll play ...

Senior Backend Engineer - Remote or Hybrid, Warehouse Data

Hiring Organisation
Jobleads-UK
Location
Cambridge, England, United Kingdom
contracts that power the product — ingestion pipelines transforming warehouse data into a model, robust APIs, durable storage, and reliable background work. You will ensure observability that lets a small team operate them confidently at 3 am. Two engines power WareBee: Physical AI and Process AI. You will build systems ...

Principal AI Infrastructure Architect - Remote

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
squads to set standards and deliver scalable systems. This role reports to the Engineering Director, focusing on real-time data pipelines, LLM-driven agents, observability, and privacy-driven data governance, enabling the next generation of Grip's event platform. #J-18808-Ljbffr ...

Senior C++ Engineer

Hiring Organisation
Randstad Digital
Location
London, United Kingdom
Employment Type
Contract
Contract Rate
£600 - £700 per day
shared library across iOS and Android layers. Evolve Protocols: Shape client-server contracts, handle network state detection, and optimize real-time messaging systems. Enhance Observability: Build deep telemetry and metrics to diagnose latency and performance anomalies across the full client-to-backend path. Kill Legacy Code: Gracefully sunset legacy networking ...

Principal Software Architect

Hiring Organisation
Spectrum It Recruitment Limited
Location
Uxbridge, London, United Kingdom
Employment Type
Permanent, Work From Home
service boundaries, ownership models and integration patterns Reviewing significant technical initiatives and providing architectural guidance across multiple engineering teams Embedding security, resilience, scalability and observability into platform design from the outset Identifying architectural risk, technical debt and platform constraints Partnering with Engineering and Product leadership on strategic technical decisions Influencing ...

Principal Software Architect

Hiring Organisation
17918
Location
London, United Kingdom
service boundaries, ownership models and integration patterns Reviewing significant technical initiatives and providing architectural guidance across multiple engineering teams Embedding security, resilience, scalability and observability into platform design from the outset Identifying architectural risk, technical debt and platform constraints Partnering with Engineering and Product leadership on strategic technical decisions Influencing ...

Lead News Architect

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
quality across the platform. Partner closely with engineering teams improving signal discoverability through normalizing metadata and implementing industry standards. Ensure data architectures support observability and lineage while providing customers with a simple, straight‐forward experience for working with News from LSEG. We encourage pragmatic experimentation with new tools, libraries ...

ServiceNow AI & Enterprise Automation Lead - Managing Consultant

Hiring Organisation
Capgemini
Location
Hampshire, United Kingdom
Employment Type
Full Time
value Translate business requirements into AI-enabled workflow solutions Solution Design & Architecture Design and support implementation of: AI Control Tower (AI lifecycle management, governance, observability) Agentic AI workflows enabling autonomous execution Now Assist/GenAI use cases across workflows Define data, integration, and workflow architectures for AI-enabled ServiceNow solutions ...

ServiceNow AI & Enterprise Automation Lead - Managing Consultant

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
driven business valueTranslate business requirements into AI-enabled workflow solutionsSolution Design & ArchitectureDesign and support implementation of:AI Control Tower (AI lifecycle management, governance, observability)Agentic AI workflows enabling autonomous executionNow Assist/GenAI use cases across workflowsDefine data, integration, and workflow architectures for AI-enabled ServiceNow solutionsConversational AI & Employee ExperienceSupport ...

ServiceNow CRM Transformation Lead - Managing Consultant

Hiring Organisation
Capgemini
Location
Hampshire, United Kingdom
Employment Type
Full Time
realisation Translate business requirements into workflow-enabled operating models Solution Design & Architecture Design and support implementation of: AI Control Tower (AI lifecycle management, governance, observability) Agentic AI workflows enabling autonomous execution Now Assist/GenAI use cases across workflows Define data, integration, and workflow architectures for AI-enabled ServiceNow solutions ...

Network Automation Engineer

Hiring Organisation
HCLTech
Location
City of London, London, United Kingdom
Modern Ops and AI-first operating model. The role focuses on Network Infrastructure as Code (NetIaC), CI/CD pipelines, AI-driven operations (AIOps), observability integration, and SRE-led reliability engineering. Key Responsibilities Develop and manage Network Infrastructure as Code (NetIaC) using Python, Ansible, and Terraform for provisioning and lifecycle … ITSM workflows. Drive AI/ML use cases such as WAN capacity forecasting, anomaly detection, predictive analytics, and self-healing networks. Integrate and manage observability platforms (SolarWinds Orion, Elastic, Grafana, ZDX) for proactive monitoring and insights. Provide engineering and support for MCP (Model Context Protocol) and AI agent integrations. Ensure ...

Head of Site Reliability Engineering (SRE)

Hiring Organisation
Jobleads-UK
Location
Bristol, England, United Kingdom
Head of SRE to define, lead and evolve our global reliability strategy. This senior leadership role is responsible for driving operational excellence, service reliability, observability, automation and continuous improvement across our technology landscape. Key Responsibilities Drive adoption of SRE principles (SLOs, error budgets, toil reduction). Establish observability and monitoring … with strong scripting and development capabilities using technologies such as Python, PowerShell, Bash, Terraform and Ansible Automation Platform. Other key skills: Robust knowledge of observability and monitoring practices, and experience implementing and managing platforms such as Dynatrace, Prometheus, Grafana, and Splunk. Good understanding of CI/CD tooling and modern ...

Growth Marketing Lead for B2B DevTools (Remote/London)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
software engineering experience Proficiency in Go, Python, or TypeScript Experience with Kubernetes or containerized applications Strong problem-solving skills Nice to have Experience with observability tools (Prometheus, Grafana, etc.) Frontend development experience with React Experience with data processing pipelines Growth & Marketing Lead #growth-marketing-lead Full-time London/Remote … relationship-building skills Ability to understand developer, platform, or DevOps workflows Nice to have Experience selling to engineering leaders or platform teams Background in observability, Kubernetes, or cloud infrastructure Experience with founder-led sales or building outbound playbooks Why Metoro Remember the last time your cluster died ...

NOC Engineer, AWS

Hiring Organisation
Spectrum IT Recruitment
Location
Reading, Berkshire, United Kingdom
Employment Type
Permanent
Salary
£60000 - £65000/annum
issues and restoring services quickly and effectively Developing automation to reduce manual operational tasks and improve platform resilience Building and improving monitoring, alerting and observability across cloud environments Working alongside Software, Platform, Cloud and Security Engineers to improve reliability and operational excellence Contributing to post-incident reviews and driving continuous … with exposure to: Linux systems administration AWS cloud infrastructure Kubernetes and Docker Production support and incident management Python, Bash or Go scripting Monitoring and observability platforms such as Grafana, Prometheus, Datadog, Splunk or CloudWatch Networking fundamentals including DNS, TCP/IP and load balancing A passion for automation, continuous improvement ...

NOC Engineer (AWS)

Hiring Organisation
Spectrum It Recruitment Limited
Location
Basingstoke, Hampshire, South East, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£60,000
issues and restoring services quickly and effectively Developing automation to reduce manual operational tasks and improve platform resilience Building and improving monitoring, alerting and observability across cloud environments Working alongside Software, Platform, Cloud and Security Engineers to improve reliability and operational excellence Contributing to post-incident reviews and driving continuous … with exposure to: Linux systems administration AWS cloud infrastructure Kubernetes and Docker Production support and incident management Python, Bash or Go scripting Monitoring and observability platforms such as Grafana, Prometheus, Datadog, Splunk or CloudWatch Networking fundamentals including DNS, TCP/IP and load balancing A passion for automation, continuous improvement ...

Azure Platform Engineer

Hiring Organisation
EMBS Engineering
Location
Newbury, West Berkshire, Berkshire, United Kingdom
Employment Type
Permanent
Salary
£65000 - £75000/annum + Benefits
Azure platform templates and engineering patterns to support consistent platform adoption. Build and improve CI/CD pipelines, deployment automation and release processes. Implement observability, monitoring, logging, resilience and operational readiness across the platform. Embed FinOps principles, improving cloud cost visibility and optimisation. Work closely with engineering teams throughout sprint … engineering patterns. Strong understanding of DevOps and DevSecOps practices including CI/CD, source control, automated testing and release management. Experience with Azure observability, monitoring, alerting, logging and platform reliability. Practical knowledge of cloud security, governance and enterprise engineering standards. Experience applying FinOps principles including tagging strategies, right-sizing ...

Principal Platform Engineer

Hiring Organisation
Sanderson Recruitment
Location
City of London, London, United Kingdom
Employment Type
Permanent
persistence platforms Provide technical leadership and architectural guidance across multiple engineering teams Define engineering standards, platform roadmaps and best practices Drive automation, resilience, observability and operational excellence initiatives Support and mentor engineers through code reviews, coaching and technical leadership Collaborate with architects and stakeholders to translate business requirements into technical … automation and DevOps practices Experience mentoring engineers and providing technical leadership Key Technologies AWS Terraform Linux Cassandra Couchbase ScyllaDB Kafka CI/CD Pipelines Observability & Monitoring Platforms Distributed Database Technologies Nice to Have Experience with additional distributed persistence technologies Background in large-scale cloud-native environments Experience defining enterprise platform ...

DevSecOps Engineering Lead CGEMJP00346044

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
management, including allowlist processes and risk acceptance where required Secrets management and identity/access management Policy enforcement for workloads, container images and infrastructure Observability, monitoring, logging and audit controls Partner with developers to embed secure-by-design engineering and ensure compliance with CLIENT security standards. Enable and govern Infrastructure … compliance tooling (e.g. Trivy scanning and vulnerability management, HashiCorp Vault, cert-manager) Containers and orchestration (e.g. Docker, AWS EKS) Infrastructure as Code (e.g. Terraform) Observability (e.g. Grafana, Loki) Scripting and automation (e.g. Python, Bash) Cloud and networking fundamentals (e.g. AWS IAM, S3, network policies) Experience delivering within the UK Government ...

Site Reliability Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
development, validation, and optimization of configuration-as-code, improving delivery speed and reducing deployment risk. Adaptable & Problem-Solver: Address complex challenges across configuration, policy, observability, and data services. Apply a data-driven approach using Prometheus and Grafana to improve reliability and performance. Ownership & Quality: Own end-to-end configuration quality … Qualifications Hands-on with Helm or Kustomize. Experience with GitOps (e.g., Argo CD). Knowledge of secrets management (e.g., HashiCorp Vault). Experience with observability (metrics/logs/tracing). CollabHiring. Why Cisco? At Cisco, we’re revolutionizing how data and infrastructure connect and protect organizations ...

Platform Engineer

Hiring Organisation
Hireful
Location
Central London, London, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£80,000
We are recruiting founding Platform Engineers on behalf of a fast-growing enterprise level (global, 500+ staff) software business with a strong engineering culture and a genuine commitment to doing things the right way. They ...

Linux Systems Engineer

Hiring Organisation
Wolviston Management Services
Location
Glasgow, City of Glasgow, United Kingdom
Employment Type
Permanent
include: Large-scale Linux operating system upgrades. Enterprise patching and vulnerability remediation. Infrastructure lifecycle management. Development of automation tooling and processes. Platform monitoring and observability improvements. Supporting the adoption of modern infrastructure engineering practices. Key Responsibilities Linux Infrastructure Engineering Administer and support large-scale enterprise Linux environments. Perform troubleshooting … workflows to improve infrastructure efficiency. Create validation and verification processes to support change activity. Reduce manual intervention through automation and engineering best practice. Monitoring & Observability Enhance monitoring, alerting and observability capabilities across infrastructure platforms. Develop dashboards and operational reporting. Work with technologies such as: Prometheus Grafana Loki Support proactive platform ...

Management Consultant - Cloud DevOps

Hiring Organisation
Capgemini
Location
Greater London, United Kingdom
Employment Type
Full Time
Salary
500000 GBP Annually
product-centric operating models. Solid understanding of hybrid/multi-cloud environments, DevOps, CI/CD, SRE, DevSecOps models, DevX, build and deployment pipelines, observability, and ITIL.Optional: Required certifications, licenses or languages Understanding of observability and monitoring platform; Experience of having collaborated with developers to implement and improve observability ...

Data Reliability Engineer

Hiring Organisation
Ashdown Group
Location
City, London, United Kingdom
Employment Type
Permanent
Salary
GBP 95,000 Annual
work from home 2 days per week. This is a high-impact role focused on improving data quality, reducing incidents, and building scalable observability across a modern enterprise data platform click apply for full job details ...

Remote SRE - Big Data, Kafka & Distributed Storage

Hiring Organisation
Jobleads-UK
Location
United Kingdom
engineer to own architecture, deployment, and capacity planning across hybrid infrastructure. You’ll optimize Kafka and Ceph pipelines, automate operational tasks, and contribute to observability tooling. Remote-friendly with US ET hours, you’ll help scale systems for multiple engineering teams. #J-18808-Ljbffr ...

AI-First Engineering Manager, Manage Squad Lead

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Reporting to the CPO, you will mentor four engineers, partner with Product and Design, own hiring, ensure quality and platform health, and drive security, observability and governance across the domain. #J-18808-Ljbffr ...