251 to 275 of 1,339 Remote/Hybrid Observability Jobs

Senior Data Engineer

Location
Greater London, England, United Kingdom
help less experienced engineers grow through pairing, code review, and knowledge sharing. Promote best practices in data engineering, including testing, CI/CD, observability, and infrastructure as code. Platform & Architecture Design, build, and maintain scalable and secure data pipelines, warehouses, and streaming systems. Ensure data is modelled and structured ...

Engineering Manager - Data

Location
Oxford, England, United Kingdom
Lead the transition from traditional DBA operations towards code-defined, automated and rebuildable database environments, including automated schema migrations and immutable infrastructure approaches. Own observability, reliability and cost management across the data estate, establishing meaningful monitoring, alerting and clear service levels for the services teams depend on. Make data quality ...

AI Platform Engineer 3194

Hiring Organisation
Allianz Commercial
Location
Surrey, United Kingdom
Employment Type
Full Time
standards (frameworks, SDKs, Spec Kits, MCP interfaces), curating the AI tool landscape and managing lifecycles from sandbox to standardization or retirement. Implement platform-wide observability (metrics, logs, traces), SLOs, cost and performance instrumentation, and incident response processes. Partner with AI Gateway, governance owners, hyperscalers, AI frontier labs for regulated industry ...

Senior Software Engineer

Location
Greater London, England, United Kingdom
data warehousing or analytical data modelling alongside transactional systems. Experience designing API-first architectures, webhook integrations or third party platform integrations. Familiarity with observability and SRE practice: monitoring, logging, tracing and incident response. We're looking for people who enjoy the buzz of change, the satisfaction of building something better ...

Lead Development Engineer

Location
Greater London, England, United Kingdom
enabled development, to reduce friction while maintaining quality and sustainable delivery Champion proportionate engineering practices, including automated testing, CI/CD, coding standards, observability, technical debt management and maintainable code Contribute hands‐on to the design and development of shared packages and services Work with Platform and Infrastructure teams ...

Senior Product Engineer (Product Manager)

Hiring Organisation
Lendable
Location
London, UK
Employment Type
Full-time
leadership. Our tech stackPlatform & backend: Kotlin, AWS, Postgres, RabbitMQ, Docker, Kubernetes. Gateway & clients: Node GraphQL gateway, React Native/TypeScript mobile app, Relay. Tooling & observability: GitHub, GitHub Actions, Jira, Confluence, Datadog, Sentry, Grafana. Interview processQuick call with the Talent Team (30 minutes)Hiring manager interview (30 minutes)Technical and Product ...

Lead DevSecOps Engineer

Location
Greater London, England, United Kingdom
secure software development lifecycle and building automated solutions to resolve them. Architecting and evolving standardized deployment patterns that bake in security, compliance, and observability, allowing product teams to focus on delivering customer functionality. Ensuring high availability, performance, and seamless integration of our self-hosted engineering suite (GitLab) with the wider ...

Technical Solutions Architect

Hiring Organisation
Bertelsmann
Location
London, UK
Employment Type
Full-time
Functions, App Services, Container Apps or AKS, Key Vault and appropriate data services. Define non-functional and engineering requirements across security, resilience, performance, scalability, observability, supportability, disaster recovery, infrastructure as code, CI/CD and automated testing. Create safe modernisation and coexistence approaches for legacy and partner platforms, including adapter ...

Senior Product Engineer (Product Manager)

Location
Greater London, England, United Kingdom
tech stack Platform & backend: Kotlin, AWS, Postgres, RabbitMQ, Docker, Kubernetes. Gateway & clients: Node GraphQL gateway, React Native/TypeScript mobile app, Relay. Tooling & observability: GitHub, GitHub Actions, Jira, Confluence, Datadog, Sentry, Grafana. Interview process Quick call with the Talent Team (30 minutes) Hiring manager interview (30 minutes) Technical and Product ...

Blockchain Software Developer - Digital Assets Platform

Location
Greater London, England, United Kingdom
immutable nature of on-chain deployments Tooling Proficiency: Working knowledge of engineering tools including Git, Jira, Jenkins, Hardhat/Foundry, Docker, Helm, and observability platforms Development Value As a Blockchain Software Developer at Citi, you will be at the forefront of one of the most exciting and strategically significant technology ...

Blockchain Software Developer - Digital Assets Platform

Hiring Organisation
Citigroup
Location
London, UK
Employment Type
Full-time
given the immutable nature of on-chain deploymentsTooling Proficiency: Working knowledge of engineering tools including Git, Jira, Jenkins, Hardhat/Foundry, Docker, Helm, and observability platformsDevelopment ValueAs a Blockchain Software Developer at Citi, you will be at the forefront of one of the most exciting and strategically significant technology domains ...

Senior SRE & DevTools Engineer (CI/CD & Observability)

Location
Ham, England, United Kingdom
Visa is seeking a Software Engineer + SRE hybrid to join its UK Cloud platform team. You will safeguard reliability, automate resolution of recurring issues, and work with developers to optimize CI/CD pipelines. ...

Azure CloudOps Engineer

Location
Greater London, England, United Kingdom
resilient, scalable, and highly automated cloud platforms. The successful candidate will be responsible for designing and operating cloud infrastructure, implementing Infrastructure as Code, enhancing observability and AIOps capabilities, and driving automation across both application and infrastructure lifecycles. This role combines Cloud Engineering, DevOps, SRE, and AIOps practices, leveraging automation … optimise CI/CD pipelines supporting both application and infrastructure deployments. Develop and maintain automation scripts, deployment tooling, and operational workflows. Implement and support observability solutions, monitoring platforms, logging frameworks, and telemetry capabilities. Develop AIOps capabilities for anomaly detection, event correlation, alert noise reduction, and proactive issue identification. Create automated ...

Senior Backend Developer

Hiring Organisation
Protein Works
Location
Liverpool, Merseyside, United Kingdom
Employment Type
Full-Time
Salary
Competitive salary
connect e-commerce, ERP, WMS, finance, and manufacturing systems. Operations, Security & AI: Own end-to-end delivery including containers, CI/CD, IaC, full observability (metrics, traces, logs), security/privacy compliance (OWASP, GDPR, pen testing), and day-to-day use of AI-assisted/agentic tools. Cross-Functional Leadership … microservices and monoliths. Data & Async Systems: Solid background in relational databases (PostgreSQL, MySQL, SQL Server) and asynchronous messaging (queues, events, retries, idempotency). DevOps & Observability: Practical experience with Docker, Linux/Windows, Git/GitHub, Jira/Agile, monitoring/alerting, and diagnostic/profiling tools in high-volume ...

Senior Site Reliability Engineer

Location
Southampton, England, United Kingdom
such as Jenkins, GitLab CI/CD, or CircleCI. Strong knowledge of containerization technologies (e.g., Docker, Kubernetes) and microservices architecture. Experience with monitoring and observability tools (e.g., Prometheus, Grafana, ELK stack, Cloudwatch). Excellent problem-solving skills and the ability to troubleshoot complex issues in distributed systems. Experience of Incident … advantage if you also have: Handson experience of working with large Kubernetes Cluster. Certification will be an added plus. Working experience of Grafana Observability Suite (Loki, Mimir, Tempo). Administration and/or development experience of standard monitoring and automation tools such as Splunk, Datadog, Pagerduty Rundeck. Familiarity with configuration ...

Senior Site Reliability Engineer

Location
Greater London, England, United Kingdom
such as Jenkins, GitLab CI/CD, or CircleCI. Strong knowledge of containerization technologies (e.g., Docker, Kubernetes) and microservices architecture. Experience with monitoring and observability tools (e.g., Prometheus, Grafana, ELK stack, Cloudwatch). Excellent problem-solving skills and the ability to troubleshoot complex issues in distributed systems. Experience of Incident … advantage if you also have: Handson experience of working with large Kubernetes Cluster. Certification will be an added plus. Working experience of Grafana Observability Suite (Loki, Mimir, Tempo). Administration and/or development experience of standard monitoring and automation tools such as Splunk, Datadog, Pagerduty Rundeck. Familiarity with configuration ...

Senior Specialist Engineer (Specialist Site Reliability Engineer SRE)

Hiring Organisation
National Health Service
Location
London, United Kingdom
Salary
£ 70 K
scalable, and perform optimally in production environments. The role will monitor and manage these aspects while taking responsibility for multiple cloud infrastructure services. Observability of systems will be key to prioritising the operational service improvements and performance improvements to meet and exceed SLOs (Service Level Objectives).Main duties … solving skills to identify bottlenecks with an engineering mindsetEnsure systems can handle current and future workloads through automation and capacity planningContinuously improve services through observability, and identify ways to improve observability practicesFollow SRE principles. Guide and educate stakeholders to adopt implemented principlesProvide technical documentation for engineers. Providing training, where appropriateWorking ...

Infrastructure Python Developer

Location
Sheffield, England, United Kingdom
deploy application/services using Docker and Kubernetes Administration for Kubernetes Resources (Pods, Ingress, Services, Secrets, CRDs, etc) Nice to have Exposure on enhancing observability with knowledge of tools such as Prometheus, Grafana, and OpenTelemetry. Advantageous to have enterprise tools knowledge (i.e., Control M, True sight, Guardium, Tenable Nessus, Delinea ...

Infrastructure Python Developer

Hiring Organisation
Experis
Location
Sheffield, South Yorkshire, United Kingdom
Employment Type
Contract
Contract Rate
£350 - £402/day
deploy application/services using Docker and Kubernetes Administration for Kubernetes Resources (Pods, Ingress, Services, Secrets, CRDs, etc) Nice to have Exposure on enhancing observability with knowledge of tools such as Prometheus, Grafana, and OpenTelemetry. Advantageous to have enterprise tools knowledge (i.e., Control M, True sight, Guardium, Tenable Nessus, Delinea ...

Java Developer

Location
Greater London, England, United Kingdom
with React or front‐end/backend integration Experience working on enterprise‐scale platforms Familiarity with non‐functional requirements such as security, resilience, and observability #J-18808-Ljbffr ...

Senior Data Engineer

Hiring Organisation
FBI &TMT
Location
Kingston Upon Thames, Surrey, South East, United Kingdom
Employment Type
Permanent
Salary
£65,000
optimise batch and streaming workflows for reliability and performance * Contribute to lakehouse architecture and medallion pattern implementation * Implement data quality checks, monitoring and observability across pipelines * Apply platform security, access control and governance standards * Support code reviews and high engineering quality standards * Identify cloud cost optimisation opportunities * Translate business requirements ...

Lead Platform Engineer

Hiring Organisation
Oscar Associates (UK) Limited
Location
Nottingham, Nottinghamshire, East Midlands, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£75,000
Support cloud migration and separation projects across Azure and GCP. Design, automate and test disaster recovery and business continuity solutions. Improve monitoring, alerting and observability across the platform. Manage cloud security, identity, secrets and certificates . Provide technical guidance on Azure architecture and platform engineering . Troubleshoot complex infrastructure issues ...

Senior Data Engineer

Hiring Organisation
Tenth Revolution Group
Location
London, South East England, United Kingdom
Employment Type
Full-Time
Salary
£70,000 - £85,000 per annum
within Databricks. Developing and optimising batch and streaming data workflows. Designing and implementing Lakehouse architectures and medallion data models. Implementing data quality, monitoring, and observability frameworks. Maintaining security, governance, and access control standards. Collaborating with Data Scientists and Analysts to deliver business-critical data products. Driving engineering best practice through ...

Head of Cyber, Platforms & IT

Location
Chester, England, United Kingdom
enforce secure‐by‐design principles, including cybersecurity standards, cloud architecture guardrails and operational controls Lead DevOps and Site Reliability Engineering (SRE) maturity, embedding monitoring, observability, automated testing and structured incident response Drive adoption of automation and AI‐enabled tooling to improve anomaly detection, incident management, vulnerability management and operational efficiency ...

Principal Engineer

Hiring Organisation
Formula Recruitment Limited
Location
London, UK
Employment Type
Full-time
large-scale API platform across both synchronous and event-driven servicesLead architectural design across a modern, cloud-native stackDrive continuous improvement across security, observability, performance, and long-term maintainabilityMentor and develop engineers and technical leads across multiple teamsWork closely with Product and Delivery leadership to keep engineering strategy aligned with ...