176 to 200 of 392 Remote/Hybrid Observability Jobs

Head of Software Engineering

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
customer demand grows. Lead on engineering security practices, particularly in the context of UK Government and defence data handling. Work with Ops to improve observability, incident response, and the reliability of a 24/7 operational system. Essential Skills and Experience Strong full-stack background: comfortable with cloud infrastructure.Proven experience ...

Head of AI Engineering

Hiring Organisation
Jobleads-UK
Location
Bexhill-on-Sea, England, United Kingdom
maintain engineering standards, AI guardrails and Responsible AI practices. Oversee model development, deployment, monitoring, evaluation and continuous improvement processes. Drive preventative controls, automation, observability and root‐cause analysis to minimise technical debt and service issues. Partner with Architecture, Risk, Compliance and business stakeholders to shape AI roadmaps and align delivery ...

Senior Cloud Architect (AWS) (Multiple)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
architecture; evaluating trends and technology Leading client CCoE transformation, establishing RACI models and operational governance Leading the client’s cloud control strategy, aligning metrics, observability, and governance across business Ensuring adherence to WellArchitected principles across the client’s enterprise solutions Overseeing the client’s cost optimisation strategy, whilst ensuring ...

Senior Product Manager - API Platform

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
product representative for API consumers (internal and external) Collaborate with the DevEx team on: Documentation and onboarding Sandbox and testing environments Observability and error reporting User developer feedback to continuously improve the API experience What success looks like New API use cases launched and adopted by partners Reduced integration time ...

Head of Integration, Data & GenAI Engineering

Hiring Organisation
Jobleads-UK
Location
Manchester, England, United Kingdom
business value, with appropriate guardrails around security, risk and responsible use. Ensuring engineering solutions are secure, scalable, resilient and supportable, with appropriate governance, standards, observability and controls built in from the outset. Working closely with product-aligned engineering teams, infrastructure, information security, architecture and business stakeholders to deliver joined ...

Engineering Manager, Connectivity - London

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
with the depth to guide architectural decisions and engage credibly in technical discussions Experience operating production distributed systems, with an understanding of what reliability, observability, and incident response require at scale Excellent communication skills, with the ability to build consensus across teams and time zones Demonstrated success building a culture ...

DevOps / Cloud / Platform Engineer (All Levels) - UK Wide

Hiring Organisation
describe.me
Location
London, South East, England, United Kingdom
Employment Type
Full-Time
Salary
£50,000 - £130,000 per annum
systems that everything else runs on. You'll work across the full platform lifecycle—from infrastructure design and provisioning through to CI/CD, observability, incident response and cost optimisation. The role suits someone who pairs strong infrastructure-as-code discipline with a genuine automation-first mindset and a real … frequently Own infrastructure-as-code (Terraform, Pulumi, CloudFormation or equivalent) and the workflows around it Operate Kubernetes clusters and supporting platform services Implement observability—metrics, logs, traces, dashboards, alerting Lead incident response, root-cause analysis and reliability improvements Drive cloud cost optimisation and capacity planning Implement security hardening, secret management ...

Security Cloud Engineer

Hiring Organisation
Stott & May Professional Search Limited
Location
Manchester, North West, United Kingdom
Employment Type
Contract
Contract Rate
£508 - £558 per day
activities, and operational workflows. * Ensure cloud environments align with security, governance, compliance, and software lifecycle standards. * Monitor cloud platforms using SIEM, cloud security, and observability tools. * Support incident, problem, and change management processes. * Work with containerised environments and infrastructure-as-code technologies. Essential Skills & Experience * Strong commercial experience with … regulated or financial services environments. * Relevant cloud certifications (AWS, Microsoft Azure, or Google Cloud). * CISSP or other recognised security certifications. * Experience with cloud observability and monitoring platforms. * Knowledge of security automation and remediation tooling. * Experience supporting enterprise-scale cloud transformation programmes. ...

Software Engineering Team Lead

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
software lifecycle from design ideation through to production and eventual decommissioning. Our engineering teams work under a true DevOps culture — with infrastructure as code, observability, automated testing, and continuous delivery treated as first-order concerns, not afterthoughts. You’ll set architectural direction, partner closely with your Product Manager counterpart … systems and microservice development - we use Azure Service Bus, and welcome experience with similar messaging technologies such as Kafka or RabbitMQ Infrastructure: Kubernetes, Docker Observability: Prometheus, Grafana Engineering culture: DevOps, infrastructure as code, automated testing across all environments including production, continuous delivery Our Engineering Approach Full ownership: Teams own their ...

Senior Platform Engineer

Hiring Organisation
REVYBE IT RECRUITMENT LIMITED
Location
City of London, London, United Kingdom
Employment Type
Permanent, Work From Home
improve developer experience, deployment processes, and platform reliability Support and optimise Azure SQL environments, ensuring performance, availability, and security Implement monitoring, logging, and observability solutions to improve platform performance and resilience Champion cloud security, governance, and automation best practices across the Azure estate Contribute to the evolution of the company … DevOps CI/CD pipelines Solid scripting skills with PowerShell Experience supporting or administering Azure SQL (or Microsoft SQL Server) Experience with monitoring and observability tooling Good understanding of cloud networking, security, identity, and platform automation Excellent communication skills with a collaborative mindset Someone who enjoys solving complex problems, improving ...

Engineering Lead

Hiring Organisation
Hays
Location
Cheshire, North West, United Kingdom
Employment Type
Contract, Work From Home
Contract Rate
Up to £500.0 per day + Inside IR35
guide implementation of scalable integration solutions across APIs, middleware, event-driven platforms, and external systems. Drive best practices across software engineering, DevOps, resiliency, observability, operational excellence, audit readiness, and governance. Conduct architecture reviews, code reviews, technical design assessments, and controls compliance reviews. Provide mentoring, coaching, and hands-on technical guidance … Services, or large-scale transformation programmes. Experience supporting regulated applications, risk technology platforms, or compliance-driven initiatives. Experience with Docker and Kubernetes. Knowledge of observability, monitoring, logging, and control monitoring frameworks. Experience working within Agile delivery environments. Experience supporting geographically distributed teams. Exposure to AI-enabled engineering tools, automation frameworks ...

Engineering Lead

Hiring Organisation
17918
Location
High Wycombe, Buckinghamshire, United Kingdom
guide implementation of scalable integration solutions across APIs, middleware, event-driven platforms, and external systems. Drive best practices across software engineering, DevOps, resiliency, observability, operational excellence, audit readiness, and governance. Conduct architecture reviews, code reviews, technical design assessments, and controls compliance reviews. Provide mentoring, coaching, and hands-on technical guidance … Services, or large-scale transformation programmes. Experience supporting regulated applications, risk technology platforms, or compliance-driven initiatives. Experience with Docker and Kubernetes. Knowledge of observability, monitoring, logging, and control monitoring frameworks. Experience working within Agile delivery environments. Experience supporting geographically distributed teams. Exposure to AI-enabled engineering tools, automation frameworks ...

Senior Site Reliability Engineer - Python

Hiring Organisation
Inspire People
Location
South West London, London, United Kingdom
Employment Type
Permanent, Part Time, Work From Home
Salary
£80,000
operating and improving cloud-based platforms and services used across DBT and wider government. Working across Python development, cloud infrastructure, CI/CD pipelines, observability and automation, you'll help improve reliability, developer experience and service performance while supporting critical business services used by thousands of users. As a Senior … code approaches. Support teams to adopt Site Reliability Engineering practices, including Service Level Indicators (SLIs), Service Level Objectives (SLOs) and error budgets. Contribute to observability across services, helping teams better understand performance, reliability and user impact. Develop and improve CI/CD pipelines to enable safe, frequent and low-risk ...

Senior Site Reliability Engineer - Python

Hiring Organisation
Inspire People
Location
Salford, Lancashire, England, United Kingdom
Employment Type
Full-Time
Salary
£63,824 - £80,158 per annum, Pro-rata, Inc benefits
operating and improving cloud-based platforms and services used across DBT and wider government. Working across Python development, cloud infrastructure, CI/CD pipelines, observability and automation, you'll help improve reliability, developer experience and service performance while supporting critical business services used by thousands of users. As a Senior … code approaches. Support teams to adopt Site Reliability Engineering practices, including Service Level Indicators (SLIs), Service Level Objectives (SLOs) and error budgets. Contribute to observability across services, helping teams better understand performance, reliability and user impact. Develop and improve CI/CD pipelines to enable safe, frequent and low-risk ...

Senior Site Reliability Engineer - Python

Hiring Organisation
Inspire People
Location
Edinburgh, Midlothian, Scotland, United Kingdom
Employment Type
Full-Time
Salary
£63,824 - £80,158 per annum, Pro-rata, Inc benefits
operating and improving cloud-based platforms and services used across DBT and wider government. Working across Python development, cloud infrastructure, CI/CD pipelines, observability and automation, you'll help improve reliability, developer experience and service performance while supporting critical business services used by thousands of users. As a Senior … code approaches. Support teams to adopt Site Reliability Engineering practices, including Service Level Indicators (SLIs), Service Level Objectives (SLOs) and error budgets. Contribute to observability across services, helping teams better understand performance, reliability and user impact. Develop and improve CI/CD pipelines to enable safe, frequent and low-risk ...

Senior Site Reliability Engineer - Python

Hiring Organisation
Inspire People
Location
Edinburgh, Midlothian, Scotland, United Kingdom
Employment Type
Permanent, Part Time, Work From Home
Salary
£80,000
operating and improving cloud-based platforms and services used across DBT and wider government. Working across Python development, cloud infrastructure, CI/CD pipelines, observability and automation, you'll help improve reliability, developer experience and service performance while supporting critical business services used by thousands of users. As a Senior … code approaches. Support teams to adopt Site Reliability Engineering practices, including Service Level Indicators (SLIs), Service Level Objectives (SLOs) and error budgets. Contribute to observability across services, helping teams better understand performance, reliability and user impact. Develop and improve CI/CD pipelines to enable safe, frequent and low-risk ...

Senior Site Reliability Engineer - Python

Hiring Organisation
Inspire People
Location
Birmingham, West Midlands, United Kingdom
Employment Type
Permanent, Part Time, Work From Home
Salary
£80,000
operating and improving cloud-based platforms and services used across DBT and wider government. Working across Python development, cloud infrastructure, CI/CD pipelines, observability and automation, you'll help improve reliability, developer experience and service performance while supporting critical business services used by thousands of users. As a Senior … code approaches. Support teams to adopt Site Reliability Engineering practices, including Service Level Indicators (SLIs), Service Level Objectives (SLOs) and error budgets. Contribute to observability across services, helping teams better understand performance, reliability and user impact. Develop and improve CI/CD pipelines to enable safe, frequent and low-risk ...

Senior Site Reliability Engineer - Python

Hiring Organisation
Inspire People
Location
Cardiff, South Glamorgan, Wales, United Kingdom
Employment Type
Permanent, Part Time, Work From Home
Salary
£80,000
operating and improving cloud-based platforms and services used across DBT and wider government. Working across Python development, cloud infrastructure, CI/CD pipelines, observability and automation, you'll help improve reliability, developer experience and service performance while supporting critical business services used by thousands of users. As a Senior … code approaches. Support teams to adopt Site Reliability Engineering practices, including Service Level Indicators (SLIs), Service Level Objectives (SLOs) and error budgets. Contribute to observability across services, helping teams better understand performance, reliability and user impact. Develop and improve CI/CD pipelines to enable safe, frequent and low-risk ...

Senior Site Reliability Engineer - Python

Hiring Organisation
Inspire People
Location
Manchester, North West, United Kingdom
Employment Type
Permanent, Part Time, Work From Home
Salary
£80,000
operating and improving cloud-based platforms and services used across DBT and wider government. Working across Python development, cloud infrastructure, CI/CD pipelines, observability and automation, you'll help improve reliability, developer experience and service performance while supporting critical business services used by thousands of users. As a Senior … code approaches. Support teams to adopt Site Reliability Engineering practices, including Service Level Indicators (SLIs), Service Level Objectives (SLOs) and error budgets. Contribute to observability across services, helping teams better understand performance, reliability and user impact. Develop and improve CI/CD pipelines to enable safe, frequent and low-risk ...

Senior Site Reliability Engineer - Python

Hiring Organisation
Inspire People
Location
Cardiff, South Glamorgan, Wales, United Kingdom
Employment Type
Full-Time
Salary
£63,824 - £80,158 per annum, Pro-rata, Inc benefits
operating and improving cloud-based platforms and services used across DBT and wider government. Working across Python development, cloud infrastructure, CI/CD pipelines, observability and automation, you'll help improve reliability, developer experience and service performance while supporting critical business services used by thousands of users. As a Senior … code approaches. Support teams to adopt Site Reliability Engineering practices, including Service Level Indicators (SLIs), Service Level Objectives (SLOs) and error budgets. Contribute to observability across services, helping teams better understand performance, reliability and user impact. Develop and improve CI/CD pipelines to enable safe, frequent and low-risk ...

Senior Site Reliability Engineer - Python

Hiring Organisation
Inspire People
Location
Darlington, County Durham, North East, United Kingdom
Employment Type
Permanent, Part Time, Work From Home
Salary
£80,000
operating and improving cloud-based platforms and services used across DBT and wider government. Working across Python development, cloud infrastructure, CI/CD pipelines, observability and automation, you'll help improve reliability, developer experience and service performance while supporting critical business services used by thousands of users. As a Senior … code approaches. Support teams to adopt Site Reliability Engineering practices, including Service Level Indicators (SLIs), Service Level Objectives (SLOs) and error budgets. Contribute to observability across services, helping teams better understand performance, reliability and user impact. Develop and improve CI/CD pipelines to enable safe, frequent and low-risk ...

Senior Site Reliability Engineer - Python

Hiring Organisation
Inspire People
Location
Belfast, County Antrim, Northern Ireland, United Kingdom
Employment Type
Permanent, Part Time, Work From Home
Salary
£80,000
operating and improving cloud-based platforms and services used across DBT and wider government. Working across Python development, cloud infrastructure, CI/CD pipelines, observability and automation, you'll help improve reliability, developer experience and service performance while supporting critical business services used by thousands of users. As a Senior … code approaches. Support teams to adopt Site Reliability Engineering practices, including Service Level Indicators (SLIs), Service Level Objectives (SLOs) and error budgets. Contribute to observability across services, helping teams better understand performance, reliability and user impact. Develop and improve CI/CD pipelines to enable safe, frequent and low-risk ...

Lead Java Developer — Real-Time Risk & Cloud (Hybrid)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
full lifecycle from design to production support, integrating new analytics and data sets across global teams. The role emphasizes scalable microservices, streaming data, and observability with ELK, Prometheus and Grafana. Hybrid work model and competitive benefits are offered. #J-18808-Ljbffr ...

Platform Engineering Lead — Remote

Hiring Organisation
Jobleads-UK
Location
Exeter, England, United Kingdom
cloud-first transformation from an Azure PaaS .NET estate to a modern AI-first platform. You will own cloud infrastructure, CI/CD, observability, security and the developer platform, guiding a DevOps team into a high‐impact capability. This hands-on leadership role combines technical delivery with people leadership, reporting ...

Hybrid AI Operations Engineer: Scale, Secure & Optimize AI

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
will maintain live agentic AI systems, manage incidents, and drive cost-efficient cloud operations at enterprise scale. The role focuses on reliability, security, observability, and performance across Python-based stacks, AWS, Postgres, and container platforms like Kubernetes or Cloud Foundry. Hybrid working with two office days per week in London ...