3,026 to 3,050 of 4,043 Permanent Observability Jobs

Lead DevOps Engineer for Low-Latency Trading Platform (Hybrid)

Location
Greater London, England, United Kingdom
infrastructure across cloud and on-prem environments. The role emphasizes reliability, security, and fast delivery, with hands-on leadership across DevOps, platform engineering, and observability initiatives. You will build CI/CD pipelines, automate provisioning and deployment, and collaborate with engineering, security, and operations teams to improve platform readiness ...

Senior Site Reliability Engineer – Cloud & Automation

Location
Christchurch, England, United Kingdom
software, ensure reliability, automate deployment and monitoring, and work with development teams to improve performance and scalability. You will apply infrastructure as code, utilize observability tools, and contribute to a culture of learning and inclusion. Strong cloud, Linux, and networking fundamentals are essential. #J-18808-Ljbffr ...

Senior AI Engineer: Architect Enterprise Multi-Agent Systems

Location
United Kingdom
agentic frameworks and evaluation pipelines. You will mentor engineers, design high-performance multi-agent systems, and ensure production-grade, LLM-backed applications with robust observability and governance. You will work with LangGraph, MCP-based orchestration, and Kubernetes in a cloud-native environment to deliver scalable, trusted enterprise solutions. #J ...

Senior SRE — Build a Global Cloud Reliability Practice

Location
Manchester, England, United Kingdom
Reliability Engineer to design and run the first reliability practice for a cloud-native SaaS platform serving hospitals. You will set SLOs, incident models, observability standards, and guide AI-driven operations initiatives. You will hands-on instrument critical services, own runbooks and dashboards, and coach engineers while scaling ...

Platform Engineer: Kubernetes & Terraform for Fintech

Location
Greater London, England, United Kingdom
weeks at £350-£500 per day Outside IR35. You will own Terraform modules, GitLab CI/CD automation, Gateway API work, and cluster observability in a high-autonomy production environment, delivering tangible platform transformation. #J-18808-Ljbffr ...

Network Automation Engineer - Low-Latency Finance Platform

Location
Greater London, England, United Kingdom
that enables rapid provisioning and reliable operations. This role focuses on network automation at scale, using Python, Ansible, Terraform, and CI/CD, plus observability, on-call duties, and collaboration with security and platform teams. #J-18808-Ljbffr ...

Cloud FinOps Analyst

Location
Tonbridge, England, United Kingdom
across Azure and Snowflake environments. A key focus of the role is leading the FinOps optimisation activities, embedding governance frameworks, and overseeing AKS cost observability using tooling such as Power BI, Kubecost etc. The FinOps Analyst partners closely with Engineering, Data, Cloud Operations, and Finance teams to enable a cost … optimisation, and waste elimination. Develop, maintain, and enforce cloud and data platform cost governance frameworks including tagging, budgeting, guardrails, and accountability processes. Oversee cost observability tooling (Kubecost, Snowflake dashboards, cloud cost portals) to ensure visibility of usage, forecasts, and budget performance. Manage budgeting, forecasting, cost allocation, and financial reporting ...

Platform Engineer - Onsite (Kubernetes/SRE)

Location
Warminster, England, United Kingdom
simulation environments. You will configure MODCloud, D2S and OpenShift, ensuring security-aligned setups and reliable operations across distributed systems. We value strong SRE practices, observability, and collaboration with engineers and modelling teams. The role requires five days on site per week in Warminster. #J-18808-Ljbffr ...

Lead SRE: AWS & Python for Scalable Reliability

Location
Glasgow, Scotland, United Kingdom
availability, performance, and resilience of production systems serving millions globally. You will embed reliability in the software lifecycle, guide cross-functional teams, and champion observability across monitoring and incident practices. The role emphasizes leadership in incident response, SLOs, and tooling, with a focus on engineering excellence and scalable, secure operations ...

Senior Backend Engineer – Cloud-Native, AWS, Scalable

Location
Manchester, England, United Kingdom
integrations and customer experiences. In this senior role you’ll own complex components, collaborate with product and other engineers, and help shape our deployment, observability and reliability practices while delivering performant, resilient systems in #J-18808-Ljbffr ...

Senior Platform Engineer – Remote UK (Cloud/SRE)

Location
Greater London, England, United Kingdom
Hudl in London, United Kingdom, is seeking a Senior Engineer to join our Platform Engineering team. You’ll work on site reliability, cloud infrastructure, observability and production operations to keep Hudl’s platform highly available, scalable and secure. You’ll lead with technical excellence, mentor engineers and drive innovation using ...

Remote NOC Engineer: Cloud Reliability & Automation (UK)

Location
West of England, England, United Kingdom
Ideal candidates will have Linux administration, AWS experience, and hands-on Terraform/Docker work, plus scripting in Python, Bash or Go and strong observability tooling knowledge. #J-18808-Ljbffr ...

Azure Platform Architect: Multi-Tenant AKS

Location
Greater London, England, United Kingdom
upskilling engineers. You will lead platform operating models, SPI communications, and best-practice cloud-native patterns. You will shape the architecture, governance, and observability stack for scalable, multi-tenant workloads in production, with a strong emphasis on IaC, GitOps, and secure, compliant design. #J-18808-Ljbffr ...

Senior Ruby on Rails Engineer | React & Python Leader

Location
Greater London, England, United Kingdom
ensure deliverables are simple, maintainable, and scalable. The role emphasizes owning code quality, API contracts, and system documentation, with responsibility for performance monitoring and observability to maintain reliability across services. #J-18808-Ljbffr ...

Remote Cloud Reliability Architect (Java/C#)

Location
United Kingdom
Bristol or London, aligning with SRE and backend engineering standards. You'll partner with Tech Operations and broader engineering teams to improve deployment safety, observability, and capacity planning, while mentoring staff and delivering maintainable code in Java or C#. #J-18808-Ljbffr ...

Cloud Data Platform Engineer — Real-Time Streaming

Location
Greater London, England, United Kingdom
focus on streaming data, cloud infrastructure, and automation. The role involves creating CI/CD pipelines, integrating data clouds with databases, and enforcing observability and governance across data pipelines. The ideal candidate has 3+ years in data platform or infrastructure roles, strong AWS and Terraform skills, and experience with Kafka ...

Software Reliability Engineer - DevOps & CI/CD Leader

Location
Greater London, England, United Kingdom
coding and configuration, guiding development teams toward rapid and secure releases. The position emphasizes practical DevOps and SRE principles, with focus on testing, automation, observability, and modern runtimes. Strong collaboration across teams is essential for success. #J-18808-Ljbffr ...

Cloud FinOps Analyst

Hiring Organisation
Manufacturing Recruitment Limited
Location
City of London, London, United Kingdom
Employment Type
Permanent
Salary
£60,000
across Azure and Snowflake environments. A key focus of the role is leading the FinOps optimisation activities, embedding governance frameworks, and overseeing AKS cost observability using tooling such as Power BI, Kubecost etc. The FinOps Analyst partners closely with Engineering, Data, Cloud Operations, and Finance teams to enable a cost … optimisation, and waste elimination. Develop, maintain, and enforce cloud and data platform cost governance frameworks including tagging, budgeting, guardrails, and accountability processes. Oversee cost observability tooling (Kubecost, Snowflake dashboards, cloud cost portals) to ensure visibility of usage, forecasts, and budget performance. Manage budgeting, forecasting, cost allocation, and financial reporting ...

Lead Backend Engineer - Java & Go for Enterprise Automation

Location
Bournemouth, England, United Kingdom
EPAS products like AaaS, Ansible Automation Platform, and AutoM8. You will lead architecture work, drive automation strategy, and mentor teams in secure coding and observability practices. The role emphasizes enterprise-grade automation, AI-assisted development, and cross-product platform engineering to support Day 2 operations and platform #J-18808-Ljbffr ...

Cloud Platform Operations Lead: Azure, Kubernetes & SRE

Location
Greater London, England, United Kingdom
will lead hands-on engineering and inspire a team of Platform Operations engineers. You'll shape platform strategy, manage incident response, and advance observability, automation, and governance across infrastructure, applications, and data. Remote/hybrid working options available with global teams. #J-18808-Ljbffr ...

Senior SRE: Front‐Office Trading Reliability

Location
Greater London, England, United Kingdom
within its Trading Technology group in London. You will embed with the software engineering team that builds front‐office trading platforms, shaping SRE patterns, observability, and resilience across globally distributed systems and AI‐accelerated incident response. You will partner with traders and senior stakeholders, lead incident responses, implement reliable code ...

Senior Data Engineer — AI-Driven Data Platform

Location
Greater London, England, United Kingdom
Product to ensure data infrastructure supports analytics, experimentation, and decision-making. This hands-on role emphasizes code quality, CI/CD, security, and observability, with a focus on knowledge sharing and mentoring across the team. #J-18808-Ljbffr ...

(senior) Devops Engineer (m/w/d)

Hiring Organisation
iVentureGroup GmbH
Location
Hammerbrook, Hamburg, Germany
Employment Type
Permanent
Salary
EUR Annual
Verantwortung für unseren operativen IT-Betrieb (24/7), während du gleichzeitig moderne Plattform-Initiativen vorantreibst. Ob Kubernetes-Cluster, CI/CD-Pipelines oder Observability - du bist in deinem Element, wenn du Systeme stabil hältst und . click apply for full job details ...

AWS Cloud Engineer: Serverless, CI/CD & Secure APIs

Location
Nottingham, England, United Kingdom
automation, and clean code. You will collaborate with product, platform, and DevOps teams, contribute to CI/CD pipelines, and ensure robust security and observability across environments. Nottingham-based with 3 days in the office weekly. #J-18808-Ljbffr ...

Automation QA Engineer (SDET) – AI-Driven CI/CD

Location
Greater London, England, United Kingdom
quality-focused software engineer to build AI-assisted quality workflows and scalable test automation. You will work across C#, TypeScript, APIs, data pipelines, and observability to reduce manual checks and improve release confidence. You will contribute to test strategy, migrate automation from Selenium to Playwright, and integrate tests into ...