3,251 to 3,275 of 3,971 Observability Jobs

Senior Security Platform Architect (SCA & Backend)

Location
Cambridge, England, United Kingdom
design and implement backend services, Python APIs, and workflow components to enable tool onboarding, analysis execution, results processing, and delivery, while improving scalability and observability across the platform. #J-18808-Ljbffr ...

Senior Platform Engineer - AWS Cloud, Automation & Self-Service

Location
Greater London, England, United Kingdom
tooling to boost developer productivity and automate key operational processes. You will lead complex platform initiatives, mentor engineers, and drive best practices for reliability, observability and governance, influencing senior stakeholders on technology choices. #J-18808-Ljbffr ...

Senior AWS Platform Engineer - Secure Cloud Automation

Location
United Kingdom
architects and consultants to deliver end-to-end platform solutions, implement secure-by-design principles, and mentor early-career colleagues while promoting automation and observability across projects. #J-18808-Ljbffr ...

Senior Backend Engineer — Scalable AI Media Platforms

Location
United Kingdom
work with DevOps on platform infrastructure. Occasional frontend touchpoints help expose backend capabilities. Responsibilities include API and system design, data processing, model integration, observability, and performance optimization to ensure low latency and high #J-18808-Ljbffr ...

Lead AI Platform Engineer - Scale Inference & Security

Location
City of Edinburgh, Scotland, United Kingdom
models and inference services. You will own the platform layer above GPU infrastructure, deploy scalable inference services with containers and Kubernetes, improve observability, and help define secure, resilient operating standards for enterprise environments. #J-18808-Ljbffr ...

Senior Cloud Infrastructure Engineer | Scale a Global Platform

Location
Greater London, England, United Kingdom
cloud-native core and payments tech to modernize banking. Senior Software Engineers in Infrastructure design and deploy scalable platform tooling, focusing on multi-cloud, observability, and data systems. You will contribute to a cloud-agnostic core platform, build automation, and integrate with open-source tools, while ensuring high reliability ...

AI Platform Architect & Infra Lead

Location
City of Edinburgh, Scotland, United Kingdom
inference services, while owning the platform layer above managed GPU infrastructure. You will deploy scalable inference services using containers, Kubernetes, and AI gateways, improve observability, and enforce security, performance and availability. #J-18808-Ljbffr ...

Senior AI/ML Solutions Architect (GenAI & MLOps)

Location
Greater London, England, United Kingdom
Field Engineering team to design production‐grade AI solutions on the Databricks platform. You will drive GenAI initiatives, RAG architectures, agentic systems, AI observability, and NLQ of structured data, while mentoring peers and influencing the platform roadmap. Some travel may be required. #J-18808-Ljbffr ...

Latency‐Focused Cloud Infrastructure Engineer

Location
Greater London, England, United Kingdom
latency-sensitive trading environment, focused on performance, resilience, and safe change management. You will work on IaC (Terraform/Terragrunt), CI/CD, and observability, collaborating with exchanges and internal stakeholders to support continuous trading operations. #J-18808-Ljbffr ...

Engineering Manager - Lead Scalable Betting Platform

Location
Greater London, England, United Kingdom
drive delivery with accountability. The role requires hands-on Java Spring and Kafka expertise, cloud-native development (AWS), and strong focus on reliability, observability, and AI-enabled tooling. #J-18808-Ljbffr ...

Lead GenAI & LLM Platform Engineer

Location
Greater London, England, United Kingdom
knowledge with strong software engineering. You will lead deployment of scalable AI systems, integrate LLMs for enterprise planning, and develop API services and observability to ensure robust performance and cost efficiency. #J-18808-Ljbffr ...

Pega Developer/Pega DevOPs Architect

Hiring Organisation
iXceed Solutions
Location
Telford, Shropshire, United Kingdom
Employment Type
Contract
Contract Rate
GBP Annual
experience with: Pega DevOps & CI/CD Automation Environment & Release Strategy Cutover Planning & Rehearsal Rollback Design & Recovery Planning Operational Acceptance Testing (OAT) Production Monitoring & Observability Security Governance & CAB Processes Experience supporting highly available, mission-critical applications and platforms. Strong stakeholder management and cross-functional leadership skills. Technical Expertise Pega DevOps … Deployment Manager CI/CD Pipeline Design & Automation Release and Environment Management Operational Readiness & Service Transition Monitoring, Alerting & Observability Cloud Operations & Platform Governance Change Advisory Board (CAB) Processes Production Support & Incident Management Certifications Pega Certified System Architect (CSA) AWS Certified DevOps Engineer - Professional ...

Senior DevOps Engineer - Platform & Automation Champion

Location
Ashford, England, United Kingdom
England is looking for an experienced professional in DevOps and Platform Engineering. You'll play a crucial role in modernizing our payment platforms, enhancing observability, and leading automation efforts across various engineering teams. The ideal candidate will have over 5 years of hands-on experience, a strong technical background ...

Remote Data Engineer for AI Data Platform

Location
United Kingdom
with a strong emphasis on privacy, security, and robust data practices. As part of a collaborative team, you’ll design data pipelines, APIs, and observability, applying IaC and container orchestration tools to keep our platform scalable, reliable, and self-serve for internal teams. #J-18808-Ljbffr ...

SRE & Operations Leader - Reliability at Scale

Location
United Kingdom
services used by internal and external customers. You will drive reliability improvements, advance automation and AI-Ops capabilities, and lead a team focused on observability, incident response, and continuous improvement. You will line manage team leaders, shape strategic direction, ensure incident management and RCAs are completed, and collaborate with ...

Principal Cloud Engineer | AI/ML FinOps & Terraform

Location
Greater London, England, United Kingdom
modular microservices, and deploy AI/ML models to optimize cloud spend. As a principal-level IC, you will own data modeling, orchestration, and observability while collaborating with FinOps analysts, cloud engineers, and platform leads to elevate the entire cost-management practice. #J-18808-Ljbffr ...

Senior Backend Engineer: Chaos & Reliability (Remote)

Location
Greater London, England, United Kingdom
guide product direction, and collaborate with friendly colleagues who live our FAITH values. You’ll design and run chaos experiments, improve load testing and observability, and introduce new tooling to boost reliability. Global teams collaborate on scalable solutions. #J-18808-Ljbffr ...

Senior Real-Time Data Engineer — AI & Streaming

Location
Greater London, England, United Kingdom
APIs, and integrations across CI/CD and development environments. Join a team that builds secure, observable backend services and contributes to AI-enabled observability platforms, enabling proactive detection and remediation of issues at scale. #J-18808-Ljbffr ...

Staff Cloud Native Engineer — AI GPU Infra Architect

Location
Greater London, England, United Kingdom
software integrations that connect AI applications and networking components at scale. In this role you’ll work on shared Kubernetes-based platforms, deployment patterns, observability foundations, infrastructure architecture, and operational tooling that help internal teams run services safely and efficiently on GPU-backed infrastructure. #J-18808-Ljbffr ...

GenAI-Driven Multi-Cloud AI Platform Architect

Location
United Kingdom
generation AI Platform architecture to enable secure, scalable, multi-cloud operations across Azure and AWS. You’ll lead design of ML pipelines, model hosting, observability, and GenAI patterns with AI safety controls. Collaborating with engineering, data science and business teams, you will produce architectural models, governance-aligned solutions, and clear ...

AI Platform Engineer: Scale AI-Native Delivery & Automation

Location
Greater London, England, United Kingdom
patterns and improving cross-team productivity. You will collaborate with AI Software Engineers to drive throughput, quality and platform adoption while embedding validation and observability into delivery workflows. #J-18808-Ljbffr ...

Senior Director, Enterprise AI Strategy & Execution

Location
Greater London, England, United Kingdom
will partner with senior leaders to translate market shifts into an actionable roadmap, prioritizing initiatives across operations, sales, finance, HR, and IT, leveraging the observability platform for measurable impact. The role requires deep expertise in LLMs, agentic systems, governance, and vendor evaluation, with a track record of driving #J ...

Senior AI Engineer - Reasoning & Decision Systems

Location
Greater London, England, United Kingdom
data and specialist models to deliver reliable software. The role offers hands-on ownership from data prep through deployment, with strong emphasis on experimentation, observability and collaboration with product and engineering teams. #J-18808-Ljbffr ...

Senior Backend Engineer (Go) — Revenue & Billing

Location
Greater London, England, United Kingdom
data flows that turn usage into invoices. You will build in Go, integrate with ERP and the Data Platform, and ensure correctness and observability as revenue scales. Joining a cross-functional team with Finance and Go To Market, you will shape the Revenue team's foundations, design robust tests ...

AWS DevOps Engineer – EKS, Terraform, GitOps

Location
Greater London, England, United Kingdom
with developers and SREs in an automation-first culture to deliver reliable, scalable systems. The role covers on-call support, incident response, cost optimization, observability, and secure workload access across AWS services, with emphasis on security and reliability. #J-18808-Ljbffr ...