876 to 900 of 1,950 Observability Jobs

AIML Software Engineer, AI for Science

Hiring Organisation
GSK
Location
Greater London, United Kingdom
Employment Type
Full Time
Salary
136125 to 226875 USD Annually
infrastructure as code. Strong problem-solving and debugging skills, and experience working in cluster settings or cloud-based environments. Experience operating production services - monitoring, observability and alerting, and diagnosing and resolving issues in live systems. Experience designing and administering SQL databases - schema design, query performance, and day-to-day operational … including defining and working to service-level objectives (SLOs/SLIs). Experience with incident response and post-incident review, and with building the observability that supports it. Infrastructure-as-code (e.g. Terraform) for provisioning and maintaining cloud environments. Experience developing and administering workloads on Kubernetes (e.g. GKE). Familiarity ...

Grafana Observability Engineer (6-month contract)

Hiring Organisation
17918
Location
London, United Kingdom
Grafana Observability Engineer (6-month contract) Location: Fully remote Our client They are delivering a company-wide Digital Transformation (DX) Programme that will modernise their technology landscape and transform how they deliver services. As part of this journey, their IT and Portfolio Delivery teams are implementing a new enterprise technology … making this an exciting opportunity to join the business and help shape their future. Role Overview Reporting to the Lead Platform Engineer, the Senior Observability Engineer will own and develop their observability capability, leading the design, implementation and continuous improvement of their monitoring and alerting platform. Working closely with infrastructure ...

Senior Cloud Engineer

Hiring Organisation
Jobleads-UK
Location
Manchester, England, United Kingdom
Implement the data pipelines, workflow orchestrations, and specialised compute footprints needed to support enterprise AI applications, using technologies like AWS Bedrock and AWS AgentCore. Observability & Reliability: Build robust monitoring and observability pipelines to ensure the health, performance, and security of distributed cloud applications and AI models. FinOps Standards: Embed automated … implementing components of data pipelines, including real-time streaming tools (AWS Kinesis, Kafka), data orchestration (dbt, Airflow), and managing vector databases for RAG architectures. Observability & Cost Management: SRE/Platform experience with the practical application of real-time monitoring and cloud cost optimisation using native CSP tools or utilities like ...

Senior Cloud Engineer

Hiring Organisation
Jobleads-UK
Location
Leeds, England, United Kingdom
Implement the data pipelines, workflow orchestrations, and specialised compute footprints needed to support enterprise AI applications, using technologies like AWS Bedrock and AWS AgentCore. Observability & Reliability: Build robust monitoring and observability pipelines to ensure the health, performance, and security of distributed cloud applications and AI models. FinOps Standards: Embed automated … implementing components of data pipelines, including real-time streaming tools (AWS Kinesis, Kafka), data orchestration (dbt, Airflow), and managing vector databases for RAG architectures. Observability & Cost Management: SRE/Platform experience with the practical application of real-time monitoring and cloud cost optimisation using native CSP tools or utilities like ...

Senior Cloud Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Implement the data pipelines, workflow orchestrations, and specialised compute footprints needed to support enterprise AI applications, using technologies like AWS Bedrock and AWS AgentCore. Observability & Reliability: Build robust monitoring and observability pipelines to ensure the health, performance, and security of distributed cloud applications and AI models. FinOps Standards: Embed automated … implementing components of data pipelines, including real-time streaming tools (AWS Kinesis, Kafka), data orchestration (dbt, Airflow), and managing vector databases for RAG architectures. Observability & Cost Management: SRE/Platform experience with the practical application of real-time monitoring and cloud cost optimisation using native CSP tools or utilities like ...

Senior Cloud Engineer

Hiring Organisation
Jobleads-UK
Location
City of Edinburgh, Scotland, United Kingdom
Implement the data pipelines, workflow orchestrations, and specialised compute footprints needed to support enterprise AI applications, using technologies like AWS Bedrock and AWS AgentCore. Observability & Reliability: Build robust monitoring and observability pipelines to ensure the health, performance, and security of distributed cloud applications and AI models. FinOps Standards: Embed automated … implementing components of data pipelines, including real-time streaming tools (AWS Kinesis, Kafka), data orchestration (dbt, Airflow), and managing vector databases for RAG architectures. Observability & Cost Management: SRE/Platform experience with the practical application of real-time monitoring and cloud cost optimisation using native CSP tools or utilities like ...

Senior Cloud Engineer

Hiring Organisation
Jobleads-UK
Location
Manchester, England, United Kingdom
Implement the data pipelines, workflow orchestrations, and specialised compute footprints needed to support enterprise AI applications, using technologies like AWS Bedrock and AWS AgentCore. Observability & Reliability: Build robust monitoring and observability pipelines to ensure the health, performance, and security of distributed cloud applications and AI models. FinOps Standards: Embed automated … implementing components of data pipelines, including real-time streaming tools (AWS Kinesis, Kafka), data orchestration (dbt, Airflow), and managing Vector Databases for RAG architectures. Observability & Cost Management: SRE/Platform experience with the practical application of real-time monitoring and cloud cost optimisation using native CSP tools or utilities like ...

Azure Platform Engineer - Scalable Infra, Terraform & CI/CD

Hiring Organisation
Jobleads-UK
Location
United Kingdom
Azure infrastructure, implement IaC with Terraform, and collaborate with developers to improve CI/CD and deployment strategies. You will mentor teammates, enhance observability and reliability, and help optimize cloud performance and security. #J-18808-Ljbffr ...

Senior Azure SRE: Cloud Reliability & Automation

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
enterprise-scale Azure platforms. The role focuses on reliability, automation, and governance to reduce operational overhead. The successful candidate will lead reliability engineering, implement observability, and drive resilient architecture with Terraform, Azure DevOps, and strong SRE practices. UK-based, with flexible same-team collaboration across regions. #J-18808-Ljbffr ...

Java Backend Engineer: AI-Driven Microservices (On-site)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
site role in London, full-time, mid-level. You will implement microservices with Spring, REST/GraphQL, and leverage Azure cloud services; focus on observability, security, and quality. Collaborate with product managers, designers, and engineers in an agile environment and contribute to hiring and onboarding. #J-18808-Ljbffr ...

Multi-Cloud SaaS Release Lead (CI/CD & DevOps)

Hiring Organisation
Jobleads-UK
Location
Manchester, England, United Kingdom
coordinating multiple teams to minimize risk and maximize cadence. The role requires strong DevOps practices, containerisation knowledge, and the ability to drive automation and observability across a SaaS platform. #J-18808-Ljbffr ...

Senior Platform Engineer: Kubernetes & CI/CD Lead

Hiring Organisation
Jobleads-UK
Location
Manchester, England, United Kingdom
modernise deployment and scale workloads. You will drive Kubernetes standardisation, containerisation of existing workloads, and the development of self-service runbooks, while owning observability #J-18808-Ljbffr ...

Senior Platform Engineer — SRE & Cloud

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
response, and mentor engineers while partnering with product and leadership to deliver scalable services. You will own platform roadmap, advance Kubernetes, Terraform, GitOps, and observability, and champion cost-aware engineering across teams. Remote options not specified. #J-18808-Ljbffr ...

Senior SRE: Kubernetes, GitOps & AI-Driven Reliability

Hiring Organisation
Jobleads-UK
Location
City Of London, England, United Kingdom
canary deployments, and secure secret management within a hybrid London-based team. Working in the Webex Engineering Group, you will advance AI-assisted configuration, observability, and reliable delivery across dev, staging, and production environments. #J-18808-Ljbffr ...

Senior Site Reliability Engineer — Cloud & Automation

Hiring Organisation
Jobleads-UK
Location
Belfast City District, Northern Ireland, United Kingdom
maintain reliable, scalable production infrastructure, implement IaC, and enhance CI/CD pipelines across our global trading platforms. You will work on container orchestration, observability, and automation, collaborating with developers to improve system resilience and performance in a high-stakes environment. #J-18808-Ljbffr ...

Senior SRE — Front-Office Trading Reliability

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
J.P. Morgan is seeking a Lead Site Reliability Engineer within the Trading Technology group to shape SRE patterns and observability across globally distributed systems. You will work directly with traders and software engineers to improve reliability, performance, and resilience in front‐office environments. You will design and implement automated remediation ...

Senior Backend Engineer – Platform Modernisation (Java/.NET)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
APIs. You will design, build, and optimize backend services, collaborate with product, operations and vendors, and champion automated testing, CI/CD and observability to ensure reliable, scalable platform delivery. This critical role requires strong Java/.NET skills and experience in financial services environments. #J-18808-Ljbffr ...

Senior MLOps & Data Platform Engineer

Hiring Organisation
Jobleads-UK
Location
England, United Kingdom
develop cloud-native data pipelines across AWS and GCP. The role focuses on turning research into reliable production systems with strong emphasis on observability, governance and collaboration with cross-functional teams. The position offers a hybrid working pattern with a competitive salary, bonus, equity and benefits, and opportunities to drive ...

Lead Compute Platform Engineer – Cloud & HPC

Hiring Organisation
Jobleads-UK
Location
City Of London, England, United Kingdom
medicine research. The Sr. Compute Platform Engineer will design, build, and operate tools and workflows across on‐prem and cloud, with emphasis on automation, observability, and scalable batch processing. You will mentor teammates and set software engineering best practices. You will collaborate with science users to optimize performance on large ...

Azure Platform Engineer - Remote, Terraform & Security Focus

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
emphasis on automation and reliability. You will design, build, and operate Azure-based infrastructure, create reusable Terraform modules, enforce security controls, and contribute to observability and SRE practices within a collaborative #J-18808-Ljbffr ...

Platform Engineer - AI-Driven Infra & CI/CD Expert

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
bottlenecks and building shared tools, infra, and automation for all developers. You’ll explore AI-enabled tooling, manage AWS, GitHub Actions, and observability, and contribute to production reliability with metrics, logs, and SLOs. #J-18808-Ljbffr ...

DevEx Pipeline Engineer: CI/CD & Platform Automation

Hiring Organisation
Jobleads-UK
Location
Cambridge, England, United Kingdom
improve developer experience across a TypeScript/Node monorepo. You’ll extend Terraform-managed infrastructure on Azure, own Helm values, and inject security and observability into the flow. The role emphasizes reducing toil, building reliable defaults, and enabling engineers to ship with confidence in a SaaS environment. #J-18808-Ljbffr ...

Senior Backend Platform Engineer - AI, Serverless

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
ideal candidate has 5+ years of backend software engineering experience, particularly with Node.js or TypeScript, and is comfortable working with AWS Lambda and observability tools. Join us to transform customer and agent experiences. #J-18808-Ljbffr ...

Senior Cloud & DevOps Engineer — SOC 2 & ISO Ready

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
focusing on cloud infrastructure. This role encompasses designing, implementing, and maintaining systems crucial for a fast-paced, AI-native workspace. You will lead DevOps, observability, and compliance tasks, ensuring audit readiness. Ideal candidates have 4-7 years of experience, strong skills in cloud technologies such as Azure, Terraform, and Kubernetes ...

Solutions Architect

Hiring Organisation
Anson McCade
Location
England, United Kingdom
implement platform governance, security and RBAC. Deliver platform maturity assessments and future operating models. Develop Infrastructure as Code using Terraform. Drive automation, monitoring, observability and FinOps best practices. Provide technical leadership across architecture, engineering and platform operations. Essential Experience Proven experience designing and delivering enterprise-scale Snowflake platforms. Strong knowledge ...