1,101 to 1,125 of 1,797 Permanent Observability Jobs

Enterprise Account Executive, EMEA Sales EMEA (Remote)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
powerful ways to monitor, troubleshoot, and optimize their AI systems. That’s where we come in. Arize AI is the leading AI & Agent Engineering observability and evaluation platform, empowering AI engineers to ship high-performing, reliable agents and applications. From first prototype to production scale, Arize AX unifies build, test ...

Chief Revenue Officer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
engineers, who work with some of the world’s largest enterprises. With a global presence, we operate across multiple industries, bringing expertise in Digital, Observability, Automation, Data, Cybersecurity, and Cloud. The group goes to market through a mix of direct enterprise selling and technology‐partner ecosystems (e.g Cribl, Crowdstrike, Pega ...

Senior Software Engineer, C++

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
knowledge to successfully translate the requirements into actual software implementation Continuously improve the stability, reliability, and performance of the trading engine Enhance monitoring and observability in collaboration with the Trading Operations team Investigate and resolve production issues such as crashes, unexpected business logic behavior, and performance bottlenecks Prepare for releases ...

Staff Forward Deployed Engineer, Google Cloud Consulting (French, German)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
customer's live infrastructure, including APIs, legacy data silos, and security perimeters as part of an expert team. Build high-performance evaluation pipelines and observability frameworks to ensure agentic systems meet requirements for accuracy, safety, and latency. Identify repeatable field patterns and friction points in Google’s AI stack, converting ...

Compliance - CCOR Risk Management Director - Executive Director

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
experience assessing and challenging designs for data/AI platforms and integrations (APIs and managed services, security gateways, IAM/least privilege, logging/observability, data residency and egress controls). Strong understanding of AI/LLM capabilities and risks across the lifecycle (model onboarding/ingestion, retrieval/ ...

Partner Forward Deployed Engineer, Google Cloud (French, German)

Hiring Organisation
Jobleads-UK
Location
City of Westminster, England, United Kingdom
customer's live infrastructure, including APIs, legacy data silos, and security perimeters as part of an expert team. Build high-performance evaluation pipelines and observability frameworks to ensure agentic systems meet requirements for accuracy, safety, and latency. Identify repeatable field patterns and friction points in Google’s AI stack, converting ...

Senior Forward Deployed Engineer, GenAI, Google Cloud

Hiring Organisation
Jobleads-UK
Location
City of Westminster, England, United Kingdom
customer's live infrastructure, including APIs, legacy data silos, and security perimeters as part of an expert team. Build high-performance evaluation pipelines and observability frameworks to ensure agentic systems meet rigorous requirements for accuracy, safety, and latency. Identify repeatable field patterns and friction points in Google\’s AI stack ...

UK Commercial Counsel Public Sector

Hiring Organisation
Jobleads-UK
Location
United Kingdom
advantage of all structured and unstructured data — securing and protecting private information more effectively — Elastic’s complete, cloud-based solutions for search, security, and observability help organizations deliver on the promise of AI. What is The Role: We are hiring a solid Corporate Counsel, with 4+ years’ experience , to join ...

Lead UI Developer – Single Dealer Platform (TypeScript / RxJS / React)

Hiring Organisation
Bank of America
Location
Greater London, United Kingdom
Employment Type
Full Time
edge. Enforce automated testing (unit, integration, and UI E2E) and CI/CD with quality gates. Instrument analytics, logging, and front end observability (e.g., Web Vitals, error tracking). Ensure secure front end practices (XSS/CSRF protection, content security policy, secrets handling, auth flows with OAuth/OIDC/ ...

Engineering Manager, Payments

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
team and drive positive change.Bonus points if:You are familiar with any one of the following technologies : Scala, Java or Typescript.You are familiar with observability, tracking and data pipeline tools and methodologies.You have previously worked in an App first business.Additional InformationHealth + Mental WellbeingPMI and cash plan healthcare access with ...

Lead Platform Engineer

Hiring Organisation
Transunion
Location
Leeds, West Yorkshire, United Kingdom
Employment Type
Permanent
Hands-on experience with CI/CD pipeline management using Harness, infrastructure automation with Terraform, and integration with Harness IaC pipelines. Experience with monitoring, observability, and security tooling, including GCP Native Observability, Prometheus, Grafana, OpenTelemetry, Wiz, HashiCorp Vault, and CheckMarxOne or similar tools. Strong working knowledge of platform and data ...

IT Platform Engineer - HPC & Linux

Hiring Organisation
Jobleads-UK
Location
Milton Keynes, England, United Kingdom
knowledge of platform security principles. Experience with configuration management and infrastructure as code tools (Ansible, Git, Vault) Knowledge of container orchestration tools, virtualisation and observability stacks e.g. Kubernetes, Grafana, Kafka, Docker, OpenStack, OLVM. Experience implementing platform and application observability solutions including monitoring, logging, tracing, alerting and telemetry. Awareness of secure ...

Backend Engineer, Platform

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
platform and product development from discovery to implementation. Contribute to high‐level architectural decisions for the core application and associated services. Improve scalability, reliability, observability, and operational safety of critical backend systems. Work on foundational platform capabilities: backend systems, service‐to‐service communication, authentication, infrastructure architecture, messaging, database scalability, deployment … Experience with PostgreSQL, MySQL, or DynamoDB. Experience building scalable web applications or backend systems handling significant traffic and data volume. Experience improving reliability, scalability, observability, and operational maturity in production systems. Experience in a DevOps culture: CI/CD pipelines, observability, incident response, and production ownership. Experience with platform capabilities ...

Network Engineer

Hiring Organisation
HCLTech
Location
London Area, United Kingdom
Modern Ops and AI-first operating model. The role focuses on Network Infrastructure as Code (NetIaC), CI/CD pipelines, AI-driven operations (AIOps), observability integration, and SRE-led reliability engineering. Key Responsibilities Develop and manage Network Infrastructure as Code (NetIaC) using Python, Ansible, and Terraform for provisioning and lifecycle … ITSM workflows. Drive AI/ML use cases such as WAN capacity forecasting, anomaly detection, predictive analytics, and self-healing networks. Integrate and manage observability platforms (SolarWinds Orion, Elastic, Grafana, ZDX) for proactive monitoring and insights. Provide engineering and support for MCP (Model Context Protocol) and AI agent integrations. Ensure ...

Principal Solutions Engineer - Observe by Snowflake

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
just to execute a function, but to help redefine the future of how work gets done. Observe by Snowflake is a high-growth SaaS observability platform built on the Snowflake AI Data Cloud, enabling businesses to troubleshoot modern distributed applications 10x faster. Now, as a core part of Snowflake … reached a major milestone in the evolution of the Snowflake platform. By bringing AI-powered observability directly into the Snowflake ecosystem, we’ve created the first truly unified platform for telemetry and business data. Based in the UK, this Principal Sales Engineer is a critical, highly-visible role within ...

Azure Platform Engineer

Hiring Organisation
Jobleads-UK
Location
Oxford, England, United Kingdom
Virtual Networks, Network Security Groups, Private Endpoints, DNS and hybrid connectivity. Ensure network architecture is secure, performant, and aligned to the Landing Zone design. Observability and Operations: Implement and maintain monitoring, alerting, and logging solutions using Azure Monitor, Log Analytics and Sentinel. Drive a culture of observability, ensuring platform health … engineering practices where they provide measurable benefit to the organisation. Experience configuring and operating Azure Monitor, Log Analytics workspaces and Sentinel for platform observability and security monitoring. Proficiency in PowerShell and/or Azure CLI for automation, scripting and configuration management. A strong security and governance mindset, with experience applying ...

Senior Backend Engineer | AI Platform

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
high degree of autonomy and ownership, as you'll be responsible for designing scalable solutions that empower multiple engineering teams while ensuring reliability, observability, and cost efficiency. What are we looking for: 5+ years of experience in Software Engineering, Backend Engineering, or Platform Engineering. Strong experience building and maintaining backend … LangChain, LangGraph, CrewAI, or similar. Experience working with cloud platforms such as Google Cloud Platform (preferred), AWS, or Azure. Strong understanding of system reliability, observability, monitoring, and incident management. Experience with Infrastructure as Code and cloud-native architectures. Previous experience working within a Platform Engineering team is a strong plus. ...

Devops Engineer

Hiring Organisation
Jobleads-UK
Location
Belfast City District, Northern Ireland, United Kingdom
platform migration to AWS ECS using Terraform, GitHub Actions and Datadog. Over the last 6 months, our team has successfully deployed Datadog as our observability platform, achieving 95% monitoring coverage across all services and significantly improving our incident response capabilities. Moving forward, we're focused on completing our ECS migration … Responsibilities Design and implement our Kubernetes (ECS) platform on AWS Maintain and optimise our CI/CD pipelines using GitHub Actions Implement and manage observability using Datadog across our platform Support and enhance our AWS cloud infrastructure Architect, review, audit, optimise and document our deployment processes Follow best practice change ...

DevOps Engineer

Hiring Organisation
Jobleads-UK
Location
Croydon, England, United Kingdom
pipelines across multiple engineering teams Automate infrastructure and deployments using Infrastructure as Code Support Azure cloud infrastructure and AKS environments Improve platform monitoring, observability, and operational efficiency Troubleshoot production and deployment issues to maintain platform reliability Collaborate closely with Software Engineers, Platform Engineers, and Security teams to improve delivery …/Kubernetes Strong Terraform or Infrastructure as Code experience Experience building and maintaining CI/CD pipelines Good understanding of monitoring, logging, and observability tools Strong troubleshooting and problem‐solving skills Experience working within Agile engineering teams Why Apply? You’ll be joining an engineering organisation operating at genuine enterprise ...

AI Native DevOps Platform Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
using AI‐first engineering practices. Working closely with Product Engineering Teams and Technical Leadership, you will build the cloud platforms, infrastructure, deployment pipelines, automation, observability frameworks, and engineering tooling that enable the rapid delivery of both AI‐powered and traditional cloud‐native applications. We are building an AI‐native engineering … platform tooling. Drive engineering productivity through AI, automation, self‐service capabilities and platform standardisation. Improve release processes and operational excellence across teams. Reliability & Observability Implement monitoring, logging, tracing, and alerting solutions. Establish platform SRE principles and operational standards. Proactively identify and resolve reliability, security, and performance issues. Lead incident response ...

Senior Platform Engineer

Hiring Organisation
Anson Mccade
Location
Manchester, North West, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£70,000
Code using Terraform across production and non-production environments Driving DevSecOps best practice and improving engineering standards across delivery teams Implementing SRE principles including observability, monitoring, SLIs/SLOs and platform reliability Supporting production environments, incident management and continuous service improvements Mentoring engineers and acting as a technical leader within … DevSecOps engineering experience within enterprise environments Proven Terraform experience building Infrastructure as Code Good understanding of Site Reliability Engineering (SRE) principles Experience with observability and monitoring tools such as Dynatrace, Grafana, Prometheus or similar Knowledge of CI/CD pipelines and modern cloud-native engineering practices Experience supporting live production ...

Site Reliability Engineer

Hiring Organisation
Jobleads-UK
Location
Belfast City District, Northern Ireland, United Kingdom
/CD pipelines to ensure efficient and reliable software delivery. Manage containerized workloads across modern orchestration platforms. Build and improve monitoring, alerting, and observability solutions for rapid issue resolution. Automate operational workflows, deployment processes, and administrative tasks. Troubleshoot complex production incidents across various systems. Collaborate with development teams to enhance … Proven track record in designing and maintaining CI/CD pipelines. Strong Linux systems administration and scripting skills (Bash, Shell, Python). Experience with observability platforms (Grafana, InfluxDB) and automation. Ability to troubleshoot complex distributed systems and applications. Familiarity with Git and modern source control workflows. Understanding of high availability ...

Senior DevOps Engineer

Hiring Organisation
Halian Technology Limited
Location
Reading, Berkshire, South East, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£85,000
reliability, and availability Implement self-service tooling to empower development teams Drive DevOps best practices across the digital product lifecycle Develop and enhance monitoring, observability, and incident response processes Support global engineering teams delivering high-traffic platforms Key Requirements Proven experience supporting digital product delivery in a DevOps or platform … with Infrastructure as Code (Terraform, Ansible, Puppet or similar) Hands-on experience with Kubernetes, Docker, and cloud platforms (AWS preferred) Experience with monitoring/observability tools (Prometheus, Grafana, ELK, APM tools) Solid understanding of system performance, scalability, and resilience Strong collaboration and communication skills within cross-functional product teams Desirable ...

DevOps Engineer

Hiring Organisation
Eligo Recruitment
Location
Stockport, Cheshire, England, United Kingdom
Employment Type
Full-Time
Salary
£70,000 - £80,000 per annum
Building and maintaining Infrastructure as Code using Terraform Automating infrastructure provisioning and deployment pipelines Managing Kubernetes and containerised workloads Implementing monitoring, logging and observability solutions Driving platform reliability, security and best practices Collaborating with engineering teams to improve developer experience Skills & Experience Essential: Strong commercial experience with Google Cloud Platform … container technologies Experience with Linux and scripting (Bash, Python or Go) Understanding of networking, IAM and cloud security principles Experience with monitoring and observability tooling Desirable: Experience with GitOps practices Knowledge of Prometheus, Grafana or similar tools Experience in a platform engineering or SRE environment Certifications in GCP are advantageous ...

Site Reliability Engineer (AWS)

Hiring Organisation
Spectrum IT Recruitment
Location
Southampton, Hampshire, United Kingdom
Employment Type
Permanent
Salary
£60000/annum Bonus, Pension, Healthcare
issues and restoring services quickly and effectively Developing automation to reduce manual operational tasks and improve platform resilience Building and improving monitoring, alerting and observability across cloud environments Working alongside Software, Platform, Cloud and Security Engineers to improve reliability and operational excellence Contributing to post-incident reviews and driving continuous … with exposure to: Linux systems administration AWS cloud infrastructure Kubernetes and Docker Production support and incident management Python, Bash or Go scripting Monitoring and observability platforms such as Grafana, Prometheus, Datadog, Splunk or CloudWatch Networking fundamentals including DNS, TCP/IP and load balancing A passion for automation, continuous improvement ...