1,301 to 1,325 of 4,709 Permanent Observability Jobs

QA Test Infrastructure Engineer (Contract)

Location
Taunton, England, United Kingdom
Work on Technology That Protects What Matters AtSiXworks, we build secure digital solutions that supportDefence and National Security missions. Our teams work on complex problems where reliability, security, and speed of innovation matter. We’re ...

Software Architect - Python & .NET.

Hiring Organisation
Computer Futures
Location
Birmingham, UK
Employment Type
Full-time
About this jobKey factsJob TitleSoftware Architect - Python & .NET.Role typePermanentStart dateASAPRemote friendlyYesLocationBirmingham, England, United KingdomSalary75000 - 90000 per annum An established and growing automation specialist is seeking an experienced Software Architect to lead the design and development ...

Senior DevOps Engineer

Hiring Organisation
Informa Connect
Location
London, UK
Employment Type
Full-time
Company DescriptionDo you want to develop your career and make an impact in the fast-growth, fast-moving B2B technology space? At Informa TechTarget, you'll collaborate and grow alongside some of the industry's ...

Developer Experience (DevEx) Engineer — Pipeline Squad

Hiring Organisation
PTC
Location
Gloucester, Gloucestershire, United Kingdom
Salary
£ 70 K
Our world is transforming, and PTC is leading the way. Our software brings the physical and digital worlds together, enabling companies to improve operations, create better products, and empower people in all aspects of their ...

Senior Technical Architect - VMware

Location
United Kingdom
Job Details: Senior Technical Architect - VMware Redcentric |Architect/Senior Architect - VMware Role Overview The role will focus on the architecture, evolution and governance of VMware platforms, with particular emphasis on VMware Cloud Foundation, VCF ...

Software Architect - Python & .NET.

Location
Birmingham, England, United Kingdom
An established and growing automation specialist is seeking an experienced Software Architect to lead the design and development of complex, integrated software platforms. The organisation delivers sophisticated automation solutions that connect enterprise applications, operational software ...

Senior Software Engineer, Git Systems

Hiring Organisation
GitHub
Location
United Kingdom
Salary
£ 70 K
About GitHubGitHub is the world’s leading platform for agentic software development — powered by Copilot to build, scale, and deliver secure software. Over 180 million developers, including more than 90% of the Fortune 100 companies ...

Software Architect - Python & .NET

Hiring Organisation
Computer Futures
Location
Birmingham, West Midlands, West Midlands (County), United Kingdom
Employment Type
Permanent
Salary
£75000 - £90000/annum
An established and growing automation specialist is seeking an experienced Software Architect to lead the design and development of complex, integrated software platforms. The organisation delivers sophisticated automation solutions that connect enterprise applications, operational software ...

Senior Cloud Infrastructure Engineer - Scale & Observability

Location
Greater London, England, United Kingdom
scale. You will own tooling in Python/Go, extend Kubernetes and Docker usage, and help advance multi-cloud readiness. You will contribute to observability through Prometheus/OpenTelemetry/Grafana, strengthen DR/BCP, and collaborate with a talented team to deliver robust, scalable infrastructure across customers. #J ...

Senior Cloud Platform Engineer – Multi-Cloud & Observability

Location
Greater London, England, United Kingdom
Thought Machine is hiring Senior Software Engineers for Infrastructure to deploy and maintain cloud-native platform infrastructure. You will work on multi-cloud orchestration, observability, and data layers to reduce cognitive load for developers and clients at scale. The role focuses on building resilient, scalable tooling with Python or Golang ...

IBM Netcool / Observability Technical Lead

Hiring Organisation
Deerfoot Recruitment Solutions Ltd
Location
London, United Kingdom
Employment Type
Full-Time
Salary
£780.00 - £830.00 per day
Netcool/Observability Technical Lead Inside IR35 Contract -up to £827pd London Hybrid - 4 Days Onsite/1 Day WFH per Week Banking Are you the person who knows exactly why an ObjectServer failover didn't behave as expected, and how to stop a flood of duplicate events before anyone … shape how thousands of infrastructure and application events are detected, correlated and actioned across EMEA, and you'll have genuine scope to modernise observability capability rather than simply keep the lights on. This is a hands-on technical leadership role with no direct reports, so your influence comes from your ...

Senior Product Manager for AI Observability

Location
Greater London, England, United Kingdom
Role Profile As part of the LSEG AI, we are hiring a Senior Product Manager–AI Observability to own the strategy, design and rollout of telemetry systems that monitor, measure and analyse how AI models, MCPs, and AI-enabled features behave across all LSEG products. This role is responsible … model performance tracking cost efficiency user experience optimisation operational reliability auditability and the long‐term evolution of our AI platform. Key Responsibilities & Accountabilities Telemetry & Observability Strategy Define the end‐to‐end telemetry vision and roadmap for LLMs, MCPs, vector stores, embeddings, inference layers and AI‐powered user experiences. Establish ...

Senior Product Manager for AI Observability

Hiring Organisation
London Stock Exchange Group
Location
London, UK
Employment Type
Full-time
Role Profile As part of the LSEG AI, we are hiring a Senior Product Manager – AI Observability to own the strategy, design and rollout of telemetry systems that monitor, measure and analyse how AI models, MCPs, and AI-enabled features behave across all LSEG products. This role is responsible … model performance tracking cost efficiency user experience optimisation operational reliability auditability and the long-term evolution of our AI platform. Key Responsibilities & Accountabilities Telemetry & Observability Strategy Define the end-to-end telemetry vision and roadmap for LLMs, MCPs, vector stores, embeddings, inference layers and AI-powered user experiences. Establish ...

Senior Developer Advocate - Data Observability

Location
Greater London, England, United Kingdom
They own projects from beginning to end, facilitating collaboration and enabling Datadog's community to solve real-world problems. With a focus on Data Observability, this role will enable our community of engineers around Datadog to be part of a movement of building better software. This is a unique opportunity … engineering expertise and advocacy skills to shape the ever‐evolving technological landscape. What You’ll Do: Act as a subject matter expert for data observability on behalf of the Datadog advocacy and engineering teams Create content in one or more mediums to build Datadog's reputation as a leader ...

Site Reliability Engineer - Banking & Finance

Location
Greater London, England, United Kingdom
software engineering and infrastructure. You'll develop internal platforms, tooling, and automation across Linux, distributed systems, and cloud-native technologies, helping improve reliability, observability, and operational efficiency across a global production environment. Responsibilities: Design and develop internal infrastructure tooling and automation. Build and maintain monitoring, observability and configuration management platforms. …/Must Have: Strong experience programming with Python, Go and/or C++ Strong Linux knowledge and understanding of distributed systems. Experience with monitoring, observability or SRE practices. Experience with CI/CD pipelines, Git and infrastructure automation. Familiarity with Kubernetes and containerised workloads. Strong analytical and troubleshooting skills. Benefits ...

Lead DevOps Engineer - Real Time Platform

Location
Greater London, England, United Kingdom
global scale. This is an individual contributor leadership role , where you’ll define infrastructure architecture, raise operational standards, and ensure resilience, security, and observability across a mission‐critical platform. AI-First Engineering This team operates with an AI-first approach. We expect hands‐on experience with AI development tooling: terminal … augmented IDEs, and automated workflows. You are ultimately accountable for production quality, security, and correctness. This means owning infrastructure review, security validation, system observability, and operational guardrails. WHAT YOU'LL DO Design and operate cloud infrastructure on AWS to support low‐latency, always‐on real‐time workloads. Own infrastructure ...

Senior / Principal Applied AI Engineer (UK / Europe, Remote)

Location
United Kingdom
production-grade AI agents, agentic workflows, and LLM integrations for internal automation and in-product features for a web-native trading platform. Own implementation, observability, and security while partnering with Product and Platform teams to deliver measurable AI systems in production. Job Description Role Senior/Principal Applied AI Engineer … integrate LLMs into internal systems and the trading product. This is a hands-on engineering role: you will write production code, put evaluation and observability on everything shipped, and operate autonomously in a lean team. Key Responsibilities Build agent loops, tool-calling, structured outputs, planning/state management, retries, guardrails ...

Observability SRE AVP: Cloud Observability & Migrations

Location
Greater London, England, United Kingdom
Citi is seeking an experienced Site Reliability Engineer to lead end-to-end observability and resiliency initiatives in a large-scale environment. You will migrate monitoring tooling to Google Cloud Observability and Grafana, implement OpenTelemetry instrumentation, and author reusable deployment solutions for OpenShift/Kubernetes and VM environments. The role ...

SRE / Platform Engineer - Remote

Hiring Organisation
Genesis10
Location
New York, United States
Employment Type
Permanent
Salary
USD 105 Hourly
infrastructure and operational problems. Engineers on the team write code every day and work across application and infrastructure layers to improve reliability, performance, scalability, observability, and system integration. A major initiative for the team is establishing a centralized observability capability across an environment where monitoring and operational data have historically … been siloed. The organization is bringing telemetry together using Datadog and enterprise data lake capabilities, creating a common observability foundation that can ultimately support AIOps, agentic AI, automated remediation, and self-healing systems. This is not a traditional operations or Solutions Architecture position. The successful candidate will be expected ...

Core Platform Developer

Location
City Of London, England, United Kingdom
reliability of internal systems. This person should be comfortable working across multiple areas of the stack, from service frameworks and API enablement to observability, governance, and developer workflows. This is a high-ownership role within a global, fast-moving engineering environment. Key Responsibilities Design and build shared backend services, frameworks … developer tooling that support internal application and service development. Develop common platform capabilities such as service templates, authentication and authorization patterns, API standards, observability integrations, error handling, and shared runtime utilities. Improve the developer experience through better tooling, automation, documentation, onboarding patterns, and paved-road workflows for engineering teams. Help ...

DevOps & Infrastructure Engineer

Location
Gloucester, England, United Kingdom
Security customers, spanning both on-premise environments and cloud-based solutions. You’ll lead hands-on DevOps and infrastructure engineering across CI/CD, observability, infrastructure-as-code and platform automation, helping teams build secure, reliable and scalable services in demanding environments. What you’ll be doing: You’ll lead … cloud-based solutions. Develop and maintain CI/CD pipelines, GitOps workflows and automated deployment approaches using tools such as ArgoCD. Implement and improve observability using Prometheus, Grafana, logging and alerting to support resilient platform operations. Use infrastructure-as-code and platform automation with Helm, Go and Terraform to deliver ...

MLOps Engineer

Hiring Organisation
DGH Recruitment
Location
City of London, London, United Kingdom
Employment Type
Permanent
platform reliability. Key Responsibilities - Design, deploy, and manage AI platforms and agent infrastructure - Build and maintain CI/CD pipelines and DevOps workflows - Implement observability, monitoring, and logging solutions - Optimise performance, scalability, and cost efficiency - Support AI teams with infrastructure, deployment, and integration - Ensure platform security, compliance, and high availability …/CD, automation, and DevOps best practices - Experience with Kubernetes/containerisation technologies - Strong programming skills (e.g. Python, Go, Node.js) - Experience with observability tools (e.g. OpenTelemetry, Datadog) - Understanding of security, performance optimisation, and scalability Desirable Skills - Experience working on AI/ML platforms or deployments - Exposure to large-scale distributed ...

Senior Software Engineer, GoLang

Location
Greater London, England, United Kingdom
continuous integration and delivery (CI/CD) Make data-guided decisions affecting core business metrics and processes Apply platform and reliability engineering practices, including observability, performance optimisation, analytics, and security best practices Facilitate collaboration between teams and promote continuous improvement Mentor junior engineers on engineering practices, coding standards, and troubleshooting … DevOps Continuous Integration and Delivery (CI/CD) Infrastructure-as-Code Hard Skills Application Development Performance Optimisation Automation Development Containerisation Development Methodologies Design Patterns Observability Analytics Security Best Practices Troubleshooting Soft Skills Mentoring Collaboration Continuous Improvement Industry Keywords Public-Facing Systems Internal Insurance Systems Tools & Technologies AWS Lambda DynamoDB Azure ...

MLOps Engineer

Hiring Organisation
Dgh Recruitment
Location
City of London, Greater London, UK
platform reliability. Key Responsibilities - Design, deploy, and manage AI platforms and agent infrastructure - Build and maintain CI/CD pipelines and DevOps workflows - Implement observability, monitoring, and logging solutions - Optimise performance, scalability, and cost efficiency - Support AI teams with infrastructure, deployment, and integration - Ensure platform security, compliance, and high availability …/CD, automation, and DevOps best practices - Experience with Kubernetes/containerisation technologies - Strong programming skills (e.g. Python, Go, Node. xehkeey js) - Experience with observability tools (e.g. OpenTelemetry, Datadog) - Understanding of security, performance optimisation, and scalability Desirable Skills - Experience working on AI/ML platforms or deployments - Exposure to large ...

Lead Product Manager AIOPs

Location
Greater London, England, United Kingdom
responsible for S&P Global's enterprise AIOps platform and strategy, driving the modernization of IT Operations and Site Reliability Engineering (SRE) through intelligent observability, event intelligence, automation, and AI-driven insights. DTS Platform & Tools – Service Enablement: We serve as thought leaders in AIOps, partnering across IT Operations, SRE, engineering … solving, prioritization, and decision‐making skills. What We’re Looking For: Basic Required Qualifications: 10+ years of experience in product management, IT operations, SRE, observability, platform engineering, or related enterprise technology roles. Strong understanding of AIOps concepts, including event correlation, anomaly detection, root cause analysis, noise reduction, predictive analytics ...