3,001 to 3,025 of 4,483 Observability Jobs

Azure SRE/ Architect Contract Dublin 6-18 Months

Hiring Organisation
Adecco
Location
Dublin, City of Dublin, Republic of Ireland
Employment Type
Contract
Contract Rate
£431 - £604/day ltd company
with horizontal teams across platform engineering, networking, security, and AI infrastructure to embed reliability into Azure-native solutions. You will be required to support observability, incident response, automation, resilience testing, and operational runbook development for Azure landing zones, ingress and egress DMZs, service integration layers, and AI infrastructure supporting Microsoft … Azure platform services with a clear focus on reliability, availability, and scalability. You will have hands-on experience with Site Reliability Engineering practices, including observability, incident response, automation, resilience testing, and operational readiness. You will be able to collaborate effectively with platform engineering, networking, security, and infrastructure teams to embed ...

Senior DevOps Engineer - ELK

Hiring Organisation
ECS Resource Group Ltd
Location
City of London, London, United Kingdom
Employment Type
Contract
Contract Rate
£500 - £700/day
become available to join one of the world's leading technology organisations as a Senior DevOps Engineer, helping to deliver and scale critical observability, automation, and platform engineering solutions across a complex enterprise environment. As a Senior DevOps Engineer, you will be responsible for: Designing, deploying, and optimising large-scale … Elasticsearch environments. Leading improvements to monitoring, logging, and observability platforms. Driving automation and infrastructure-as-code best practices. Supporting and enhancing Kubernetes and GitOps deployments. Troubleshooting complex performance and reliability issues. Collaborating with technical teams to improve platform scalability, resilience, and security. Promoting DevOps best practices across engineering teams. Requirements ...

Senior QA Engineer

Hiring Organisation
SRG
Location
Warrington, Cheshire, United Kingdom
Employment Type
Full-Time
Salary
£45,000 - £50,000 per annum
both fast and reliable. You'll help move quality earlier into the process (shift-left), while also using real production insights to improve decisions (observability-led quality). There's strong scope to influence how QA operates within the team, from testing strategy through to continuous improvement. What … considered early in design and development Leading exploratory testing to uncover issues real users might experience Using monitoring, metrics, and logs to drive observability-led quality and improve production outcomes Identifying risks early and helping the team make informed decisions Improving QA processes, standards, and ways of working across ...

Senior Infrastructure & Operations Engineer (Kubernetes / Platform Reliability)

Location
Greater London, England, United Kingdom
production systems at scale and who focuses on making infrastructure predictable and stable. You’ll work across Kubernetes, networking, CI/CD, Cloudflare, and observability to create a platform engineers can trust. What You’ll Do Design, deploy, and maintain production Kubernetes clusters. Own cluster reliability, upgrades, security, and performance. … Build and operate monitoring, logging, and alerting pipelines. Ensure full-stack observability across infrastructure and services. Design and maintain CI/CD pipelines that are fast, reproducible, and safe. Improve deployment strategies (rollouts, canaries, rollbacks). Automate infrastructure provisioning and configuration. Investigate and resolve production incidents. Improve system resilience, redundancy ...

Senior Software Engineer

Location
United Kingdom
environment Work with AWS serverless technologies, including Lambda, SQS, EventBridge and API Gateway Work with MongoDB and document databases Contribute to CI/CD, observability, incident management and production support Work closely with Product, Engineering and Operations to turn business requirements into scalable solutions Mentor engineers and contribute to engineering … standards and best practice Strong Node.js & TypeScript Strong system design and architecture AWS & Serverless MongoDB/document databases Distributed systems, observability and incident management CI/CD and production operations Experience working in a regulated environment, ideally FinTech or financial services Comfortable owning services within a build-and-run environment ...

Senior Product Manager (SaaS)

Hiring Organisation
LinuxRecruit
Location
London, UK
Employment Type
Full-time
comfort with technical details are must-haves. You'll be on the technical side too, having experience with containerised platforms using Kubernetes, databases, and observability tools such as Prometheus and OpenTelemetry. This is a chance to shape the future of observability and security, build products people count ...

Remote Senior Mobile Engineer - Instrumentation SDK (iOS) UK Remote

Location
Clevedon, Somerset, United Kingdom
Grafana Labs is the company behind Grafana Cloud, the fully managed observability platform trusted by more than 10,000 organizations to ensure reliability, resolve incidents faster, and optimize telemetry at scale. Built on open source and open standards and designed for interoperability across any stack, Grafana Cloud brings … observability and observability to AI, giving teams (and their agents) unified visibility so they can see, understand, and act on all their disparate data, wherever it lives, and move at the speed of their ambitions. Customers, including Anthropic, Bloomberg, NVIDIA, Microsoft, and Salesforce, rely on Grafana Labs. ...

Remote Senior Mobile Engineer - Instrumentation SDK (iOS) UK Remote

Location
Bakewell, Derbyshire, United Kingdom
Grafana Labs is the company behind Grafana Cloud, the fully managed observability platform trusted by more than 10,000 organizations to ensure reliability, resolve incidents faster, and optimize telemetry at scale. Built on open source and open standards and designed for interoperability across any stack, Grafana Cloud brings … observability and observability to AI, giving teams (and their agents) unified visibility so they can see, understand, and act on all their disparate data, wherever it lives, and move at the speed of their ambitions. Customers, including Anthropic, Bloomberg, NVIDIA, Microsoft, and Salesforce, rely on Grafana Labs. ...

Platform Engineer, Kubernetes & Automation — AI Cloud

Location
United Kingdom
seeking a Platform Engineer to operate and advance a cloud-native platform powering AI workloads. You will work on Kubernetes clusters, automation, and observability to deliver reliable services across our AI infrastructure. You will collaborate with software, infra, and SRE teams, mentor peers, and help define standards for deployment ...

Senior Platform & Cloud Engineer – Azure, DevOps

Location
Greater London, England, United Kingdom
cloud solutions. You will partner with architects and other engineers to deliver cloud adoption, environment design, and operational readiness, while embedding security, compliance and observability throughout. #J-18808-Ljbffr ...

Senior Software Engineer

Hiring Organisation
IO Associates
Location
Bath, Somerset, South West, United Kingdom
Employment Type
Contract, Work From Home
Contract Rate
£550 - 600 per day
environment Work with AWS serverless technologies , including Lambda, SQS, EventBridge and API Gateway Work with MongoDB and document databases Contribute to CI/CD, observability, incident management and production support Work closely with Product, Engineering and Operations to turn business requirements into scalable solutions Mentor engineers and contribute to engineering … standards and best practice Key Skills Strong Node.js & TypeScript Strong system design and architecture AWS & Serverless MongoDB/document databases Distributed systems, observability and incident management CI/CD and production operations Experience working in a regulated environment , ideally FinTech or financial services Comfortable owning services within a build ...

Platform Engineer

Location
Greater London, England, United Kingdom
direction and build the systems, tooling and processes that the wider engineering team relies on. You’ll work across production infrastructure, Linux performance, observability, deployments, developer experience and internal tooling, with significant freedom to decide what needs improving and take ownership of delivering it. Responsibilities Own and improve infrastructure supporting … real-time, 24/7 production systems Build reliable deployment, rollback and operational workflows Improve observability across metrics, logging, dashboards, tracing and alerting Develop tooling and automation that allows engineers to ship faster and more safely Improve CI/CD, build processes, test environments and configuration workflows Work ...

Cloud-Native Backend Engineer for Data Processing

Location
Greater London, England, United Kingdom
with a focus on reliability and performance. You will collaborate with product managers and researchers to design scalable systems, use IaC, and contribute to observability with Prometheus and Loki. The team values curiosity and ownership, shipping robust software from Canary Wharf. #J-18808-Ljbffr ...

Senior Python Data Platform Engineer for Finance Data Lakes

Location
Greater London, England, United Kingdom
components with Python in our London office. You will own ETL pipelines, data lakes/lakehouses, and distributed systems while improving CI/CD, observability, and infrastructure. Experience with financial/market data is essential, as is a strong background in databases and data-intensive systems. #J-18808-Ljbffr ...

DataOps Engineer - Secure, Automated Data Pipelines

Location
United Kingdom
architect and deliver DataOps capabilities, focusing on repeatability, governance, and scalable data applications in air-gapped environments. You will implement and manage monitoring and observability to ensure data quality across the data flow, using tools like Airflow, Docker, Terraform and Kubernetes, with a strong emphasis on CI/ ...

Fullstack Engineer

Location
Cheltenham, England, United Kingdom
looking for a Fullstack Engineer with the following skills: The Fundamentals: Modern JavaScript GitLab CI/CD Observability Technologies Java GitLab Version Control Database Technologies Web API Integration Apache NiFi Message Broker Technologies OpenShift Web Auth Distributed Architecture Kubernetes SecDevOps Cloud Native Technologies Required Clearance: DV or recently active ...

Platform Engineer: Azure Cloud Infra, CI/CD & Kubernetes

Location
Greater London, England, United Kingdom
release systems, primarily Azure, Terraform and Ansible. This hands-on role focuses on reliable, scalable, and secure cloud environments, continuous integration and deployment, observability, cost optimisation, and collaboration with engineering teams to enable rapid, safe software delivery. #J-18808-Ljbffr ...

Remote Senior Mobile Engineer - Instrumentation SDK (iOS) UK Remote

Location
Loanhead, Midlothian, United Kingdom
Grafana Labs is the company behind Grafana Cloud, the fully managed observability platform trusted by more than 10,000 organizations to ensure reliability, resolve incidents faster, and optimize telemetry at scale. Built on open source and open standards and designed for interoperability across any stack, Grafana Cloud brings … observability and observability to AI, giving teams (and their agents) unified visibility so they can see, understand, and act on all their disparate data, wherever it lives, and move at the speed of their ambitions. Customers, including Anthropic, Bloomberg, NVIDIA, Microsoft, and Salesforce, rely on Grafana Labs. ...

Automation-First QA Engineer for API & Backend

Location
England, United Kingdom
design and maintain automated test suites (Karate, Cucumber with RestAssured) in Java, integrate tests into CI/CD, and contribute to contract testing and observability with New Relic. #J-18808-Ljbffr ...

Hybrid AI Platform Engineer — MLOps/DevOps

Location
City Of London, England, United Kingdom
collaborate with data science and engineering teams to deploy AI workloads and ensure reliable infrastructure. The role focuses on building CI/CD pipelines, observability, and IaC automation, while optimizing performance and cost. You will ensure security and high availability for production systems. #J-18808-Ljbffr ...

Lead AI Infra SRE: Scale, Reliability & Mentorship

Location
Gloucester, England, United Kingdom
improving automation to support AI/HPC workloads. You will configure and operate resilient Linux systems (Ubuntu), refine performance, and contribute to the observability stack with Prometheus and Grafana. #J-18808-Ljbffr ...

MLOps & DevOps Engineer for Production Platform

Location
Manchester, England, United Kingdom
specialist software team in Manchester. You will build and operate delivery pipelines, environments, and the model-serving path with IaC, CI/CD, observability, cost control, and production support across the platform. You will bring production experience, strong Terraform and Kubernetes skills, CI/CD ownership, and the ability ...

Platform Engineer: Core Backend & Developer Tools

Location
Greater London, England, United Kingdom
shared backend services, frameworks, and tooling that improve how internal teams develop, deploy, and operate software. The role emphasizes platform engineering, API standards, and observability across services. The ideal candidate has extensive backend experience, strong Python skills, and familiarity with FastAPI, Docker, Kubernetes, and CI/CD in cloud contexts. ...

Product Engineering Environment Lead

Hiring Organisation
Randstad Technologies
Location
London, South East England, United Kingdom
Employment Type
Full-Time
Salary
£600.00 - £650.00 per day
tested recovery capabilities with defensible RTO/RPO metrics. Automation & Self-Service: Drive Infrastructure/Environments as Code (IaC/EaC), automated provisioning, and observability to enable on-demand instantiation. Your Experience Professional Background: Proven track record leading enterprise-scale environment estates, spanning hybrid cloud and physical/network infrastructure. … Automation & Tooling: Hands-on experience with Infrastructure-as-Code and Environment-as-Code tools (e.g., Terraform, Ansible), CI/CD pipelines, and observability frameworks. SRE, SDLC & DR Depth: Strong understanding of SRE principles (SLIs/SLOs, toil reduction), release management, and High Availability/Disaster Recovery design and testing. ...

Product Engineering Environment Lead

Hiring Organisation
Randstad Technologies Recruitment
Location
London, United Kingdom
Employment Type
Contract
Contract Rate
£600 - £650/day
tested recovery capabilities with defensible RTO/RPO metrics. Automation & Self-Service: Drive Infrastructure/Environments as Code (IaC/EaC), automated provisioning, and observability to enable on-demand instantiation. Your Experience Professional Background: Proven track record leading enterprise-scale environment estates, spanning hybrid cloud and physical/network infrastructure. … Automation & Tooling: Hands-on experience with Infrastructure-as-Code and Environment-as-Code tools (e.g., Terraform, Ansible), CI/CD pipelines, and observability frameworks. SRE, SDLC & DR Depth: Strong understanding of SRE principles (SLIs/SLOs, toil reduction), release management, and High Availability/Disaster Recovery design and testing. ...