126 to 150 of 166 Prometheus Jobs in London

Sr. Observability Engineer – Kings Cross, London

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
hybrid and cloud-native environments.* Innovate & Automate: Spearhead the evaluation, selection, and implementation of cutting-edge observability tools and platforms (e.g., Dynatrace, OpenTelemetry, Prometheus, Grafana). Architect and build robust, automated observability pipelines. Take an active part in documenting and defining processes and best practice.* Optimize & Analyze: Conduct deep-dive … experience in architecting and designing large-scale monitoring and observability solutions.* Expert-Level Tooling: Deep expertise with modern observability platforms (e.g., Dynatrace, AWS Cloudwatch, Prometheus, Grafana, ELK Stack, Splunk, OpenTelemetry).* Cloud & Infrastructure: Advanced knowledge of major cloud platforms (AWS, Azure, GCP), containerization (Docker, Kubernetes), and Infrastructure as Code (Terraform ...

Senior SRE - Kubernetes, GitOps & AI-Driven Resilience

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
coding, architecture, and operations across deployment pipelines. Join a team shaping scalable, secure delivery in a modern collaboration platform, leveraging Argo CD, Vault, and Prometheus/Grafana in a hybrid work model. #J-18808-Ljbffr ...

Kubernetes SRE: DevOps, Observability & Reliability

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
platform. In this role, you will implement GitOps workflows with Argo CD, manage canary releases and secret management with Vault, and monitor reliability with Prometheus and Grafana, collaborating across teams in a hybrid London office. #J-18808-Ljbffr ...

Infrastructure Engineer-Hyper-V

Hiring Organisation
NEEV LIMITED
Location
London, United Kingdom
Employment Type
Contract
Contract Rate
From £400 to £450 per day 400 - 450 GBP/day InsideIR35
Altaro, or native Hyper-V Replica. Exposure to hybrid cloud and private cloud platforms. Familiarity with monitoring and observability platforms such as Azure Monitor, Prometheus, Grafana, Splunk, or similar tools. Experience supporting enterprise VDI environments. Understanding of ITIL Incident, Problem, Change, and Release Management. Experience working in regulated industries such ...

Staff HFT Rust Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
conference attendance Technology Stack Primary language: Rust Additional tools: C++, Python, Kubernetes, Kafka, ClickHouse Infrastructure: AWS, bare metal colocation in strategic locations Monitoring: Prometheus, Grafana, custom observability tools #J-18808-Ljbffr ...

Software Engineering III - AI/ML Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
resolution. Experience in observability such as white and black box monitoring, service level objective alerting, and telemetry collection using tools such as Grafana, Dynatrace, Prometheus, Datadog, Splunk, and others Strong understanding of SLI/SLO/SLA and Error Budgets Hands‐on experience using enterprise‐authorized AI‐assisted software development ...

Senior Software Engineer

Hiring Organisation
Hackajob Ltd
Location
South West London, London, United Kingdom
Employment Type
Permanent
Salary
£95,000
maintenance, to retirement Designing systems that scale Expertise in some of our main programming languages - TypeScript, Java, Golang, Rust, Python Desirable Experience with Kubernetes, Prometheus, Terraform, NoSQL or GCP Perks of joining us: Company pension contributions at 5% Individualised training budget for you to learn on the job and level ...

Senior Software Engineer London, Greater London, England, United Kingdom

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
wrk2, and JVM profiling to identify and fix performance bottlenecks Hands‐on experience with instrumentation and analysis of production metrics using tools like Prometheus, Grafana, InfluxDB, or the ELK stack to identify performance bottlenecks and ensure system health About the Team You will join a team that helps build ...

ai engineer for defence systems

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
database systems Nice to have: Experience with Rust and Python, container-based, cloud-native, and edge application architectures, maintaining and operating production systems with Prometheus, Grafana, ELK, or similar, SQL, NoSQL, and streaming database systems, production ML systems, professional experience in a defence context Условия Competitive compensation and VSOP options ...

Lead Backend Engineer (Routing Squad)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
languages: Python, Go Tech infrastructure: AWS, CDK TypeScript, Lambda, SQS, EventBridge, RDS, DynamoDB Data tooling: GCP, BigQuery, Looker, Looker Studio Observability: Loki, Tempo, Grafana, Prometheus Event‐driven architecture and domain‐driven design Our interviewing process Intro call with the hiring manager Live coding challenge solving a problem A take home ...

SRE: Kubernetes + GitOps for Reliable Microservices

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Kustomize, Argo CD, and strong GitOps practices. The role emphasizes AI-assisted tooling, secure secret management, and a data-driven approach to reliability using Prometheus and Grafana within a hybrid London office setting. #J-18808-Ljbffr ...

Site Reliability Engineer – 24/7 On-Call & Automation Focus

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
building tools to automate repetitive tasks, with collaboration across development teams to design resilient, scalable services. Applicants should bring experience with Grafana/Prometheus, Docker, OpenShift or Kubernetes, and strong #J-18808-Ljbffr ...

Python Developer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
part of childhood spent hacking away in 8-bit assembly language Python 3.10+, JavaScript, TypeScript, and Go RabbitMQ, Kafka, PostgreSQL, Redis, Linux servers, OpenTelemetry, Prometheus, Grafana, and Zabbix Nice to have: Interest in functional programming and its application in the real world Условия: Extremely lucrative salary and significant bonus Greenfield ...

Senior Engineering Manager, Developer Experience

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
track record of owning a technical domain end-to-end. You bring strong technical foundation across the DevEx stack: CI/CD, observability (Prometheus, Grafana, or equivalent), Kubernetes-based platforms - sufficient to make sound architectural decisions and earn engineer trust. You know how to lead through ambiguity and organisational change ...

Machine Learning Systems & Infrastructure Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
/CD pipelines, including self‐hosted GPU runners. Observability and reliability: Monitoring, logging, and alerting for job performance, data‐pipeline health, and cost (e.g., Prometheus/Grafana, OpenTelemetry); define SLOs and incident response for the systems you own. Security and access: Manage secrets, IAM, and network boundaries (e.g., Tailscale, cloud … storage with caching layers. Familiarity with ML workflow orchestration and experiment tracking (e.g., Kubeflow Pipelines, MLflow). Experience with monitoring and observability tooling (e.g., Prometheus/Grafana, OpenTelemetry) and CI/CD for infra and ML workflows (e.g., GitHub Actions). At SpAItial, we are committed to creating a diverse ...

Lead Java Developer — Real-Time Risk & Cloud (Hybrid)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
design to production support, integrating new analytics and data sets across global teams. The role emphasizes scalable microservices, streaming data, and observability with ELK, Prometheus and Grafana. Hybrid work model and competitive benefits are offered. #J-18808-Ljbffr ...

Head of Engineering

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
polishing what nobody needs. Our tech stack: Primary: TypeScript, React, Node, SQL Frameworks: Express, React‐hooks, Redux, RTK, Mantine Infrastructure: Docker, GCP, Kubernetes, Tracing, Prometheus While familiarity with our stack is helpful, we value your ability to learn and adapt over specific technical experience. We’re deliberately looking for ambitious ...

DevOps Engineer

Hiring Organisation
CBSbutler Holdings Limited trading as CBSbutler
Location
Croydon, London, United Kingdom
Employment Type
Contract
Contract Rate
£470 - £500/day
maintaining Infrastructure as Code using Terraform. Managing AWS infrastructure including RDS, Lambda and S3. Supporting Docker and Helm deployments. Monitoring platform health using Grafana, Prometheus and Kibana. Managing Linux infrastructure. Administering secrets management and security certificates. Troubleshooting infrastructure, deployment and platform issues. Working closely with development and platform teams … Looking For You'll have strong experience across most of the following technologies: Kubernetes Docker Helm AWS Terraform Jenkins Linux Amazon RDS Bitbucket Grafana, Prometheus and Kibana CI/CD pipeline management and automation Infrastructure as Code Cloud platform support Experience with Vault or similar secrets management tools ...

Senior Java Developer

Hiring Organisation
CODEVERSE LIMITED
Location
London, South East, England, United Kingdom
Employment Type
Contractor
Contract Rate
£300 - £400 per day
modern API design practices. Implement resilient systems with appropriate retry, timeout and error-handling mechanisms. Work with monitoring and observability tools including Grafana, Prometheus and Dynatrace . Provide hands-on production support, troubleshooting and root cause analysis. Support P1/P2 incidents within a 24x7 support environment. Contribute to technical … experience. Experience with AWS Aurora/PostgreSQL . Production experience with Docker and Kubernetes . Strong REST API and microservices experience. Experience with Grafana, Prometheus or Dynatrace . Strong debugging, troubleshooting and problem-solving skills. Experience working with highly available and scalable systems. Experience with API resilience, retries, timeouts ...

Site Reliability Engineer- Spacetime UK

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
roadmap to mature our observability stack, moving from cloud-native tools to a robust, scalable, and insightful platform built on best-in-class technologies (Prometheus, OpenTelemetry, etc.). If you are an SRE who thrives on platform-building challenges and wants to be relied upon to build a production-grade … this role includes on-call responsibilities. Key Responsibilities Help design and build Aalyria's centralized observability platform, integrating and scaling tools for metrics (e.g. Prometheus), logging (e.g. Loki), and distributed tracing (e.g. Tempo/OpenTelemetry). Define, implement, and manage a robust framework of Service Level Objectives (SLOs), Service Level ...

DevOps and Automation Engineer (Contract)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
towards self-service environment provisioning using Infrastructure-as-Code (Terraform) and pipeline-driven automation.* Implement advanced observability and monitoring: Use platforms such as Datadog, Prometheus, Grafana, and OpenTelemetry to provide real-time insights into system health, deployments, and business metrics.* Embed security and compliance by design: Integrate security into every … generate new automation ideas and create user stories for rapid prototyping.* Use technologies like Terraform, Ansible, Azure DevOps, Github, OctopusDeploy, Kubernetes, OpenTelemetry, Datadog, Grafana, Prometheus, low-code automation platforms (e.g., Power Automate, UiPath) to help evolve the team's capabilities.* Coordinate planned outage and environment refreshes in collaboration with project ...

Cloud Engineering & Architecture - Senior Platform Engineer AI - Vice President

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Temporal, or custom agentic loops) to coordinate multi-step diagnostic and remediation tasks. AIOps & Intelligent Observability: Ability to integrate traditional observability stacks (e.g., Datadog, Prometheus, OpenTelemetry) with AI/ML models to automate root-cause analysis, anomaly detection, and semantic log clustering. Self-healing Infrastructure Engineering: Experience designing closed-loop … Experience working in regulated industries is a plus. Preferred Qualifications Experience building self-service platforms for development teams. Familiarity with observability and monitoring tools (Prometheus, Grafana, Datadog, CloudWatch). Background in financial services or other highly regulated environments. About Goldman Sachs At Goldman Sachs, we commit our people, capital ...

Cloud Engineering & Architecture - Senior Platform Engineer AI - Vice President

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Temporal, or custom agentic loops) to coordinate multi‐step diagnostic and remediation tasks. AIOps & Intelligent Observability: Ability to integrate traditional observability stacks (e.g., Datadog, Prometheus, OpenTelemetry) with AI/ML models to automate root‐cause analysis, anomaly detection, and semantic log clustering. Self‐Healing Infrastructure Engineering: Experience designing closed‐loop … Qualifications: AWS certifications (Solutions Architect Professional, DevOps Engineer, etc.). Experience building self‐service platforms for development teams. Familiarity with observability and monitoring tools (Prometheus, Grafana, Datadog, CloudWatch). Background in financial services or other highly regulated environments. About the company At the company, we commit our people, capital ...

Cloud Infra Engineer (Go/Kubernetes) for AI Platform

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
seeks a Confluent Cloud Infrastructure Software Engineer to design, implement and operate cloud foundation services. You will work on Kubernetes operators, Terraform, Datadog and Prometheus, focusing on high availability and scalable infrastructure within the Confluent Cloud Platform. Experience with large-scale systems and public cloud environments is essential. You will ...

Director, Observability & AIOps Engineering

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
AIOps innovation. The ideal candidate will have extensive experience in software engineering, possess leadership skills, and have a strong knowledge of observability platforms like Prometheus and Grafana. A commitment to promoting diversity and inclusion is also essential. #J-18808-Ljbffr ...