351 to 375 of 497 Prometheus Jobs in the UK

Lead SRE - Chase UK

Location
London, United Kingdom
discovery, ingress, networking, and load balancing. Experience with Kubernetes. Experience with cloud computing services. Familiarity with common observability and reliability toolchains such as Grafana, Prometheus, Elasticsearch, Kibana, or Jaeger. Ability to use AI-assisted engineering tools responsibly, including validating outputs, understanding failure modes, and applying secure handling of sensitive information. ...

Lead Software Engineer - LLM Ops Platform Reliability

Hiring Organisation
Hackajob Ltd
Location
Glasgow, Lanarkshire, Scotland, United Kingdom
Employment Type
Permanent
local GPU clusters, using reproducible infrastructure as code and continuous delivery pipelines Implement observability (logs, metrics, traces) with dashboards and actionable alerting, including Prometheus metrics and Grafana/Alertmanager integration for LLM and GPU workloads Tune GPU and accelerator capacity, autoscaling, and cost efficiency for LLM inference workloads using performance ...

Lead SRE - Chase UK

Hiring Organisation
Hackajob Ltd
Location
South West London, London, United Kingdom
Employment Type
Permanent
discovery, ingress, networking, and load balancing. Experience with Kubernetes. Experience with cloud computing services. Familiarity with common observability and reliability toolchains such as Grafana, Prometheus, Elasticsearch, Kibana, or Jaeger. Ability to use AI-assisted engineering tools responsibly, including validating outputs, understanding failure modes, and applying secure handling of sensitive information. ...

Software Engineering Specialist

Hiring Organisation
BT Group
Location
Cheltenham, Gloucestershire, United Kingdom
Salary
£ 70 K
with the Elastic stack technologies and Kibana plugin developmentHave used DevOps tools and principles like Git, Jenkins, GitOpsUsed system monitoring tools such as CheckMK, Prometheus, Grafana or LokiOur PackageTailored benefits make a real difference. That’s why we offer a comprehensive range to support your growth, wellbeing, and everyday life. ...

Platform Engineer

Location
Greater London, England, United Kingdom
hands-on with Kubernetes, and enough Terraform to have opinions about how to structure it. Familiarity with CI/CD pipelines and observability tooling (Prometheus, Grafana, Datadog, or similar). Strong technical foundation: Demonstrated ability to write production-quality code and solve hard technical problems (experience with Python, TypeScript/ ...

SRE Observability Technical Lead - Vice President

Location
Belfast City District, Northern Ireland, United Kingdom
Experience in SRE, Observability Engineering, or platform infrastructure roles focused on operational telemetry. Hands‐on experience in observability tools and stacks such as Grafana, Prometheus, OpenTelemetry, ELK, Splunk, and similar platforms. Deep understanding of SLIs, SLOs, Error Budgets, and telemetry best practices in high‐availability environments. Proven ability to troubleshoot ...

Lead SRE - Chase UK

Hiring Organisation
JP Morgan Chase
Location
London, United Kingdom
Salary
£ 80 K
components, including service discovery, ingress, networking, and load balancing.Experience with Kubernetes.Experience with cloud computing services.Familiarity with common observability and reliability toolchains such as Grafana, Prometheus, Elasticsearch, Kibana, or Jaeger.Ability to use AI-assisted engineering tools responsibly, including validating outputs, understanding failure modes, and applying secure handling of sensitive information. ...

SRE Observability Technical Lead - Vice President

Hiring Organisation
Citigroup
Location
Belfast, Down, United Kingdom
Salary
£ 60 K
stack.Qualifications:Experience in SRE, Observability Engineering, or platform infrastructure roles focused on operational telemetry.Hands-on experience in observability tools and stacks such as Grafana, Prometheus, OpenTelemetry, ELK, Splunk, and similar platforms.Deep understanding of SLIs, SLOs, Error Budgets, and telemetry best practices in high-availability environments.Proven ability to troubleshoot integration issues ...

Platform Engineer (10x Openings)

Location
Greater London, England, United Kingdom
Ceph or similar) at an engineering level. Background building Kubernetes operators using frameworks such as Kopf, controller‐runtime, or similar. Experience with observability tooling: Prometheus, Grafana, OpenTelemetry, or structured logging in distributed systems. Experience building SaaS or PaaS layers on top of an IaaS platform. Exposure to serverless or inference ...

Senior Software Engineer – Live & VOD Video Infrastructure

Hiring Organisation
Roku
Location
Cambridge, Cambridgeshire, United Kingdom
Salary
£ 80 K
similar technologiesExperience with GPU-accelerated encoding or hardware media pipelinesFamiliarity with Kubernetes, ECS, Nomad, or other orchestration platformsExperience with observability stacks such as Prometheus, Grafana, OpenTelemetry, ELK, or DatadogExperience building fault-tolerant ingest or transcoding platforms operating across multiple regions#LI-JC5What's Roku's approach to hybrid working Roku fosters ...

SRE Managing Consultant - Cloud Operating Model

Location
Manchester, England, United Kingdom
SLIs/SLOs, incident management, observability, and continuous improvement across cloud and hybrid platforms.* Exposure to modern observability tooling and ecosystems (e.g. Datadog, Dynatrace, Prometheus, OpenTelemetry, Loki), with a strong understanding of how metrics, logs, and traces are applied to inform reliability strategy, incident management, and operational decision‐making.## **Security ...

SRE Managing Consultant - Cloud Operating Model

Location
Greater London, England, United Kingdom
SLIs/SLOs, incident management, observability, and continuous improvement across cloud and hybrid platforms.* Exposure to modern observability tooling and ecosystems (e.g. Datadog, Dynatrace, Prometheus, OpenTelemetry, Loki), with a strong understanding of how metrics, logs, and traces are applied to inform reliability strategy, incident management, and operational decision‐making.## **Security ...

Senior Infrastructure Engineer, Research Singapore

Location
Greater London, England, United Kingdom
consume data Experience building model serving infrastructure with latency and throughput requirements Familiarity with experiment tracking tools (Weights & Biases, MLflow) and observability stacks (Prometheus, Grafana) What we offer Build what actually matters Help shape an AI-native engineering company at a formative stage, tackling problems that genuinely matter for industry ...

SC Cleared DevOps Engineer (EKS AWS)

Hiring Organisation
Hays Technology
Location
London, United Kingdom
Employment Type
Contract
Contract Rate
£420/day £420 p/d Inside IR35
Building and maintaining CI/CD pipelines using GitLab CI Containerising and deploying applications with Docker and Helm Implementing monitoring, alerting and logging using Prometheus and Grafana Supporting microservices-based environments and platform modernisation initiatives Driving automation, infrastructure as code and DevSecOps best practices Providing incident response, RCA and production … looking for Strong AWS experience (EC2, S3, RDS, IAM, VPC) Kubernetes (EKS), Docker and Helm Terraform and Ansible GitLab CI/CD and Git Prometheus and Grafana Kafka and relational/non-relational databases (MongoDB, MySQL, PostgreSQL) Linux administration (RHEL/Ubuntu) Experience supporting highly available production platforms and microservices ...

SC Cleared DevOps Engineer (EKS AWS)

Hiring Organisation
Hays Specialist Recruitment Limited
Location
London, South East England, United Kingdom
Employment Type
Full-Time
Salary
£420.00 per day
Building and maintaining CI/CD pipelines using GitLab CI Containerising and deploying applications with Docker and Helm Implementing monitoring, alerting and logging using Prometheus and Grafana Supporting microservices-based environments and platform modernisation initiatives Driving automation, infrastructure as code and DevSecOps best practices Providing incident response, RCA and production … looking for Strong AWS experience (EC2, S3, RDS, IAM, VPC) Kubernetes (EKS), Docker and Helm Terraform and Ansible GitLab CI/CD and Git Prometheus and Grafana Kafka and relational/non-relational databases (MongoDB, MySQL, PostgreSQL) Linux administration (RHEL/Ubuntu) Experience supporting highly available production platforms and microservices ...

Senior Lead Site Reliability / DevOps Engineer

Hiring Organisation
JP Morgan Chase
Location
Glasgow, Lanarkshire, United Kingdom
Salary
£ 80 K
reliability for large-scale OpenTelemetry pipelines on hybrid on-prem/cloud environments, supporting telemetry ingestion, processing, and export to backends such as InfluxDB, Prometheus, Elasticsearch, and OpenSearchDrives the assessment, refactoring, and incremental migration of custom legacy telemetry collection code to standardized OpenTelemetry instrumentation, reducing technical debt while maintaining system … Advanced proficiency and experience in observability such as white and black box monitoring, SLO alerting, and telemetry collection using tools such as Grafana, Dynatrace, Prometheus, Datadog, Splunk, Elasticsearch, etc.Proficiency in continuous integration and continuous delivery tools (e.g., Jenkins, GitLab, Terraform, etc.)Experience with container and container orchestration (e.g., ECS, Kubernetes ...

DevOps Engineer

Hiring Organisation
G.R.E. Recruitment Limited
Location
Guildford, Surrey, South East, United Kingdom
Employment Type
Permanent
Salary
£50,000
have an understanding of automation and help them automate some of their manual processes. Other technologies they use are: Scripting Languages: Bash & Python Monitoring: Prometheus & Grafana Databases: Maria DB (open to SQL) CI: Jenkins & GIT This role is fully onsite 5 days per week so all candidates must live commutable ...

Automation Engineer

Location
Glasgow, Scotland, United Kingdom
automation and observability experience with Python, PyKit (Perl/PowerShell a plus) . Familiarity with CI/CD pipelines (Jenkins), Git, and technologies like Prometheus, Grafana, and Ansible . Experience with REST APIs and debugging complex issues beyond documentation. Background in systems administration (UNIX/Windows) and knowledge of backup ...

Software Engineer – Infrastructure and Automation

Hiring Organisation
Quant Capital
Location
London, United Kingdom
Salary
£ 80 K
firmwideComfortable working across technical domains and collaborating with peersNice to HaveExperience with Docker, KVM, or other container/virtualisation toolsFamiliar with observability stacks (Prometheus, Grafana)TerraformPractical knowledge of networking protocols and hardware environmentsUnderstanding of low-latency or post-trade systemsWhy Apply Work on the internal infrastructure that makes the firm ...

Linux Systems Engineer – Trading

Hiring Organisation
Quant Capital
Location
London, United Kingdom
Salary
£ 80 K
responsible for designing and supporting highly available systems across a diverse technology stack.My client leverages the latest technologies such as Dock, Kubernetes, Prometheus and Grafana as well as CI/CD to meet the growing demands of the business.A fantastic opportunity has arisen for a candidate looking to further their ...

Lead Network Operations Engineer

Hiring Organisation
G Research
Location
London, United Kingdom
Salary
£ 80 K
relationships Familiarity with automation and CI/CD tooling (e.g. Ansible, Python, Jenkins) and Infrastructure as Code Experience with observability and monitoring tools (e.g. Prometheus, Grafana, OpenTelemetry, ELK) Strong ability to analyse and troubleshoot distributed systems end-to-end Proactive, self-starting mindset with a strong sense of ownership Comfortable ...

Senior Software Engineer

Hiring Organisation
Hackajob Ltd
Location
South West London, London, United Kingdom
Employment Type
Permanent
Salary
£95,000
maintenance, to retirement Designing systems that scale Expertise in some of our main programming languages - TypeScript, Java, Golang, Rust, Python Desirable Experience with Kubernetes, Prometheus, Terraform, NoSQL or GCP Perks of joining us: Company pension contributions at 5% Individualised training budget for you to learn on the job and level ...

Senior Software Engineer London, Greater London, England, United Kingdom

Location
Greater London, England, United Kingdom
wrk2, and JVM profiling to identify and fix performance bottlenecks Hands‐on experience with instrumentation and analysis of production metrics using tools like Prometheus, Grafana, InfluxDB, or the ELK stack to identify performance bottlenecks and ensure system health About the Team You will join a team that helps build ...

Senior MySQL Database Administrator (DBA)

Hiring Organisation
Aristocrat Technologies
Location
London, United Kingdom
Salary
£ 80 K
Database provisioningo Backup and recoveryo Replication, failover, and disaster recoveryo Schema migrations and deployments· Enhance observability through monitoring, alerting, and tooling (e.g. PMM, Prometheus, Grafana).· Shift performance and security detection earlier in the development lifecycle via tooling and guidelines.Database Governance & Security· Define and enforce database security standards, access controls ...

Engineer II, Site Reliability (Remote, GBR)

Hiring Organisation
CrowdStrike
Location
United Kingdom
Salary
£ 60 K
intrinsic drive to make things betterBias towards small development projects and the occasional larger projectsHave experience with modern monitoring and telemetry stacks (ELK, Prometheus, Grafana, Zabbix)Gather and analyze metrics from both operating systems and applications to assist in performance tuning and fault findingAbility to lead incident analysis for incidents ...