426 to 450 of 608 Prometheus Jobs in the UK

Lead SRE

Location
Westminster, West End, United Kingdom
discovery, ingress, networking, and load balancing. Experience with Kubernetes. Experience with cloud computing services. Familiarity with common observability and reliability toolchains such as Grafana, Prometheus, Elasticsearch, Kibana, or Jaeger. Ability to use AI-assisted engineering tools responsibly, including validating outputs, understanding failure modes, and applying secure handling of sensitive information. ...

Lead Software Engineer - LLM Ops Platform Reliability

Location
Paisley, Scotland, United Kingdom
local GPU clusters, using reproducible infrastructure as code and continuous delivery pipelines Implement observability (logs, metrics, traces) with dashboards and actionable alerting, including Prometheus metrics and Grafana/Alertmanager integration for LLM and GPU workloads Tune GPU and accelerator capacity, autoscaling, and cost efficiency for LLM inference workloads using performance ...

Software Engineering Specialist

Hiring Organisation
BT Group
Location
Cheltenham, Gloucestershire, United Kingdom
Salary
£ 70 K
with the Elastic stack technologies and Kibana plugin developmentHave used DevOps tools and principles like Git, Jenkins, GitOpsUsed system monitoring tools such as CheckMK, Prometheus, Grafana or LokiOur PackageTailored benefits make a real difference. That’s why we offer a comprehensive range to support your growth, wellbeing, and everyday life. ...

Platform Engineer

Location
Greater London, England, United Kingdom
hands-on with Kubernetes, and enough Terraform to have opinions about how to structure it. Familiarity with CI/CD pipelines and observability tooling (Prometheus, Grafana, Datadog, or similar). Strong technical foundation: Demonstrated ability to write production-quality code and solve hard technical problems (experience with Python, TypeScript/ ...

Principal Machine Learning Infrastructure Engineer London, United Kingdom

Location
Greater London, England, United Kingdom
consume data Experience building model serving infrastructure with latency and throughput requirements Familiarity with experiment tracking tools (Weights & Biases, MLflow) and observability stacks (Prometheus, Grafana) What we offer Equity options – share in our success and growth. 10% employer pension contribution – invest in your future. Free office lunches – great food ...

Lead SRE - Chase UK

Location
Greater London, England, United Kingdom
discovery, ingress, networking, and load balancing. Experience with Kubernetes. Experience with cloud computing services. Familiarity with common observability and reliability toolchains such as Grafana, Prometheus, Elasticsearch, Kibana, or Jaeger. Ability to use AI-assisted engineering tools responsibly, including validating outputs, understanding failure modes, and applying secure handling of sensitive information. ...

Software Engineer (ML Projects)

Location
Greater London, England, United Kingdom
cloud‐native TeamCity for CI/CD (lots of teams are releasing code 15-20 times per day!) Terraform Prometheus and Grafana If you have built and deployed complex Python applications or have hands‐on experience with generative AI and LLMs, we would be especially keen to talk. ...

Platform Engineer (10x Openings)

Location
Greater London, England, United Kingdom
Ceph or similar) at an engineering level. Background building Kubernetes operators using frameworks such as Kopf, controller‐runtime, or similar. Experience with observability tooling: Prometheus, Grafana, OpenTelemetry, or structured logging in distributed systems. Experience building SaaS or PaaS layers on top of an IaaS platform. Exposure to serverless or inference ...

Senior Software Engineer - Live & VOD Video Infrastructure

Hiring Organisation
Roku
Location
Cambridge, Cambridgeshire, UK
Employment Type
Full-time
similar technologiesExperience with GPU-accelerated encoding or hardware media pipelinesFamiliarity with Kubernetes, ECS, Nomad, or other orchestration platformsExperience with observability stacks such as Prometheus, Grafana, OpenTelemetry, ELK, or DatadogExperience building fault-tolerant ingest or transcoding platforms operating across multiple regions#LI-JC5What's Roku's approach to hybrid working? Roku fosters ...

Database Reliability Engineer

Location
Manchester, England, United Kingdom
multi-cloud—while ensuring rigorous data integrity and mobility A Security & Observability Mindset: You believe security is paramount. You focus on building deep observability (Prometheus/Grafana/OpenTelemetry/Humio) and automated guardrails so the fleet is secure by design without requiring manual intervention Engineering via Code: While ...

Test Environment Manager (10105)

Location
Greater London, England, United Kingdom
Improvement Monitor environment availability, health, performance and utilisation. Develop appropriate metrics and dashboards for environment reporting. Work with monitoring and logging technologies such as Prometheus, Grafana and Splunk. Identify opportunities to improve environment reliability, automation, scalability and cost efficiency. Drive continuous improvement across Test Environment Management processes. Essential Experience 5+ … Operations teams. Highly Desirable Experience Azure/Azure DevOps Jenkins/GitLab Terraform or other IaC technologies HP NonStop infrastructure Java-based environments Prometheus/Grafana/Splunk UFT/Selenium/Cucumber Linux shell scripting Large-scale financial services or similarly complex regulated environments *Rates depend on experience ...

Senior DevOps Engineer (Trade Surveillance Platform)

Location
Belfast, Northern Ireland, United Kingdom
Build and maintain CI/CD pipelines, driving automation and release consistency Administer Kubernetes clusters, managing orchestration, scaling, and performance Implement monitoring and observability (Prometheus, Grafana, ELK) to ensure system reliability Troubleshoot complex infrastructure, Kubernetes, and database issues with clear root cause analysis Configure AWS networking and implement secure access … Code (Terraform, CloudFormation, Ansible) Proven CI/CD and automation experience Strong Linux administration and troubleshooting skills Experience with monitoring/logging tools (Prometheus, Grafana, ELK) and root cause analysis Knowledge of storage (NFS/EFS), databases (PostgreSQL), and web servers (Nginx, Apache, Tomcat) Experience with SSO and security (SAML ...

Test Environment Manager (10105)

Hiring Organisation
Salt Search
Location
London, South East England, United Kingdom
Employment Type
Full-Time
Salary
£500.00 - £850.00 per day
Improvement Monitor environment availability, health, performance and utilisation. Develop appropriate metrics and dashboards for environment reporting. Work with monitoring and logging technologies such as Prometheus, Grafana and Splunk . Identify opportunities to improve environment reliability, automation, scalability and cost efficiency. Drive continuous improvement across Test Environment Management processes. Essential Experience … following would be advantageous: Azure/Azure DevOps Jenkins/GitLab Terraform or other IaC technologies HP NonStop infrastructure Java-based environments Prometheus/Grafana/Splunk UFT/Selenium/Cucumber Linux shell scripting Large-scale financial services or similarly complex regulated environments *Rates depend on experience and client ...

Senior Lead Site Reliability / DevOps Engineer

Location
Auchentibber, Scotland, United Kingdom
reliability for large-scale OpenTelemetry pipelines on hybrid on-prem/cloud environments, supporting telemetry ingestion, processing, and export to backends such as InfluxDB, Prometheus, Elasticsearch, and OpenSearch Drives the assessment, refactoring, and incremental migration of custom legacy telemetry collection code to standardized OpenTelemetry instrumentation, reducing technical debt while maintaining … Advanced proficiency and experience in observability such as white and black box monitoring, SLO alerting, and telemetry collection using tools such as Grafana, Dynatrace, Prometheus, Datadog, Splunk, Elasticsearch, etc. Proficiency in continuous integration and continuous delivery tools (e.g., Jenkins, GitLab, Terraform, etc.) Experience with container and container orchestration (e.g. ...

Senior Lead Site Reliability / DevOps Engineer

Hiring Organisation
JP Morgan Chase
Location
Glasgow, UK
Employment Type
Full-time
reliability for large-scale OpenTelemetry pipelines on hybrid on-prem/cloud environments, supporting telemetry ingestion, processing, and export to backends such as InfluxDB, Prometheus, Elasticsearch, and OpenSearchDrives the assessment, refactoring, and incremental migration of custom legacy telemetry collection code to standardized OpenTelemetry instrumentation, reducing technical debt while maintaining system … Advanced proficiency and experience in observability such as white and black box monitoring, SLO alerting, and telemetry collection using tools such as Grafana, Dynatrace, Prometheus, Datadog, Splunk, Elasticsearch, etc. Proficiency in continuous integration and continuous delivery tools (e.g., Jenkins, GitLab, Terraform, etc.)Experience with container and container orchestration (e.g. ...

LiveOps Engineer

Hiring Organisation
Civica
Location
United Kingdom
Salary
£ 70 K
standards for reliability, performance, and customer experience.What you will be doing: Monitoring and incident response: Operate our live environments using modern observability tools (Grafana, Prometheus, DataDog, Azure Monitor). Respond quickly to alerts, triage incidents, and coordinate with SRE and Platform teams to resolve issues and restore service. Automation …/CD pipelines and fluent with version control (GitHub Actions, Azure DevOps, Jenkins, or similar). Experienced with monitoring and alerting solutions (Prometheus, Grafana, DataDog, Elastic, Azure Monitor). Strong analytical and problem-solving skills, with the ability to stay calm during incidents. Collaborative communicator who thrives in cross-functional ...

GCP Data Engineer

Hiring Organisation
Teksystems
Location
Sheffield, South Yorkshire, United Kingdom
Employment Type
Contract
Contract Rate
£400/day
DevOps practices, CI/CD pipelines, and automation tools to support continuous delivery. experience with monitoring, logging, and observability tools such as CloudWatch, Prometheus, and Grafana. Strong understanding of security best practices in cloud environments, including IAM, encryption, and network security. Significant experience working with relational databases and data integration … activities. You will work in an Agile and DevOps-oriented environment, using tools such as Terraform, CloudFormation, AWS CDK, Docker, Kubernetes/EKS, CloudWatch, Prometheus, Grafana, and a range of AWS and GCP services. The work involves close collaboration with developers, traders, and business stakeholders across multiple regions, focusing ...

Senior Specialist Engineer (Specialist Site Reliability Engineer SRE)

Hiring Organisation
National Health Service
Location
London, United Kingdom
Salary
£ 70 K
programming/scripting languages such as Python, PowerShell or BashUnderstanding of Linux/Unix & Windows systems, networking, and distributed systemsExperience with observability tools (e.g., Prometheus, Grafana, Datadog) and alerting systemsUnderstanding of infrastructure automation (e.g., Terraform, Ansible, PowerShell, Helm)Excellent communication and collaboration skillsPossesses problem solving skills and the ability … programming/scripting languages such as Python, PowerShell or BashUnderstanding of Linux/Unix & Windows systems, networking, and distributed systemsExperience with observability tools (e.g., Prometheus, Grafana, Datadog) and alerting systemsUnderstanding of infrastructure automation (e.g., Terraform, Ansible, PowerShell, Helm)Excellent communication and collaboration skillsPossesses problem solving skills and the ability ...

DevOps Engineer

Hiring Organisation
Halian Technology Limited
Location
Reading, Berkshire, South East, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£85,000
Our client is an innovative healthcare technology company whose SaaS platform supports healthcare professionals across the UK. They're looking for a DevOps Engineer to help scale, maintain, and enhance a highly available cloud-based ...

Senior Software Engineer

Location
Greater London, England, United Kingdom
Job Title: Java Kafka EngineerLocation: NorthamptonAbout the Job you are considering:Join the Financial Services Business Unit within Capgemini’s Cloud & Custom Applications (C&CA) practice, where we Consult with Purpose, Continuously Evolve, and Architect ...

Lead DevOps Engineer — Secure Cloud & On-Prem Platform

Location
Gloucester, England, United Kingdom
driving secure, scalable services. You'll mentor engineers, implement CI/CD, GitOps and IaC using ArgoCD, Terraform and Helm, and ensure observability with Prometheus and Grafana in demanding environments. #J-18808-Ljbffr ...

Senior AI Platform Engineer — Multi-Cloud Infra & Reliability

Location
Greater London, England, United Kingdom
Platform/DevOps engineer to own our infrastructure as code, CI/CD pipelines, and multi-cloud reliability. You’ll work across Terraform, Prometheus, Grafana, and Python to keep services fast, auditable, and secure for enterprise customers. You’ll join a small, fast-moving team building the V7 Go platform ...

Database Engineer

Hiring Organisation
firstpointgroup
Location
United Kingdom
Salary
£ 70 K
queries and configurations for high traffic, data intensive workloadsOwn backup, restore, PITR and HA strategy across globally distributed systemsMonitor platform health using Ops Manager, Prometheus and GrafanaWork alongside software and infrastructure teams on schema design, automation and application performanceWhat they are looking forA degree in Computer Science or a closely ...

Database Engineer

Hiring Organisation
firstpointgroup
Location
Moffat, Dumfries & Galloway, UK
Employment Type
Full-time
queries and configurations for high traffic, data intensive workloadsOwn backup, restore, PITR and HA strategy across globally distributed systemsMonitor platform health using Ops Manager, Prometheus and GrafanaWork alongside software and infrastructure teams on schema design, automation and application performanceWhat they are looking forA degree in Computer Science or a closely ...

Platform / System Engineer

Location
City Of London, England, United Kingdom
Demonstrable experience of external vendor relationship management Nice to haves Containerization (Docker/Kubernetes) in a production environment Monitoring tools in a production environment (Prometheus/Grafana/ELK stack/Splunk) If you are interested, submit your application now. #J-18808-Ljbffr ...