76 to 100 of 346 Prometheus Jobs in London

Site Reliability Engineer - Core

Hiring Organisation
Blockchain
Location
London, United Kingdom
Salary
£ 70 K
allocation, network and/or internals.Experience working with cloud solutions (GCP or AWS).Deep understanding and demonstrable experience with modern monitoring tools such as Prometheus, Datadog, Grafana, TelegrafExperience with infrastructure as code tools. Experience with complex Terraform deployments is a plus.Solid background with configuration management tools. Experience with Saltstack ...

Cloud Support Engineer

Location
Greater London, England, United Kingdom
systems knowledge, including filesystems, networking and system internals Programming skills in Golang and Python and experience with infrastructure tools or observability stacks (e.g. Grafana, Prometheus, EFK) Confidence in working with cloud-native platforms and tools (e.g. Kubernetes, Terraform, AWS/GCP, Docker) Excellent communication skills under pressure, being able ...

Manager, System and Platform Operations

Location
Greater London, England, United Kingdom
networking, security, and system architecture. Proficient in scripting languages (Java, Golang, Python, Bash, or similar). Experience with monitoring and observability tools (DataDog, Prometheus, Grafana). Knowledge of database management systems (PostgreSQL, Bigtable). Understanding of API and microservices architecture. Strong people leadership skills with at least a year ...

DevOps & Environment Lead (AWS / Cloud Platforms)

Hiring Organisation
London Stock Exchange Group
Location
London, United Kingdom
Salary
£ 100 K
environments meet security and compliance standards.Lead root-cause analysis and implement preventative automation.Tech StackAWS, Terraform, CloudFormation, GitHub Actions, Jenkins, Docker, Kubernetes/EKS, CloudWatch, Prometheus/Grafana, Geneos, ELK/Opensearch, Python, Shell.Experience We ValueStrong experience managing and automating cloud environments on AWS.Expertise in CI/CD and DevOps practices.Experience ...

Infrastructure Developer

Hiring Organisation
SQUAREPOINT CAPITAL
Location
London, United Kingdom
Salary
£ 120 K
containers, GPU runtime, HashiVaultCMDB/API Platform: FastAPI, PostgreSQL, SQLAlchemy, Pydantic v2, Knative Eventing, PythonCloud: GCP (primary), AWS; BigQuery, GKE, cloud storage, Sagemaker HyperpodObservability: Prometheus, Grafana, ELK/Loki, AkvoradoSource Control & CI/CD: GitLab, Jira, ConfluenceRequired Qualifications:Bachelor’s degree in computer science, Software Engineering, or a related technical ...

Senior Dev Ops Engineer

Location
Greater London, England, United Kingdom
manual Kubernetes, etc.) Extensive experience designing and building high‐availability, resilient infrastructure Experience setting up and maintaining monitoring, alerting, log collection (rsyslog, Grafana, Prometheus, ELK) Automation and scripting experience (Terraform, Packer, Ansible, Bash) Extensive Linux administration experience (Ubuntu LTS) Bachelor’s Degree in Computer Science or equivalent experience Experience ...

Product Support Engineer

Location
Greater London, England, United Kingdom
. Exposure to SRE and DevOps practices , CI/CD and modern production environments. Experience with monitoring and observability tools such as Grafana, Prometheus, Datadog, Splunk, New Relic or similar . Experience troubleshooting distributed, event-driven or highly available systems . Strong understanding of incident response, RCA and production reliability ...

Analytics Services Platform Engineer

Hiring Organisation
G Research
Location
London, United Kingdom
Salary
£ 80 K
understanding of AWS analytics technologies including EMR, MSK, Athena, Redshift, Glue and MWAAExperience with CI/CD and observability tools such as Jenkins, ArgoCD, Prometheus, Grafana and OpenTelemetryStrong problem-solving skills and a systematic approach to diagnosing and resolving issuesHighly desirable skillsExperience with streaming frameworks such as Flink, Kafka Streams ...

Site Reliability Engineer

Location
City of Westminster, England, United Kingdom
Amazon Web Services and Google Cloud Platform Experience supporting Kubernetes (EKS) environments and service mesh technologies such as Istio Knowledge of observability tooling including Prometheus, Grafana or Coralogix Experience with PostgreSQL, MongoDB or HashiCorp Vault Experience using GitLab, Flux or Helm within CI/CD pipelines Knowledge of PCI-compliant ...

Senior Platform Engineer

Location
Greater London, England, United Kingdom
freedom while keeping us in control. What you'll be doing from day one Owning our infrastructure as code in Terraform, plus alerting, observability (Prometheus, Grafana) and reliability, including load testing and disaster recovery exercises. Making CI/CD faster (GitHub Actions, ArgoCD) and taking obstacles out of engineers ...

Senior Platform Engineer

Hiring Organisation
V7 Labs
Location
London, United Kingdom
Salary
£ 100 K
people freedom while keeping us in control.What you’ll be doing from day one:Owning our infrastructure as code in Terraform, plus alerting, observability (Prometheus, Grafana) and reliability, including load testing and disaster recovery exercises.Making CI/CD faster (GitHub Actions, ArgoCD) and taking obstacles out of engineers’ way, from ...

Database Platform Engineer

Hiring Organisation
G Research
Location
London, United Kingdom
Salary
£ 80 K
Kubernetes environmentsUnderstanding of AWS services relevant to data platforms, such as RDS, Aurora, S3, EC2 or EKSFamiliarity with modern observability stacks, such as Prometheus, Grafana, Elk or OTelDesirable Skills:Experience with cloud-native and distributed SQL database, such as Aurora, YugabyteDB or TiDBKnowledge of data streaming and integration tools, including ...

Database Platform Engineer

Location
Greater London, England, United Kingdom
environments Understanding of AWS services relevant to data platforms such as RDS, Aurora, S3, EC2 or EKS Familiarity with modern observability stacks such as Prometheus, Grafana, Elk or OTel Desirable: experience with cloud‐native and distributed SQL databases such as Aurora, YugabyteDB or TiDB Desirable: knowledge of data streaming ...

Senior MLOps Engineer

Hiring Organisation
MFK Recruitment
Location
London, United Kingdom
Salary
£ 80 K
skills, including multi-stage or multi-architecture builds.Experience building CI/CD pipelines for Machine Learning systems.Experience with monitoring and observability tools such as Prometheus, Grafana, Datadog or similar.Linux systems administration and shell-scripting experience.Strong software engineering practices, including Git, testing and code reviews.Experience delivering AI or Machine Learning systems ...

Site Reliability Engineer

Location
Greater London, England, United Kingdom
Amazon Web Services and Google Cloud Platform Experience supporting Kubernetes (EKS) environments and service mesh technologies such as Istio Knowledge of observability tooling including Prometheus, Grafana or Coralogix Experience with PostgreSQL, MongoDB or HashiCorp Vault Experience using GitLab, Flux or Helm within CI/CD pipelines Knowledge of PCI-compliant ...

Senior Cloud Engineer (K8S)

Location
Greater London, England, United Kingdom
focus on end-user availability. Desirable but notrequired: Experience with Openstack cloud platform s. Experience with solutions for monitoring and observability. e.g. Grafana, Prometheus, OpenSearch/ElasticSearch, Loki. Experience with High Performance Computing (HPC) environments using SLURM or similar batch workload solutions. Programming experience with Python3 utilising classes and inheritance. ...

DevOps Team Lead - Development Platform

Hiring Organisation
US Bank
Location
London, United Kingdom
Salary
£ 100 K
grade reliability experienceKubernetes workload management, Helm/Helmfile chart authoring, Docker image best practicesTerraform module development and state management for AWS infrastructureMonitoring with CloudWatch, Prometheus, Grafana, or ELK; incident investigation skillsExperience implementing DevSecOps practices across pipelines and infrastructure.Familiarity with container and IaC vulnerability scanning.Proficiency with secrets management using External Secrets ...

Senior Platform Engineer

Hiring Organisation
JP Morgan Chase
Location
London, United Kingdom
Salary
£ 100 K
infrastructure managementExperience with server hardware selection, BIOS/firmware tuning, and bare-metal provisioning for performance-sensitive workloadsFamiliarity with observability and telemetry stacks (e.g., Prometheus, Grafana, ELK)Experience with market-data and messaging technologies common to electronic trading (e.g., multicast market data, low-latency messaging buses)Performing research, development ...

Staff SRE, AI Infrastructure

Hiring Organisation
wayve
Location
London, United Kingdom
Salary
£ 80 K
C++) with a bias toward automation.Deep troubleshooting skills across networking, storage, distributed systems, and performance at scale.Experience designing and operating observability stacks (e.g. Datadog, Prometheus, Grafana, OpenTelemetry).Clear communication skills, including leading incidents, writing postmortems, and influencing teams to prioritise reliability improvements.Desirable skillsFamiliarity with infrastructure-as-code (e.g. Terraform ...

Staff SRE, AI Infrastructure

Hiring Organisation
wayve
Location
London, UK
Employment Type
Full-time
bias toward automation. Deep troubleshooting skills across networking, storage, distributed systems, and performance at scale. Experience designing and operating observability stacks (e.g. Datadog, Prometheus, Grafana, OpenTelemetry).Clear communication skills, including leading incidents, writing postmortems, and influencing teams to prioritise reliability improvements. Desirable skillsFamiliarity with infrastructure-as-code (e.g. Terraform ...

Principal Site Reliability Engineer, Infrastructure Observability

Location
Greater London, England, United Kingdom
observability, APM and infrastructure monitoring, and application‐specific logging Knowledge/experience with observability tools such as New Relic, SolarWinds DPA, Elastic Stack, Prometheus, Grafana, Splunk, and cloud native tools Knowledge/experience with cloud management tools such as Ansible, Terraform, Vault, and Vagrant Works independently, with guidance in only ...

Principal Platform Engineer

Location
Greater London, England, United Kingdom
modern platform and engineering tooling such as GitHub Actions, Apigee, Airflow, and related cloud‐native technologies. Expertise with observability platforms such as Data Dog, Prometheus, Grafana, ELK, Splunk or equivalent, including monitoring, logging, tracing, reliability engineering, and incident management. Sound technical judgement with the ability to balance innovation, risk, operational ...

Fastly: Senior SRE – Networks

Location
Greater London, England, United Kingdom
analyze internet traffic patterns across multiple dimensions using flow-based tools. Experience working with alerting, monitoring and visibility tools (such as Graphite/Grafana, Prometheus, or Splunk). Knowledge across cloud hosting solutions (i.e., GCP, AWS and Azure). Knowledge of DevOps practices and CI/CD pipelines (ie. ...

Software Engineering Specialist

Hiring Organisation
BT Group
Location
London, United Kingdom
Salary
£ 80 K
management & IaC frameworks (Terraform, Ansible, Chef, Puppet)Containerisation & container platforms (Kubernetes, Helm-based management & deployment)Artifact & package management (Nexus/Artifactory/Harbor)Observability (Prometheus, Grafana, ELK stack, Fleet, Defend)Identity & access management (IAM, OAuth/OIDC e.g. Keycloak, Secrets Mgmt e.g. Vault)Encryption & certificate lifecycle management (TLS/ ...

Software Engineering Specialist

Location
Greater London, England, United Kingdom
management & IaC frameworks (Terraform, Ansible, Chef, Puppet) Containerisation & container platforms (Kubernetes, Helm-based management & deployment) Artifact & package management (Nexus/Artifactory/Harbor) Observability (Prometheus, Grafana, ELK stack, Fleet, Defend) Identity & access management (IAM, OAuth/OIDC e.g. Keycloak, Secrets Mgmt e.g. Vault) Encryption & certificate lifecycle management (TLS/ ...