12 of 12 Prometheus Jobs in the City of London

Senior AWS Site Reliability Engineer

Hiring Organisation
Spectrum IT Recruitment Limited
Location
City of London, London, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£70,000
/CD, or CircleCI Strong understanding of containerization (e.g., Docker, Kubernetes) and microservices architecture Skilled in using observability and monitoring tools such as Prometheus, Grafana, ELK stack, or AWS CloudWatch Excellent analytical and troubleshooting abilities, especially within complex distributed systems Proven experience handling incident management and conducting blameless postmortems, including ...

AI Platform Engineer

Location
City Of London, England, United Kingdom
with AWS and/or Azure, Kubernetes, and Docker . Experience with GitOps, ideally Argo CD or Flux. Familiarity with observability tooling such as Prometheus, Grafana, Datadog, Splunk, ELK, or OpenTelemetry. Understanding of platform security, secrets management, policy-as-code, and SRE practices. Experience working in Legal, professional services ...

Platform Engineer

Location
City Of London, England, United Kingdom
cloud‐native infrastructure. Experience with Kubernetes, including Amazon EKS, and Docker or other container runtimes. Experience with monitoring and observability tools such as Grafana, Prometheus, Loki or Datadog. Experience with artifact‐management platforms such as JFrog Artifactory. Knowledge of AWS Batch, AWS Step Functions, AWS Identity and Access Management ...

Principal Engineer I — Prepurchase Platform

Location
City Of London, England, United Kingdom
tools and techniques - using LLMs, AI-assisted development, and automation to accelerate engineering workflows and improve system intelligence. Familiarity with observability tooling: Grafana, Splunk, Prometheus,OpenTracing, or equivalent. Strong understanding of security best practices - OAuth/OIDC, input validation,secretsmanagement. Track recordof leading technical initiatives across multiple teams without direct ...

Senior Linux DevOps Engineer

Hiring Organisation
RedTech Recruitment Ltd
Location
City of London, London, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£90,000
performance troubleshooting within Linux environments Hands-on experience operating containerised workloads using Docker and Kubernetes Experience with monitoring, logging and observability technologies such as Prometheus, Grafana, Loki, OpenTelemetry or the ELK Stack Experience with Infrastructure as Code and automation tooling such as Terraform and Ansible Experience building and managing … Engineer/Linux/Bash/Shell Scripting/Python/Kubernetes/Docker/Terraform/Ansible/Microsoft Azure/Azure/Prometheus/Grafana/Loki/OpenTelemetry/ELK Stack/GitLab CI/CD/GitHub Actions/Jenkins/ArgoCD/Helm/ ...

Site Reliability Engineer

Hiring Organisation
SR2 | Socially Responsible Recruitment | Certified B Corporation™
Location
City of London, London, United Kingdom
demo environments Automate infrastructure provisioning and deployment workflows (Terraform, GitHub Actions, GitOps) Package and deploy applications to customer environments Implement and optimise observability tooling (Prometheus, Grafana, Loki) Support incident response, monitoring, and backup/recovery planning Mentor project teams in DevSecOps practices and environment management Ensure cloud environments are secure … cost-optimised Tech Environment & Skills: Cloud Engineering: AWS/Azure/GCP, Linux, Terraform (IaC) Containers: Kubernetes, Docker, Helm (OpenShift a plus) Observability: Prometheus, Grafana, Loki (network visualisation desirable) CI/CD & GitOps: GitHub Actions, ArgoCD/Flux Security: Cloud access models, Zero Trust principles Programming: Python or Golang preferred ...

Software engineering specialist

Location
City Of London, England, United Kingdom
deployments, and tool setups directly into developer workflows. Tech Stack Delivery: Automate the provisioning and configuration of our operational ecosystem, including tools across observability (Prometheus, Elastic, Checkmk), security/compliance (Tenable, Red Hat Satellite), service registry/IPAM (NetBox), container orchestration (ArgoCD), and artifact management (Artifactory, GitLab). Engineering Standards ...

Site Reliability Engineer

Location
City Of London, England, United Kingdom
with Kubernetes, containerisation technologies, and cloud infrastructure. Strong understanding of networking, distributed systems, and infrastructure automation. Experience with monitoring and observability tools such as Prometheus, Grafana, Splunk, or similar. Proven track record of solving complex reliability, scalability, or performance challenges. Excellent problem-solving skills and ability to operate effectively ...

Senior Site Reliability Engineer

Location
City Of London, England, United Kingdom
speed and reducing deployment risk. Adaptable & Problem-Solver : Address complex challenges across configuration, policy, observability, and data services. Apply a data-driven approach using Prometheus and Grafana to improve reliability and performance. Ownership & Quality : Own end-to-end configuration quality, enforcing governance with Open Policy Agent. Ensure secure, compliant deployments ...

Platform / System Engineer

Location
City Of London, England, United Kingdom
Demonstrable experience of external vendor relationship management Nice to haves Containerization (Docker/Kubernetes) in a production environment Monitoring tools in a production environment (Prometheus/Grafana/ELK stack/Splunk) If you are interested, submit your application now. #J-18808-Ljbffr ...

Staff Python Engineer (ML)

Location
City Of London, England, United Kingdom
clear, and easy to test Developing observability for new and existing ML applications and GenAI/LLM integrations , making use of the Grafana Stack (Prometheus, Loki, Tempo) Develop integrations and services that communicate with Google Services. Working closely with Data Scientists and ML Engineers throughout the lifecycle of productionising their ...

Director of Software Engineering (AIOps) - Executive Director

Location
City Of London, England, United Kingdom
Preferred qualifications, capabilities, and skills Usage of large scale Observability Platforms (example products IBM Netcool, (Watson CloudPak AiOps), Tivoli, SCOM, SMARTS, Dynatrace, Splunk, Elastic, Prometheus, Grafana or Messaging systems for telemetry transport like Kafka, OTEL) Formal training or certification on software engineering concepts and expert applied experience. In addition, advanced … Cycle Solid understanding of agile methodologies such as CI/CD, Application Resiliency, and Security Experience with one or more observability platforms such as Prometheus, Grafana, Dynatrace, Datadog, Splunk etc. About Us J.P. Morgan is a global leader in financial services, providing strategic advice and products to the world ...