8 of 8 Prometheus Jobs in the City of London

Principal Engineer I — Prepurchase Platform

Location
City Of London, England, United Kingdom
tools and techniques - using LLMs, AI-assisted development, and automation to accelerate engineering workflows and improve system intelligence. Familiarity with observability tooling: Grafana, Splunk, Prometheus,OpenTracing, or equivalent. Strong understanding of security best practices - OAuth/OIDC, input validation,secretsmanagement. Track recordof leading technical initiatives across multiple teams without direct ...

Senior Linux DevOps Engineer

Hiring Organisation
RedTech Recruitment Ltd
Location
City of London, London, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£90,000
performance troubleshooting within Linux environments Hands-on experience operating containerised workloads using Docker and Kubernetes Experience with monitoring, logging and observability technologies such as Prometheus, Grafana, Loki, OpenTelemetry or the ELK Stack Experience with Infrastructure as Code and automation tooling such as Terraform and Ansible Experience building and managing … Engineer/Linux/Bash/Shell Scripting/Python/Kubernetes/Docker/Terraform/Ansible/Microsoft Azure/Azure/Prometheus/Grafana/Loki/OpenTelemetry/ELK Stack/GitLab CI/CD/GitHub Actions/Jenkins/ArgoCD/Helm/ ...

Senior Forward Deployment Engineer

Hiring Organisation
Luxoft
Location
City of London, London, United Kingdom
controlled failover, and production-recovery exercises, highlighting skills in system reliability and continuity planning. Experience with enterprise observability tools such as Splunk, ELK, Grafana, Prometheus, OpenTelemetry, AppDynamics, or Dynatrace, reflecting proficiency in monitoring and diagnostics. Experience modernizing monolithic or legacy enterprise applications into maintain ...

Site Reliability Engineer

Hiring Organisation
SR2 | Socially Responsible Recruitment | Certified B Corporation™
Location
City of London, London, United Kingdom
demo environments Automate infrastructure provisioning and deployment workflows (Terraform, GitHub Actions, GitOps) Package and deploy applications to customer environments Implement and optimise observability tooling (Prometheus, Grafana, Loki) Support incident response, monitoring, and backup/recovery planning Mentor project teams in DevSecOps practices and environment management Ensure cloud environments are secure … cost-optimised Tech Environment & Skills: Cloud Engineering: AWS/Azure/GCP, Linux, Terraform (IaC) Containers: Kubernetes, Docker, Helm (OpenShift a plus) Observability: Prometheus, Grafana, Loki (network visualisation desirable) CI/CD & GitOps: GitHub Actions, ArgoCD/Flux Security: Cloud access models, Zero Trust principles Programming: Python or Golang preferred ...

Software engineering specialist

Location
City Of London, England, United Kingdom
deployments, and tool setups directly into developer workflows. Tech Stack Delivery: Automate the provisioning and configuration of our operational ecosystem, including tools across observability (Prometheus, Elastic, Checkmk), security/compliance (Tenable, Red Hat Satellite), service registry/IPAM (NetBox), container orchestration (ArgoCD), and artifact management (Artifactory, GitLab). Engineering Standards ...

Platform Engineer AWS IaC

Hiring Organisation
Client Server
Location
City of London, London, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£75,000
seeking to improve operational efficiency through scripting, tooling and automation. You'll monitor and optimise AWS environments for performance, reliability and cost using CloudWatch, Prometheus, Grafana and other observability tools, implement AWS security best practice, including IAM, roles, policies and least privilege access. As a senior member of the team ...

Site Reliability Engineer

Location
City Of London, England, United Kingdom
with Kubernetes, containerisation technologies, and cloud infrastructure. Strong understanding of networking, distributed systems, and infrastructure automation. Experience with monitoring and observability tools such as Prometheus, Grafana, Splunk, or similar. Proven track record of solving complex reliability, scalability, or performance challenges. Excellent problem-solving skills and ability to operate effectively ...

Senior Site Reliability Engineer

Location
City Of London, England, United Kingdom
speed and reducing deployment risk. Adaptable & Problem-Solver : Address complex challenges across configuration, policy, observability, and data services. Apply a data-driven approach using Prometheus and Grafana to improve reliability and performance. Ownership & Quality : Own end-to-end configuration quality, enforcing governance with Open Policy Agent. Ensure secure, compliant deployments ...