4 of 4 Permanent Grafana Jobs in Westminster

Security Platform Engineer, UK Security Operations

Location
Westminster, West End, United Kingdom
Experience with Kubernetes security, including workload isolation, Role-Based Access Control (RBAC), and network policies, containerisation, orchestration, and Kubernetes observability tools (e.g., Falco, Prometheus, Grafana). Experience with infrastructure-as-code and configuration management tools (e.g., Terraform, Helm, ArgoCD). Active, or the ability to obtain, a Developed Vetting ...

Principal Software Engineer - Platform Engineering - Accelerator Business

Location
Westminster, West End, United Kingdom
knowledge ofCI/CD, application resiliency, and secure delivery (e.g., SLSA framework and GitOps). Deep experience with Observability and Monitoring tools (e.g., Prometheus, Grafana, OTEL). Expertise in performance optimisation of distributed systems (e.g., caching, network latency). Practical experience with Service Mesh technologies (e.g., Istio, Linkerd, Cillium). ...

Lead Site Reliability Engineer

Location
Westminster, West End, United Kingdom
experience in front office trading environments or similarly high pressure, low latency domains. Proficiency with SRE tooling and techniques, including FIX messaging, Kafka, Grafana, Splunk, ITRS Geneos, Dynatrace, InfluxDB, MQ (IBM MQ or similar), Oracle DB Demonstrated experience using enterprise-authorized AI capabilities within the work environment to improve ...

Vice President, Site Reliability Engineering

Location
Westminster, West End, United Kingdom
Objectives, and service health measures aligned to operational and business priorities. Build and optimize monitoring, observability, and alerting capabilities using tools such as Prometheus, Grafana, AppDynamics, and Splunk. Apply AIOps capabilities to improve event correlation, anomaly detection, root cause analysis, predictive insights, and proactive issue prevention. Partner with engineering, infrastructure … Demonstrated ability to define and operationalize SLIs, SLOs, dashboards, alerts, and health indicators. Hands-on experience with enterprise monitoring and observability platforms including Prometheus, Grafana, AppDynamics, and Splunk. Strong troubleshooting, analytical, and problem-solving skills in complex distributed or production environments. Strong verbal and written communication skills, with the ability ...