5 of 5 Permanent Grafana Jobs in Hertfordshire

Site Reliability Engineer

Hiring Organisation
Jobleads-UK
Location
Watford, England, United Kingdom
incident reviews and implement remediation actions Maintain and improve runbooks and operational documentation Observability Implement and maintain monitoring using: + Splunk + CloudWatch + Grafana Improve: + Logging quality + Metrics coverage + Alerting accuracy Contribute to linking system performance to user experience signals Automation & engineering Develop scripts and tooling ...

Senior Site Reliability Engineer: Lead Resilience & Incidents

Hiring Organisation
Jobleads-UK
Location
Watford, England, United Kingdom
will own SLOs/SLIs, push automation with Terraform, mentor engineers, and coordinate incident response with a focus on observability using Splunk, CloudWatch, Grafana and Quantum Metric. #J-18808-Ljbffr ...

Performance and Monitoring Engineer

Hiring Organisation
Solus Accident Repair Centres
Location
Birchanger, Hertfordshire, United Kingdom
Employment Type
Permanent
Salary
GBP 40,000 - 50,000 Annual
both technical and non-technical teams Desirable qualifications Microsoft certifications (AZ-900, AZ-104, AZ-305, AZ-500) or similar Experience with LogicMonitor admin, Grafana or other observability tools Familiarity with SRE concepts (SLIs, SLOs, error budgets) Understanding of ITIL processes Who are Solus? Solus, who are owned by Aviva ...

Senior / Lead Site Reliability Engineer

Hiring Organisation
Jobleads-UK
Location
Watford, England, United Kingdom
review Ensure blameless post‐mortems with clear remediation ownership Observability & service insight Define and evolve observability strategy using: Splunk (log analytics) CloudWatch (AWS telemetry) Grafana (metrics visualisation) Quantum Metric (user behaviour insight) Standardise: Alerting quality and signal‐to‐noise ratio Dashboards aligned to SLOs and customer impact Drive correlation between … environments, ideally AWS (ECS, with exposure or experience in EKS/Kubernetes) Hands‐on experience with: Terraform (Infrastructure as Code) Observability stacks (Splunk, CloudWatch, Grafana) Strong programming skills (Python, Go, or similar) SRE practices Proven experience implementing: SLOs, SLIs, error budgets Incident management frameworks Observability strategies Strong experience in distributed ...

Senior / Lead Site Reliability Engineer

Hiring Organisation
Jobleads-UK
Location
Watford, England, United Kingdom
review Ensure blameless post-mortems with clear remediation ownership Observability & service insight Define and evolve observability strategy using: Splunk (log analytics) CloudWatch (AWS telemetry) Grafana (metrics visualisation) Quantum Metric (user behaviour insight) Standardise: Alerting quality and signal-to-noise ratio Dashboards aligned to SLOs and customer impact Drive correlation between … environments, ideally AWS (ECS, with exposure or experience in EKS/Kubernetes) Hands-on experience with: Terraform (Infrastructure as Code) Observability stacks (Splunk, CloudWatch, Grafana) Strong programming skills (Python, Go, or similar) SRE practices Proven experience implementing: SLOs, SLIs, error budgets Incident management frameworks Observability strategies Strong experience in distributed ...