Site Reliability Engineer
- Hiring Organisation
- Hackajob Ltd
- Location
- Manchester, North West, United Kingdom
- Employment Type
- Permanent, Work From Home
delivery lifecycles. An understanding of SRE principles, including SLIs, SLOs, reliability measurement and incident management. Hands-on experience with observability tools such as OpenTelemetry, Splunk, New Relic, Grafana or PagerDuty. Proficiency in shell scripting for automation and system management. Experience with Infrastructure as Code, including Terraform and Ansible. Knowledge … consistency. Write and contribute to code, telemetry and instrumentation that improve service reliability and observability. Build dashboards and operational views using telemetry from Grafana, Splunk, New Relic and related platforms. Configure and manage Cloudflare edge services using Infrastructure as Code and integrate edge telemetry with observability platforms. Diagnose incidents ...