Site Reliability Engineer
- Hiring Organisation
- Hackajob Ltd
- Location
- Manchester, North West, United Kingdom
- Employment Type
- Permanent, Work From Home
/7 enterprise where uptime, performance and stability are critical. Practical experience using LLM platforms and coding assistants safely to improve productivity, quality and root-cause analysis. Additional Information Develop and maintain resilient tools, operational APIs and automation for effective system management. Use orchestration and scripting to remove … trace issues from the edge through to origin systems and coordinate effective remediation. Participate in live incident response, post-mortems and root-cause analysis to prevent recurrence. Maintain and administer monitoring, alerting, APM and analytics toolsets, including PagerDuty workflows. Drive initiatives that improve reliability, observability, performance ...