Site Reliability Engineer, Studios
- Location
- Greater London, England, United Kingdom
rapid diagnosis, mitigation, communication, and post-incident follow-up. Drive root cause analysis and corrective actions following incidents, with a focus on prevention and continuous improvement. Support the design, testing, and documentation of high availability, backup, failover, and disaster recovery arrangements. Help enforce security, access control, patching, and operational … languages. Experience working in high-availability, live production, or other business-critical operational environments. Strong troubleshooting skills, calm decision-making under pressure, and a continuous improvement mindset. Excellent communication and collaboration skills, including the ability to work effectively with technical and non-technical stakeholders. Desirable Experience Supporting ...