Site Reliability Engineer, Studios
- Location
- Uxbridge, England, United Kingdom
systems that support live, business‐critical environments. The successful candidate will play a key role in improving service reliability, observability, incident response, automation, and disaster recovery readiness across IMG platforms, while working closely with engineering, operations, and project stakeholders. Key Responsibilities And Accountabilities Design, build, and maintain reliable … dashboards, and service health indicators. Define and maintain SLIs, SLOs, alerting standards, and operational runbooks for critical services. Automate infrastructure provisioning, configuration, deployment, and recovery processes using Infrastructure as Code and scripting. Partner with software, platform, broadcast engineering, and operational teams to improve release quality, resilience, and supportability. ...