Site Reliability Engineer (Mid / Senior)
- Hiring Organisation
- Reed
- Location
- South West London, London, England, United Kingdom
- Employment Type
- Full-Time
- Salary
- Salary negotiable
strong focus on Ubuntu-based systems . The Role As an SRE, you will play a key role in ensuring the availability, performance, security and resilience of production systems. Working in a small, collaborative team, you’ll take ownership of day-to-day platform operations, incident response and continuous … Python Monitor system health using tools such as Prometheus, Grafana, Zabbix or Nagios Investigate and resolve production incidents (on-call rota involved) Implement security hardening and infrastructure best practices Manage backup and disaster recovery processes and regular testing Support and improve CI/CD pipelines and deployment processes ...