Site Reliability Engineer, Infrastructure - ThousandEyes
- Location
- City Of London, England, United Kingdom
Collaboration, and Observability portfolios. Our distributed Site Reliability Engineering team of approximately nine engineers owns the availability, latency, performance, efficiency, monitoring, emergency response, and capacity planning of the platform while partnering closely with application development teams. We believe in operations, infrastructure, and everything as code, creating a collaborative … related technical knowledge. Experience designing and implementing scalable, resilient, and well-tested distributed systems. Experience with Service Level Objectives, Service Level Agreements, monitoring, alerting, capacity planning, incident management, or disaster-recovery testing. Experience building automation that reduces repetitive work, improves release safety, or increases infrastructure efficiency. Strong communication ...