Senior Site Reliability Engineer
Site Reliability Engineer (L3) – Windows Infrastructure
Location: London, UK
Corporate Title: Assistant Vice President / Vice President
Client Profile: Global Tier-1 Financial Institution
Position Type: Full-time, Permanent
Role Overview
On behalf of our client - a leading global investment bank and financial services institution - we are seeking an experienced L3 Site Reliability Engineer to join their London-based Infrastructure SRE team. In this role, you will play a key part in driving operational excellence, automation, and continuous service improvement across the firm's core banking and securities platforms. You will act as a senior technical escalation point, responsible for eliminating operational toil, optimizing low-latency Windows systems, and implementing robust Infrastructure as Code (IaC) solutions within a highly regulated enterprise environment.
Key Responsibilities
- Automation & IaC: Eliminate manual, repetitive tasks by developing software, self-service tools, and Infrastructure as Code using PowerShell, Python, Ansible, and Terraform.
- System Optimization & Reliability: Deep-dive into system internals to tune low-latency Windows infrastructure, manage resource distribution, and establish actionable Service Level Objectives (SLOs).
- L3 Escalation & RCA: Serve as an ultimate technical escalation point for production incidents, leading detailed Root Cause Analyses (RCAs) to build permanent remediations.
- Monitoring & Tooling: Build proactive monitoring solutions focused on early symptom detection rather than post-outage alerts.
- Collaboration & Security: Partner with application, development, and risk teams to execute safe release procedures, remediate technical debt, and ensure adherence to CIS security benchmarks and audit controls.
Essential Skills & Qualifications
- Windows Internals: Deep expertise in Windows Server internals, Active Directory, DNS, DHCP, Kerberos, and performance tuning for low-latency systems.
- Scripting & IaC: Strong hands-on experience with PowerShell, Python, or C#, alongside Infrastructure as Code frameworks (Ansible, Terraform) and version control (Git).
- CI/CD & DevOps: Proven background in CI/CD pipelines (Jenkins, TeamCity) and modern SRE methodologies.
Desirable Experience
- Container orchestration tools such as Docker and Kubernetes.
- Public/hybrid cloud platforms (Azure, AWS, or GCP).
- Enterprise database systems (SQL Server, Oracle) or virtualized platforms (Nutanix, VMware).
If this is of interest to you, do not hesitate and apply!