1 to 25 of 39 Site Reliability Engineering Jobs in Central London

Lead Site Reliability Engineer

Hiring Organisation
Inspire People
Location
City of London, London, United Kingdom
Employment Type
Permanent, Part Time, Work From Home
Salary
£80,000
support economic growth across the UK. The Department for Business, Innovation, Science and Trade (BIST), in partnership with Inspire People, is seeking a Senior SRE Squad Lead with experience leading and developing engineers, strong DevOps and Site Reliability Engineering expertise, cloud platform experience, infrastructure-as-code capability … UK. BIST's Digital, Data and Technology (DDaT) directorate develops and operates the tools and services that enable this mission. As a Senior SRE Squad Lead, you will play a key role in leading engineers while remaining hands-on in the design, delivery and continuous improvement of reliable, secure ...

Site Reliability Engineer

Hiring Organisation
REVYBE IT RECRUITMENT LIMITED
Location
City of London, London, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£85,000
period of growth and investing heavily in its engineering and platform capabilities. They're looking for an experienced Site Reliability Engineer (SRE) to join the team and play a key role in building highly reliable, scalable, and observable infrastructure. This is a hands-on role focused … experience Help improve platform resilience, scalability, and disaster recovery capabilities Contribute to capacity planning and performance optimisation as the platform scales Establish and champion SRE best practices across the wider engineering function What We're Looking For Proven commercial experience working as an SRE, DevOps Engineer, Platform Engineer ...

Production Engineering Manager

Location
City of Westminster, England, United Kingdom
Meta is seeking a Production Engineering Manager to lead a team responsible for the reliability, scalability, and operational excellence of Meta's production infrastructure and services. In this role, you will manage a team of production engineers who own the full lifecycle of systems — from capacity planning … performance optimization to incident response and automation. You will drive technical strategy, champion AI-augmented workflows, and partner closely with software engineering, infrastructure, and product teams to ensure Meta's services operate at global scale with high availability and efficiency.Production Engineering Manager Responsibilities:Manage a team of production ...

Senior DevSecOps Engineer

Location
City Of London, England, United Kingdom
operating the software delivery infrastructure required to develop and deploy advanced autonomous systems for defence applications. This role sits at the intersection of software engineering, platform engineering, cyber security, and defence systems engineering. The DevSecOps Engineer works alongside autonomy, software, systems, integration, and test engineers to create secure … delivery pipelines that enable teams to rapidly develop, integrate, test, and deploy mission critical software. The ideal candidate has a strong software and platform engineering background combined with significant experience operating within UK defence environments. They have a strong understanding of the UK Ministry of Defence/NATO approach ...

Lead Site Reliability Engineer

Location
Westminster, West End, United Kingdom
trading technology stack is undergoing a multi year convergence and modernization journey. You will play a pivotal role in shaping our next generation SRE patterns, reliability frameworks, observability strategy, and performance engineering capabilities across globally distributed systems. This role is ideal for an SRE specialist who thrives … codebase (Java, Kotlin, Python) to implement reliability improvements, performance optimisations, bug fixes, and automation. Lead the design and rollout of modern SRE patterns across trading systems, including automated remediation, self healing workflows, and resilience engineering. Uses enterprise-authorized AI capabilities within the work environment to accelerate major-incident triage ...

Lead SRE - Chase UK

Location
Westminster, West End, United Kingdom
building the bank of the future from the ground up, offering you the chance to join us and make a significant impact. As a Site Reliability Engineer at JPMorgan Chase within the International Consumer Bank, you will play a crucial role in this initiative, dedicated to delivering … oriented and possess an interest in the financial sector and focus on addressing our customer needs. We work in teams focused on improving the reliability, resilience, observability, and operability of customer-facing digital banking services. We build automation, define measurable reliability practices, reduce operational friction, and partner with ...

Senior Principal Software Engineer Dev O

Location
City Of London, England, United Kingdom
reduce duplication, improve efficiency and close capability gaps. WHY JOIN THE TEAM Working across engineering teams and with architecture, service management, DevOps/SRE and security partners, the SPSE will combine hands‐on technical leadership with strategic influence. They will shape and deliver cross‐cutting improvements spanning operational maturity … directly with engineering teams on complex operational and security challenges; organise and lead work across engineering, product, architecture, service management, DevOps/SRE, security and other enabling functions; and create clear technical guidance, reference patterns and learning that help teams adopt better practices. YOUR SKILLS AND EXPERIENCE Significant ...

Platform Engineer

Location
City Of London, England, United Kingdom
Platform Engineer Department: Technology Employment Type: Permanent - Full Time Location: London Reporting To: Segun Ikuesan Description This is a hands‐on engineering role within the Platform Engineering team, which forms part of Technology Operations. Platform Engineering is responsible for building and operating the infrastructure, platforms and developer … tooling that enable our engineering and quantitative research teams to deliver software reliably, securely and at scale. The role will contribute to the design, build, automation and operation of a hybrid production platform across AWS and on‐premises environments, with a particular focus on the HashiCorp platform, including Nomad ...

Product Associate - SRE Team - Chase UK

Location
Westminster, West End, United Kingdom
oriented and possess an interest in the financial sector and focus on addressing our customer needs. We work in teams focused on improving the reliability, resilience, observability, and operability of customer-facing digital banking services. We build automation, define measurable reliability practices, reduce operational friction, and partner with … engineering teams to ensure services are designed, delivered, and operated with reliability in mind. Job responsibilities Support the product strategy and delivery of reliability capabilities, including standards, observability, incident practices, automation, and developer experience improvements. Partner with engineers, site reliability engineers, and cross-functional teams ...

Senior Linux DevOps Engineer

Hiring Organisation
RedTech Recruitment Ltd
Location
City of London, London, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£90,000
annum + excellent benefits Requirements for Senior Linux DevOps Engineer: Strong commercial experience working as a Senior DevOps Engineer, Linux Engineer, Platform Engineer, Site Reliability Engineer or similar Excellent Linux systems administration and command line skills, with experience operating and troubleshooting large-scale production environments Strong scripting … Linux Engineer/Linux Systems Engineer/Linux Infrastructure Engineer/Senior Platform Engineer/Platform Engineer/Site Reliability Engineer/SRE/Infrastructure Engineer/DevSecOps Engineer/Linux/Bash/Shell Scripting/Python/Kubernetes/Docker/Terraform/Ansible/Microsoft ...

Production Engineer

Location
City Of London, England, United Kingdom
issues across our trading platform. You will leverage deep expertise in FIX, Linux, Windows Server, DevOps, databases, networking, and cloud technologies to ensure platform reliability and performance. This is a hands-on leadership role involving complex troubleshooting across cross-platform market-leading technologies, driving automation and tooling improvements … approach to the day-to-day, with the resilience to handle high-pressures production incidents Desired Experience with Site Reliability Engineering (SRE) practices, including monitoring, incident response, and post-mortem analysis Proven experience applying AI or machine-learning models to optimise workflows, identify patterns, and drive intelligent ...

Site Reliability Engineer, Infrastructure - ThousandEyes

Location
City Of London, England, United Kingdom
deeply integrated across the Cisco technology portfolio, delivering AI-powered assurance insights within Cisco’s Networking, Security, Collaboration, and Observability portfolios. Our distributed Site Reliability Engineering team of approximately nine engineers owns the availability, latency, performance, efficiency, monitoring, emergency response, and capacity planning of the platform while … operational on-call rotation. Hands-on experience with infrastructure-as-code tooling and codebases, preferably Terraform. Hands-on experienceleveraging AIas a force multiplier of SRE activities, such as automati ng toil away and improving operational efficiency. Professional experience administering and troubleshooting GNU/Linux systems, including system libraries, file systems ...

Scala Engineer

Location
City Of London, England, United Kingdom
Experience working within Continuous Integration environments Strong understanding of Agile methodologies Experience with testing and automation Awareness of Site Reliability Engineering (SRE) principles and support Experience troubleshooting incidents and restoring services following outages Experience working in a you build it, you run it environment Strong collaborative ...

Scala Engineer

Location
City Of London, England, United Kingdom
services. Support incremental re-architecting initiatives to reduce technical complexity and improve maintainability. Develop clean, testable and maintainable code using Scala and modern engineering practices. Design, build and maintain secure APIs, databases and applications. Collaborate with Product Owners, Business Analysts, Data Engineers and wider technical teams to deliver effective … design and development experience. Experience working with databases and SQL. Hands-on AWS cloud experience. Understanding of Site Reliability Engineering (SRE) principles. Experience supporting and restoring production services during incidents. Strong appreciation of testing, automation and software quality practices. Experience working within Agile environments. Experience with Continuous ...

Site Reliability Engineer

Hiring Organisation
REVYBE IT RECRUITMENT LIMITED
Location
City, London, United Kingdom
Employment Type
Permanent
Salary
GBP 85,000 Annual
Site Reliability Engineer Up to £85,000 + Benefits Central London Hybrid (2/3 days a week in the office) Build, Scale & Improve the Reliability of a Fast-Growing SaaS Platform We're partnering with a fast-growing SaaS company that's going through an exciting … period of growth and investing heavily in its engineering and platform capabilities click apply for full job details ...

ML Compute SRE Lead: Scale, Uptime & Automation

Location
City of Westminster, England, United Kingdom
Google London is seeking a Systems Engineering Manager for Site Reliability Engineering in ML Compute. You will lead a multi-disciplinary team, own uptime, and shape reliability strategy for large-scale services. You will mentor engineers, drive end-to-end availability, and collaborate with cross ...

Senior SRE Engineer — Cloud Reliability & Automation

Location
City of Westminster, England, United Kingdom
Google London, UK is seeking a Software Engineer III in Site Reliability Engineering for the GCE AI team. This mid-level role focuses on building reliable, scalable systems, code development, and mentoring junior team members. The position emphasizes deep expertise in distributed systems, problem solving, and collaboration ...

AI & SRE Consultant

Hiring Organisation
Akkodis
Location
City of London, London, United Kingdom
Employment Type
Permanent
Salary
£88000 - £96000/annum
Platform & Site Reliability Engineering Senior Consultant Cloud Operating Model Transformation | AI, Cloud & Automation Ready to help organisations redefine how they operate in the age of AI? We are partnering with a leading global consulting organisation seeking a Senior Consultant to join a fast-growing Cloud Advisory practice. ...

Head of Production Management- J.P. Morgan Personal Investing

Location
Westminster, West End, United Kingdom
powered solutions and intelligent automation to reduce manual intervention, fast-track resolution, and continuously improve operational efficiency. Champion an automation-first, shift-left SRE cultureleveraging shared tooling and automation to ensure consistency, reduce duplication, and maintain alignment with firmwide standards. Oversee capacity management and planning, ensuring infrastructure scales to meet … management standards change, incident, capacity, and automation across multiple engineering teams operating in a you-build-it-you-run-it model, underpinned by SRE principles and disaster recovery planning. Composure, decisiveness, and authority during incidents, vendor failure, or regulatory escalation, with a proven ability to protect business lines under ...

Fractional DevOps Engineer

Hiring Organisation
Elliot Marsh
Location
Central London, London, United Kingdom
Employment Type
Part Time
lead in shaping the companys Google Cloud Platform environment, deployment processes, security controls and operational practices. Working closely with the technical leadership and engineering team, the successful candidate will build robust foundations for secure software delivery, monitoring, resilience and future scale. The role will suit a pragmatic engineer … jobs, queues, webhooks and critical integrations, documenting procedures and transferring knowledge to the internal team Fractional DevOps Engineer You: - Strong commercial experience across DevOps, Site Reliability Engineering, cloud infrastructure and/or cloud security - Significant hands-on production experience with Google Cloud Platform and Terraform. - Excellent knowledge ...

Senior Scala Engineer CGEMJP00355784

Hiring Organisation
Experis
Location
West End, London, Stratford and New Town, United Kingdom
Employment Type
Contract
Contract Rate
£590 - £637/day
discipline rather than adhering to tightly defined roles. Knowledge & experience Security clearance (SC-level) API design Data analysis Databases AWS suite experience Awareness of Site Reliability Engineering and support Skilled at returning services to good states in outage situations Understands the importance of testing and automation Working ...

Principal Platform Engineer

Hiring Organisation
Sanderson Recruitment
Location
City of London, London, United Kingdom
Employment Type
Permanent
large-scale distributed systems and database platforms? We're looking for a hands-on technical leader to help shape the future of our platform engineering capability. This is an opportunity to lead complex engineering initiatives, define technical strategy, and act as a subject matter expert across AWS infrastructure … technical authority for distributed database and persistence technologies Required Experience 8+ years' experience in Platform Engineering, Infrastructure Engineering, DevOps, SRE or Software Engineering Expert-level AWS infrastructure experience Strong Infrastructure as Code expertise with Terraform Strong Linux systems administration and networking knowledge Experience designing and operating distributed ...

Staff Platform Engineer - AI Native SaaS Platform

Location
City Of London, England, United Kingdom
2025. Their platform is redefining the sector, and with revenues nearly 10x since the start of last year, they're continuing to expand their Engineering team to match the ambition of their product and customers. The product is real-time, data-rich and AI-native, creating complex engineering … infrastructure as code Strong understanding of networking, IAM, managed services and secure cloud architecture Experience owning observability, logging, APM or distributed tracing tooling An SRE mindset across SLOs, reliability, incident response and toil reduction A track record of technical leadership, mentoring and delivering complex projects through others Strong systems ...

Principal DevOps Engineer- SC Cleared

Location
City Of London, England, United Kingdom
Clearance or SC Clearance eligibility - London Based or ability to travel to London. - Experience as in a Senior/Principal DevOps/Platform engineering role PRINCIPAL DEVOPS ENGINEER ESSENTIAL SKILLS - Strong Linux expertise (RHEL/CentOS) with experience supporting production systems - Git for version control and collaborative development (GitHub … understanding of modern platform architecture, microservices and cloud-native systems - Good communication skills and ability to provide technical leadership. KEY SKILLS DEVOPS ENGINEER, DEVOPS, SITE RELIABILITY ENGINEER, PLATFROM ENGINEER, PLATFROM, CLOUD, AWS, AWS CLOUD, AWS ENGINEER, ANSIBLE, TERRAFORM, CLOUD SUPPORT, CLOUD INFRASTRUCTURE, DEFENCE, NATIONAL SECURITY, DV CLEARED ...

Principal Platform Engineer

Hiring Organisation
Searchability NS&D
Location
City of London, London, United Kingdom
consultancies, delivering large-scale cloud platforms for critical Public Sector programmes. You'll be the technical authority responsible for helping shape platform strategy, defining engineering standards, influencing architecture decisions and providing hands-on technical leadership across multiple delivery teams. We're looking for someone who enjoys solving difficult engineering … hands-on! Due to the sensitive nature of the work, SC Clearance eligibility is required. PRINCIPAL PLATFORM ENGINEER ESSENTIAL EXPERIENCE Principal or Senior Platform Engineering experience Large scale AWS environments Linux Kubernetes Git Terraform Cloud platform architecture Technical leadership across multiple teams Experience leading platform engineering within large ...