1 to 25 of 50 Site Reliability Engineering Jobs in Central London

Systems Engineering Manager, Site Reliability Engineering, ML Compute

Location
City of Westminster, England, United Kingdom
. Track record of mentoring technical leads. Proven success leading and influencing multiple technical teams. About the job Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed, fault-tolerant systems. SRE ensures that Google's services—both … internally critical and our externally-visible systems—have reliability, uptime appropriate to users' needs and a fast rate of improvement. Additionally SRE’s will keep an ever-watchful eye on our systems capacity and performance. Much of our software development focuses on optimizing existing systems, building infrastructure and eliminating ...

Software Engineer III, Site Reliability Engineering, GCE AI

Location
City of Westminster, England, United Kingdom
Science or Engineering. 2 years of experience designing, analyzing, and troubleshooting large-scale distributed systems. About the job Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed, fault-tolerant systems. SRE ensures that Google Cloud's services—both … internally critical and our externally-visible systems—have reliability, uptime appropriate to customer's needs and a fast rate of improvement. Additionally SRE’s will keep an ever-watchful eye on our systems capacity and performance. Much of our software development focuses on optimizing existing systems, building infrastructure ...

Lead Site Reliability Engineer

Hiring Organisation
Inspire People
Location
City of London, London, United Kingdom
Employment Type
Permanent, Part Time, Work From Home
Salary
£80,000
support economic growth across the UK. The Department for Business, Innovation, Science and Trade (BIST), in partnership with Inspire People, is seeking a Senior SRE Squad Lead with experience leading and developing engineers, strong DevOps and Site Reliability Engineering expertise, cloud platform experience, infrastructure-as-code capability … UK. BIST's Digital, Data and Technology (DDaT) directorate develops and operates the tools and services that enable this mission. As a Senior SRE Squad Lead, you will play a key role in leading engineers while remaining hands-on in the design, delivery and continuous improvement of reliable, secure ...

Site Reliability Engineer

Hiring Organisation
REVYBE IT RECRUITMENT LIMITED
Location
City of London, London, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£85,000
period of growth and investing heavily in its engineering and platform capabilities. They're looking for an experienced Site Reliability Engineer (SRE) to join the team and play a key role in building highly reliable, scalable, and observable infrastructure. This is a hands-on role focused … experience Help improve platform resilience, scalability, and disaster recovery capabilities Contribute to capacity planning and performance optimisation as the platform scales Establish and champion SRE best practices across the wider engineering function What We're Looking For Proven commercial experience working as an SRE, DevOps Engineer, Platform Engineer ...

Site Reliability Engineer

Hiring Organisation
SR2 | Socially Responsible Recruitment | Certified B Corporation™
Location
City of London, London, United Kingdom
Site Reliability Engineer (SRE) DevSecOps | Cloud Engineering | Observability | Production Environments | London SR2 is supporting a major 3-year programme and looking for an experienced Site Reliability Engineer (SRE) to join the Production Engineering team. This function underpins the reliability, security, and performance … likely) IR35: Inside Location: London twice a week (hybrid model) Clearance: SC level may be required depending on deployment If you’re an experienced SRE who thrives on building reliable, secure, and cost-efficient production systems, apply now for immediate consideration. ...

Production Engineering Manager

Location
City of Westminster, England, United Kingdom
Meta is seeking a Production Engineering Manager to lead a team responsible for the reliability, scalability, and operational excellence of Meta's production infrastructure and services. In this role, you will manage a team of production engineers who own the full lifecycle of systems — from capacity planning … performance optimization to incident response and automation. You will drive technical strategy, champion AI-augmented workflows, and partner closely with software engineering, infrastructure, and product teams to ensure Meta's services operate at global scale with high availability and efficiency.Production Engineering Manager Responsibilities:Manage a team of production ...

Senior DevSecOps Engineer

Location
City Of London, England, United Kingdom
operating the software delivery infrastructure required to develop and deploy advanced autonomous systems for defence applications. This role sits at the intersection of software engineering, platform engineering, cyber security, and defence systems engineering. The DevSecOps Engineer works alongside autonomy, software, systems, integration, and test engineers to create secure … delivery pipelines that enable teams to rapidly develop, integrate, test, and deploy mission critical software. The ideal candidate has a strong software and platform engineering background combined with significant experience operating within UK defence environments. They have a strong understanding of the UK Ministry of Defence/NATO approach ...

DevSecOps Engineer

Location
City of Westminster, England, United Kingdom
operating the software delivery infrastructure required to develop and deploy advanced autonomous systems for defence applications. This role sits at the intersection of software engineering, platform engineering, cyber security, and defence systems engineering. The DevSecOps Engineer works alongside autonomy, software, systems, integration, and test engineers to create secure … delivery pipelines that enable teams to rapidly develop, integrate, test, and deploy mission critical software. The ideal candidate has a strong software and platform engineering background combined with significant experience operating within UK defence environments. They have a strong understanding of the UK Ministry of Defence/NATO approach ...

Software Engineer Lead - Site Reliability

Location
City of Westminster, England, United Kingdom
strategic partners and third parties., As a Lead DevOps Engineer, you will be a hands-on contributor and technical lead focused on improving the engineering foundations of the Customer Digital Platform. You will work across internal squads and with strategic partners, third parties, service, architecture and security teams … GitHub/GitHub Actions, Azure DevOps, Terraform and automated testing. Improve deployment safety, release readiness and operational readiness for customer-facing digital services. Apply SRE principles pragmatically to improve availability, recoverability, monitoring and incident learning. Strengthen monitoring, logging, tracing, alerting and service-health dashboards across digitally connected workloads. Reduce single ...

Lead Site Reliability Engineer

Location
Westminster, West End, United Kingdom
trading technology stack is undergoing a multi year convergence and modernization journey. You will play a pivotal role in shaping our next generation SRE patterns, reliability frameworks, observability strategy, and performance engineering capabilities across globally distributed systems. This role is ideal for an SRE specialist who thrives … codebase (Java, Kotlin, Python) to implement reliability improvements, performance optimisations, bug fixes, and automation. Lead the design and rollout of modern SRE patterns across trading systems, including automated remediation, self healing workflows, and resilience engineering. Uses enterprise-authorized AI capabilities within the work environment to accelerate major-incident triage ...

Lead SRE - Chase UK

Location
Westminster, West End, United Kingdom
building the bank of the future from the ground up, offering you the chance to join us and make a significant impact. As a Site Reliability Engineer at JPMorgan Chase within the International Consumer Bank, you will play a crucial role in this initiative, dedicated to delivering … oriented and possess an interest in the financial sector and focus on addressing our customer needs. We work in teams focused on improving the reliability, resilience, observability, and operability of customer-facing digital banking services. We build automation, define measurable reliability practices, reduce operational friction, and partner with ...

Senior Principal Software Engineer Dev O

Location
City Of London, England, United Kingdom
reduce duplication, improve efficiency and close capability gaps. WHY JOIN THE TEAM Working across engineering teams and with architecture, service management, DevOps/SRE and security partners, the SPSE will combine hands‐on technical leadership with strategic influence. They will shape and deliver cross‐cutting improvements spanning operational maturity … directly with engineering teams on complex operational and security challenges; organise and lead work across engineering, product, architecture, service management, DevOps/SRE, security and other enabling functions; and create clear technical guidance, reference patterns and learning that help teams adopt better practices. YOUR SKILLS AND EXPERIENCE Significant ...

Full-Stack Engineer - ML Platform, Applied AI - Vice President

Location
City of Westminster, England, United Kingdom
Investment Banking, you will design and deliver production architectures for AI-powered products and services. You will work at the intersection of software engineering and applied research to translate innovative ideas into scalable, enterprise-grade solutions. You will collaborate closely with cloud and site reliability engineering … distributed, multi-threaded, and scalable applications Build, test, and deploy highly secure automated pipelines for cloud systems, desktop applications, and ML solutions Apply software engineering and computer science best practices Develop and deploy business-critical, data-intensive applications Leverage foundational libraries and services for reuse across teams Utilize MLOps ...

Product Associate - SRE Team - Chase UK

Location
Westminster, West End, United Kingdom
oriented and possess an interest in the financial sector and focus on addressing our customer needs. We work in teams focused on improving the reliability, resilience, observability, and operability of customer-facing digital banking services. We build automation, define measurable reliability practices, reduce operational friction, and partner with … engineering teams to ensure services are designed, delivered, and operated with reliability in mind. Job responsibilities Support the product strategy and delivery of reliability capabilities, including standards, observability, incident practices, automation, and developer experience improvements. Partner with engineers, site reliability engineers, and cross-functional teams ...

DevOps Engineer

Location
City Of London, England, United Kingdom
London (4-5 days a week in office) This is an opportunity for a DevOps Engineer to work at the intersection of infrastructure engineering and AI technology within a high-performance environment. You will play a key role in building and scaling modern infrastructure platforms, with a particular focus … premise environments that support business-critical workloads. THE COMPANY They are a globally operating investment and technology-driven organisation with a strong engineering culture. Their teams work closely with technical and business stakeholders to deliver robust, scalable infrastructure across a complex environment. This role offers exposure to cutting-edge ...

Senior Linux DevOps Engineer

Hiring Organisation
RedTech Recruitment Ltd
Location
City of London, London, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£90,000
annum + excellent benefits Requirements for Senior Linux DevOps Engineer: Strong commercial experience working as a Senior DevOps Engineer, Linux Engineer, Platform Engineer, Site Reliability Engineer or similar Excellent Linux systems administration and command line skills, with experience operating and troubleshooting large-scale production environments Strong scripting … Linux Engineer/Linux Systems Engineer/Linux Infrastructure Engineer/Senior Platform Engineer/Platform Engineer/Site Reliability Engineer/SRE/Infrastructure Engineer/DevSecOps Engineer/Linux/Bash/Shell Scripting/Python/Kubernetes/Docker/Terraform/Ansible/Microsoft ...

Production Engineer

Location
City Of London, England, United Kingdom
issues across our trading platform. You will leverage deep expertise in FIX, Linux, Windows Server, DevOps, databases, networking, and cloud technologies to ensure platform reliability and performance.This is a hands-on leadership role involving complex troubleshooting across cross-platform market-leading technologies, driving automation and tooling improvements, and acting … working hoursPositive approach to the day-to-day, with the resilience to handle high-pressure production incidentsDesiredExperience with Site Reliability Engineering (SRE) practices, including monitoring, incident response, and post-mortem analysisProven experience applying AI or machine-learning models to optimise workflows, identify patterns, and drive intelligent automation ...

Production Engineer

Location
City Of London, England, United Kingdom
issues across our trading platform. You will leverage deep expertise in FIX, Linux, Windows Server, DevOps, databases, networking, and cloud technologies to ensure platform reliability and performance. This is a hands-on leadership role involving complex troubleshooting across cross-platform market-leading technologies, driving automation and tooling improvements … approach to the day-to-day, with the resilience to handle high-pressures production incidents Desired Experience with Site Reliability Engineering (SRE) practices, including monitoring, incident response, and post-mortem analysis Proven experience applying AI or machine-learning models to optimise workflows, identify patterns, and drive intelligent ...

Site Reliability Engineer, Infrastructure - ThousandEyes

Location
City Of London, England, United Kingdom
deeply integrated across the Cisco technology portfolio, delivering AI-powered assurance insights within Cisco’s Networking, Security, Collaboration, and Observability portfolios. Our distributed Site Reliability Engineering team of approximately nine engineers owns the availability, latency, performance, efficiency, monitoring, emergency response, and capacity planning of the platform while … operational on-call rotation. Hands-on experience with infrastructure-as-code tooling and codebases, preferably Terraform. Hands-on experienceleveraging AIas a force multiplier of SRE activities, such as automati ng toil away and improving operational efficiency. Professional experience administering and troubleshooting GNU/Linux systems, including system libraries, file systems ...

Scala Engineer

Location
City Of London, England, United Kingdom
Experience working within Continuous Integration environments Strong understanding of Agile methodologies Experience with testing and automation Awareness of Site Reliability Engineering (SRE) principles and support Experience troubleshooting incidents and restoring services following outages Experience working in a you build it, you run it environment Strong collaborative ...

Scala Engineer

Location
City Of London, England, United Kingdom
services. Support incremental re-architecting initiatives to reduce technical complexity and improve maintainability. Develop clean, testable and maintainable code using Scala and modern engineering practices. Design, build and maintain secure APIs, databases and applications. Collaborate with Product Owners, Business Analysts, Data Engineers and wider technical teams to deliver effective … design and development experience. Experience working with databases and SQL. Hands-on AWS cloud experience. Understanding of Site Reliability Engineering (SRE) principles. Experience supporting and restoring production services during incidents. Strong appreciation of testing, automation and software quality practices. Experience working within Agile environments. Experience with Continuous ...

Site Reliability Engineer

Hiring Organisation
REVYBE IT RECRUITMENT LIMITED
Location
City, London, United Kingdom
Employment Type
Permanent
Salary
GBP 85,000 Annual
Site Reliability Engineer Up to £85,000 + Benefits Central London Hybrid (2/3 days a week in the office) Build, Scale & Improve the Reliability of a Fast-Growing SaaS Platform We're partnering with a fast-growing SaaS company that's going through an exciting … period of growth and investing heavily in its engineering and platform capabilities click apply for full job details ...

ML Compute SRE Lead: Scale, Uptime & Automation

Location
City of Westminster, England, United Kingdom
Google London is seeking a Systems Engineering Manager for Site Reliability Engineering in ML Compute. You will lead a multi-disciplinary team, own uptime, and shape reliability strategy for large-scale services. You will mentor engineers, drive end-to-end availability, and collaborate with cross ...

Senior SRE Engineer — Cloud Reliability & Automation

Location
City of Westminster, England, United Kingdom
Google London, UK is seeking a Software Engineer III in Site Reliability Engineering for the GCE AI team. This mid-level role focuses on building reliable, scalable systems, code development, and mentoring junior team members. The position emphasizes deep expertise in distributed systems, problem solving, and collaboration ...

AI & SRE Consultant

Hiring Organisation
Akkodis
Location
City of London, London, United Kingdom
Employment Type
Permanent
Salary
£88000 - £96000/annum
Platform & Site Reliability Engineering Senior Consultant Cloud Operating Model Transformation | AI, Cloud & Automation Ready to help organisations redefine how they operate in the age of AI? We are partnering with a leading global consulting organisation seeking a Senior Consultant to join a fast-growing Cloud Advisory practice. ...