1 to 25 of 108 Site Reliability Engineer Jobs in London

Senior Specialist Engineer (Specialist Site Reliability Engineer SRE)

Hiring Organisation
UK Health Security Agency
Location
Birmingham, Chilton, Leeds, Liverpool, London, Porton, E14 4PU, United Kingdom
Salary
£41983.00 to £52113.00
summary An SRE engineer will apply engineering principles to remediate infrastructure and operational problems. The primary focus will be on automation and CI/CD; ensuring our services run reliably, are scalable, and perform optimally in production environments. The role will monitor and manage these aspects while taking responsibility … operational service improvements and performance improvements to meet and exceed SLOs (Service Level Objectives). Main duties of the job Working with the HPC & SRE Team to: Ensure services are stable, scalable, performant and automated Respond to incidents, troubleshooting issues, and restoring services as quickly as possible Prioritise operational service ...

Site Reliability Engineer

Hiring Organisation
REVYBE IT RECRUITMENT LIMITED
Location
City of London, London, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£85,000
period of growth and investing heavily in its engineering and platform capabilities. They're looking for an experienced Site Reliability Engineer (SRE) to join the team and play a key role in building highly reliable, scalable, and observable infrastructure. This is a hands-on role focused … experience Help improve platform resilience, scalability, and disaster recovery capabilities Contribute to capacity planning and performance optimisation as the platform scales Establish and champion SRE best practices across the wider engineering function What We're Looking For Proven commercial experience working as an SRE, DevOps Engineer, Platform Engineer ...

Principal Site Reliability Engineer (we have office locations in Cambridge, Leeds and London)

Location
Greater London, England, United Kingdom
Reliability Engineer As Principal Site Reliability Engineer, you will help establish and grow Genomics England's organisation-wide SRE capability. You will be responsible for: Building and leading a small team of Site Reliability Engineers (initially 3 people) Identifying the highest-value … operate reliable services Developing services and capabilities that reduce operational toil and improve engineering effectiveness Partnering with product and engineering teams to embed SRE principles into their ways of working Influencing engineering strategy, governance and technical direction across the organisation Helping to build a culture of reliability, continuous improvement ...

Azure Infrastructure / SRE Engineer

Hiring Organisation
Teksystems
Location
Central London, London, United Kingdom
Employment Type
Contract
Azure tenant/infrastructure. Required Skills The successful candidate will be a student of the Google Site Reliability Engineering (SRE) philosophy as applied to managing large-scale cloud infrastructure, possess skills and xp within one or more of the following areas, and demonstrate a willingness to learn additional … skills via certification and/or on-the-job learning where required. Programming, Software & Network Principles xp with SRE and Azure DevOps Ability to script (Bash/PowerShell, Azure CLI), code (Python, C#, Java), query (SQL, Kusto query language) coupled with xp with software versioning control systems (e.g., GitHub ...

Senior Site Reliability Engineer

Location
Greater London, England, United Kingdom
Site Reliability EngineerApplylocations: London (82)time type: Full timeposted on: Posted Todayjob requisition id: JR101516**Senior Site Reliability Engineer (SRE) - GCP/Kubernetes****About the Role**We are seeking an experienced and highly motivated Senior Site Reliability Engineer (SRE) to join … small, agile engineering team. This role offers the unique opportunity to drive the reliability, scalability, and performance of our core platform with a high degree of autonomy and ownership.The successful candidate will split their time between providing expert operational support for our critical systems and leading exciting new infrastructure ...

Site Reliability Engineering Lead

Location
City Of London, England, United Kingdom
lead a team of SREs responsible for ensuring the reliability, scalability, security, performance, and operational excellence of mission-critical applications and platforms. The SRE Lead will partner closely with engineering, architecture, security, operations, and business stakeholders to drive cloud modernization, operational maturity, observability excellence, automation, and continuous improvement. This … mentor a team of Site Reliability Engineers, fostering a culture of ownership, operational excellence, collaboration, and continuous learning. Define and drive SRE strategy, standards, best practices, and operational frameworks across engineering organizations. Partner with product and platform teams to improve application reliability, scalability, security, performance, and resilience. ...

Senior Site Reliability Engineer

Location
Greater London, England, United Kingdom
Senior Site Reliability Engineer (SRE) - GCP/Kubernetes We are seeking an experienced and highly motivated Senior Site Reliability Engineer (SRE) to join our small, agile engineering team. This role offers the unique opportunity to drive the reliability, scalability, and performance … Kubernetes application deployment. Monitoring & Observability: Implement and manage robust monitoring, alerting, and logging solutions to ensure clear system visibility and proactive issue identification. Reliability & Performance: Define, measure, and enforce Service Level Objectives (SLOs) and Service Level Indicators (SLIs). Participate in on-call rotation (if applicable) and lead post ...

Site Reliability Engineer – 11863CF

Location
Greater London, England, United Kingdom
11863CF £500 – 580 per day Site Reliability Engineer – DV Cleared Our client is urgently looking for an experienced Site Reliability Engineer to join their team on a contract basis, initially for 6 months with a view to extend. Please note, the role is OUTSIDE … IR35. You must hold live UKIC DV. The role is on-site 4 days per week in Gloucester. Site Reliability Engineer – Key Skills: Must hold a live UKIC DV Modern configuration management tools (such as Ansible, Chef or similar) Experience working with Terraform Docker containers & container ...

Engineer - Site Reliability

Location
Greater London, England, United Kingdom
early‐session coverage from London that ensures continuous, high‐availability operations across Cboe's real‐time low‐latency trading platforms. The London‐based SRE provides technical support to Cboe Trade Desk and Operations Support Center staff across time zones, and works closely with Software Engineering, Systems Engineering, and Network Engineering … spoken — is required. This role demands clear, precise, and unambiguous communication at all times. As the operational bridge between Cboe's European and APAC SRE teams and its US‐based leadership, the ability to communicate with clarity across time zones, cultures, and technical disciplines is fundamental to the success ...

Site Reliability Software Engineer (Hybrid)

Location
Greater London, England, United Kingdom
shapes without surgery. With over 600,000+ successful outcomes, EarWell® is a proven, non-invasive treatment option for families. We are looking for a Site Reliability Engineer to join our growing team. The ideal candidate has a strong technical background in software development and systems operations, with … healthcare security, privacy, and compliance requirements in mind. Document architecture, workflows, troubleshooting steps, deployment processes, and support procedures. Required Qualifications Experience as a Software Engineer, Site Reliability Engineer, DevOps Engineer, Systems Engineer, or similar technical role. Experience supporting production applications or infrastructure ...

Site Reliability Engineer, Studios

Location
Uxbridge, England, United Kingdom
rotations, to support live operations and critical systems. Occasional travel may be required depending on project and client needs. IMG is looking for a Site Reliability Engineer to help design, build, operate, and continuously improve resilient, secure, and highly available platforms that underpin our digital, cloud … services. This role is suited to someone who combines strong infrastructure and software engineering capability with an operational mindset, and who can help embed reliability engineering practices across systems that support live, business‐critical environments. The successful candidate will play a key role in improving service reliability, observability ...

Site Reliability Engineer, Studios

Location
Greater London, England, United Kingdom
rotations, to support live operations and critical systems. Occasional travel may be required depending on project and client needs. IMG is looking for a Site Reliability Engineer to help design, build, operate, and continuously improve resilient, secure, and highly available platforms that underpin our digital, cloud … services. This role is suited to someone who combines strong infrastructure and software engineering capability with an operational mindset, and who can help embed reliability engineering practices across systems that support live, business-critical environments. The successful candidate will play a key role in improving service reliability, observability ...

Systems Engineering Manager, Site Reliability Engineering, ML Compute

Location
Greater London, England, United Kingdom
mathematics). Track record of mentoring technical leads. Proven success leading and influencing multiple technical teams. About the job Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed, fault-tolerant systems. SRE ensures that Google's services—both our internally … critical and our externally-visible systems—have reliability, uptime appropriate to users' needs and a fast rate of improvement. Additionally SRE’s will keep an ever-watchful eye on our systems capacity and performance. Much of our software development focuses on optimizing existing systems, building infrastructure and eliminating work ...

Site Reliability Engineer (Chinese speaking, £90k, Financial Services,Technology, London)

Location
Greater London, England, United Kingdom
have an exciting opportunity for a Site Reliability Engineer role in a Singapore‐based financial services company. The company is a fast‐growing global trading platform looking for a Site Reliability Engineer to bridge development and operations using software engineering. Job Requirements A bachelor … degree in computer science or a closely related field is required. Three or more years of DevOps, Site Reliability, and Infrastructure Support experience. Linux technical abilities (CentOS preferred) including shell scripting, Docker, and Ansible. Familiarity with Python (preferred) and/or Go or Ruby. Experience with cloud infrastructure ...

Site Reliability Engineer

Location
Greater London, England, United Kingdom
Site Reliability Engineer Reports to: Labs Team Lead Looper Insights Remote-first, with regular visits to our Byfleet and Hounslow data centres The company Looper Insights builds analytics products that help the world’s leading media and entertainment companies understand how their content is performing across digital … collaborate closely to produce a valuable service for an industry about which we are all passionate. The role We’re looking for a Site Reliability Engineer to keep the global LooperBox fleet running, the physical backbone behind every piece of data Looper Insights produces. LooperBoxes ...

Software Engineer III, Site Reliability Engineering, GCE AI

Location
Greater London, England, United Kingdom
Computer Science or Engineering. 2 years of experience designing, analyzing, and troubleshooting large-scale distributed systems. About the job Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed, fault-tolerant systems. SRE ensures that Google Cloud's services—both … internally critical and our externally-visible systems—have reliability, uptime appropriate to customer's needs and a fast rate of improvement. Additionally SRE’s will keep an ever-watchful eye on our systems capacity and performance. Much of our software development focuses on optimizing existing systems, building infrastructure ...

Vice President, Site Reliability Engineering

Hiring Organisation
Hackajob Ltd
Location
South West London, London, United Kingdom
Employment Type
Permanent
Were seeking a future team member for the role of Vice President - Site Reliability Engineer to join our team. This role is located in London. Role Summary BNY is seeking a Vice President - Site Reliability Engineer to design, build, deploy, and scale resilient, automated … driven automation, and modern software delivery practices. Experience supporting distributed systems, cloud-native platforms, or container-based architectures. Knowledge of Agile, DevOps, and SRE operating models, including continuous improvement and blameless post-incident practices. Ability to influence engineering standards and drive adoption of common tooling and automation patterns across teams. ...

Jobshare - Sr Lead Software Engineer - Site Reliability Engineer, Python & Infrastructure management - Part time/Jobshare

Location
Greater London, England, United Kingdom
domains, and advise others on the technical and business issues facing them. You will will set the vision, strategy, and operating model for our SRE transformation - enabling our business-aligned support teams to deliver higher reliability, stronger resilience, and a measurably better end-user experience across the board. … responsibilities Defines the SRE vision, north-star outcomes, and multi-year roadmap for the Production Management team, aligned to both CIB and JPM Global Technology priorities. Establishes the SRE operating model across global regions (ways of working, intake, prioritization, engagement with engineering teams and production support). Partners with business ...

Senior SRE Engineer - ASE Traffic & Secure Services Network

Location
Greater London, England, United Kingdom
Senior SRE Engineer - ASE Traffic & Secure Services Network Shanghai, Shanghai, China Software and Services At Apple, we build systems that power services used by hundreds of millions of people around the world, and every second counts. The Services Engineering organization is at the heart of this mission, ensuring … platforms are performant, secure, and always available. We're seeking a technically strong Site Reliability Engineer (SRE) to join our growing London team, focused on the future of traffic management, load balancing, and secure networking infrastructure.You’ll play a key role in shaping the next generation ...

Senior Network Site Reliability Engineer

Location
Greater London, England, United Kingdom
business-critical collaboration platform used by companies around the world. As our product and infrastructure scale, we are looking for a Senior Network Site Reliability Engineer to help strengthen the reliability, availability, and scalability of our production environment. In this role, you will focus on cloud … infrastructure What you’ll need 8+ years of professional experience in infrastructure, reliability, networking, or software engineering 6+ years of experience as an SRE, DevOps Engineer, Network Engineer, Software Engineer, or similar Hands-on experience with AWS infrastructure, including EC2, VPC, ALB, S3, Route ...

Lead Site Reliability Engineer

Hiring Organisation
Inspire People
Location
City of London, London, United Kingdom
Employment Type
Permanent, Part Time, Work From Home
Salary
£80,000
support economic growth across the UK. The Department for Business, Innovation, Science and Trade (BIST), in partnership with Inspire People, is seeking a Senior SRE Squad Lead with experience leading and developing engineers, strong DevOps and Site Reliability Engineering expertise, cloud platform experience, infrastructure-as-code capability … UK. BIST's Digital, Data and Technology (DDaT) directorate develops and operates the tools and services that enable this mission. As a Senior SRE Squad Lead, you will play a key role in leading engineers while remaining hands-on in the design, delivery and continuous improvement of reliable, secure ...

Principal Site Reliability Engineer

Location
Greater London, England, United Kingdom
driven workflows, trusted data, and seamless collaboration, to deliver the insight and context needed for confident, competitive decision-making. The Opportunity As a Principal Site Reliability Engineer at Veson Nautical, you will design, build, monitor, and support the cloud infrastructure that underpins our rapidly growing SaaS platform. … observable, and easier to operate. You will have significant influence over the architectural direction of the platform. The Team You'll join a global Site Reliability Engineering team with members in the United States and the United Kingdom. This is a senior individual contributor role without direct reports ...

Principal Site Reliability Engineer, Infrastructure Observability

Location
Greater London, England, United Kingdom
toolchain and systems, code build and deployment, incident response, and 24x7 monitoring and support. The candidate will also have extensive experience operating within a SRE function within a complex, distributed environment. They will have a demonstrated ability to work horizontally and vertically within an organization with diverse partners and sponsor … learning through blameless post-mortems to improve the shared goal of reliability across services Transform operations teams by facilitating internal change to adopt SRE standard methodologies across the organization and driving strategic growth in this area within Global Technology Analyzes incidents impacting technology availability for high-level trends across ...

Site Reliability Engineer - Service Assurance Systems

Location
Greater London, England, United Kingdom
enrich raw data, making it readily consumable by our stakeholders across operations, engineering and management. As a Site Reliability Engineer (SRE), you will play a key role in bridging the gap between software development and operational reliability. You will be responsible for supporting the applications built ...

Vice President, Site Reliability Engineering

Hiring Organisation
Hackajob Ltd
Location
London, United Kingdom
Employment Type
Permanent
Salary
GBP Annual
Were seeking a future team member for the role of Vice President - Site Reliability Engineer to join our team. This role is located in London. Role Summary BNY is seeking a Vice President - Site Reliability Engineer to design, build, deploy, and scale resilient, automated ...