176 to 200 of 518 Site Reliability Engineering Jobs in London

Site Reliability Engineer - Live Ops & Cloud Resilience

Location
Greater London, England, United Kingdom
seeking an experienced Site Reliability Engineer to design, build, and operate resilient, secure platforms underpinning our digital and live operations. You’ll focus on reliability, observability, automation, and disaster recovery across hybrid environments, collaborating with engineering, operations, and project stakeholders. The role emphasizes improving service availability ...

Head of Production Management- J.P. Morgan Personal Investing

Hiring Organisation
Hackajob Ltd
Location
South West London, London, United Kingdom
Employment Type
Permanent
powered solutions and intelligent automation to reduce manual intervention, fast-track resolution, and continuously improve operational efficiency. Champion an automation-first, shift-left SRE cultureleveraging shared tooling and automation to ensure consistency, reduce duplication, and maintain alignment with firmwide standards. Oversee capacity management and planning, ensuring infrastructure scales to meet … management standards change, incident, capacity, and automation across multiple engineering teams operating in a you-build-it-you-run-it model, underpinned by SRE principles and disaster recovery planning. Composure, decisiveness, and authority during incidents, vendor failure, or regulatory escalation, with a proven ability to protect business lines under ...

Head of Production Management- J.P. Morgan Personal Investing

Location
Westminster, West End, United Kingdom
powered solutions and intelligent automation to reduce manual intervention, fast-track resolution, and continuously improve operational efficiency. Champion an automation-first, shift-left SRE cultureleveraging shared tooling and automation to ensure consistency, reduce duplication, and maintain alignment with firmwide standards. Oversee capacity management and planning, ensuring infrastructure scales to meet … management standards change, incident, capacity, and automation across multiple engineering teams operating in a you-build-it-you-run-it model, underpinned by SRE principles and disaster recovery planning. Composure, decisiveness, and authority during incidents, vendor failure, or regulatory escalation, with a proven ability to protect business lines under ...

Site Reliability Engineer — Production & Incident Response

Location
Greater London, England, United Kingdom
慨正橡扯 is looking for a Site Reliability Engineer to manage incident response and oversee the offshore Production Support team … London. This hybrid position will require you to ensure system reliability and maintain high observability standards. Ideal candidates will bring significant experience in SRE, embracing automation tools to enhance performance. You will be instrumental in protecting client interests and contributing to the company during a vital stage of growth. ...

Staff Site Reliability Engineer

Location
Greater London, England, United Kingdom
that usable, and MystraAI is the agentic layer we are building on top of it. This is a Staff-level role that owns the reliability, performance, security and integrity of that infrastructure end-to-end — and sets the technical direction that other teams build on. You will lead … source level rather than as a black box — and ideally have contributed code upstream. Reliability engineering for data platforms. You bring true SRE discipline — SLOs, observability, capacity planning and incident response — to analytical data systems and pipelines. Data-as-a-Service productisation. You think in terms of data ...

Site Reliability Engineer — Scale, Automation & Observability

Location
Greater London, England, United Kingdom
Apple Inc. in London is seeking a Site Reliability Engineer to help manage and optimize scalable services that power Apple’s media and services. You will own the reliability of distributed systems across data centers, collaborating with software engineers to improve performance and resilience. The role requires ...

Market Data Engineer

Location
Greater London, England, United Kingdom
users of the market data platform and address evolving business needs Help shape the team’s technical direction by driving platform enhancements and new engineering initiatives Collaborate with hardware and software engineering teams across the firm to build real-time market data processing and distribution systems Optimize platform … performance through network and systems programming techniques designed to reduce latency and improve reliability Design and build tools that automate support and platform management, including monitoring, real-time and historical metrics, visualization, and self-service administrative capabilities Strengthen operational processes and workflows across release management, incident response, remediation, exchange ...

Network SRE

Hiring Organisation
Quant Capital
Location
London, UK
Employment Type
Full-time
Network SRE – 250,000-300,000 total compensation – 4 days in officeQuant Capital is urgently looking Network SRE for our high profile client. Our client is a leading quantitative trading company and liquidity provider. Their focus on technology has allowed them to deeply penetrate the market and gain market share. … Shared Engineering team that focuses on designing, developing, and maintaining infrastructure and tools. The team requires a Network Site Reliability Engineer (SRE) with strong network fundamentals, problem-solving skills, and a keen interest in diverse tools and techniques. The role involves collaborative work across various teams, exploring ...

Remote Principal SRE - Healthcare Platform Reliability Lead

Location
Greater London, England, United Kingdom
MediSolution in London is seeking a Site Reliability Engineer (SRE) to ensure the reliability of healthcare platforms. The candidate will lead efforts in automating operations and improving service availability. With a focus on troubleshooting and incident management, applicants should have 7+ years of experience in enterprise applications ...

AI Platform & Site Reliability Engineering Consultant/Senior Consultant - Cloud Operating Model

Location
Greater London, England, United Kingdom
address the entire breadth of their business needs. It delivers end-to-end services and solutions leveraging strengths from strategy and design to engineering, all fueled by its market leading capabilities in AI, cloud and data, combined with its deep industry expertise and partner ecosystem. The Group reported ...

Senior Cloud SRE - Kubernetes, GCP & CI/CD

Location
Greater London, England, United Kingdom
leading technology firm in Greater London is seeking a Senior Cloud Engineer to contribute technical expertise within a cloud engineering team. This role involves architecting scalable Kubernetes environments on Google Cloud Platform (GCP) and ensuring robust security measures. The ideal candidate will have extensive experience in DevOps or Site Reliability Engineering, deployment of production-grade Kubernetes clusters, and proficiency in CI/CD pipelines, alongside programming skills in Python, Go, and Bash. Join this dynamic team to drive innovative cloud solutions. #J-18808-Ljbffr ...

Observability SRE

Location
Greater London, England, United Kingdom
find your spark. Because that’s what drives you to be better, be more and ultimately, be more fulfilled. Job Title: Observability SRE Location: London, UK Employment Type: Fixed term contract (12 months duration) Job type: Onsite Job/Group Overview: SRE within the Group Platform Services & Engineering division … which provides the common services to Development, Infrastructure and Production Services. This is an SRE/support position responsible for administering and supporting Production environment as well as engineering reliability into the products/services we support i.e. monitoring & observability platform. The successful candidate will have a vital ...

Scala Engineer | SC Cleared | London | £620 pd

Hiring Organisation
Hays Technology
Location
London, United Kingdom
Employment Type
Contract
Contract Rate
£620/day £620 per day (inside IR35)
teams to develop effective solutions Contribute to database design and application development with strong security and data protection principles Support operational excellence through monitoring, reliability and incident resolution Work within a continuous integration environment and contribute to automation and testing practices Essential skills and experience Active SC clearance Strong … Agile delivery environments Experience within a "you build it, you run it" operating model Knowledge of data analysis and database technologies Awareness of Site Reliability Engineering principles Experience restoring services during outage situations Desirable skills and experience Docker and containerisation principles Jenkins Kibana Grafana Airflow Desired behaviours ...

Sofware Delivery Manager

Location
Greater London, England, United Kingdom
agile Scrum teams that support the organisation's broadcast, media, digital and operational technology capabilities. Working closely with Product Managers, Business Analysts, Development and Engineering Leads, QA teams and business stakeholders, the Software Delivery Manager ensures that cross-functional teams deliver high quality software solutions efficiently, predictably … Experience managing multiple software delivery teams simultaneously. Strong understanding of Agile, Scrum, Kanban and Lean delivery principles. Experience working closely with Product Management and Engineering teams. Excellent stakeholder management and communication skills. Proven ability to manage risks, dependencies and competing priorities. Experience with delivery tooling such as: o Jira ...

Sofware Delivery Manager

Hiring Organisation
UKTV
Location
Greater London, United Kingdom
Employment Type
Full Time
agile Scrum teams that support the organisation's broadcast, media, digital and operational technology capabilities. Working closely with Product Managers, Business Analysts, Development and Engineering Leads, QA teams and business stakeholders, the Software Delivery Manager ensures that cross-functional teams deliver high quality software solutions efficiently, predictably … Experience managing multiple software delivery teams simultaneously. Strong understanding of Agile, Scrum, Kanban and Lean delivery principles. Experience working closely with Product Management and Engineering teams. Excellent stakeholder management and communication skills. Proven ability to manage risks, dependencies and competing priorities. Experience with delivery tooling such as: o Jira ...

Site Reliability Engineer — Cloud & Live Ops (Hybrid)

Location
Uxbridge, England, United Kingdom
United Kingdom is seeking a Site Reliability Engineer to design, build, operate and continuously improve resilient, secure, and highly available platforms underpinning live, broadcast‐adjacent services. Based at Stockley Park in Uxbridge with hybrid options, the role involves improving observability, incident response, automation, disaster recovery, and collaborating with … engineering, operations and project stakeholders. #J-18808-Ljbffr ...

Senior Scala Engineer

Location
Greater London, England, United Kingdom
discipline rather than adhering to tightly defined roles. Knowledge & experience Security clearance (SC-level) API design Data analysis Databases AWS suite experience Awareness of Site Reliability Engineering and support Skilled at returning services to good states in outage situations Understands the importance of testing and automation Working ...

Senior Scala Engineer

Location
Greater London, England, United Kingdom
discipline rather than adhering to tightly defined roles. Knowledge & experience Security clearance (SC-level) API design Data analysis Databases AWS suite experience Awareness of Site Reliability Engineering and support Skilled at returning services to good states in outage situations Understands the importance of testing and automation Working ...

Senior Scala Engineer CGEMJP

Location
Greater London, England, United Kingdom
discipline rather than adhering to tightly defined roles. Knowledge & experience Security clearance (SC-level) API design Data analysis Databases AWS suite experience Awareness of Site Reliability Engineering and support Skilled at returning services to good states in outage situationsUnderstands the importance of testing and automation Working ...

Senior Scala Engineer CGEMJP

Location
City Of London, England, United Kingdom
discipline rather than adhering to tightly defined roles. Knowledge & experience Security clearance (SC-level) API design Data analysis Databases AWS suite experience Awareness of Site Reliability Engineering and support Skilled at returning services to good states in outage situations Understands the importance of testing and automation Working ...

Senior Scala Engineer CGEMJP00355784

Hiring Organisation
Experis
Location
West End, London, Stratford and New Town, United Kingdom
Employment Type
Contract
Contract Rate
£590 - £637/day
discipline rather than adhering to tightly defined roles. Knowledge & experience Security clearance (SC-level) API design Data analysis Databases AWS suite experience Awareness of Site Reliability Engineering and support Skilled at returning services to good states in outage situations Understands the importance of testing and automation Working ...

Hybrid Site Reliability Engineer — Healthcare Infra & Automation

Location
Greater London, England, United Kingdom
Cranial Technologies is seeking a Site Reliability Engineer to help build and maintain scalable, secure healthcare technology systems. The role focuses on automation, observability, and reducing manual operations to keep clinical and business processes reliable. You will support production apps, APIs, and integrations, while collaborating with IT, security ...

Cloud Operations Engineer

Hiring Organisation
Quant Capital
Location
London, UK
Employment Type
Full-time
Cloud Operations Engineer – FintechUp to 90,000 + 12% pension, bonuses and benefitsQuant Capital is urgently looking for a Site Reliability Engineer to join or well-known Fintech50 client who produces software disrupting the wealth management market. My client is a market leading SAAS provider to financial advisory … remediate with stable solutionsJenkinsLinux, CentOMS SQLFamiliarity with development languages, such as .NET, Java or PythonThis role suits a senior Engineer from a DevOps or SRE background who is real technologist and cloud specialist interested in the latest tooling and technologies that support software development and infrastructure. The firm ...

Solutions Architect, Studios

Location
Uxbridge, England, United Kingdom
procurement activity, with particular emphasis on pre sales engagement, proof of concept activity, and technical validation. This role will work across commercial, technical, operational, engineering, procurement, and transformation teams to translate business opportunities into clear solution options, technical direction, and delivery‐ready architecture. A core part of the role … conversations, and leading or coordinating proof of concepts that demonstrate feasibility, value, and operational fit. The role will also work closely with DevOps and Site Reliability Engineering teams to ensure proposed solutions are automatable, observable, supportable, resilient, and aligned to modern operational practices. The role will play ...

Solutions Architect, Studios

Location
Greater London, England, United Kingdom
procurement activity, with particular emphasis on pre sales engagement, proof of concept activity, and technical validation. This role will work across commercial, technical, operational, engineering, procurement, and transformation teams to translate business opportunities into clear solution options, technical direction, and delivery-ready architecture. A core part of the role … conversations, and leading or coordinating proof of concepts that demonstrate feasibility, value, and operational fit. The role will also work closely with DevOps and Site Reliability Engineering teams to ensure proposed solutions are automatable, observable, supportable, resilient, and aligned to modern operational practices. The role will play ...