251 to 275 of 663 Site Reliability Engineering Jobs in the UK

Senior Site Reliability Engineer - Python

Hiring Organisation
Inspire People
Location
Edinburgh, Midlothian, United Kingdom
Employment Type
Permanent
Salary
GBP 80,000 Annual
digital services that support businesses across the UK. The Department for Business and Trade (DBT), in partnership with Inspire People, is seeking a Senior Site Reliability Engineer with strong software engineering capability (either experience of Python development or a desire to develop Python expertise), experience building applications ...

Platform Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
significant contributions to project goals, focusing on cloud platform engineering. This individual demonstrates a solid understanding of the principles and practices of platform engineering, especially within cloud environments, and shows proficiency in specific technical areas related to cloud infrastructure, automation, and scalability.Key responsibilities include at least … following: Cloud platforms:Community Involvement: Actively participating in the professional cloud platform engineering community, contributing insights, and staying abreast of the latest trends and best practices.Team Contribution: Making significant contributions to team objectives, particularly in designing, building, and maintaining cloud based platforms and infrastructure.Technical Proficiency: Exhibiting a good grasp ...

Lead Product Manager AIOPs

Hiring Organisation
Jobleads-UK
Location
City Of London, England, United Kingdom
operations through intelligent observability, event correlation, anomaly detection, predictive insights, and automation. They partner with AIOps vendors, IT Operations, infrastructure, platform engineering, SRE, service management, and application teams to reduce operational noise, improve service reliability, and accelerate incident response across a complex enterprise environment. Responsibilities Execute the enterprise … enterprise AIOps roadmap, aligning delivery plans and priorities with reliability goals, operational maturity, and business outcomes. Lead cross‐functional collaboration across infrastructure, SRE, platform engineering, service management, and application teams to operationalize AIOps capabilities at scale. Establish governance, success metrics, and operating rhythms to measure improvements in alert ...

Full Stack / Data Engineer · London ·

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
hands‐on Developer/Data Engineer who thrives in a fast‐paced startup environment. You’ll work across application development, data architecture, and site reliability — solving complex problems, supporting production systems, and building the foundations for scalable internal and external data self‐service through Microsoft Fabric … experience setting up Microsoft Fabric , Data Lakehouse , or Medallion architecture . Demonstrated ability to troubleshoot complex distributed systems and identify root causes. Exposure to SRE principles (monitoring, incident management, resilience). Familiarity with AI/LLMs (e.g., using OpenAI, Azure OpenAI, or similar APIs). Excellent problem‐solving skills ...

Graduate Cloud Operations Engineer - Newcastle

Hiring Organisation
Hackajob Ltd
Location
Newcastle Upon Tyne, Tyne and Wear, North East, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£35,000
skills and breadth of experience needed to launch a career in enterprise cloud operations. You will spend significant time in our Cloud Operations Engineering team and will be involved in: Building hands-on skills deploying, administering, monitoring, and automating infrastructure across AWS and Azure Ensuring service excellence by maintaining … standards. Tooling & Automation Seeing how strategic tooling initiatives and automation solutions enhance productivity, and how emerging technologies and AI capabilities are evaluated and adopted. Site Reliability Engineering Understanding how SLOs, error budgets, and data-driven reliability practices are used to build self-healing, resilient systems. Service ...

SRE Security engineer

Hiring Organisation
FBI &TMT
Location
London, United Kingdom
Employment Type
Contract, Work From Home
Contract Rate
£500 - £594 per day
About the Role We are looking for an experienced Site Reliability Engineer (SRE) Security Engineer to join a major client programme on a long-term contract through to March 2027 . This is a primarily remote opportunity, although candidates must be willing and able to travel to client … with engineering, operations, and security teams. Key Responsibilities Design, implement, and maintain secure, resilient, and highly available platforms. Embed security best practices within SRE and DevOps processes. Monitor system performance, availability, and security posture across environments. Automate operational and security controls using Infrastructure as Code and CI/ ...

SRE Security engineer

Hiring Organisation
Matchtech
Location
London, South East, England, United Kingdom
Employment Type
Contractor
Contract Rate
£500 - £594 per day
About the Role We are looking for an experienced Site Reliability Engineer (SRE) Security Engineer to join a major client programme on a long-term contract through to March 2027 . This is a primarily remote opportunity, although candidates must be willing and able to travel to client … with engineering, operations, and security teams. Key Responsibilities Design, implement, and maintain secure, resilient, and highly available platforms. Embed security best practices within SRE and DevOps processes. Monitor system performance, availability, and security posture across environments. Automate operational and security controls using Infrastructure as Code and CI/ ...

Senior Site Reliability Engineer, Scalable Infra

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Cisco ThousandEyes is seeking a seasoned Site Reliability Engineer to design, operate and scale large‐scale distributed systems that process telemetry data at high volumes. You will work across AWS, Kubernetes, and infrastructure‐as‐code, applying AI to automate toil … improve reliability. Collaboration with software engineers and on‐call incident management are core parts of the role. Ideal candidates have 5+ years in SRE/DevOps, strong coding skills in Python or Go, and demonstrable #J-18808-Ljbffr ...

Staff Site Reliability Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
that usable, and MystraAI is the agentic layer we are building on top of it. This is a Staff-level role that owns the reliability, performance, security and integrity of that infrastructure end-to-end — and sets the technical direction that other teams build on. You will lead … source level rather than as a black box — and ideally have contributed code upstream. Reliability engineering for data platforms. You bring true SRE discipline — SLOs, observability, capacity planning and incident response — to analytical data systems and pipelines. Data-as-a-Service productisation. You think in terms of data ...

Site Reliability Engineer - Cloud, Automation & Resilience

Hiring Organisation
Jobleads-UK
Location
Glasgow, Scotland, United Kingdom
Barclays is seeking a Site Reliability Engineer to ensure the resilience of critical banking systems in Glasgow. The role focuses on incident response, automation, and scalable operations across a global L2 team. You will work with Java applications, Oracle databases, and cloud platforms to maintain peak availability. Applicants ...

Cloud Native DevOps Engineer- SC Cleared

Hiring Organisation
Searchability NS&D
Location
London, South East, England, United Kingdom
Employment Type
Full-Time
Salary
£45,000 - £100,000 per annum, Negotiable
CLEARED CLOUD NATIVE DEVOPS ENGINEER- Permanent opportunity for a Cloud Native DevOps Engineer with SC Clearance. - Salary up to £100,000 DOE - On-site opportunity with London based offices - To apply, please call Laura Jackson on , or email with an up-to-date CV. … process and submit (subject to required skills) your application to our client in conjunction with this vacancy only. KEY SKILLSDEVOPS ENGINEER, DEVOPS, SITE RELIABILITY ENGINEER, PLATFROM ENGINEER, PLATFROM, CLOUD, CLOUD NATIVE, AWS, AWS CLOUD, AWS ENGINEER, ANSIBLE, TERRAFORM, CLOUD SUPPORT, CLOUD INFRASTRUCTURE, DEFENCE, NATIONAL SECURITY, DV CLEARED ...

Database Site Reliability Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Responsibilities Design, configure, and operate highly available database and data storage systems Drive reliability, performance, scalability, and operational excellence across database environments Lead platform improvement initiatives Monitor, optimise, and automate database operations Identify and address technical debt and performance bottlenecks Support reliability of data pipelines and services Participate … ability to own services end-to-end Strong technical expertise Excellent problem-solving and collaboration skills Solid understanding of database design, performance optimisation, and reliability engineering principles Experience with monitoring, observability, alerting, and incident management practices Understanding of data platforms and data pipelines #J-18808-Ljbffr ...

Remote Principal SRE - Healthcare Platform Reliability Lead

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
MediSolution in London is seeking a Site Reliability Engineer (SRE) to ensure the reliability of healthcare platforms. The candidate will lead efforts in automating operations and improving service availability. With a focus on troubleshooting and incident management, applicants should have 7+ years of experience in enterprise applications ...

Director, Head of Technology Resilience and Production Operations

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
vision to become ‘the worlds most trusted Bank’ and ensuring continued regulatory confidence* The role holder will work with the Head of Digital Engineering Services and Solutions Department Head to deliver transformation through a reliable, robust, sustainable, scalable and efficient operating model by leveraging best practices* The role holder … Resilience related projects and programmes measuring the effectiveness of these services delivered* This is a Leadership position and an integral part of the Digital Engineering Solutions and Services Leadership team maintain compliance and regulatory obligations.**NUMBER OF DIRECT REPORTS**TBC - Team Size circa 25**KEY RESPONSIBILITIES****Planning & Strategy ...

Senior SRE: AI Infra Reliability & Scaling (Remote)

Hiring Organisation
Jobleads-UK
Location
United Kingdom
Runware is hiring a Site Reliability Engineer to ensure the reliability, performance and resilience of its growing platform. You will work … across software, infrastructure and production operations to reduce toil and drive lasting improvements in complex distributed systems. You will own production reliability, define SRE practices, investigate issues across APIs, queues and databases, and lead incident reviews. #J-18808-Ljbffr ...

CDS Clear IT Site Reliability Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
role is for a business focused Site Reliability Engineer with strong experience in IT environments, tools, and technologies. The individual filling this role will be a business focused problem solver with a desire to learn and closely partner with highly engaged production Risk, Operations and IT Development teams ...

Senior Cloud SRE - Kubernetes, GCP & CI/CD

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
leading technology firm in Greater London is seeking a Senior Cloud Engineer to contribute technical expertise within a cloud engineering team. This role involves architecting scalable Kubernetes environments on Google Cloud Platform (GCP) and ensuring robust security measures. The ideal candidate will have extensive experience in DevOps or Site Reliability Engineering, deployment of production-grade Kubernetes clusters, and proficiency in CI/CD pipelines, alongside programming skills in Python, Go, and Bash. Join this dynamic team to drive innovative cloud solutions. #J-18808-Ljbffr ...

Software Engineer III, Full Stack, Publisher Inventory

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
control and optimize ad placement, policy compliance, and publisher revenue. Leverage Google's state-of-the-art AI tooling and LLM APIs to improve engineering velocity and build AI-augmented product features. Collaborate with cross-functional partners including PMs, UX, and cross-sites with other engineering partners … deliver seamless end-to-end features. Write robust, well-tested, and high-performance code, ensuring high quality and reliability across our platform.### Skills Required* Kotlin* C++* TypeScript* Dart* Angular* Java* APIs* SDKs* LLM APIs* Generative AI### Tags:UKLondonAngular DeveloperFlutter DeveloperShare Job:Application planning## What to evaluate before applying### Visa ...

Lead Site Reliability Engineer: Architect Resilience & AI Ops

Hiring Organisation
Jobleads-UK
Location
City of Westminster, England, United Kingdom
JPMorgan Chase & Co. in London seeks a Lead Site Reliability Engineer to shape the future for a globally recognized firm. You will lead resiliency reviews, break complex problems into actionable work for engineers, and serve as technical lead for medium to large products. As part of the Infrastructure … Platforms team, you will guide incident response, mentor peers, and drive AI-enabled reliability workflows across the SDLC while ensuring security and traceability throughout. #J-18808-Ljbffr ...

SRE Engineering Manager: Traffic Steering & DNS

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
United States Digital Space LLC is seeking a Site Reliability Engineer to lead a seasoned team ensuring the availability and performance of mission-critical services. You will automate responses, oversee incident management, and drive reliability across globally distributed systems. Join a culture of curiosity and collaboration, mentoring ...

Site Reliability Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
iGaming company based in London. Estimated Benefits Health Insurance Pension Stock Options Benefits estimated based on industry standards We’re hiring a Site Reliability Engineer to join our London team This is a fantastic opportunity for someone passionate about reliability, scalability and automation. You’ll be pivotal ...

Cloud Platform Engineer (Senior / Lead)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
identity, network, workloads and data for both human and machine/agent identities; secure by default, least privilege, secrets management and continuous compliance. Apply SRE practices—SLOs/SLIs, observability, capacity planning, resilience and blameless incident management—to keep the platform reliable and cost‐efficient. Partner with data engineering … ability (e.g. Python, Go) and strong observability, reliability and cost‐optimisation practices. Desirable requirements: Experience working as a Site Reliability Engineer (SRE) with SLOs/SLIs, error budgets and incident management. A third top‐tier cloud certification, or specialist security/Kubernetes certifications (e.g. CKA/ ...

Principal Platform Engineer

Hiring Organisation
Sanderson Recruitment
Location
City of London, London, United Kingdom
Employment Type
Permanent
large-scale distributed systems and database platforms? We're looking for a hands-on technical leader to help shape the future of our platform engineering capability. This is an opportunity to lead complex engineering initiatives, define technical strategy, and act as a subject matter expert across AWS infrastructure … technical authority for distributed database and persistence technologies Required Experience 8+ years' experience in Platform Engineering, Infrastructure Engineering, DevOps, SRE or Software Engineering Expert-level AWS infrastructure experience Strong Infrastructure as Code expertise with Terraform Strong Linux systems administration and networking knowledge Experience designing and operating distributed ...

Senior AVP, Custody Applications & SRE

Hiring Organisation
Jobleads-UK
Location
Belfast City District, Northern Ireland, United Kingdom
Citigroup Inc. is seeking an experienced professional to join the Global Custody Production Support team in Belfast. The role focuses on reliability, automation, and operability across distributed custody applications and settlement platforms. You will lead production incident response, drive improvements with cross-functional partners, and apply Site Reliability Engineering principles to ensure high availability and resilience in a global, hybrid work environment. #J-18808-Ljbffr ...

Senior Platform Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
diversified range of trading strategies. We employ over 130 colleagues in Jersey, Geneva, London, Singapore, New York and Shanghai. This is a hands‐on engineering role within the Platform Engineering team, part of Technology Operations. Platform Engineering is responsible for building and operating the infrastructure, platforms … comply with all organisational, statutory and regulatory policies and procedures. Experience, Knowledge & Skills Five or more years of experience in platform engineering, DevOps, SRE, infrastructure engineering or a closely related role. Strong experience operating production or production‐like infrastructure, ideally across hybrid cloud and on‐premises environments. Hands ...