351 to 375 of 925 Site Reliability Engineering Jobs in the UK

Principal Software Engineer

Hiring Organisation
Jobleads-UK
Location
Reigate and Banstead, England, United Kingdom
production practices, define standards for handoff, establish quality gates for each conversion layer, and efficiently turn validated prototypes into production systems. ARTIFICIAL INTELLIGENCE AI Engineering Strategy: Pioneer AI integration patterns, publish on AI‐augmented development, and lead AI in engineering. AI Evaluation & Observability: Shape AI observability and evaluation practices … practices, optimize portfolios, balance investments, and train service thinking. Site Reliability Engineering: Define standards, drive incident improvements, design capacity, and mentor SRE practices. Technical Writing: Define standards, create systems and templates, train on spec‐driven development, and ensure quality. WORKING‐LEVEL SKILLS Cloud Platforms: Design solutions, manage ...

AWS SRE: Reliability, Observability & Cost Optimisation

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Source Technology Limited is seeking a Site Reliability Engineer in London for a hybrid role, requiring 3+ years of SRE experience, especially in Kubernetes. Responsibilities include improving system reliability, observability, and cost efficiency. The ideal candidate will work closely with development and platform teams, and should ...

Head of Production Management- J.P. Morgan Personal Investing

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
powered solutions and intelligent automation to reduce manual intervention, fast‐track resolution, and continuously improve operational efficiency. Champion an automation‐first, shift‐left SRE culture leveraging shared tooling and automation to ensure consistency, reduce duplication, and maintain alignment with firm‐wide standards. Oversee capacity management and planning, ensuring infrastructure scales … management standards — change, incident, capacity, and automation — across multiple engineering teams operating in a you‐build‐it‐you‐run‐it model, underpinned by SRE principles and disaster recovery planning. Composure, decisiveness, and authority during incidents, vendor failure, or regulatory escalation, with a proven ability to protect business lines under ...

Head of Production Management - FinTech Reliability Leader

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
powered solutions and intelligent automation to reduce manual intervention, fast-track resolution, and continuously improve operational efficiency. Champion an automation-first, shift-left SRE cultureleveraging shared tooling and automation to ensure consistency, reduce duplication, and maintain alignment with firmwide standards. Oversee capacity management and planning, ensuring infrastructure scales to meet … management standards — change, incident, capacity, and automation — across multiple engineering teams operating in a you-build-it-you-run-it model, underpinned by SRE principles and disaster recovery planning. Composure, decisiveness, and authority during incidents, vendor failure, or regulatory escalation, with a proven ability to protect business lines under ...

SRE: Cloud Infra, Kubernetes & Go — Hybrid

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Nominet, a world-leading domain name registry operating at the heart of the UK internet, is hiring a Site Reliability Engineer to join our Reliability Engineering team. You will design, deploy, and operate scalable cloud infrastructure on AWS and Kubernetes, with a focus on infrastructure … GitOps, CI/CD, and self-service tooling. You’ll build reusable tooling in Go, manage production services, and work with developers to improve reliability and security while reducing toil. #J-18808-Ljbffr ...

Site Reliability Engineer - Gloucester - NS West

Hiring Organisation
Hackajob Ltd
Location
Gloucester, Gloucestershire, South West, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£70,000
community engagement and outreach activities to help build tech and cyber skills in the region. What you could be doing for us: As an SRE, fundamentally you will be doing work that has historically been done by an operations team, but using software and systems engineering expertise … automation to reduce human labour, limiting traditional manual operations work (incident tickets, on-call etc.) to no more than half of the SRE team's time. Core role accountabilities include: Supporting and maintaining essential service that support core mission applications, proactively enhancing their availability, performance and stability; Finding innovative solutions ...

SRE Lead - AI/ML Data Platform & Reliability

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
United States Digital Space LLC seeks a Site Reliability Engineer for AI/ML Data Platforms to lead the development of scalable and resilient data solutions. You will engage in root cause analysis and mentor team members while managing complex production environments. Essential qualifications include proficiency in site reliability culture, incident management, and tools like AWS and Databricks. Candidates should also demonstrate strong skills in Python or PySpark, as well as AI-assisted software development. #J-18808-Ljbffr ...

Global SRE Manager - Real-Time Trading Platform Reliability

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Cares, Inc is seeking a Sr. Manager of Site Reliability Engineering in London. The role entails overseeing a distributed team of SRE technologists to ensure high availability of trading platforms. Candidates should possess a Bachelor's Degree in a related field and deep expertise in technical operations ...

Lead Site Reliability Engineer - Observability & Resilience

Hiring Organisation
Jobleads-UK
Location
Glasgow, Scotland, United Kingdom
JPMorgan Chase & Co. seeks a Lead Site Reliability Engineer to define the future of reliability for a global firm. You will lead critical resiliency design reviews, break complex problems into actionable work, and mentor engineers across large-scale OpenTelemetry pipelines in hybrid environments. You will guide incident ...

Senior Lead SRE - Reliability & Observability Leader

Hiring Organisation
Jobleads-UK
Location
Glasgow, Scotland, United Kingdom
JPMorgan Chase in the United Kingdom is seeking a Senior Lead Site Reliability Engineer … join an agile team focused on reliability, observability, and performance across critical platforms. You will mentor engineers, lead incident response, and shape SRE strategy while delivering scalable, secure production systems. The role demands deep expertise in cloud, automation, OpenTelemetry, and instrumentation, with a track record of improving service levels ...

Senior Site Reliability Engineer - Lead Resilient Platform

Hiring Organisation
Jobleads-UK
Location
Watford, England, United Kingdom
Allwyn in the UK is seeking a Senior/Lead Site Reliability Engineer to provide technical leadership for reliability across the digital estate, ensuring high availability, performance, and resilience of customer-facing systems during normal operation and peak lottery events. The role blends hands-on engineering ...

Lead Site Reliability Engineer: Drive Stability & Observability

Hiring Organisation
Jobleads-UK
Location
Glasgow, Scotland, United Kingdom
JPMorgan Chase is seeking a Lead Site Reliability Engineer in the UK to shape the reliability strategy for large-scale hybrid environments. You will guide incident response, drive service level objectives, and mentor a team of engineers across multiple domains. The role emphasizes deep expertise in observability ...

Site Reliability Engineer - Cloud & AI Network Infra

Hiring Organisation
Jobleads-UK
Location
United Kingdom
Jobgether is seeking a Site Reliability Engineer in Network Infrastructure based in the United Kingdom to strengthen reliability and scalability of critical network systems. You will design, automate, and operate systems enabling high-performance digital services, focusing on resilient environments, observability, and safe change processes. You will … define SLIs/SLOs, drive reliability improvements, own incident responses, and collaborate with network and platform teams to implement scalable #J-18808-Ljbffr ...

Senior Site Reliability Engineer - Python

Hiring Organisation
17918
Location
Manchester, Lancashire, United Kingdom
digital services that support businesses across the UK. The Department for Business and Trade (DBT), in partnership with Inspire People, is seeking a Senior Site Reliability Engineer with strong software engineering capability (either experience of Python development or a desire to develop Python expertise), experience building applications ...

Senior Site Reliability Engineer - Python

Hiring Organisation
17918
Location
Birmingham, Warwickshire, United Kingdom
digital services that support businesses across the UK. The Department for Business and Trade (DBT), in partnership with Inspire People, is seeking a Senior Site Reliability Engineer with strong software engineering capability (either experience of Python development or a desire to develop Python expertise), experience building applications ...

Senior Site Reliability Engineer - Python

Hiring Organisation
17918
Location
Cardiff, Glamorgan, United Kingdom
digital services that support businesses across the UK. The Department for Business and Trade (DBT), in partnership with Inspire People, is seeking a Senior Site Reliability Engineer with strong software engineering capability (either experience of Python development or a desire to develop Python expertise), experience building applications ...

Senior Site Reliability Engineer - Python

Hiring Organisation
17918
Location
Belfast, County Antrim, United Kingdom
digital services that support businesses across the UK. The Department for Business and Trade (DBT), in partnership with Inspire People, is seeking a Senior Site Reliability Engineer with strong software engineering capability (either experience of Python development or a desire to develop Python expertise), experience building applications ...

Senior Site Reliability Engineer - Python

Hiring Organisation
Inspire People
Location
Edinburgh, Midlothian, United Kingdom
Employment Type
Permanent
Salary
GBP 80,000 Annual
digital services that support businesses across the UK. The Department for Business and Trade (DBT), in partnership with Inspire People, is seeking a Senior Site Reliability Engineer with strong software engineering capability (either experience of Python development or a desire to develop Python expertise), experience building applications ...

Platform Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
significant contributions to project goals, focusing on cloud platform engineering. This individual demonstrates a solid understanding of the principles and practices of platform engineering, especially within cloud environments, and shows proficiency in specific technical areas related to cloud infrastructure, automation, and scalability.Key responsibilities include at least … following: Cloud platforms:Community Involvement: Actively participating in the professional cloud platform engineering community, contributing insights, and staying abreast of the latest trends and best practices.Team Contribution: Making significant contributions to team objectives, particularly in designing, building, and maintaining cloud based platforms and infrastructure.Technical Proficiency: Exhibiting a good grasp ...

Consultant - Manager, Backend Developer - Tech Consultancy - Defence & Security

Hiring Organisation
Deloitte
Location
London, United Kingdom
Salary
£ 80 K
Agile ecosystem supported by tooling including Atlassian, Jenkins, GitLab, OWASP and AWS componentry.Ensure your solution works in a reliable and resilient way using Site Reliability Engineering methods to increase availability while reducing costs and callouts.Help the client and end users to understand trade-offs when making product … great community of architects and back-end developers who run workshops together, share the best articles they find on Slack, and contribute to the engineering culture.Go beyond standard duties and responsibilities to champion small details, spot opportunities and add extra value for our clients.Connect to your skills and professional ...

Consultant - Manager, Backend Developer - Tech Consultancy - Defence & Security

Hiring Organisation
Deloitte
Location
Manchester, Greater Manchester, United Kingdom
Salary
£ 70 K
Agile ecosystem supported by tooling including Atlassian, Jenkins, GitLab, OWASP and AWS componentry.Ensure your solution works in a reliable and resilient way using Site Reliability Engineering methods to increase availability while reducing costs and callouts.Help the client and end users to understand trade-offs when making product … great community of architects and back-end developers who run workshops together, share the best articles they find on Slack, and contribute to the engineering culture.Go beyond standard duties and responsibilities to champion small details, spot opportunities and add extra value for our clients.Connect to your skills and professional ...

Consultant - Manager, Backend Developer - Tech Consultancy - Defence & Security

Hiring Organisation
Deloitte
Location
Cambridge, Cambridgeshire, United Kingdom
Salary
£ 80 K
Agile ecosystem supported by tooling including Atlassian, Jenkins, GitLab, OWASP and AWS componentry.Ensure your solution works in a reliable and resilient way using Site Reliability Engineering methods to increase availability while reducing costs and callouts.Help the client and end users to understand trade-offs when making product … great community of architects and back-end developers who run workshops together, share the best articles they find on Slack, and contribute to the engineering culture.Go beyond standard duties and responsibilities to champion small details, spot opportunities and add extra value for our clients.Connect to your skills and professional ...

Consultant - Manager, Backend Developer - Tech Consultancy - Defence & Security

Hiring Organisation
Deloitte
Location
Bristol, Gloucestershire, United Kingdom
Salary
£ 70 K
Agile ecosystem supported by tooling including Atlassian, Jenkins, GitLab, OWASP and AWS componentry.Ensure your solution works in a reliable and resilient way using Site Reliability Engineering methods to increase availability while reducing costs and callouts.Help the client and end users to understand trade-offs when making product … great community of architects and back-end developers who run workshops together, share the best articles they find on Slack, and contribute to the engineering culture.Go beyond standard duties and responsibilities to champion small details, spot opportunities and add extra value for our clients.Connect to your skills and professional ...

Lead Product Manager AIOPs

Hiring Organisation
Jobleads-UK
Location
City Of London, England, United Kingdom
operations through intelligent observability, event correlation, anomaly detection, predictive insights, and automation. They partner with AIOps vendors, IT Operations, infrastructure, platform engineering, SRE, service management, and application teams to reduce operational noise, improve service reliability, and accelerate incident response across a complex enterprise environment. Responsibilities Execute the enterprise … enterprise AIOps roadmap, aligning delivery plans and priorities with reliability goals, operational maturity, and business outcomes. Lead cross‐functional collaboration across infrastructure, SRE, platform engineering, service management, and application teams to operationalize AIOps capabilities at scale. Establish governance, success metrics, and operating rhythms to measure improvements in alert ...

Full Stack / Data Engineer · London ·

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
hands‐on Developer/Data Engineer who thrives in a fast‐paced startup environment. You’ll work across application development, data architecture, and site reliability — solving complex problems, supporting production systems, and building the foundations for scalable internal and external data self‐service through Microsoft Fabric … experience setting up Microsoft Fabric , Data Lakehouse , or Medallion architecture . Demonstrated ability to troubleshoot complex distributed systems and identify root causes. Exposure to SRE principles (monitoring, incident management, resilience). Familiarity with AI/LLMs (e.g., using OpenAI, Azure OpenAI, or similar APIs). Excellent problem‐solving skills ...