251 to 275 of 1,104 Site Reliability Engineering Jobs in the UK

Product Associate - SRE Team - Chase UK

Hiring Organisation
Hackajob Ltd
Location
South West London, London, United Kingdom
Employment Type
Permanent
oriented and possess an interest in the financial sector and focus on addressing our customer needs. We work in teams focused on improving the reliability, resilience, observability, and operability of customer-facing digital banking services. We build automation, define measurable reliability practices, reduce operational friction, and partner with … engineering teams to ensure services are designed, delivered, and operated with reliability in mind. Job responsibilities Support the product strategy and delivery of reliability capabilities, including standards, observability, incident practices, automation, and developer experience improvements. Partner with engineers, site reliability engineers, and cross-functional teams ...

HPC Infrastructure Site Reliability Engineer

Location
Gloucester, England, United Kingdom
experience operating large‐scale distributed systems and recent hands‐on expertise in high‐performance computing (HPC) and AI infrastructure. This is an operations‐first SRE role, working in a 24/7/365 on‐call environment, responsible for ensuring reliability, performance, and continuous improvement of mission‐critical infrastructure. … This role sits within a cross‐functional organisation spanning network engineering, infrastructure SRE, Platform SRE, infrastructure tooling engineers (software) and data centre operations. The ideal candidate has progressed through large‐scale, globally distributed or multi‐site infrastructure environments and has more recently specialised in GPU‐accelerated HPC systems. ...

Manager, Platform Engineer, Solutions, Engineering, AI & Data

Hiring Organisation
Deloitte
Location
Manchester, United Kingdom
# 23180 Job description Connect to your IndustryDo you want to be at the heart of some of the biggest and most ambitious engineering projects Technology & Transformation: How do you make intelligent, future-proof decisions in a world where change is the one constant? That's exactly what … premise).Exposure to service mesh technologies (e.g., Istio, Linkerd) and configuration management tools (e.g., Chef, Puppet).Knowledge of site reliability engineering (SRE) principles and practices.GCP certifications (e.g., Cloud Digital Leader, Associate Cloud Engineer, DevOps Engineer).ITILv4 Foundation certification.Familiarity with GCP AI/ML services (AI Platform, AutoML ...

Manager, Platform Engineer, Solutions, Engineering, AI & Data

Hiring Organisation
Deloitte
Location
Holywood, United Kingdom
# 23180 Job description Connect to your IndustryDo you want to be at the heart of some of the biggest and most ambitious engineering projects Technology & Transformation: How do you make intelligent, future-proof decisions in a world where change is the one constant? That's exactly what … premise).Exposure to service mesh technologies (e.g., Istio, Linkerd) and configuration management tools (e.g., Chef, Puppet).Knowledge of site reliability engineering (SRE) principles and practices.GCP certifications (e.g., Cloud Digital Leader, Associate Cloud Engineer, DevOps Engineer).ITILv4 Foundation certification.Familiarity with GCP AI/ML services (AI Platform, AutoML ...

Cloud DevOps Engineer

Hiring Organisation
Sanderson Recruitment
Location
London, United Kingdom
Employment Type
Contract
Contract Rate
£550 - £600 per day + Outside IR-35
Clearance (Mandatory) The Opportunity We are supporting a major government programme seeking an experienced DevOps/Cloud Engineer to join a high-performing cloud engineering team. This role will focus on the design, automation and support of secure cloud infrastructure, enabling the delivery of modern digital services at scale. … sector experience GDS-aligned delivery experience Helm Prometheus Grafana ELK/Elastic Stack Azure or Azure DevOps exposure Site Reliability Engineering (SRE) experience Contract Details £600 per day Outside IR35 Initial 6-month engagement Strong likelihood of extension One day per week onsite in London Four days ...

Cloud Network Engineer

Hiring Organisation
AMS CWS
Location
London, United Kingdom
Employment Type
Contract
services. Develop scripts and tools (e.g., Python, Go, Bash) to streamline network operations, ensure consistency, and improve efficiency Site Reliability Engineering (SRE) for Networks: Embrace a 'you build it, you run it' mindset for network services. Take ownership of the reliability, performance, and availability … Jenkins . Automated testing experience using Terratest, Cucumber, Pytest-BDD, AWS Fault Injection Simulator or Chaos Mesh . Experience applying DevOps, agile and SRE practices to cloud networks, including monitoring, logging, alerting, incident response and performance optimisation. Strong communication skills with strategic thinking and adaptability Next steps Next steps This ...

DevOps Engineer (Security Cleared)

Location
Greater London, England, United Kingdom
Solirius Reply, part of the Reply Group, is a technology consultancy and digital transformation partner that helps organisations solve complex challenges through strategy, design, engineering, and delivery. We work closely with our clients to deliver secure, accessible, user-focused services that evolve with their needs. By combining deep technical … Ministry of Housing, Communities and Local Government, UEFA, International Olympic Committee, and Mercedes-Benz. Our services span the full digital delivery lifecycle, including architecture, engineering, delivery management, user-centred design, business analysis, data, DevOps, and AI. We operate as a collaborative and inclusive organisation that empowers our people ...

DV Cleared Site Reliability Engineer - Cheltenham - Outside IR35

Hiring Organisation
Halian Technology Limited
Location
Cheltenham, Gloucestershire, South West, United Kingdom
Employment Type
Contract
Cleared Site Reliability Engineer- Cheltenham - Outside IR35 Halian's leading Government client is looking for a DV Cleared Site Reliability Engineer to work a long term contract in Cheltenham, this role is majority onsite (4 days a week) and Outside IR35 - Great Day rates available! Skills ...

Senior DevOps / Platform Engineer (Google Cloud)

Location
Greater London, England, United Kingdom
Cloud's premier partner in AI, driving transformation for world-class businesses. We push the boundaries of technology with expertise in machine learning, data engineering, and analytics on Google Cloud Platform. By partnering with us, clients future-proof their operations, unlock actionable insights, and stay ahead of the curve … Experience: Previous experience working in a start-up or scale-up environment Containerisation/Virtualisation Expertise: Proficiency with technologies such as Terraform and Kubernetes SRE Principles: Experience in implementing Site Reliability Engineering (SRE) principles Cloud Native Architecture: Hands-on experience with cloud-native architectures, ideally on Google ...

Site Reliability Software Engineer (Hybrid)

Location
Greater London, England, United Kingdom
shapes without surgery. With over 600,000+ successful outcomes, EarWell® is a proven, non-invasive treatment option for families. We are looking for a Site Reliability Engineer to join our growing team. The ideal candidate has a strong technical background in software development and systems operations, with experience … security, privacy, and compliance requirements in mind. Document architecture, workflows, troubleshooting steps, deployment processes, and support procedures. Required Qualifications Experience as a Software Engineer, Site Reliability Engineer, DevOps Engineer, Systems Engineer, or similar technical role. Experience supporting production applications or infrastructure in a business-critical environment. Strong scripting ...

Network Engineer

Hiring Organisation
Third Nexus Group Limited
Location
Cambridge, Cambridgeshire, United Kingdom
Employment Type
Contract
Contract Rate
£375 - £400/annum
governance, ITSM integration, automation expansion, documentation standards, platform optimisation, network baselining and development of a strategic roadmap towards Site Reliability Engineering (SRE) practices. Key Responsibilities Review the existing network data landscape, including current data sources, data quality, and data flows between network management, monitoring, automation … automation workflows, operational runbooks and controlled implementation processes. Implement new platform features and reduce repetitive manual activity. Define and develop a roadmap towards Network SRE practices. Propose service reliability measures, including relevant KPIs, SLIs and SLOs. Identify opportunities for proactive monitoring, fault prevention and self-healing automation. Work with ...

Exploitation Engineers

Location
Cheltenham, England, United Kingdom
help deliver real-world mission outcomes. If you have a passion for technology and are keen to build a career using innovative engineering to support GCHQ’s mission, this role could be for you. As a Exploitation Engineer, you’ll work as part of an Agile team to design … develop and deploy technical solutions that support mission requirements. Using a combination of software, firmware and hardware engineering, you’ll contribute across the full engineering lifecycle, from understanding user needs to development, testing and deployment. Many of the challenges you’ll encounter will be complex, with no obvious ...

DV Cleared Site Reliability Engineer - Cheltenham - Outside IR35

Location
Cheltenham, Gloucestershire, United Kingdom
Cleared Site Reliability Engineer - Cheltenham - Outside IR35 Check out the role overview below If you are confident you have got the right skills and experience, apply today. Halian's leading Government client is looking for a DV Cleared Site Reliability Engineer to work a long term ...

Lead Site Reliability Engineer

Location
Southampton, England, United Kingdom
service for multi-media evidence management and Emergency Contact Centres to a worldwide customer base. We are currently expanding our Cloud Platform Engineering team to ensure we continue to offer exemplary service to our customers. This is a very hands‐on role. You will be involved in ensuring … cloud platforms are observable, measurable, reliable, scalable, and maintainable. It’s likely that the successful candidate will have significant experience in a DevOps, SRE, Cloud Engineer, or Cloud Development role. How will you make an impact? Act as part of a team of SREs that act as the ‘gatekeepers ...

VodafoneThree - SRE III

Location
Greater London, England, United Kingdom
connected future with technologies like Cloud, AI and big data. What you’ll do In this role, within VodafoneThree's Performance and Chaos Engineering (PaCE) team, you will play a key role in enabling engineering teams to deliver scalable, resilient, and high-performing digital services. PaCE is responsible … observability, capacity planning, and operational readiness to improve service reliability and customer experience. You will collaborate closely Product, Engineering, Platform, Architecture, and SRE teams to design resilient, scalable solutions and proactively identify performance and reliability risks. Drive continuous improvement through data-driven insights, experimentation, and chaos engineering ...

Staff SRE, AI Infrastructure

Hiring Organisation
wayve
Location
London, UK
Employment Type
Full-time
each other to deliver impact. Make Wayve the experience that defines your career! The roleThis is a rare opportunity to be a founding Staff SRE shaping the reliability of large-scale AI systems and GPU compute infrastructure from the ground up. As a Staff Cloud Site Reliability … Compute platform (large-scale, multi-tenant GPU fleets and scheduling systems driving model training and inference at scale).This is a founding Cloud SRE role. You won't inherit a mature SRE function, you'll help create it. You will define the frameworks, automation, and operational standards that ensure ...

DevOps Platform Engineer

Hiring Organisation
Appcast
Location
Remote, UK
With teams across North America, Europe, and the Middle East, First Orion operates in a highly collaborative, distributed environment that brings together world-class engineering, operations, and customer-focused talent. This role will work closely with colleagues across multiple regions, contributing to products and services that help millions … support CI/CD workflows using Gradle and Bitbucket Pipelines Ensure platform deployments are reproducible, consistent, and cloud-agnostic Monitor platform health, performance, and reliability using Prometheus and Grafana Manage secrets and secure platform access using Vault Troubleshoot production issues and support deployment activities across environments Partner closely with ...

DevOps Platform Engineer

Hiring Organisation
First Orion
Location
United Kingdom
With teams across North America, Europe, and the Middle East, First Orion operates in a highly collaborative, distributed environment that brings together world-class engineering, operations, and customer-focused talent. This role will work closely with colleagues across multiple regions, contributing to products and services that help millions … support CI/CD workflows using Gradle and Bitbucket Pipelines Ensure platform deployments are reproducible, consistent, and cloud-agnostic Monitor platform health, performance, and reliability using Prometheus and Grafana Manage secrets and secure platform access using Vault Troubleshoot production issues and support deployment activities across environments Partner closely with ...

DV Cleared Site Reliability Engineer - Cheltenham - Outside IR35

Hiring Organisation
Halian Technology Limited
Location
Cheltenham, Gloucestershire, United Kingdom
Employment Type
Contract
Contract Rate
GBP Annual
Cleared Site Reliability Engineer- Cheltenham - Outside IR35 Halian's leading Government client is looking for a DV Cleared Site Reliability Engineer to work a long term contract in Cheltenham, this role is majority onsite (4 days a week) and Outside IR35 - Great Day rates available! Skills ...

Platform Lead: Fintech Architecture & SRE

Location
United Kingdom
supports rapid growth, and upholds the trust expected in financial services. They will design a resilient architecture that accelerates product development, and deliver exceptional reliability as we grow.Partnering closely with product and engineering teams, they will combine hands-on building with strategic technical leadership to ensure our platform … making process for platform technologies, balancing in-house development with best-in-class third-party solutions* Drive a Site Reliability Engineering (SRE) culture, ensuring high availability, low latency, and robust disaster recovery capabilities* Manage and optimize our cloud infrastructure, focusing on Infrastructure-as-Code (e.g., Terraform), containerization ...

Platform Lead - UK

Location
United Kingdom
supports rapid growth, and upholds the trust expected in financial services. They will design a resilient architecture that accelerates product development, and deliver exceptional reliability as we grow.Partnering closely with product and engineering teams, they will combine hands-on building with strategic technical leadership to ensure our platform … making process for platform technologies, balancing in-house development with best-in-class third-party solutions* Drive a Site Reliability Engineering (SRE) culture, ensuring high availability, low latency, and robust disaster recovery capabilities* Manage and optimize our cloud infrastructure, focusing on Infrastructure-as-Code (e.g., Terraform), containerization ...

DevOps Engineer

Hiring Organisation
Harnham - Data & Analytics Recruitment
Location
London, South East England, United Kingdom
Employment Type
Full-Time
Salary
£100,000 - £120,000 per annum
London (4-5 days a week in office) This is an opportunity for a DevOps Engineer to work at the intersection of infrastructure engineering and AI technology within a high-performance environment. You will play a key role in building and scaling modern infrastructure platforms, with a particular focus … premise environments that support business-critical workloads. THE COMPANY They are a globally operating investment and technology-driven organisation with a strong engineering culture. Their teams work closely with technical and business stakeholders to deliver robust, scalable infrastructure across a complex environment. This role offers exposure to cutting-edge ...

VP, Site Reliability Engineering: Scale & Resilience

Location
Birmingham, England, United Kingdom
Goldman Sachs is seeking a VP-level Site Reliability Engineer to architect and operate highly reliable platforms that support critical … services at scale. You will collaborate across engineering teams to improve production systems and enable rapid delivery of new services. The role emphasizes SRE principles such as SLOs, error budgets, and blameless post-mortems, with leadership opportunities in incident response and on-call design within a financial services context. ...

Senior Linux DevOps Engineer

Hiring Organisation
RedTech Recruitment Ltd
Location
City of London, London, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£90,000
annum + excellent benefits Requirements for Senior Linux DevOps Engineer: Strong commercial experience working as a Senior DevOps Engineer, Linux Engineer, Platform Engineer, Site Reliability Engineer or similar Excellent Linux systems administration and command line skills, with experience operating and troubleshooting large-scale production environments Strong scripting … Linux Engineer/Linux Systems Engineer/Linux Infrastructure Engineer/Senior Platform Engineer/Platform Engineer/Site Reliability Engineer/SRE/Infrastructure Engineer/DevSecOps Engineer/Linux/Bash/Shell Scripting/Python/Kubernetes/Docker/Terraform/Ansible/Microsoft ...

AI Platform & SRE Transformation Consultant

Location
Manchester, England, United Kingdom
Site Reliability Engineering Managing Consultant to help clients design, build and scale secure AI platforms. You will blend platform engineering, SRE, observability and intelligent operations to move from experimentation to production-grade services. You will engage with CIO/CTO and engineering leaders, shape platform ...