276 to 300 of 663 Site Reliability Engineering Jobs in the UK

Service Delivery Manager – HPC/Super Computing

Hiring Organisation
Jobleads-UK
Location
United Kingdom
Microsoft technology with solid overview of the Microsoft cloud servicesProject Management/Prosci Change Management/ITIL certification is a plusUnderstanding DevOps, Site Reliability Engineering, Continuous Improvement is a plusSupercomputing/HPC experience is a plus.Ability to meet Microsoft, customer and/or government security screening requirements ...

Platform Operations Manager (DevOps & Site Reliability Engineering)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Working Student - Data Migration (all genders) - Fixed-term 2 months E-commerce parttime the company is a Hamburg-based B2B software company connecting brands and retailers across the fashion industry. We build the digital infrastructure ...

Senior AWS SRE Lead — Cloud Reliability & Automation

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
company in the UK is seeking a Lead Site Reliability Engineer to strengthen reliability, availability and performance of global digital platforms. The role emphasizes AWS expertise, automation, monitoring, and incident management within a hybrid working model. You will mentor engineers, collaborate with cross‐functional teams, and drive … best practices in cloud engineering, security, and reliability across multiple regions. #J-18808-Ljbffr ...

AI-Driven Senior SRE Leader for Large-Scale Systems

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
United States Digital Space LLC is looking for a seasoned engineer to join their Site Reliability Engineering team in London, UK. This pivotal role focuses on ensuring reliable and robust Google Ads services while utilizing AI-driven automation. Candidates should have a Bachelor's degree in Computer … Science and substantial software development experience. The position requires expertise in problem-solving, innovative engineering solutions, and mentoring skills to guide a team of engineers within a dynamic environment. #J-18808-Ljbffr ...

SRE Leadership: Reliability, Automation & Observability

Hiring Organisation
Jobleads-UK
Location
Bristol, England, United Kingdom
Jobtailor in Bristol, UK, is seeking a senior Site Reliability Engineer to define and evolve the reliability strategy across … global technology stack. You will drive observability, automation, and continuous improvement, partnering with Engineering, Infrastructure, Security, and Technology Operations to embed modern SRE practices while promoting shared ownership. The role requires leading cross-functional teams, deep expertise in SRE concepts (SLOs, SLIs, error budgets), and hands-on automation #J ...

DevOps Engineer

Hiring Organisation
ISR Recruitment
Location
United Kingdom
major UK Government Agency on a large-scale cloud transformation programme and is seeking an experienced Senior DevOps Engineer to join a multidisciplinary engineering team. This is an exciting opportunity to contribute to the design, deployment and operation of secure, cloud-native platforms spanning AWS, Microsoft Azure and Google … enterprise scale. Skills and Experience: Essential Proven commercial experience as a DevOps Engineer, Platform Engineer, Cloud Engineer or Site Reliability Engineer (SRE). Strong experience deploying and managing AWS infrastructure using Terraform. Commercial experience operating cloud-native platforms across AWS, Microsoft Azure and Google Cloud Platform. Excellent knowledge ...

Developer Enablement, Technical Architect – Release on Demand (SVP)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
effort to make RoD elastically scalable and enterprise-grade reliable. Design for high availability, graceful degradation, and zero-downtime deployments. Define SLOs, own the SRE practice for the platform, and be accountable when things need fixing. Define the Observability Strategy. Establish a comprehensive observability framework — distributed tracing, structured logging, metrics … development language. Designing, building and consuming domain specific RESTful services & APIs Proficiency with relational and/or NoSQL databases: PostgreSQL, MongoDB or Couchbase Demonstrated SRE or platform engineering experience — SLOs, incident management, reliability engineering at scale Experience defining and implementing observability strategies: distributed tracing, structured logging, metrics ...

Site Reliability Engineer (SRE), London

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
approach to cloud intelligence, extending the security and privacy of Apple devices into the cloud to unlock even more intelligence for our users. This SRE team is responsible for the availability and automation of the critical systems and services that enable PCC to deliver cloud intelligence without compromising user privacy. … future of privacy-preserving cloud infrastructure at scale, this is the opportunity for you! Description We\'re looking for a hardworking and passionate SRE Engineer to join this amazing team. You will be an accomplished builder and problem-solver, eager to tackle challenging technical problems. You have a deep understanding ...

Principal Platform Engineer

Hiring Organisation
Jobleads-UK
Location
City Of London, England, United Kingdom
impact Build and maintain infrastructure-as-code (Terraform or similar), CI/CD pipelines, and Kubernetes platforms as products consumed by engineering teams. SRE & reliability Define and drive SLOs, error budgets, and observability standards (metrics, logging, tracing) across the platform Take part in post-incident reviews and help … technical roadmap for the platform function, balancing reliability investment against delivery. What we're looking for Must have A background in Operations or SRE running highly available, redundant production platforms — you understand failure domains, graceful degradation, and what "five nines" costs Deep hands-on experience with AWS (VPC design ...

Principal Platform Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
impact***** ****Build and maintain infrastructure-as-code (Terraform or similar), CI/CD pipelines, and Kubernetes platforms as products consumed by engineering teams********SRE & reliability***** ****Define and drive SLOs, error budgets, and observability standards (metrics, logging, tracing) across the platform***** ****Take part in post-incident reviews and help … roadmap for the platform function, balancing reliability investment against delivery****## ****What we're looking for********Must have***** ****A background in Operations or SRE running highly available, redundant production platforms — you understand failure domains, graceful degradation, and what "five nines" costs***** ****Deep hands-on experience with AWS (VPC design ...

Senior Site Reliability Engineer

Hiring Organisation
Staffworx Limited
Location
London, United Kingdom
Employment Type
Permanent, Work From Home
team is expanding and hiring engineers now The role: Build, operate and maintain high-performance, scalable, reliable services across UK Government deployments Own production reliability: monitoring, alerting, configuration management and upgrades Lead automation to reduce manual operations, using modern platforms including LLM/AI tooling Deploy new products into … need: UK SC clearance, or eligibility to obtain it (active SC strongly preferred; no visa sponsorship) 1-5 years in infrastructure engineering or SRE, building and deploying production systems - not just debugging Hands-on Kubernetes/Docker in production; Terraform, Ansible or CI/CD pipeline experience Proficiency ...

Site Reliability Engineer- Spacetime UK

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Role Overview This isn't a "keep the lights on" SRE role. This is a strategic, high-impact opportunity to build the nervous system for a platform that transforms how networks of satellites, ground stations, and fleets are interconnected and orchestrated. You will be building the core observability stack that … cloud-native tools to a robust, scalable, and insightful platform built on best-in-class technologies (Prometheus, OpenTelemetry, etc.). If you are an SRE who thrives on platform-building challenges and wants to be relied upon to build a production-grade observability stack from the ground up, this role ...

SRE: AI-Driven Infra, Reliability & Observability

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Signal AI is seeking a Site Reliability Engineer (SRE) to join their Infrastructure team in Greater London. In this role, you will help evolve the infrastructure behind Signal AI's decision intelligence platform and lead efforts to enhance operational performance. We’re looking for candidates experienced ...

AWS SRE Lead: Reliability & Automation

Hiring Organisation
Jobleads-UK
Location
United Kingdom
British Council seeks a Lead Site Reliability Engineer to ensure reliability and performance of global digital platforms. You will design, implement, and improve resilient systems with AWS and Azure, automate infrastructure, and mentor teams. You will lead cross-functional collaboration, promote best practices, and contribute ...

Staff Software Engineer

Hiring Organisation
Jobleads-UK
Location
Belfast, Northern Ireland, United Kingdom
pace of innovation while improving the developer experience. The Harness Software Delivery Platform includes modules for CI, CD, Cloud Cost Management, Feature Flags, Service Reliability Management, Security Testing Orchestration, Chaos Engineering, Software Engineering Insights, and continues to expand at an incredibly fast pace. About The Role Design … technical leadership. Perform peer reviews of specifications, designs, and code. Identify technical debt & scaling issues in the base code and drive improvement. Work alongside Site Reliability Engineers and cross‐functional teams to diagnose/troubleshoot any production performance related issues. Use Java, Golang, and Python. Build systems ...

Senior SRE: GCP, Kubernetes & Automation Leader

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Brevan Howard CFD LTD is seeking a Senior Site Reliability Engineer (SRE) to enhance the reliability, scalability, and performance of its core platform. The successful candidate will provide operational support and lead infrastructure projects. This role requires hands-on experience with Google Cloud Platform (GCP) and Kubernetes ...

Remote SRE Security Engineer with SC Clearance

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Matchtech is seeking an experienced Site Reliability Engineer (SRE) Security Engineer for a major client programme. This long-term contract runs through March 2027, with a primarily remote setup and occasional travel to client offices as required. You will help ensure the reliability, security, and resilience … critical platforms, working closely with engineering, operations, and security teams to embed security into SRE and DevSecOps practices. #J-18808-Ljbffr ...

Kubernetes SRE: DevOps, Observability & Reliability

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Cisco's Webex Engineering Group in London is seeking a Site Reliability Engineer to own the design, deployment, and operation of Kubernetes-based microservices. You will build reusable configurations and advance a robust, scalable platform. In this role, you will implement GitOps workflows with Argo CD, manage … canary releases and secret management with Vault, and monitor reliability with Prometheus and Grafana, collaborating across teams in a hybrid London office. #J-18808-Ljbffr ...

Senior DevOps Platform Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
DevOps Platform Engineer London We are currently supporting a leading organisation looking to hire an experienced Senior DevOps Platform Engineer to join their cloud engineering function. This is an excellent opportunity for a highly skilled DevOps professional to work on complex cloud-native environments, designing and automating scalable platforms … Implement GitOps delivery models using tools such as ArgoCD and Helm Develop reusable automation frameworks for infrastructure provisioning, configuration management and deployments Improve platform reliability, scalability and security through automation and engineering best practices Support cloud migrations from traditional data centre environments into modern cloud-native platforms Implement ...

Senior Lead SRE & DevOps - Reliability Architect

Hiring Organisation
Jobleads-UK
Location
Glasgow, Scotland, United Kingdom
JPMorgan Chase & Co. in Glasgow seeks a Senior Lead Site Reliability/DevOps Engineer to enhance reliability, observability, and performance of critical platforms. You will lead OpenTelemetry pipelines, guide incidents, and drive migrations across hybrid environments to secure scalable systems. Role requires deep cloud-native experience, mastery ...

Senior AWS Site Reliability Engineer

Hiring Organisation
Spectrum IT Recruitment
Location
London, United Kingdom
Employment Type
Permanent
Salary
£60000 - £70000/annum Bonus, Medical Care
credentials Do You Have What It Takes? 3-6 years of hands-on experience in a similar role, with a strong emphasis on systems engineering, automation, and service reliability Proficient in at least one programming language such as Python, Go, Java, or C#, along with scripting skills … PowerShell Solid grasp of cloud platforms like AWS, including an understanding of how core services like EC2, ECS, Lambda, and DynamoDB operate under reliability constraints Practical experience using infrastructure-as-code tools like CloudFormation or Terraform In-depth knowledge of CI/CD principles and hands-on experience with ...

Azure SRE Engineer: Reliability, Automation & Scaling

Hiring Organisation
Jobleads-UK
Location
United Kingdom
Redcentric plc is seeking a Site Reliability Engineer to join the Cloud Services Group. You will design and operate Azure-based platforms with a focus on availability, reliability, and scalability, collaborating with development and operations teams. You will implement automated tooling, CI/CD pipelines, IaC with ...

Azure SRE Engineer

Hiring Organisation
Teksystems
Location
Yorkshire, United Kingdom
Employment Type
Contract, Work From Home
seeking a permanent Azure Site Reliability Engineer (SRE) to join our growing development team. The successful candidate will be instrumental in building, improving, and evolving our current Azure technology stack. This is a remote role with the expectation of occasional office visits for team collaboration and personal development. … development team to enhance and maintain our Azure infrastructure. Implement and manage Infrastructure as Code (IaC) and CI/CD pipelines. Focus on typical SRE tasks such as monitoring, scaling, and continuous improvement of our systems. Handle some service-related queries (this is not a 24/7 on-call ...

Azure SRE Engineer

Hiring Organisation
17918
Location
London, United Kingdom
seeking a permanent Azure Site Reliability Engineer (SRE) to join our growing development team. The successful candidate will be instrumental in building, improving, and evolving our current Azure technology stack. This is a remote role with the expectation of occasional office visits for team collaboration and personal development. … development team to enhance and maintain our Azure infrastructure. Implement and manage Infrastructure as Code (IaC) and CI/CD pipelines. Focus on typical SRE tasks such as monitoring, scaling, and continuous improvement of our systems. Handle some service-related queries (this is not a 24/7 on-call ...

Front-Office SRE Lead: Observability & AI-Driven Reliability

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
JPMorgan Chase & Co. in London is seeking a Lead Site Reliability Engineer to shape next‐gen SRE patterns, observability, and reliability across globally distributed trading systems. You will partner with front‐office traders, contribute production code (Java/Python/Kotlin), drive incident response, and lead … assisted reliability initiatives while collaborating with infrastructure, cloud, and security teams. #J-18808-Ljbffr ...