326 to 350 of 925 Site Reliability Engineering Jobs in the UK

Remote Senior Azure SRE - Reliability & Automation

Hiring Organisation
Jobleads-UK
Location
United Kingdom
Spectris Group is seeking a Senior Site Reliability Engineer to ensure the reliability, performance, and scalability of Azure-based cloud platforms. You will lead reliability initiatives, strengthen incident response, and drive automation across deployments and runbooks while partnering with engineering, product, and operations teams. The role emphasizes SRE best practices, platform resilience, and operational excellence, with remote UK flexibility and opportunities for professional growth and #J-18808-Ljbffr ...

Lead Infrastructure Engineer

Hiring Organisation
Jobleads-UK
Location
City of Edinburgh, Scotland, United Kingdom
Lead Cloud Engineer within the Agriculture and Rural Economy (ARE) Digital and Data Division, you will play a pivotal role in shaping the cloud engineering capability that underpins critical digital services across the Scottish Government. Our mission is to provide the technical expertise and capacity needed to deliver secure … support sustainable economic growth in agriculture, the food industry and rural communities. You will lead the development and implementation of cloud strategy, architecture and engineering standards across our AWS estate, working closely with Product, Platform, Software Engineering, Data, Site Reliability Engineering and Cyber Security teams. ...

Site Reliability Engineer, K8s (Remote International)

Hiring Organisation
PulsePoint
Location
United Kingdom
Salary
£ 60 K
Build the platform behind PulsePointPulsePoint operates large-scale data and advertising platforms that power business-critical services used every day across the company.Our Platform Engineering team owns the foundation that enables engineering teams to move quickly and safely. We build and operate the infrastructure that supports Kubernetes workloads … networking to Kubernetes, observability and developer experience.The environment supports large-scale Kubernetes workloads, multi-petabyte data systems and business-critical services used across multiple engineering organizations.We're looking for an experienced engineer to help shape its next stage of growth.This is a hands-on role focused on architecture, reliability ...

Site Reliability Engineer-GCP

Hiring Organisation
Wipro
Location
Bristol, Gloucestershire, United Kingdom
Salary
£ 70 K
consulting company focused on building innovative solutions that address clients’ most complex digital transformation needs. Leveraging our holistic portfolio of capabilities in consulting, design, engineering, and operations, we help clients realize their boldest ambitions and build future-ready, sustainable businesses. With over 230,000 employees and business partners across … Hands-on)Kubernetes (Production experience)CI/CD + Terraform (IaC)Observability (Dynatrace preferred/equivalent acceptable with alignment)Good-to-Have Skills:SRE practices (SLO/SLI, incident management, RCA)Python/Bash/scriptingPrometheus, Grafana, ELK, SplunkCloud networking & securityMulti-cloud exposure (AWS/Azure)Banking/financial services ...

Staff Engineering Support Engineer

Hiring Organisation
Dragos
Location
United Kingdom
Salary
£ 70 K
kind of work that matters to you, you are in the right place. About the Role: Dragos is looking for a Staff Engineering Support Engineer to join the Site Reliability and Support team, based in the United Kingdom. This role is a critical anchor for EMEA coverage … global team that handles Tier-3 customer escalation triage and cloud operations for the Dragos platform. You will report to the Engineering SRE & Support Manager and work closely with Customer Experience, Engineering, and Product Development teams across time zones to resolve high-impact issues and keep cloud-hosted ...

Platform & Workplace Engineering Director

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Overview Leading and developing two engineering teams — Cloud Platform (cloud infrastructure, observability, SRE, incident response) and Enterprise Services (IT support, IAM, Okta, Slack and core SaaS) — setting technical direction, hiring and performance standards across both. Delivering the reliability, performance and cost-efficiency of the cloud platform underpinning production … post-incident learning, in line with the broader infrastructure strategy. Executing the roadmap for enterprise services and identity, treating internal IT as an engineering discipline — automation-first, self-service where possible, and measured on employee productivity and security outcomes. Tracking and reporting meaningful operational metrics — availability, MTTR, change failure ...

Senior Infrastructure Engineer

Hiring Organisation
Community Fibre Limited
Location
London, United Kingdom
Salary
£ 80 K
building new servers/systems as well as ensuring that all our existing ones are maintained, reliable and resilient.The role covers backend systems engineering, infrastructure, and site reliability engineering within the Network Technology group.You should be hungry for hands-on experience, have a desire to think … WSGIDeep understanding of key service health metricsKnowledge of current security best practices including encryption and systems hardeningExperience implementing technologies with a focus on operability, reliability, and scalability.Network aptitude. The skill to take a new technology, research it, test it in the lab and deploy it live.Excellent understanding of telecommunications ...

Lead Azure SRE: Reliability, Observability & Automation

Hiring Organisation
Jobleads-UK
Location
Nottingham, England, United Kingdom
Building Society is seeking a Lead Azure Site Reliability Engineer to drive resilience, performance, and availability of Azure platforms. You will champion SRE practices, guide reliability improvements, and lead modern operating standards to support 500+ users and 50+ sites. You will mentor engineers, shape cloud strategy ...

Senior Site Reliability Engineer

Hiring Organisation
Spectrum IT Recruitment
Location
Southampton, Hampshire, United Kingdom
Salary
£ 60 K
annumSouthampton, HampshirePermanentSenior Site Reliability EngineerSouthampton HQ - 2 Times a week in OfficeCloud, SaaS, AWS, The company deliver cutting-edge enterprise software solutions across both cloud and on-premises environments, empowering organisations to enhance customer experiences, maintain regulatory compliance, and proactively fight fraud. The company are trusted by businesses … similar credentialsDo You Have What It Takes 3-6 years of hands-on experience in a similar role, with a strong emphasis on systems engineering, automation, and service reliabilityProficient in at least one programming language such as Python, Go, Java, or C#, along with scripting skills in Bash ...

Senior Site Reliability Engineer - Cloud-Native Leader

Hiring Organisation
Jobleads-UK
Location
Colchester, England, United Kingdom
Source Group International Ltd in the United Kingdom is seeking a Senior Site Reliability Engineer to enhance reliability, availability, and scalability of our SaaS platform. You will blend software engineering with cloud infrastructure, automation, and operational leadership to deliver resilient systems that enable rapid product delivery. … will own production reliability, define SLI/SLO, lead blameless postmortems, and automate toil. #J-18808-Ljbffr ...

Site Reliability Engineer - Scale, Observe, Automate

Hiring Organisation
Jobleads-UK
Location
Watford, England, United Kingdom
Allwyn is seeking a Site Reliability Engineer to support the reliability and performance of our digital services. You will work with production systems, automation, and observability to keep services stable, scalable and well-instrumented across peak events. You will collaborate with senior SREs and engineering teams … drive automation, incident response, and improvement of runbooks, while contributing to platform reliability and user experience signals. #J-18808-Ljbffr ...

Senior Engineer- Platform Engineering

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Databricks, covering environment provisioning, CI/CD integration, and access‐control models. Implement observability frameworks, including monitoring, logging, and alerting, and contribute to SRE practices such as SLIs/SLOs, reliability engineering, and incident management. Embed security and compliance standards into all platform components, ensuring auditability, policy enforcement … developer experience improvements through platform standardisation, self‐service tooling, templates, and AI‐enabled capabilities (e.g. Copilot, intelligent automation). Collaborate with Architecture, Cloud COE, SRE, and engineering teams to deliver consistent and governed platform capabilities across the organisation. Mentor junior engineers and contribute to technical leadership, standards definition ...

DevOps Engineer, Studios

Hiring Organisation
iMG world
Location
London, United Kingdom
Salary
£ 60 K
available infrastructure and delivery pipelines across our digital and broadcast platforms.The successful candidate will play a key role in enabling continuous delivery, improving system reliability, and supporting high-profile clients and live event services, ensuring optimal performance and resilience across all environments.Key Responsibilities and AccountabilitiesDesign, build, and maintain scalable … efficient, reliable software delivery across multiple teams.Automate infrastructure provisioning using Infrastructure as Code tools such as Terraform, CloudFormation, or similar.Monitor system performance, availability, and reliability using observability tools such as Prometheus, Grafana, and ELK stack.Ensure high availability and disaster recovery strategies are in place and tested regularly.Collaborate closely with ...

DevOps Engineer, Studios

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
infrastructure and delivery pipelines across our digital and broadcast platforms. The successful candidate will play a key role in enabling continuous delivery, improving system reliability, and supporting high-profile clients and live event services, ensuring optimal performance and resilience across all environments. Key Responsibilities and Accountabilities Design, build … software delivery across multiple teams. Automate infrastructure provisioning using Infrastructure as Code tools such as Terraform, CloudFormation, or similar. Monitor system performance, availability, and reliability using observability tools such as Prometheus, Grafana, and ELK stack. Ensure high availability and disaster recovery strategies are in place and tested regularly. Collaborate ...

Head of Production Management- J.P. Morgan Personal Investing

Hiring Organisation
JP Morgan Chase
Location
London, United Kingdom
Salary
£ 120 K
powered solutions and intelligent automation to reduce manual intervention, fast-track resolution, and continuously improve operational efficiency. Champion an automation-first, shift-left SRE culture leveraging shared tooling and automation to ensure consistency, reduce duplication, and maintain alignment with firmwide standards. Oversee capacity management and planning, ensuring infrastructure scales … management standards — change, incident, capacity, and automation — across multiple engineering teams operating in a you-build-it-you-run-it model, underpinned by SRE principles and disaster recovery planning. Composure, decisiveness, and authority during incidents, vendor failure, or regulatory escalation, with a proven ability to protect business lines under ...

DevOps Engineer

Hiring Organisation
Capgemini
Location
City of Bristol, United Kingdom
Employment Type
Full Time
customer’s outcomes as we group around activities in agile type sprints. Continue to strengthen and bolster your existing capabilities in platform engineering through a mix of professional training, certifications, and experiences. Collaborate and influence our iterative Capgemini offerings. You can bring your whole self to work. At Capgemini … everyone. Your skills and experience We’d particularly welcome applications from people with experience in Databricks or Splunk A proven track record in platform engineering, particularly with: Azure/AWS/GCP, Terraform, Ansible, CI/CD pipelines, Kubernetes, and VCS (GitHub). Can demonstrate tangible experiences ...

Principal Database Platform Engineer

Hiring Organisation
Jobleads-UK
Location
Sheffield, England, United Kingdom
Month (Extendable)****We are currently seeking an experienced Principal Engineer whose main area of expertise is Database but is complemented with strong engineering skills.****As a Principal Engineer you will be at the forefront of our technology, influencing the strategy and direction of our services. This is a leadership … bank’s shared database solutions (DBaaS/PaaS), followed by their implementation.• Stay informed about industry trends, emerging technologies and advancements in engineering practices, evaluating and recommending innovative solutions as appropriate.• Support the Platform Lead and identify solutions to engineering gaps/challenges.• Facilitate development of cross-functional ...

Senior Platform Engineer

Hiring Organisation
Talent Locker
Location
Farnborough, Hampshire, South East, United Kingdom
Employment Type
Permanent
Salary
£75,000
Senior Platform Engineer - Farnborough, Hybrid - Security Cleared - Up to £79,000 We're looking for a Senior Platform Engineer to join a growing engineering team delivering secure, resilient, and scalable platforms across cloud and on-premise environments. This is an opportunity to work on complex, mission-critical projects where … technology supporting the Defence and National Security sector. Working alongside highly skilled engineers, architects and technical specialists, you'll help shape platform strategy, drive engineering best practices, and deliver robust infrastructure that enables modern software development at scale. What you'll be doing Designing, building and maintaining secure, scalable ...

SRE - Site Reliability Engineer - Observability & Performance

Hiring Organisation
Sanderson Recruitment
Location
Bristol, Somerset, United Kingdom
Employment Type
Contract
Contract Rate
GBP 550 - 600 Daily
SRE - Observability and Performance Up to £600 per day outside IR35 6 month initial contract Bristol - Largely remote I'm currently working with a client who is looking for an SRE to implement and enhance observability across Java applications, middleware and Linux infrastructure using Grafana click apply for full ...

Software Engineer – Production Automation & CI/CD – 72096

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
detects broken builds and identifies the root-cause change. Build solutions that recommend or implement fixes for build and production failures. Monitor production quality, reliability, capacity and latency. Investigate and respond to production regressions, incidents and outages. Reduce recurring operational and on-call work by approximately 80-90% through … automation. Maintain cloud-rendering, build and deployment pipelines. Complete infrastructure and dependency migrations without disrupting downstream CI/CD. Collaborate with runtime, production engineering and operational teams. Essential experience Eight or more years of professional software engineering experience, or equivalent expertise. Proven experience building and operating CI/ ...

Site Reliability Engineer – 24/7 On-Call & Automation Focus

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Systems Digital Intelligence is seeking a Site Reliability Engineer to blend software engineering with operations, automating support and improving system health for a key national security customer. The role focuses on monitoring, incident response, and building tools to automate repetitive tasks, with collaboration across development teams ...

Lead Site Reliability Engineer - Scale & Resilience

Hiring Organisation
Jobleads-UK
Location
Glasgow, Scotland, United Kingdom
JPMorgan Chase & Co. in the United Kingdom seeks a Lead Software Engineer focused on Site Reliability and Operations Excellence to strengthen incident response, drive reliability standards, and guide durable improvements across a complex technology stack. You will own the incident management lifecycle, serve as Incident Commander … major incidents, and lead drills and readiness efforts. The role partners with Engineering, Product, Infrastructure, and Security to deliver measurable reliability #J-18808-Ljbffr ...

Senior Site Reliability Engineer — Hybrid Kubernetes & GitOps

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Cisco’s Webex Engineering Group in London is seeking a Senior Site Reliability Engineer to own the design, deployment, and operation of Kubernetes-based microservices, delivering reliability and scalable deployments in a hybrid work environment. You will drive GitOps workflows with Argo CD, use Helm ...

Sr. Technical Support Engineer - Confluent

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
resolve complex Apache Kafka and Confluent Platform issues. They manage high-priority production cases, lead difficult escalations, build customer trust, and collaborate with Engineering, Product, and other internal teams What You’ll Do In this role you will: Own and manage complex customer cases across production and non-production … manage resolution plans, including priorities, dependencies, risks, mitigations, and timelines. Communicate clearly with customers and set expectations during high-pressure situations. Collaborate with Engineering and Product to identify bugs, improve product quality, and provide field feedback. Create and review technical documentation, knowledge articles, and training material. Mentor and coach ...

Principal AI-Driven Site Reliability Engineer

Hiring Organisation
Jobleads-UK
Location
Glasgow, Scotland, United Kingdom
JPMorgan Chase is seeking a Principal Site Reliability Engineer to help advance reliability across the software lifecycle, shaping infrastructure design and implementing improvements that boost operational efficiency. You will leverage AI-enabled capabilities to analyze incidents, validate recommendations, and ensure data sensitivity is respected while collaborating with ...