101 to 125 of 1,094 Site Reliability Engineering Jobs in the UK

Azure CloudOps Engineer

Location
Greater London, England, United Kingdom
Infrastructure as Code, enhancing observability and AIOps capabilities, and driving automation across both application and infrastructure lifecycles. This role combines Cloud Engineering, DevOps, SRE, and AIOps practices, leveraging automation, AI-assisted operations, and self-healing capabilities to improve platform reliability, operational efficiency, and service availability. Key Responsibilities: Design … cloud and hybrid environments. Essential Skills & Experience 3+ years' experience in Cloud Engineering, DevOps, Platform Engineering, Site Reliability Engineering (SRE), or Cloud Operations roles. Strong experience with Microsoft Azure services and cloud infrastructure environments. Strong experience with Terraform is preferred. Experience with Helm, CloudFormation ...

Site Reliability Engineer

Hiring Organisation
Proactive Appointments
Location
Gloucester, Gloucestershire, United Kingdom
Employment Type
Contract
Contract Rate
GBP 500 - 580 Daily
Site Reliability Engineer - DV Cleared Our client is urgently looking for an experienced Site Reliability Engineer to join their team on a contract basis, initially for 6 months with a view to extend. Please note, the role is OUTSIDE of IR35. You must hold live UKIC … DV. The role is on-site 4 days per week in Gloucester. Site Reliability Engineer - Key Skills: Must hold a live UKIC DV Modern configuration management tools (such as Ansible, Chef or similar) Experience working with Terraform Docker containers & container orchestration tools (such as Kubernetes, OpenShift ...

Site Reliability Engineer

Hiring Organisation
Proactive Appointments
Location
Gloucester, Gloucestershire, United Kingdom
Employment Type
Full-Time
Salary
£500.00 - £580.00 per day
Site Reliability Engineer – DV Cleared Our client is urgently looking for an experienced Site Reliability Engineer to join their team on a contract basis, initially for 6 months with a view to extend. Please note, the role is OUTSIDE of IR35. You must hold live UKIC … DV. The role is on-site 4 days per week in Gloucester. Site Reliability Engineer – Key Skills: Must hold a live UKIC DV Modern configuration management tools (such as Ansible, Chef or similar) Experience working with Terraform Docker containers & container orchestration tools (such as Kubernetes, OpenShift ...

Senior Lead Site Reliability / DevOps Engineer

Hiring Organisation
JP Morgan Chase
Location
Glasgow, UK
Employment Type
Full-time
integral part of an agile team that's constantly pushing the envelope to enhance, build, and deliver top-notch reliability and observability for our most critical platforms. As a Senior Lead Site Reliability/DevOps Engineer at JPMorgan Chase within the Commercial & Investment Bank … Drive significant business impact through your capabilities and contributions, and apply deep technical expertise and problem-solving methodologies to tackle a diverse array of reliability, observability, and performance challenges that span multiple technologies and applications. Job responsibilitiesRegularly provides technical guidance and direction on site reliability practices ...

Junior Software Engineer (with SRE Influences)

Location
Greater London, England, United Kingdom
Junior Software Engineer (with SRE Influences) About Fogsphere Fogsphere is a London‐based innovator focused on transforming workplace and urban safety through advanced AI, Computer Vision, and Industrial IoT. Built on a principled “Edge‐to‐Fog‐to‐Cloud” architecture, our platform turns passive CCTV cameras and sensors into proactive hazard … reliability of Fogsphere’s AI platform . This role blends software engineering with elements of site reliability engineering (SRE) — giving you exposure not only to building robust software but also to ensuring its smooth, automated, and scalable deployment in production environments. We are looking ...

Senior Site Reliability Engineer

Location
Greater London, England, United Kingdom
Senior Site Reliability EngineerApplylocations: London (82)time type: Full timeposted on: Posted Todayjob requisition id: JR101516**Senior Site Reliability Engineer (SRE) - GCP/Kubernetes****About the Role**We are seeking an experienced and highly motivated Senior Site Reliability Engineer (SRE) to join our small … agile engineering team. This role offers the unique opportunity to drive the reliability, scalability, and performance of our core platform with a high degree of autonomy and ownership.The successful candidate will split their time between providing expert operational support for our critical systems and leading exciting new infrastructure ...

Site Reliability Engineer

Hiring Organisation
Quant Capital
Location
London, UK
Employment Type
Full-time
Site Reliability Engineer – FintechQuant Capital is urgently looking for a Site Reliability Engineer to join or well known Fintech50 client who produces software disrupting the wealth management market. My client is a market leading SAAS provider to financial advisory business nationwide. They are currently in growth … stable solutions. Familiarity with development languages, such as .NET, Java or PythonRedisDocker/KubernetesDatabase experiencesThis role suits a senior Engineer from a DevOps or SRE background who is real technologist interested in the latets tooling and technologies that support software development and infrasturtcure. The firm has a corporate feel ...

Vice President - Site Reliability Engineering (SRE) - The Core Engineering - Birmingham

Hiring Organisation
Goldman Sachs
Location
Birmingham, United Kingdom
downtime.This role is for software engineers who enjoy solving complex distributed system problems, building tools and platforms that make teams more effective, and championing SRE principles (such as SLOs, error budgets, and blameless post-mortems) across a large engineering organization.Key ResponsibilitiesPartner with engineering leadership to establish service level … holistically and understand how individual components interact under load.Strong interpersonal skills to collaborate with product developers, influence architectural decisions, prioritize toil reduction, and drive SRE adoption without direct authority.Ability to translate complex technical issues into clear, actionable insights for both technical and non-technical stakeholders.Highly motivated, pro-active and capable ...

principal engineer- international technology & Starbucks digital solutions

Location
Greater London, England, United Kingdom
technical excellence across EMEA while aligning to global technology strategy and leading the Starbucks Digital Solutions technical direction for International markets. It will set engineering direction, raise standards and guide decisions across internally developed and third-party platforms that matter most to our business, customers, partners, baristas and shareholders.As … Management organisations in a product-led operating model.• Knowledge of modern engineering practices including Platform Engineering, Site Reliability Engineering (SRE), AI-assisted development and Developer Experience (DevEx).What else should you know?• We have a flexible working policy. Meaning 50% of the time we collaborate ...

Site Reliability Engineer, Studios

Location
Uxbridge, England, United Kingdom
rotations, to support live operations and critical systems. Occasional travel may be required depending on project and client needs. IMG is looking for a Site Reliability Engineer to help design, build, operate, and continuously improve resilient, secure, and highly available platforms that underpin our digital, cloud, and broadcast … adjacent services. This role is suited to someone who combines strong infrastructure and software engineering capability with an operational mindset, and who can help embed reliability engineering practices across systems that support live, business‐critical environments. The successful candidate will play a key role in improving service ...

Site Reliability Engineer

Location
West of England, England, United Kingdom
Bristol office or London. Are you ready to make your mark? About the Role We are looking for a Site Reliability Engineer (SRE) to provide confidence to our customers that our systems will be available at all times. As we continue the global expansion of our product deployments … this role will help strengthen our core SRE function and support the integration of Voice/Unified Communications (UC) services with AI-powered capabilities. You will work across production infrastructure, real-time communications, and AI-dependent services to improve reliability, scalability, observability, and customer experience while collaborating with global ...

Site Reliability Engineer - Fintech

Hiring Organisation
Quant Capital
Location
London, UK
Employment Type
Full-time
Site Reliability Engineer – FintechQuant Capital is urgently looking for a Site Reliability Engineer to join or well known Fintech50 client who produces software disrupting the wealth management market. My client is a market leading SAAS provider to financial advisory business nationwide. They are currently in growth … with development languages, such as .NET, Java or Python·Redis·Docker/Kubernetes·Database experiencesThis role suits a senior Engineer from a DevOps or SRE background who is real technologist interested in the latets tooling and technologies that support software development and infrasturtcure. The firm has a corporate feel ...

Site Reliability Engineer - Fintech / Linux

Hiring Organisation
Quant Capital
Location
London, UK
Employment Type
Full-time
Site Reliability Engineer – Fintech/Linux Site Reliability Engineer – Fintech/Linux85,000 Plus BonusQuant Capital is urgently looking for a Site Reliability Engineer to join our high profile client. Our client is a major global financial exchange, driven by technology. They … trading environment for their clients. They have grown massively and recently were voted in the top 50 fintech firms globally. Day to Day the Site Reliability Engineer will: Analyzing and optimizing trading platform Monitoring development activities, change management tickets Monitoring U.S. production, disaster recovery, and certification systems ...

Site Reliability Engineer

Location
Slough, England, United Kingdom
Site Reliability Engineer (SRE) DevSecOps | Cloud Engineering | Observability | Production Environments | London SR2 is supporting a major 3-year programme and looking for an experienced Site Reliability Engineer (SRE) to join the Production Engineering team. This function underpins the reliability, security, and performance … likely) IR35: Inside Location: London twice a week (hybrid model) Clearance: SC level may be required depending on deployment If you’re an experienced SRE who thrives on building reliable, secure, and cost-efficient production systems. #J-18808-Ljbffr ...

Site Reliability Engineer, Studios

Hiring Organisation
iMG world
Location
London, United Kingdom
call rotations, to support live operations and critical systems.Occasional travel may be required depending on project and client needs.IMG is looking for a Site Reliability Engineer to help design, build, operate, and continuously improve resilient, secure, and highly available platforms that underpin our digital, cloud, and broadcast-adjacent … services. This role is suited to someone who combines strong infrastructure and software engineering capability with an operational mindset, and who can help embed reliability engineering practices across systems that support live, business-critical environments.The successful candidate will play a key role in improving service reliability ...

Site Reliability Engineer, Studios

Location
Greater London, England, United Kingdom
rotations, to support live operations and critical systems. Occasional travel may be required depending on project and client needs. IMG is looking for a Site Reliability Engineer to help design, build, operate, and continuously improve resilient, secure, and highly available platforms that underpin our digital, cloud, and broadcast … adjacent services. This role is suited to someone who combines strong infrastructure and software engineering capability with an operational mindset, and who can help embed reliability engineering practices across systems that support live, business-critical environments. The successful candidate will play a key role in improving service ...

Production Engineering Manager

Location
City of Westminster, England, United Kingdom
Meta is seeking a Production Engineering Manager to lead a team responsible for the reliability, scalability, and operational excellence of Meta's production infrastructure and services. In this role, you will manage a team of production engineers who own the full lifecycle of systems — from capacity planning … performance optimization to incident response and automation. You will drive technical strategy, champion AI-augmented workflows, and partner closely with software engineering, infrastructure, and product teams to ensure Meta's services operate at global scale with high availability and efficiency.Production Engineering Manager Responsibilities:Manage a team of production ...

Senior DevOps Engineer - AWS - Manchester

Hiring Organisation
Circle Group
Location
Manchester, North West, United Kingdom
Employment Type
Permanent
Salary
£70,000
also assist with CloudOps activities. Are you an experienced IT professional with a strong background in DevOps and Site Reliability Engineering (SRE)? Are you passionate about working with cutting-edge technologies, driving agile methodologies, and implementing CI/CD practices? Do you have knowledge of infrastructure … code? Experience required: - Solid experience in a similar role, working on DevOps or SRE initiatives within complex IT environments with Software Engineering - AWS environment - Proficiency in DevOps practices and related technologies, such as CI/CD pipelines & infrastructure as code tools such as Terraform, Ansible, Puppet or Bicep. - Strong ...

Jobshare - Sr Lead Software Engineer - Site Reliability Engineer, Python & Infrastructure management - Part time/Jobshare

Location
Greater London, England, United Kingdom
domains, and advise others on the technical and business issues facing them. You will will set the vision, strategy, and operating model for our SRE transformation - enabling our business-aligned support teams to deliver higher reliability, stronger resilience, and a measurably better end-user experience across the board. … responsibilities Defines the SRE vision, north-star outcomes, and multi-year roadmap for the Production Management team, aligned to both CIB and JPM Global Technology priorities. Establishes the SRE operating model across global regions (ways of working, intake, prioritization, engagement with engineering teams and production support). Partners with ...

Software Engineer, SRE

Location
United Kingdom
Site Reliability Engineer, you will enhance system reliability, observability and performance through a strong engineering approach and assist with incident resolution and best practices. Full-time Closes 30/09/2026 You will have strong software engineering skills, approaching system reliability and observability … policy. Preferred Skills and Experience Knowledge and experience of modern software development techniques and lifecycles. Excellent knowledge of Site Reliability Engineering (SRE) principles, including the creation and management of effective Service Level Indicators (SLI's) and Service Level Objectives (SLO's) for reliability and customer satisfaction. ...

Software Engineer, SRE

Location
Manchester, England, United Kingdom
Site Reliability Engineer, you will enhance system reliability, observability and performance through a strong engineering approach and assist with incident resolution and best practices. Full-time Closes 30/09/2026 You will have strong software engineering skills, approaching system reliability and observability … policy. Preferred Skills and Experience Knowledge and experience of modern software development techniques and lifecycles. Excellent knowledge of Site Reliability Engineering (SRE) principles, including the creation and management of effective Service Level Indicators (SLI's) and Service Level Objectives (SLO's) for reliability and customer satisfaction. ...

Site Reliability Engineer

Location
Belfast City District, Northern Ireland, United Kingdom
respect and fairness. If you're ready to push boundaries and challenge the status quo in security, we want to hear from you. Site Reliability Engineer Summary: Our growing technology company is seeking an experienced Site Reliability Engineer with deep Azure expertise to help maintain … availability, performance, and reliability of our critical SaaS applications. In this role, you will own and drive automation, monitoring, incident response, and infrastructure improvements across our multi-cloud, multi-region environment, working closely with senior engineering and cross-functional teams. What You'll Do: Own the availability ...

Software Engineer, GPU Infrastructure- ChatGPT Engineering

Location
Greater London, England, United Kingdom
About the Team ChatGPT Engineering builds and operates the compute platform powering one of the world's largest AI products. Every ChatGPT conversation relies on massive GPU clusters serving inference workloads with high reliability, efficiency, and performance. As our GPU fleet continues to grow, we're investing … production infrastructure, preferably GPU clusters or other compute-intensive distributed systems. Have a background in Production Engineering, Site Reliability Engineering (SRE), Infrastructure Engineering, or Platform Engineering. Have built software that automates operational workflows rather than relying on manual processes. Have experience with Kubernetes, Linux systems ...

Lead Product Manager AIOPs

Hiring Organisation
S&P Global
Location
London, UK
Employment Type
Full-time
responsible for S&P Global's enterprise AIOps platform and strategy, driving the modernization of IT Operations and Site Reliability Engineering (SRE) through intelligent observability, event intelligence, automation, and AI-driven insights. DTS Platform & Tools – Service Enablement: We serve as thought leaders in AIOps, partnering across … Operations, SRE, engineering, infrastructure, service management, and application teams to solve enterprise operational challenges. Our mission is to improve reliability, reduce operational complexity, optimize technology investments, and enable more proactive and resilient technology operations by applying AI.Responsibilities and Impact: Own and execute the AIOps product roadmap, aligning priorities ...

Lead Product Manager AIOPs

Location
Greater London, England, United Kingdom
responsible for S&P Global's enterprise AIOps platform and strategy, driving the modernization of IT Operations and Site Reliability Engineering (SRE) through intelligent observability, event intelligence, automation, and AI-driven insights. DTS Platform & Tools – Service Enablement: We serve as thought leaders in AIOps, partnering across … Operations, SRE, engineering, infrastructure, service management, and application teams to solve enterprise operational challenges. Our mission is to improve reliability, reduce operational complexity, optimize technology investments, and enable more proactive and resilient technology operations by applying AI. Responsibilities and Impact: Own and execute the AIOps product roadmap, aligning ...