401 to 425 of 1,104 Site Reliability Engineering Jobs in the UK

Principal Database Platform Engineer

Location
Sheffield, England, United Kingdom
Month (Extendable)****We are currently seeking an experienced Principal Engineer whose main area of expertise is Database but is complemented with strong engineering skills.****As a Principal Engineer you will be at the forefront of our technology, influencing the strategy and direction of our services. This is a leadership … bank’s shared database solutions (DBaaS/PaaS), followed by their implementation.• Stay informed about industry trends, emerging technologies and advancements in engineering practices, evaluating and recommending innovative solutions as appropriate.• Support the Platform Lead and identify solutions to engineering gaps/challenges.• Facilitate development of cross-functional ...

Site Reliability Engineer - Data Platform & Scale

Location
East Hagbourne, England, United Kingdom
Open Cosmos Ltd is seeking a Site Reliability Engineer to enhance the reliability and scalability of our data platform. You will be responsible for monitoring, troubleshooting, and improving our systems while collaborating with engineering teams to design robust infrastructure. The ideal candidate will have expertise ...

Senior Sales Engineer London, England, United Kingdom

Location
Greater London, England, United Kingdom
relationship that will provide continuous value to our customers. Finally, you will have the opportunity to work cross-functionally with our Product Management and Engineering teams to share your knowledge and experiences to ultimately improve our business and our customers’ success. We seek talent who wants to leverage their … technical validations during the Proof of Value phase Be successful working with all levels of an organization, from executives down to individual developers and Site Reliability Engineers Deliver product and technical demonstrations of the Sumo Logic service Work cross functionally with Product Management and Engineering to improve ...

Site Reliability Engineer, London

Location
Greater London, England, United Kingdom
BIG. Operating at our scale, across multiple geographically dispersed data centers and servicing hundreds of millions of users presents unique challenges. As an SRE at Apple, you'll need to solve these problems using data, teamwork, and your own expertise. SREs at Apple own the full infrastructure stack; from device … close partnership with our development teams and aim to design & build new services together. We're passionate about software and automation in SRE and develop a variety of tooling and infrastructure. Our services run on mixed & hybrid platforms. Responsibilities Create outstanding customer experience, and help developers write better code faster ...

DevOps Engineer

Location
Greater London, England, United Kingdom
institutions, and market data vendors. We provide robust technical expertise and adopt a pragmatic, flexible approach to delivering and supporting business‐critical IT systems. Engineering and technology is at the heart of everything we do, and we're looking for a DevOps Engineer to join our Technology Operations team … availability, and capacity. Investigate and resolve platform, infrastructure, and application‐related incidents. Work closely with Development, DevOps, and Application Support teams to improve system reliability and deployment processes. Collaborating with developers work closely with software engineers and operations teams to understand business requirements and translate into technical solutions Contribute ...

DevOps Engineer

Location
Greater London, England, United Kingdom
troubleshooting complex application and infrastructure issues. Knowledge of networking principles including DNS, load balancing, firewalls, and TLS/SSL. Experience working within DevOps or Site Reliability Engineering practices. Experience with Amazon EKS. Exposure to GitHub Actions, GitLab CI, Azure DevOps, or Jenkins. Knowledge of security best practices … availability, and capacity. Investigate and resolve platform, infrastructure, and application‐related incidents. Work closely with Development, DevOps, and Application Support teams to improve system reliability and deployment processes. Collaborating with developers work closely with software engineers and operations teams to understand business requirements and translate into technical solutions Contribute ...

NOC Engineer / SRE

Location
United Kingdom
offer you the ultimate career opportunity that will light a fire within you. So, what’s the role all about? The SRE – NOC role sits at the intersection of traditional Network Operations Center (NOC) responsibilities and engineering‐driven reliability practices . This role focuses on 24/… service reliability, incident response, operational automation, and observability , while actively reducing operational toil through software and automation. Unlike a traditional NOC analyst, an SRE‐NOC is expected to engineer problems away , not just respond to alerts. How will you make an impact? Incident Response & Operations Act as a primary ...

Google Cloud Network DevOps Engineer

Location
Manchester, England, United Kingdom
pivotal role in developing, optimising and operating the networking foundations that underpin our transformation This is a fantastic opportunity to join a forward‐thinking engineering team helping Lloyds Banking Group progress towards our ambition of becoming the UK’s largest FinTech. You’ll bring hands‐on expertise, influence technical … innovative Cloud‐focused DevOps Engineer to join our Google Cloud team. You’ll be an inspiring, hands‐on engineer with experience in DevOps or Site Reliability Engineering and a history of driving the delivery of technical platforms in enterprise environmentsExperience reviewing and providing feedback on peer code ...

Platform Engineer

Location
Greater London, England, United Kingdom
Full-Time Working Pattern: Monday to Friday Start Date: ASAP About the role As a Platform Engineer, you will be responsible for the management, reliability, performance, and continuous improvement of our Kubernetes infrastructure. You’ll play a key role in operating and enhancing our Kubernetes environments, containerised workloads … troubleshooting complex application and infrastructure issues. Knowledge of networking principles including DNS, load balancing, firewalls, and TLS/SSL. Experience working within DevOps or Site Reliability Engineering practices. Experience with Amazon EKS. Exposure to GitHub Actions, GitLab CI, Azure DevOps, or Jenkins. Knowledge of security best practices ...

AI & SRE Consultant

Hiring Organisation
Akkodis
Location
City of London, London, United Kingdom
Employment Type
Permanent
Salary
£88000 - £96000/annum
Platform & Site Reliability Engineering Senior Consultant Cloud Operating Model Transformation | AI, Cloud & Automation Ready to help organisations redefine how they operate in the age of AI? We are partnering with a leading global consulting organisation seeking a Senior Consultant to join a fast-growing Cloud Advisory practice. ...

Senior Backend Engineer (Go), Tenant Scale: Gitaly

Hiring Organisation
Appcast
Location
Remote, UK
including new approaches to scaling Git storage and improving how repository data is managed and served.You’ll balance roadmap delivery, technical improvements, customer and reliability work, and operational support. We value pragmatic engineering, thoughtful trade-offs, continuous improvement, and sustainable ways of working while operating a service that … stakeholders, communicating progress and trade-offs, identifying risks, and keeping execution moving with minimal guidance.Lead technical design for distributed storage, Git repository management, performance, reliability, and scalability problems, using data and benchmarking to guide decisions.Partner with engineers and teams across GitLab, including Git, Infrastructure, Site Reliability Engineering ...

Senior Backend Engineer (Go), Tenant Scale: Gitaly

Hiring Organisation
GitLab
Location
United Kingdom
including new approaches to scaling Git storage and improving how repository data is managed and served.You’ll balance roadmap delivery, technical improvements, customer and reliability work, and operational support. We value pragmatic engineering, thoughtful trade-offs, continuous improvement, and sustainable ways of working while operating a service that … stakeholders, communicating progress and trade-offs, identifying risks, and keeping execution moving with minimal guidance.Lead technical design for distributed storage, Git repository management, performance, reliability, and scalability problems, using data and benchmarking to guide decisions.Partner with engineers and teams across GitLab, including Git, Infrastructure, Site Reliability Engineering ...

Senior DevOps & Cloud SRE Lead (AWS / K8s)

Location
Greater London, England, United Kingdom
About the Role Own our global cloud infrastructure, Kubernetes clusters, infrastructure-as-code (Terraform), zero-downtime deployment pipelines, and 99.99% uptime SRE standards. Key Responsibilities Manage multi-region AWS infrastructure using Terraform and CloudFormation Maintain production EKS/Kubernetes clusters, ingress controllers, and service meshes Build automated CI/… pipelines using GitHub Actions, Docker, and Helm Set up observability stack (Prometheus, Grafana, Datadog) and incident management Requirements 4+ years in DevOps, Site Reliability Engineering, or Cloud Architecture Deep expertise in AWS, Kubernetes, Terraform, Docker, and Linux administration Strong scripting skills in Python, Bash ...

Technical Site Reliability Engineer

Hiring Organisation
Anduril Industries
Location
London, United Kingdom
operate together in future contested multi-domain environments. You'll join a small, multinational team of engineers spanning multiple disciplines such as wargaming, game engineering, HPC simulations, LLM agents and VR environments; all to give our warfighters and researchers the leverage to explore faster, test more ideas, and better … current, and trustworthy. A failed scenario run or a silent regression after a software release costs operators and engineers’ real time.As our founding Site Reliability Engineer, you will design, build, and operate the infrastructure that makes this possible. You'll work at the intersection of hardware, simulation software ...

Data Platform Infra Site Reliability Engineer

Location
Greater London, England, United Kingdom
Selection changes the language of the page/content Data Platform Infra Site Reliability Engineer London, England, United Kingdom Software and Services At Apple, we believe that innovation flourishes in an environment where ideas are challenged, collaboration is encouraged, and technology is pushed to its limits. This environment … inspire innovation in everything we do. Imagine what you could accomplish here! Join Apple and help us make the world a better place.As an SRE on our team, you'll own the reliability, performance, and scale of the distributed storage and data platform systems that power Apple's services. ...

Cloud-Scale Site Reliability Engineer: Automation & Resilience

Location
Birmingham, England, United Kingdom
prominent technology firm in Birmingham is seeking a Site Reliability Engineer tasked with ensuring the reliability, performance, and scalability of critical enterprise systems. This role combines software and systems engineering to enhance automation and support cloud transformations. The ideal candidate will have expertise in Unix, Windows ...

Site Reliability Engineer: Automation & Observability

Location
Greater London, England, United Kingdom
Apple Inc. is seeking a Site Reliability Engineer in London to join the Apple Services Engineering team. You will help sustain large-scale services powering the App Store, Apple Music, TV, Podcasts and Books for users worldwide. As an SRE, you’ll work with Linux, open source tools and internal software to manage configuration, deployment, logging and monitoring. You’ll collaborate with development teams ...

Site Reliability Engineer - Live Ops & Cloud Resilience

Location
Greater London, England, United Kingdom
seeking an experienced Site Reliability Engineer to design, build, and operate resilient, secure platforms underpinning our digital and live operations. You’ll focus on reliability, observability, automation, and disaster recovery across hybrid environments, collaborating with engineering, operations, and project stakeholders. The role emphasizes improving service availability ...

Head of Production Management- J.P. Morgan Personal Investing

Hiring Organisation
Hackajob Ltd
Location
South West London, London, United Kingdom
Employment Type
Permanent
powered solutions and intelligent automation to reduce manual intervention, fast-track resolution, and continuously improve operational efficiency. Champion an automation-first, shift-left SRE cultureleveraging shared tooling and automation to ensure consistency, reduce duplication, and maintain alignment with firmwide standards. Oversee capacity management and planning, ensuring infrastructure scales to meet … management standards change, incident, capacity, and automation across multiple engineering teams operating in a you-build-it-you-run-it model, underpinned by SRE principles and disaster recovery planning. Composure, decisiveness, and authority during incidents, vendor failure, or regulatory escalation, with a proven ability to protect business lines under ...

Head of Production Management- J.P. Morgan Personal Investing

Location
Westminster, West End, United Kingdom
powered solutions and intelligent automation to reduce manual intervention, fast-track resolution, and continuously improve operational efficiency. Champion an automation-first, shift-left SRE culture leveraging shared tooling and automation to ensure consistency, reduce duplication, and maintain alignment with firmwide standards. Oversee capacity management and planning, ensuring infrastructure scales … management standards change, incident, capacity, and automation across multiple engineering teams operating in a you-build-it-you-run-it model, underpinned by SRE principles and disaster recovery planning. Composure, decisiveness, and authority during incidents, vendor failure, or regulatory escalation, with a proven ability to protect business lines under ...

Site Reliability Engineer I: Cloud-Native & Observability

Location
Greater London, England, United Kingdom
Axon is seeking an experienced Site Reliability Engineer for its Real Time Operations in London. You will contribute to building reliable cloud-native services, collaborate with RTO engineering teams, and enable product teams to scale features globally. We value engineers who write clean code, champion reliability, and create self-service tooling for rapid provisioning and incident response. This role is based in London with on-site work expectations. #J-18808-Ljbffr ...

Site Reliability Engineer

Location
Birmingham, England, United Kingdom
Overview Our client is seeking a high-impact Site Reliability Engineer to join a team responsible for ensuring the reliability, performance, and scalability of critical enterprise systems. This role blends software and systems engineering to drive automation, prevent service-impacting incidents, and support transformative cloud initiatives. … work for any US Employer without sponsorship. Benefits & Extras Work on cutting-edge distributed systems and cloud transformations Solve challenging performance, scalability, and reliability problems Collaborate with teams driving automation and monitoring initiatives Exposure to enterprise-scale network and fault-tolerant architectures High-impact role with visibility across technical ...

Manufacturing Site Reliability Engineer — Drive Lasting Improvements

Location
United Kingdom
Whitworths is seeking a Site Reliability Engineer to drive long-term reliability of manufacturing assets. You’ll lead RCA, analyze downtime data, and develop practical, lasting solutions that reduce repeat failures and improve maintenance strategies. You’ll work with Engineering and Operations to implement medium ...

Senior Lead SRE - Reliability & Observability Leader

Location
Glasgow, Scotland, United Kingdom
JPMorgan Chase in the United Kingdom is seeking a Senior Lead Site Reliability Engineer … join an agile team focused on reliability, observability, and performance across critical platforms. You will mentor engineers, lead incident response, and shape SRE strategy while delivering scalable, secure production systems. The role demands deep expertise in cloud, automation, OpenTelemetry, and instrumentation, with a track record of improving service levels ...

Lead Site Reliability Engineer - Observability & Resilience

Location
Glasgow, Scotland, United Kingdom
JPMorgan Chase & Co. seeks a Lead Site Reliability Engineer to define the future of reliability for a global firm. You will lead critical resiliency design reviews, break complex problems into actionable work, and mentor engineers across large-scale OpenTelemetry pipelines in hybrid environments. You will guide incident ...