151 to 175 of 792 Site Reliability Engineering Jobs in England

Sr Site Reliability Engineer I

Location
Greater London, England, United Kingdom
company where you matter. Your Impact Join us in shaping the future of infrastructure automation for mission‐critical law enforcement systems. As a Senior Site Reliability Engineer, you will drive the development of a next‐generation infrastructure provisioning and automation platform. This platform empowers engineering teams … background, fluency in languages such as Go or Python, and deep experience designing and operating cloud platforms, who is motivated to improve developer velocity, reliability, and platform resilience. Some roles may also require legal eligibility to work in a firearms environment. Work Location: This role is based ...

Site Reliability / Infrastructure Engineer - Defence & Government

Hiring Organisation
JLA Resourcing Ltd
Location
London, United Kingdom
Employment Type
Permanent, Work From Home
foundation who enjoys building systems, solving complex production problems and improving infrastructure through code and automation. This would particularly suit someone from an SRE, Software Engineering or Infrastructure Engineering background who wants to work on technically challenging, high-impact projects. The Role You'll work across software … purely operational DevOps experience. Ideally, you'll have: Active UK SC clearance - this is a key requirement Around 1-3+ years' experience within SRE, Infrastructure, Platform or Software Engineering A Computer Science degree or similarly strong technical/software engineering background Hands-on Kubernetes experience AWS/ ...

Site Reliability Engineer - Frontend

Hiring Organisation
Capital On Tap
Location
London, United Kingdom
Salary
£ 80 K
just getting started! ðLondon, Old Street | ð 2 Days in OfficeSRE at Capital On Tap ðAt Capital On Tap, we run a hybrid embedded SRE model - We aim to work closely with the teams to provide them the best support. As a Site Reliability Engineer (SRE) you will … apply.Interview process ðFirst stage: 30 minute intro and values call with Talent PartnerSecond stage: 60 minute CV overview and technical chat with the SRE team lead and the SRE & Platform Engineering ManagerThird stage: 75 minute technical exercise & questions with the SRE lead Final stage: 30 minute chat with ...

Devops SRE

Location
Greater London, England, United Kingdom
Cloud Engineering team is seeking a seasoned and passionate Senior Cloud Engineer with deep hands‐on development and cloud engineering expertise. In this role, you will serve as a key technical contributor within a cloud‐focused engineering team, working on one of the Group’s flagship initiatives … best practices and business goals. Required Skills & Experience Core Cloud & DevOps Competencies Extensive experience in DevOps or Site Reliability Engineering (SRE) roles across consumer or SaaS environments. Strong expertise in deploying and managing production‐grade Kubernetes clusters and containerised services. Hands‐on experience with Kubernetes ...

Platform Engineer

Location
City Of London, England, United Kingdom
Platform Engineer Department: Technology Employment Type: Permanent - Full Time Location: London Reporting To: Segun Ikuesan Description This is a hands‐on engineering role within the Platform Engineering team, which forms part of Technology Operations. Platform Engineering is responsible for building and operating the infrastructure, platforms and developer … tooling that enable our engineering and quantitative research teams to deliver software reliably, securely and at scale. The role will contribute to the design, build, automation and operation of a hybrid production platform across AWS and on‐premises environments, with a particular focus on the HashiCorp platform, including Nomad ...

Venue & Studio Deployments System Engineer, Event Productions

Hiring Organisation
Amazon
Location
London, United Kingdom
Salary
£ 80 K
Bachelor's degree in Systems Engineering, Computer Science, or related field or relevant work experience- Experience in site reliability engineering (SRE), systems engineering, systems administration, DevOps, security administration, or network administration- Experience working with Linux- Experience in systems engineering- Experience ...

Staff Software Engineer, AI Reliability Engineering

Hiring Organisation
Humanloop
Location
London, United Kingdom
Salary
> £ 150 K
beneficial AI systems.About the RoleClaude has your back. AIRE has Claude's. Help us keep Claude reliable for everyone who depends on it.AIRE (AI Reliability Engineering) partners with teams across Anthropic to improve reliability across our most critical serving paths -- every hop from the SDK through … strength comes from people who've built product stacks, scaled databases, run massive distributed systems, and everything in between.Strong candidates may alsoHave been an SRE, Production Engineer, or in similar reliability-focused roles on large scale systemsHave experience operating large-scale model serving or training infrastructure (>1000 GPUs).Have ...

Senior Observability Engineer

Location
City Of London, England, United Kingdom
Title Senior Observability Engineer Job Description Senior Observability Engineer Location London Employment type Permanent, Full Time Reporting into Senior Engineering Manager - SRE and Observability About IG Group IG Group (LSE: IGG) is a leading global fintech company, established in 1974 and headquartered in London. As a constituent … teams to turn telemetry data into better, faster software. About the team This role sits within the Observability team, part of IG's broader SRE and Platform Engineering function. The team is responsible for the tools, platforms, and standards that enable engineering teams across IG to understand ...

Head of Cyber, Platforms & IT

Location
Chester, England, United Kingdom
enforce secure‐by‐design principles, including cybersecurity standards, cloud architecture guardrails and operational controls Lead DevOps and Site Reliability Engineering (SRE) maturity, embedding monitoring, observability, automated testing and structured incident response Drive adoption of automation and AI‐enabled tooling to improve anomaly detection, incident management, vulnerability management … , cybersecurity or reliability teams in complex SaaS environments Deep expertise in cloud‐native architecture (Azure, AWS or equivalent) Strong understanding of DevOps, SRE principles and SaaS production operations Experience implementing secure‐by‐design frameworks and managing cybersecurity governance Experience owning uptime, reliability and systemic operational performance Strong ...

Senior Observability Engineer

Location
Greater London, England, United Kingdom
Title**Senior Observability Engineer**Job Description****Senior Observability Engineer****Location: London****Employment type:** Permanent, Full Time**Reporting into:** Senior Engineering Manager – SRE and Observability**About IG Group**IG Group (LSE: IGG) is a leading global fintech company, established in 1974 and headquartered in London. As a constituent … teams to turn telemetry data into better, faster software.**About the team**This role sits within the Observability team, part of IG’s broader SRE and Platform Engineering function. The team is responsible for the tools, platforms, and standards that enable engineering teams across IG to understand ...

Site Reliability Engineer - NS London

Hiring Organisation
BAE SYSTEMS
Location
London, United Kingdom
Salary
£ 70 K
maintained. This role blends operational product support with software engineering to create applications to understand the overall health of our systems. The SRE team sits within a wider programme at the core of the customer mission.The role holder:As an SRE, fundamentally you will be doing work that … human labour, with the objective of limiting traditional manual operations work (incident tickets, on-call etc.) to no more than half of the SRE team's time (and aiming for considerably less). You will have an enthusiasm to learn and experiment, to develop tools to understand application health ...

AWS Cloud Architect, Technology Consulting- London, Leeds, Manchester or Newcastle

Hiring Organisation
Momentum Worldwide
Location
London, United Kingdom
Salary
£ 70 K
security and automation technologies. You will help design scalable, secure and resilient cloud solutions for our clients, enabling digital transformation through modern architecture, platform engineering and DevOps practices.You should bring practical experience of cloud technologies and infrastructure modernisation, alongside an understanding of how AI-enabled tooling and modern engineering … capabilities- Strong understanding of cloud networking, identity, security and resilience patterns- Experience supporting cloud engineering, DevOps or Site Reliability Engineering (SRE) teams- Cloud architecture certifications in additional public cloud platforms (Azure or GCP)- Experience using AI or agentic techniques within the role through GitHub Copilot, Claude ...

Systems Engineer, Cryptography, Access and Identity Services

Hiring Organisation
AmazonWebServices
Location
London, United Kingdom
Salary
£ 80 K
availability environment, building and operating critical Cryptography, Access and Identity services for our customers. This exciting role is designed for someone with a strong engineering background and a passion for driving efficiency, quality, and process improvements within our service operations. As a Systems Engineer at Amazon you will utilize … supported in the workplace and at home, there’s nothing we can’t achieve. Basic qualifications- Experience in site reliability engineering (SRE), systems engineering, systems administration, DevOps, security administration, or network administration- Experience working with Linux- Experience in systems engineering- Experience ...

VP, SRE: Architect Resilient, Scalable Systems

Location
Birmingham, England, United Kingdom
Goldman Sachs is seeking a Vice President in Site Reliability Engineering (SRE) within Core Engineering at the Birmingham location. You will lead reliability efforts across distributed systems, driving SLOs, observability, and incident response while shaping scalable, automated platforms. This role emphasizes design reviews, reliability ...

ML Ops Engineer

Hiring Organisation
Anaplan
Location
London, United Kingdom
Salary
£ 80 K
welcome; join us and let’s build what’s next - together!Role OverviewWe are seeking a ML Ops Engineer to join our Platform Engineering team at Anaplan. In this role, you will design, scale, and maintain high-performance MLOps and LLMOps infrastructure supporting our cutting-edge AI-infused scenario … observability using tools like Prometheus, Grafana, OpenTelemetry, and Weights & Biases or MLflow.Your SkillsHands-on production experience in DevOps, Site Reliability Engineering (SRE), or Platform Engineering, with some experience dedicated to AI/ML infrastructure.Proven track record of deploying, scaling, and operationalising machine learning models and LLMs ...

ML Ops Engineer

Location
Greater London, England, United Kingdom
welcome; join us and let’s build what’s next - together! Role Overview We are seeking a ML Ops Engineer to join our Platform Engineering team at Anaplan. In this role, you will design, scale, and maintain high-performance MLOps and LLMOps infrastructure supporting our cutting-edge AI-infused … tools like Prometheus, Grafana, OpenTelemetry, and Weights & Biases or MLflow. Your Skills Hands-on production experience in DevOps, Site Reliability Engineering (SRE), or Platform Engineering, with some experience dedicated to AI/ML infrastructure. Proven track record of deploying, scaling, and operationalising machine learning models ...

Principal AWS Cloud Architect, Technology Consulting- London, Leeds, Manchester or Newcastle

Hiring Organisation
Momentum Worldwide
Location
London, United Kingdom
Salary
£ 70 K
cloud solutions that accelerate digital transformation and operational excellence. You will help clients realise the full value of cloud adoption through modern architecture, platform engineering, DevOps practices and infrastructure automation, delivering sustainable outcomes in complex and regulated environments.This is a permanent within our Technology Solutions function. At Credera … models- Strong understanding of cloud networking, identity, security and resilience patterns- Experience supporting cloud engineering, DevOps or Site Reliability Engineering (SRE) transformations- Cloud architecture certifications in additional public cloud platforms (Azure or GCP)- Experience using AI or agentic techniques within the role through GitHub Copilot, Claude ...

Principal AWS Cloud Architect, Technology Consulting- London, Leeds, Manchester or Newcastle

Hiring Organisation
Momentum Worldwide
Location
Manchester, Greater Manchester, United Kingdom
Salary
£ 60 K
cloud solutions that accelerate digital transformation and operational excellence. You will help clients realise the full value of cloud adoption through modern architecture, platform engineering, DevOps practices and infrastructure automation, delivering sustainable outcomes in complex and regulated environments.This is a permanent within our Technology Solutions function. At Credera … models- Strong understanding of cloud networking, identity, security and resilience patterns- Experience supporting cloud engineering, DevOps or Site Reliability Engineering (SRE) transformations- Cloud architecture certifications in additional public cloud platforms (Azure or GCP)- Experience using AI or agentic techniques within the role through GitHub Copilot, Claude ...

Systems Engineer, Database Services

Hiring Organisation
AmazonWebServices
Location
London, United Kingdom
Salary
£ 80 K
supported in the workplace and at home, there’s nothing we can’t achieve. Basic qualifications- Experience in site reliability engineering (SRE), systems engineering, systems administration, DevOps, security administration, or network administration- Experience working with Linux- Experience in any of the following: Python, Java, Perl ...

Site Reliability Engineer

Location
Greater London, England, United Kingdom
Site Reliability Engineer Reports to: Labs Team Lead Looper Insights Remote-first, with regular visits to our Byfleet and Hounslow data centres The company Looper Insights builds analytics products that help the world’s leading media and entertainment companies understand how their content is performing across digital platforms … collaborate closely to produce a valuable service for an industry about which we are all passionate. The role We’re looking for a Site Reliability Engineer to keep the global LooperBox fleet running, the physical backbone behind every piece of data Looper Insights produces. LooperBoxes sit in front ...

Sr. Observability Engineer – Kings Cross, London

Location
Greater London, England, United Kingdom
data for swift root cause identification. Drive post-incident reviews and implement long-term solutions to enhance system resilience.* Collaborate & Influence: Partner with Development, SRE, and Infrastructure leaders to embed observability into the entire technology lifecycle. Influence and drive the adoption of observability best practices across the global organization. Champion … this.**Job Requirements:**Essential Qualifications* Experience: 5-7+ years of hands-on experience in an Observability, Site Reliability Engineering (SRE), or DevOps role, with a proven track record of leading complex projects.* Technical Leadership: Demonstrated experience in architecting and designing large-scale monitoring and observability solutions. ...

Product Associate - SRE Team - Chase UK

Hiring Organisation
JP Morgan Chase
Location
London, United Kingdom
Salary
£ 100 K
oriented and possess an interest in the financial sector and focus on addressing our customer needs. We work in teams focused on improving the reliability, resilience, observability, and operability of customer-facing digital banking services. We build automation, define measurable reliability practices, reduce operational friction, and partner with … engineering teams to ensure services are designed, delivered, and operated with reliability in mind.Job responsibilitiesSupport the product strategy and delivery of reliability capabilities, including standards, observability, incident practices, automation, and developer experience improvements.Partner with engineers, site reliability engineers, and cross-functional teams to understand problems ...

Senior Site Reliability Engineer

Hiring Organisation
VIQU IT Recruitment
Location
Milton Keynes, Buckinghamshire, United Kingdom
Employment Type
Permanent
Salary
GBP 75,000 Annual
Senior Site Reliability Engineer Up to £75,000 plus bonus and on call allowance Milton Keynes (2 days on site a week) VIQU have partnered with a well-established B2B SaaS company who are going through a significant platform transformation. and so are hiring for a Senior … Site Reliability Engineer to build stability, respond to live incidents, and assist with system upkeep click apply for full job details ...

Site Reliability Engineer with Python

Hiring Organisation
Nexus Jobs
Location
London, United Kingdom
Salary
£ 80 K
000Sector: I.T. & CommunicationsJob Type: PermanentWork Hours: Full TimeContact: Jas GujralEmail: cv@nexusjobs.comTelephone: 020 7488 6900Apply for this job nowJob DescriptionSite Reliability Engineer with PythonOur Client looking to bring on a site reliability engineer to help deploy, manage, troubleshoot, and enhance our complex cloud-based set of internal … variety of users across our wide-ranging organization.You will have at least 7 to 10 years hands-on expertise working as a Site Reliability Engineer.You will work closely with IT, product, and engineering to extend and maintain this set of tools and services and to help debug ...

Principal Site Reliability Engineer, Infrastructure Observability

Location
Greater London, England, United Kingdom
toolchain and systems, code build and deployment, incident response, and 24x7 monitoring and support. The candidate will also have extensive experience operating within a SRE function within a complex, distributed environment. They will have a demonstrated ability to work horizontally and vertically within an organization with diverse partners and sponsor … learning through blameless post‐mortems to improve the shared goal of reliability across services Transform operations teams by facilitating internal change to adopt SRE standard methodologies across the organization and driving strategic growth in this area within Global Technology Analyzes incidents impacting technology availability for high‐level trends across ...