26 to 50 of 52 Senior Reliability Engineer Jobs in England

Senior Software Engineer, Reliability & Large-Scale Infra

Location
Greater London, England, United Kingdom
Google is seeking software engineers to join the Dependency Tracing Team within Platform Reliability Engineering to help minimize outages and improve cloud infrastructure reliability. The role emphasizes building scalable systems, collaborating across teams, and driving reliability across Google Cloud Platform. As part of Google’s technical workforce ...

Senior Cloud Reliability Engineer - Hybrid Role

Location
Southampton, England, United Kingdom
NiCE Ltd. is seeking an experienced DevOps/SRE to manage production systems, automate platform infrastructure, and improve reliability across a suite of distributed applications. You will monitor health, tune performance, and partner with development teams to streamline releases in a hybrid work model. Candidates should have ...

Senior Cloud Reliability Engineer - Hybrid Role

Location
Greater London, England, United Kingdom
NiCE Ltd. is seeking an experienced DevOps/SRE to manage production systems, automate platform infrastructure, and improve reliability across a suite of distributed applications. You will monitor health, tune performance, and partner with development teams to streamline releases in a hybrid work model. Candidates should have ...

Senior Go Engineer: Scale, Automate & Improve Reliability

Location
Greater London, England, United Kingdom
Jobtailor is seeking a senior software engineer in London to drive the evolution of large-scale systems using Go, AWS, and Azure. You will contribute to platform reliability, observability, and scalable delivery while mentoring junior engineers and collaborating across teams. The role emphasizes CI/CD, automation ...

Senior SRE: Hybrid Cloud Reliability Engineer

Location
Horsell, England, United Kingdom
Capgemini is seeking Site Reliability Engineers to grow their careers in a hybrid setup, delivering and maintaining services for UK clients. You’ll work within Pods, focusing on reusable platform blueprints, GitOps, and multi-cloud operations while engaging in on-call rotations and on-site client engagements. You will ...

Senior Performance QA Engineer: Scale & Reliability

Location
Greater London, England, United Kingdom
Chip UK is seeking a Senior QA Engineer specializing in Performance & Scalability to own performance across 30+ microservices and legacy dependencies. You will define pass/fail thresholds, embed performance in delivery plans, and lead performance initiatives beyond gate testing. You’ll design tests, automate, and monitor across ...

Senior Full Stack Engineer — Reliability & On-Call Lead

Location
Leeds, England, United Kingdom
Burendo is seeking an experienced Full Stack Engineer focused on operating and improving production systems. You will join an L2/L3 managed service team supporting business-critical applications across cloud platforms, participating in on-call rotations and leading technical investigations. The role emphasizes production stability, incident response ...

Senior Software Engineer, Dependency Tracing & Reliability

Location
City of Westminster, England, United Kingdom
Google's Dependency Tracing team seeks a Senior Software Engineer to design and optimize large-scale distributed systems in a cloud context. You will translate complex requirements into robust architectures, mentor juniors, and drive quality through rigorous testing. You will collaborate with cross-functional teams to minimize outages ...

Senior Lead Software Engineer - LLM Ops Platform Reliability

Hiring Organisation
Hackajob Ltd
Location
Milton, Cambridgeshire, UK
reliably in production at scale. In this role, you'll build and operate large language model serving infrastructure, bringing strong engineering fundamentals and site reliability practices to cutting-edge AI platforms. You'll work hands-on with cloud and Kubernetes-based deployments, deep observability, and cost-aware performance tuning. … enjoy solving hard production problems and making platforms measurably better, you'll find meaningful impact and growth here. As a Senior Lead Software Engineer at our client within the AI and Machine Learning Platform team, you will build and scale AI infrastructure that modernizes traditional infrastructure management ...

Senior Data & MLOps Engineer - AI Reliability Platform

Location
Greater London, England, United Kingdom
CoreWeave is recruiting a Senior Data & MLOps Engineer to design and scale the GPU Intelligence Platform infrastructure, building pipelines for data, features, and model training while delivering insights for system health and optimization. You will transition prototypes to production across a fleet, focusing on scalable distributed services, separating ...

Senior Software Engineer, Cloud & Reliability

Location
England, United Kingdom
JPMorgan Chase is seeking a Software Engineer III for its Corporate and Investment Bank Payments Technology – Account Services to strengthen reliability, performance, and automation of mission-critical systems. You will design robust software, build automation, monitor health, and push for scalable, secure solutions in a fast-paced environment. ...

Senior Lead Data Engineer AI-Driven Reliability Dashboards

Location
Greater London, England, United Kingdom
JPMorganChase in the United Kingdom seeks a Senior Lead Data Engineer within the Behavioral Insights Team to turn operational signals and platform data into actionable insights that improve reliability, risk/control health, and delivery efficiency. You will own the reliability and performance of reporting ...

Senior Infrastructure & Operations Engineer (Kubernetes / Platform Reliability)

Location
Greater London, England, United Kingdom
looking for a senior infrastructure and operations engineer to own and evolve our platform reliability. You’ll design, operate, and maintain our Kubernetes-based infrastructure, build reliable monitoring and alerting pipelines, and ensure our systems remain stable under real-world load and failure conditions. This is a hands … Cloudflare, and observability to create a platform engineers can trust. What You’ll Do Design, deploy, and maintain production Kubernetes clusters. Own cluster reliability, upgrades, security, and performance. Build and operate monitoring, logging, and alerting pipelines. Ensure full-stack observability across infrastructure and services. Design and maintain CI/ ...

Senior Platform Engineer – Production Reliability & On-Call

Location
Greater London, England, United Kingdom
Heidi is hiring for a Platform/SRE role in London. You will own production reliability, participate in on‐call and incident response, and drive improvements across observability, deployments, and operational tooling. This is an ops‐heavy, hands‐on position suitable for mid‐level to senior SREs ...

Senior Systems Engineer: Automation & Reliability

Location
Greater London, England, United Kingdom
LexisNexis Legal & Professional is seeking a Systems Engineer to support enterprise systems through design, maintenance and change management. You will troubleshoot issues, automate tasks, and collaborate with development, support, and vendor teams to deliver reliable technology solutions. The role focuses on system stability, incident response, and ongoing improvements across ...

Senior Full-Stack Engineer: Platform Reliability (Onsite)

Location
Birmingham, England, United Kingdom
Infused Solutions in Birmingham is seeking a Senior Full Stack Developer to tackle complex production and platform reliability challenges. You will collaborate with engineering teams to improve performance, observability, and code quality across a large-scale SaaS platform. The role offers onsite work in Birmingham, a salary ...

Senior DevOps Engineer: Cloud Automation & Reliability

Location
Slough, England, United Kingdom
technology company in Slough is seeking a DevOps Engineer to build functional systems that enhance customer experience. The role involves deploying product updates and identifying production issues while implementing integrations. Candidates should have a strong background in software engineering, particularly with Java, Ruby, or Python, and experience in public ...

SNR Reliability Maintenance Engineer

Hiring Organisation
Amazon
Location
Matlock, Derbyshire, United Kingdom
Employment Type
Permanent
Salary
GBP Annual
Reliability Maintenance Engineering (RME) team is central to Amazon's commitment to innovation. As Amazon evolves and adapts, this team makes sure that the tools and technologies we use do as well. As a Senior RME Technician, you'll help us stay one step ahead, adopting the latest ...

Senior FPGA Engineer - High-Reliability Space-Grade Design

Location
East Hagbourne, England, United Kingdom
American Society of Civil Engineers is seeking an experienced FPGA Engineer to join our engineering team across the United Kingdom. You will contribute to high-reliability systems, spanning FPGA design, architecture and verification, in a multi-disciplinary environment. The role involves working on technically complex projects at multiple ...

Senior Dependency Tracing Engineer - Cloud Reliability

Location
Greater London, England, United Kingdom
United States Digital Space LLC seeks an experienced software engineer to join our team building next-generation technologies. You will work on high-impact projects that affect billions of users, spanning distributed systems, data storage, and security, with opportunities to switch teams as our fast-paced business grows. ...

Senior Software Engineer - Dependency Tracing & Reliability

Location
Greater London, England, United Kingdom
Google’s software engineers develop the next-generation technologies that change how billions of users connect, explore, and interact with information and one another. Our products need to handle information at massive scale, and extend ...

Senior Backend Engineer — Real-Money Execution & Reliability

Location
Tees Valley, England, United Kingdom
Xavvy Limited in the United Kingdom is seeking a trusted engineer to work on guardrails_engine, the executors, broker adapters, the scheduler—load-bearing, permanent full-time. You should know Python, FastAPI, SQLAlchemy, Postgres at production scale, and Async scheduling with APScheduler or equivalent; you will own execution-path ...

Senior FPGA Engineer – VHDL/SoC for High-Reliability

Location
Portsmouth, England, United Kingdom
Octagon Group in Southampton, UK is seeking an experienced FPGA Engineer to join a specialist engineering team. You will own the FPGA lifecycle from design to verification for high-reliability products, with hands-on responsibility across development, testing, and documentation. The role requires 5+ years in FPGA, strong ...

Senior ML Engineer - Financial AI & Model Reliability

Location
Greater London, England, United Kingdom
will build rigorous benchmarks, fine-tune open models for financial guidance, and investigate why models fail on numbers and layered rules to improve reliability and auditability. #J-18808-Ljbffr ...

Senior Laravel Backend Engineer - APIs, Jobs & Reliability

Location
Greater London, England, United Kingdom
Jodie AI is a UK company building an AI receptionist for busy businesses. We are looking for a backend engineer to own Laravel services, APIs and background jobs, ensuring reliable webhook processing and scalable databases. You will collaborate with London-based engineers and contribute to customer-facing workflows. ...