326 to 350 of 1,104 Site Reliability Engineering Jobs in the UK

Sr. Network Site Reliability Engineer (SREs)

Location
Greater London, England, United Kingdom
/ML Technologies and Professional services in the UK and EU market. Job Description Overview We are seeking a highly experienced Senior Network SRE with deep expertise across multi-vendor network infrastructure, automation, and reliability engineering. The ideal candidate will possess strong technical leadership, hands‐on engineering capabilities … resilient, scalable, and observable network environments. Key Responsibilities Design, implement, and maintain highly available network solutions across routing, switching, firewalling, and wireless technologies. Apply SRE principles to improve network reliability, scalability, and performance. Develop and maintain automation workflows using Ansible, Salt, and related frameworks to reduce operational toil. Build ...

Senior Site Reliability Engineer

Hiring Organisation
Imanage
Location
London, United Kingdom
SRE is part of a global organization that leverages the latest technology to communicate with our colleagues across the globe. We organize ourselves into distributed teams -- SRE teams are anchored to iManage offices across the globe. Tuesdays and Thursdays are dedicated to in-office collaboration, rapid innovation, and developing … engage in and often lead architectural discussions, reduce toil, and deliver scalable, resilient platforms that support our customers and organization. As a Senior SRE, you’ll help scale our cloud platform, collaborate across teams to promote standardization and resiliency, and participate in on-call rotations. ...

Software Engineers Ref. 3852

Location
Manchester, England, United Kingdom
other engineers to deliver innovative and scalable systems. You’ll have the opportunity to contribute to large-scale projects, share knowledge with the wider engineering community and learn from experienced colleagues. We encourage engineers to take ownership of their work and provide the support and autonomy needed to develop … their expertise. If you’re successful in securing a Software Engineer role, please note that occasional travel to another GCHQ site may be required. Where overnight stays are necessary, GCHQ will fully cover travel and accommodation expenses. Beyond delivering impactful work, you’ll be part of a culture that ...

SRE Engineer with .Net C# - Glasgow, UK

Location
Glasgow, Scotland, United Kingdom
About the Job you are considering:We are seeking an experienced Site Reliability Engineer SRE AWS DevOps Engineer with strong expertise in AWS Cloud DevOps practices automation and operational support This role is primarily focused on maintaining supporting and enhancing critical production systems while driving reliability scalability … preferred candidates with exceptional AWS and DevOps expertise and moderate development experience will also be consideredThis position sits at the intersection of Operations Engineering and Cloud Infrastructure requiring a proactive individual who can automate processes improve system reliability troubleshoot production issues and contribute to application enhancements when requiredYour ...

Senior Site Reliability Engineer (SRE)

Hiring Organisation
fortice
Location
London, UK
Employment Type
Full-time
Hybrid | London80,000 – 110,000/annum plus benefitsRole: As a Senior Site Reliability Engineer for a global consultancy, you'll initially be aligned to a Defence-related project, where you'll lead a team in helping to relocate data to a new cloud platform. Working hybrid … would expect to be on site 2 days/month in Central London. This will require you hold an active UK Government Security Clearance, which you would be sponsored through, if not currently held. This role will see you be 50% operations-focused, 50% automation-focused – from systems builds ...

Lead Site Reliability Engineer (Dynatrace)

Hiring Organisation
SF Partners Admin
Location
United Kingdom
Employment Type
Permanent
looking for Strong hands-on Dynatrace implementation and administration experience Experience designing and implementing observability/monitoring solutions end-to-end Strong SRE and production engineering background Experience configuring instrumentation, metrics, alerting and monitoring Understanding of technologies such as OneAgent, ActiveGate, distributed tracing and application/infrastructure monitoring Experience … mentoring other engineers The opportunity You'll join a sizeable engineering capability working across complex, large-scale environments, taking a leading role in SRE and observability engineering. There is flexibility around some of the wider cloud/platform technology stack for candidates with genuinely strong Dynatrace and SRE expertise. ...

Lead Site Reliability Engineer (Dynatrace)

Location
London, United Kingdom
looking for Strong hands-on Dynatrace implementation and administration experience Experience designing and implementing observability/monitoring solutions end-to-end Strong SRE and production engineering background Experience configuring instrumentation, metrics, alerting and monitoring Understanding of technologies such as OneAgent, ActiveGate, distributed tracing and application/infrastructure monitoring Experience … mentoring other engineers The opportunity You'll join a sizeable engineering capability working across complex, large-scale environments, taking a leading role in SRE and observability engineering. There is flexibility around some of the wider cloud/platform technology stack for candidates with genuinely strong Dynatrace and SRE expertise. ...

Lead Site Reliability Engineer (Dynatrace)

Hiring Organisation
SF Partners
Location
South West England, United Kingdom
Employment Type
Full-Time
Salary
£80,000 - £100,000 per annum
looking for Strong hands-on Dynatrace implementation and administration experience Experience designing and implementing observability/monitoring solutions end-to-end Strong SRE and production engineering background Experience configuring instrumentation, metrics, alerting and monitoring Understanding of technologies such as OneAgent, ActiveGate, distributed tracing and application/infrastructure monitoring Experience … mentoring other engineers The opportunity You'll join a sizeable engineering capability working across complex, large-scale environments, taking a leading role in SRE and observability engineering. There is flexibility around some of the wider cloud/platform technology stack for candidates with genuinely strong Dynatrace and SRE expertise. ...

Sr Lead AI Platform Engineer

Hiring Organisation
JP Morgan Chase
Location
Glasgow, United Kingdom
team moves from shipping individual use cases to running multiple production platforms and the production tail of new use cases, you will set the engineering standard for deployment, scalability, security, and reliability, and lead the practices that keep our services running.This is a Senior VP-level role … degree in Computer Science, Engineering, or a related technical field (or equivalent applied experience)Preferred qualifications, capabilities, and skillsSite Reliability Engineering (SRE) experience and familiarity with reliability practices (SLOs, error budgets)Experience within financial services technologyFamiliarity with JPM-internal platform, cloud, and AI/ML infrastructure ...

Site Reliability Engineer

Hiring Organisation
Capital On Tap
Location
London, UK
Employment Type
Full-time
just getting started! ðLondon, Old Street | ð 2 Days in OfficeSRE at Capital On Tap ðAt Capital On Tap, we run a hybrid embedded SRE model. We aim to work closely with the teams within Capital On Tap to provide them the best support. Our main objective currently … much visibility into our platform's health while offering scalable solutions. What You'll be doing: As a Site Reliability Engineer (SRE) you will help ensure our platforms are fast, reliable, and scalable. You'll design, build, and monitor systems, prevent issues before they happen. Using SLAs, SLIs ...

Platform Reliability Engineer

Location
Greater London, England, United Kingdom
## Platform Reliability EngineerApply: Fully Onsite: London, UK: Full time: Posted Today: REQ-006568Welcome to ConocoPhillips, where innovation and excellence create a platform for opportunity and growth. Come realize your full potential here.**Who We Are**We are one of the world’s largest independent exploration and production companies … every individual. Wherever possible, we use these differences to drive competitive business advantage, personal growth and, ultimately, create business success.**Job Summary**The Platform Reliability Engineer supports business-critical commercial trading platforms and transformation solutions across hybrid and cloud environments. The role ensures reliable production operations, rapid incident response ...

DV Cleared Software Developers

Hiring Organisation
Flint UK Technology Services
Location
Cheltenham, Gloucestershire, United Kingdom
Employment Type
Contract
Contract Rate
GBP Daily
operate resilient, cloud-native infrastructure that underpins some of the UK's most sensitive environments. The Opportunity Working as part of a highly skilled engineering team, you'll develop and support modern platforms using automation, cloud technologies and Infrastructure-as-Code. You'll play a key role in creating … deployment pipelines, and embedding security throughout the software delivery life cycle. This is an opportunity to work with cutting-edge technologies on projects where reliability, innovation and security are paramount. What You'll Bring Strong software engineering experience using languages such as Java, Python or C# Hands ...

Senior Site Reliability Engineer – Cloud & Platform Automation

Location
Belfast City District, Northern Ireland, United Kingdom
Group is seeking a Site Reliability Engineer (SRE) III to strengthen our Google Cloud platform and middleware stack. You will help design resilient, low-latency systems across CME’s core derivatives applications and mentor junior engineers. In addition to hands-on engineering, you will drive cloud transformation … participate in disaster recovery planning, and lead reliability improvements across cross-functional teams in a hybrid UK-based role. #J-18808-Ljbffr ...

eDV-Cleared Site Reliability Engineer – Platform Reliability

Location
Cheltenham, England, United Kingdom
Forward Role Recruitment is seeking a Site Reliability Engineer (eDV) for national security environments. You will own production reliability, instrument services, and drive improvements in deployment, monitoring and performance. You'll work across cloud, platform and DevOps domains with AWS, Kubernetes, Terraform, Linux, CI/… scripting in Python or Bash. Active eDV clearance is required; role is on-site at national security hubs. #J-18808-Ljbffr ...

Site Reliability Engineer - Core

Hiring Organisation
Blockchain
Location
London, United Kingdom
distributed financial platform tackles some of the most interesting problems in the crypto for millions of our customers and continues to grow rapidly. The SRE team at blockchain combines software and systems engineering to provide a platform that abstracts complexity for increased security, reliability and rapid product delivery.The … SRE organization at Blockchain is a work in progress - our focus is always on how to make our existing systems better. We pride ourselves on having created an environment where individuals have a high degree of freedom in proposing, discussing, designing and implementing changes. We are a team that places ...

Core AI Engineer

Location
Greater London, England, United Kingdom
Artificial Intelligence, Automation and Intelligent Engineering. We are building enterprise-scale AI capabilities that improve service resilience, automate operational workflows, accelerate engineering productivity and enhance customer outcomes.As a Core AI Engineer, you will play a leading technical role in the design, development and deployment of AI solutions across … systems design.* AI observability, evaluation and governance frameworks.Desirable Experience* Experience within Financial Services or highly regulated environments.* Knowledge of Service Reliability Engineering (SRE) principles.* Experience developing AI-powered operational tooling.* Experience building internal AI platforms or developer enablement capabilities.* Familiarity with Microsoft AI ecosystem, Copilot technologies and Azure ...

Senior SRE - Cloud Reliability & Automation

Location
Knutsford, England, United Kingdom
Senior Site Reliability Engineer to drive reliability, scalability, and performance across our core banking platforms. The role combines hands-on SRE work with software engineering, embedding SRE practices and maturity across diverse stakeholder groups. You will apply advanced programming, automation, and data‐driven approaches to reduce ...

SRE Engineer: Core Systems & Automation

Location
Birmingham, England, United Kingdom
Goldman Sachs is seeking a Site Reliability Engineer to join the Compliance Engineering SRE team. You will ensure production services remain healthy, automated, and scalable across cloud-native platforms. Collaborate with engineering to improve reliability, implement monitoring, and reduce downtime while balancing feature velocity with … stability. A strong background in SRE, programming, and ownership is essential. #J-18808-Ljbffr ...

Infrastructure Site Reliability Engineer

Location
Gloucester, England, United Kingdom
cloud. We are building purpose-built AI infrastructure - from powered land, to compute, to software . As we scale our platform and expand our engineering organisation, we are looking for leaders who can build strong teams, uphold high standards, and deliver reliably at pace. Job Summary We’re looking … clearly to non-technical stakeholders and customers Uphold a culture of: do, document, automate Willingness to cross train with Platform Engineering/Platform SRE to fully support both our infrastructure and platform stacks. Willingness to cross train with HPC Engineering, supported by NVIDIA to enhance our HPC supportability ...

Senior Site Reliability Engineer — Reliability Lead

Location
Milton Keynes, England, United Kingdom
VIQU IT Recruitment is partnering with a well-established B2B SaaS company to hire a Senior Site Reliability Engineer in Milton Keynes (2 days on-site per week). You will build stability, respond to live incidents, and help with the platform transformation, owning how the cloud ...

Incident Manager

Location
Leeds, England, United Kingdom
execute rollback/failover/service-degradation procedures. Send initial stakeholder notifications within the SLA and maintain a consistent cadence: technical details for engineering; business impact for leadership. Coordinate with Compliance and Regulatory teams for impact notifications and initiate customer-facing communications (app banners, status page updates) when required. … building enhanced dashboards, lowering detection thresholds, and adding event‐specific synthetic monitors per the readiness plan. Report on noisy alert sources monthly and drive engineering teams to fix the root causes generating non-actionable alerts. Partner with product and engineering teams to agree on monitoring thresholds and ensure ...

Principal SRE (AWS, Azure, Terraforms, Kubernetes)

Location
Greater London, England, United Kingdom
China, Australia, and UAE. Interested in joining our smart, fun, and talented team? Position Overview Fourth is actively seeking an experienced and pragmatic Principal SRE to join our worldwide team. We are progressing rapidly in developing automated, highly reliable, and zero-downtime infrastructure pipelines that are becoming the standard across … valuable and achievable chunks. You have excellent written and verbal communication skills, allowing you to work effectively with our worldwide development teams and SRE community to select the right patterns and practices. You understand the importance of standardisation of technology and practices and have experience of implementing these ...

Platform Engineer

Location
Greater London, England, United Kingdom
operate AI workload infrastructure, including model gateways, retrieval services, orchestration components, and supporting cloud or Kubernetes resources. Observability, Monitoring & Site Reliability (SRE) Instrument services and implement monitoring, logging, and alerting as code using standard tooling (Prometheus, Grafana, OpenTelemetry). Participate in the on‐call rotation, responding to incidents … least one major cloud platform (AWS or Azure) and Kubernetes/Docker. Familiarity with observability tooling (Grafana, Datadog, Splunk, ELK, OpenTelemetry) and basic SRE practices. Exposure to test automation, policy-as-code, and platform security practices. Familiarity with ITIL best practices (incident, change, and problem management) preferred. Experience with Lean ...

Platform Engineer

Hiring Organisation
Kaluza
Location
London, United Kingdom
title: Platform Engineer (Senior level)Location: London or Bristol (Including Hybrid)Salary: 56,000 - 84,000Team: Platform EngineeringReporting To: Engineering ManagerThis role is based in the UK and requires an existing right to work in the UK. At this time, we are not able to offer visa sponsorship … global team and our sponsorship policy is evaluated on a role-by-role basis. We encourage you to keep an eye on our careers site to stay informed about future opportunities where we are able to offer visa sponsorship.Kaluza is the Energy Intelligence Platform, turning energy complexity into seamless ...

Trainee DevOps Engineer | No experience needed (Ref: 7501)

Hiring Organisation
Qualify Nation Recruitment
Location
Plymouth, Devon, United Kingdom
Employment Type
Full-Time
Salary
£28,000 - £38,000 per annum
Platforms (AWS, Microsoft Azure and Google Cloud) Configuration Management Monitoring and Logging Security Best Practices (DevSecOps) Networking Fundamentals Automation and Scripting Incident Management and Reliability Engineering Practical Experience You will work on realistic DevOps projects that may include: Building CI/CD pipelines Deploying applications to cloud environments … completion, learners may pursue roles such as: Junior DevOps Engineer DevOps Engineer Cloud Support Engineer Platform Engineer Infrastructure Engineer Site Reliability Engineer (SRE) Build and Release Engineer Cloud Operations Engineer Systems Administrator Cloud Infrastructure Engineer Apply Today If you are looking to start a career in DevOps ...