176 to 200 of 278 Site Reliability Engineer Jobs

Vice President - Site Reliability Engineering (SRE) - The Core Engineering - Birmingham

Location
West Midlands, England, United Kingdom
This role is for software engineers who enjoy solving complex distributed system problems, building tools and platforms that make teams more effective, and championing SRE principles (such as SLOs, error budgets, and blameless post-mortems) across a large engineering organization. Key Responsibilities Partner with engineering leadership to establish service level … understand how individual components interact under load. Strong interpersonal skills to collaborate with product developers, influence architectural decisions, prioritize toil reduction, and drive SRE adoption without direct authority. Ability to translate complex technical issues into clear, actionable insights for both technical and non-technical stakeholders. Highly motivated, pro-active ...

Lead Site Reliability / DevOps Engineer

Location
Auchentibber, Scotland, United Kingdom
defining the future of a globally recognized firm and have a direct and significant effect in a realm tailored for top achievers in site reliability. As a Software Engineer at JPMorgan Chase within the Commercial & Investment Bank, youhold a leadership role in your team, demonstrate strong knowledge across … technical lead for medium to large-sized products, and provide advice and mentoring to other engineers. Job responsibilities Demonstrates and champions site reliability culture and practices and exerts technical influence throughout your team Leads initiatives to improve the reliability and stability of your team’s applications ...

Engineer - Site Reliability Engineering

Location
Greater London, England, United Kingdom
TeamWe are evolving our Reliability Engineering team to move beyond support and operations. As a Senior Engineer in Site Reliability, you will be part of a diverse and inclusive organization that has full ownership of the availability, performance, and scalability of one of the most critical … purpose.* Write automation to scale systems sustainably, prevent service issues, or when they occur, quickly recover service.* Partner with development teams to improve system reliability, observability, and release velocity.* Participate in on-call rotations, incident response, postmortems, and root cause analysis and resolution.* Be a vocal advocate of strong ...

Site Reliability Engineer

Location
Leeds, England, United Kingdom
Description At evoke, Site Reliability Engineering (SRE) is all about delivering exceptional customer experiences through reliable, scalable and high-performing technology. Operating at the heart of our betting and gaming platforms, our SRE team combines observability, automation and engineering excellence to ensure our systems perform when it matters … systems can scale effectively to meet changing customer demand. Drive automation initiatives that improve operational efficiency, reduce manual effort and enhance service reliability. Promote SRE best practices throughout the organisation, influencing teams through data-driven recommendations and continuous improvement initiatives. Who we are looking for We are committed to responsible ...

AWS SRE Devops engineer

Location
Glasgow, Scotland, United Kingdom
established technology-driven organisation is seeking an experienced Site Reliability Engineer (SRE) to strengthen and scale their cloud-native data platform, utilising AWS, Snowflake, and Databricks. This position offers the opportunity to drive automation, resilience, and operational excellence across critical data services. Key Responsibilities: Automate infrastructure provisioning … root causes to improve availability and performance. Provide operational support, drive incident resolution, and implement automated fixes for recurring issues. Requirements: Strong knowledge of SRE principles and practical experience defining SLAs, SLOs, and error budgets. Demonstrated AWS expertise (e.g., EC2, S3, IAM, VPC, CloudWatch) in production environments. Experience with observability ...

Lead Site Reliability Engineer

Location
Southampton, England, United Kingdom
cloud platforms are observable, measurable, reliable, scalable, and maintainable. It’s likely that the successful candidate will have significant experience in a DevOps, SRE, Cloud Engineer, or Cloud Development role. How will you make an impact? Act as part of a team of SREs that act as the ‘gatekeepers … production and actively manage the work backlog and develop reliability improvements. Lead investigations into root cause outages, performance, and cost issues. Lead initiatives to develop the automation of low-value tasks balanced against project delivery demands. You will provide technical leadership and to wider Cloud Operations and Support teams ...

SRE & Cloud Engineer — Multi-Cloud, Kubernetes, Terraform

Location
Glasgow, Scotland, United Kingdom
Dianaduggan is seeking an experienced Site Reliability Engineer (SRE) to join a high-performing engineering team in Glasgow as a Cloud Engineer. You will enhance cloud reliability, scalability, automation and operational excellence across enterprise cloud platforms. Based in Glasgow, the role is on-site ...

SRE Engineer – FinTech Reliability, Observability & Cloud

Location
Greater London, England, United Kingdom
Hamilton Barnes Associates Limited is seeking a Site Reliability Engineer to work at the intersection of software engineering and infrastructure. You'll develop internal platforms, tooling, and automation across Linux, distributed systems, and cloud-native technologies to improve reliability and operational efficiency in a global production ...

Senior SRE Engineer: Reliability, Cloud & Automation

Location
Greater London, England, United Kingdom
London Stock Exchange Group is looking for a Senior Engineer in Site Reliability who will join a driven team focused on system availability, performance, and scalability. Responsibilities include maintaining service level objectives, writing automation for system resilience, and partnering with development teams. Required qualifications include a Bachelor … computer science, experience in Object Oriented programming and cloud systems, and DevOps familiarity. The role is pivotal in ensuring 24/7 system reliability and promoting engineering best practices. #J-18808-Ljbffr ...

Site Reliability Engineer — Scale, Automate & Improve Resilience

Location
Swindon, England, United Kingdom
Edenred Finland Oy Swindonissa hakee Site Reliability Engineeria (SRE) osaksi Infrastructure Engineering -tiimiä. Tavoitteena on varmistaa järjestelmien luotettavuus, skaalautuvuus ja liiketoiminnan prioriteettien mukaisuus. Rooli tarjoaa mahdollisuuden vaikuttaa järjestelmien suorituskykyyn, vahvistaa resilienssiä sekä kehittää automatisoituja ratkaisuja nykyaikaisessa, globaalissa ympäristössä. Valinta pohjautuu kumppanuuteen tiimien kanssa ja jatkuvan parantamisen ilmapiiriin. #J ...

Software Engineer Lead - Site Reliability

Location
Telford, England, United Kingdom
proactive, self-starting engineer who enjoys getting things done and improving the reliability of live digital services, Standard Life could be the place for you. We’re looking for a Lead DevOps Engineer to join our Digital Engineering team. This role is focused on making immediate, practical … GitHub/GitHub Actions, Azure DevOps, Terraform and automated testing. Improve deployment safety, release readiness and operational readiness for customer-facing digital services. Apply SRE principles pragmatically to improve availability, recoverability, monitoring and incident learning. Strengthen monitoring, logging, tracing, alerting and service-health dashboards across digitally connected workloads. Reduce single ...

Senior Site Reliability Engineer

Location
Greater London, England, United Kingdom
more about life at DeepL on LinkedIn, Instagram, and our Blog. Meet the team behind this journey We currently have two teams in the SRE track, and some SRE peers in other parts of the business. The in-track teams work closely, often collaborating. The first team SRE: Excellence focussed … making it easier to use our systems and provide tools and services to help run and monitor our products. The second team SRE: Accelerate works closely with our product development teams to use these as effectively as possible and embed better practice in teams. Both teams help ensure the services ...

Director of Site Reliability Engineering

Location
Greater London, England, United Kingdom
will influence engineering standards, enhance operational frameworks, and foster a culture of continuous improvement across mission‐critical environments. Responsibilities Lead and scale a global SRE organization, focusing on engineering excellence and team empowerment Collaborate with product, platform, operations, and security teams to embed reliability within SDLC practices Define … deliver systemic improvements across production environments Establish observability strategies with standardized tooling for metrics, logs, and tracing to support distributed systems Adopt and enforce SRE practices, including SLIs, SLOs, SLAs, and error budgets across services Drive resilience strategies with highly available architectures and disaster recovery readiness Champion an automation‐first ...

Entry Level - Site Reliability Engineer - (Remote - United Kingdom)

Location
Belfast City District, Northern Ireland, United Kingdom
growing as engineers, and having fun in a collaborative environment. Do you want to build and manage scaleable, self-healing, globally-distributed systems? Our Site Reliability engineers keep Yelp fast, available, and growing, connecting users to great local businesses. No matter how many times we get searched, scraped … solutions don’t work at our scale and contribute upstream to open source projects. Participate in light on-call rotations - we have geographically distributed SRE teams for follow-the-sun support, which means nobody needs to be on-call 24h a day! What It Takes To Succeed Familiarity with Linux ...

Entry Level - Site Reliability Engineer - (Remote - United Kingdom)

Location
City of Edinburgh, Scotland, United Kingdom
growing as engineers, and having fun in a collaborative environment. Do you want to build and manage scaleable, self-healing, globally-distributed systems? Our Site Reliability engineers keep Yelp fast, available, and growing, connecting users to great local businesses. No matter how many times we get searched, scraped … solutions don’t work at our scale and contribute upstream to open source projects. Participate in light on-call rotations - we have geographically distributed SRE teams for follow-the-sun support, which means nobody needs to be on-call 24h a day! What It Takes To Succeed Familiarity with Linux ...

Site Reliability Engineer III: Scale & Automation

Location
Greater London, England, United Kingdom
Google London, UK, is seeking a Mid-level Software Engineer III in Site Reliability Engineering for the GCE AI team. You will help build and run large‐scale, fault‐tolerant systems, with emphasis on reliability, uptime, and performance. Expect collaboration across software and systems teams, mentoring ...

DevOps / SRE Engineer (London)

Location
Greater London, England, United Kingdom
Bumper continues to scale across the UK and Europe — and we’re excited to welcome a passionate and experienced DevOps/SRE Engineer , based in our London office, to help us take our platform reliability and engineering excellence to the next level. Our Head Office is in Sheffield … with the vision of being the leading automotive payment and insights platform! A bit about the role... We’re looking for a DevOps/SRE Engineer to join our growing Infrastructure team. In this role, you’ll help shape and strengthen the reliability, scalability and observability ...

DevOps / SRE Engineer (Sheffield)

Location
Sheffield, England, United Kingdom
Bumper continues to scale across the UK and Europe — and we’re excited to welcome a passionate and experienced DevOps/SRE Engineer , based in our Sheffield office, to help us take our platform reliability and engineering excellence to the next level. Our Head Office is in Sheffield … with the vision of being the leading automotive payment and insights platform! A bit about the role... We’re looking for a DevOps/SRE Engineer to join our growing Infrastructure team. In this role, you’ll help shape and strengthen the reliability, scalability and observability ...

SRE Engineer: Core Systems & Automation

Location
Birmingham, England, United Kingdom
Goldman Sachs is seeking a Site Reliability Engineer to join the Compliance Engineering SRE team. You will ensure production services remain healthy, automated, and scalable across cloud-native platforms. Collaborate with engineering to improve reliability, implement monitoring, and reduce downtime while balancing feature velocity with stability. … strong background in SRE, programming, and ownership is essential. #J-18808-Ljbffr ...

AVP, Observability & SRE Engineer

Location
Greater London, England, United Kingdom
Citi is seeking a Site Reliability Engineer - Assistant Vice President in London to drive end‐to‐end observability, migrate legacy monitoring to Google Cloud Observability and Grafana, and implement OpenTelemetry instrumentation across OpenShift/Kubernetes environments. The role emphasizes hands‐on deployment, automation (Ansible/Terraform ...

Senior Site Reliability Engineer – Global Platform Ops

Location
Ipswich, England, United Kingdom
Group is seeking an Site Reliability Engineering Professional to help operate BT International's global core platforms. You will ensure reliability, security, and scalability while driving incident resolution and continual service improvement across international networks and digital platforms. You will collaborate with engineering, product, suppliers ...

Site Reliability Engineer, Infrastructure - ThousandEyes

Location
City Of London, England, United Kingdom
deeply integrated across the Cisco technology portfolio, delivering AI-powered assurance insights within Cisco’s Networking, Security, Collaboration, and Observability portfolios. Our distributed Site Reliability Engineering team of approximately nine engineers owns the availability, latency, performance, efficiency, monitoring, emergency response, and capacity planning of the platform while partnering … operational on-call rotation. Hands-on experience with infrastructure-as-code tooling and codebases, preferably Terraform. Hands-on experienceleveraging AIas a force multiplier of SRE activities, such as automati ng toil away and improving operational efficiency. Professional experience administering and troubleshooting GNU/Linux systems, including system libraries, file systems ...

Site Reliability Engineer I

Location
Greater London, England, United Kingdom
drive real change. Constantly grow as you work hard for a mission that matters at a company where you matter. Your Impact As an SRE contributor in Axon's Real Time Operations organization, you are passionate about delivering solutions to the real-time problems our mission-critical cloud native services … encounter. You are also obsessed about achieving the high quality and reliability our customers demand. You will work closely not only with your peers, but also the RTO engineering teams, allowing your technical deliverables to reach the entire engineering organization, enabling product teams to continuously deliver features ...

Site Reliability Engineer - Private Cloud Compute

Location
Greater London, England, United Kingdom
approach to cloud intelligence, extending the security and privacy of Apple devices into the cloud to unlock even more intelligence for our users. This SRE team is responsible for the availability and automation of the critical systems and services that enable PCC to deliver cloud intelligence without compromising user privacy. … future of privacy-preserving cloud infrastructure at scale, this is the opportunity for you! Description We're looking for a hardworking and passionate SRE Engineer to join this amazing team. You will be an accomplished builder and problem-solver, eager to tackle challenging technical problems. You have a deep ...

SRE Engineer - AWS, .NET & Automation Focus

Location
Glasgow, Scotland, United Kingdom
Capgemini is seeking an experienced Site Reliability Engineer with strong AWS and DevOps expertise to maintain and enhance critical production systems across hybrid cloud environments. You will automate operations, improve reliability, and support deployments, working at the intersection of operations engineering and cloud infrastructure. The role ...