76 to 100 of 115 Remote Site Reliability Engineer Jobs

Site Reliability Engineer

Location
Leeds, England, United Kingdom
evoke, Site Reliability Engineering (SRE) is all about delivering exceptional customer experiences through reliable, scalable and high-performing technology. Operating at the heart of our betting and gaming platforms, our SRE team combines observability, automation and engineering excellence to ensure our systems perform when it matters most. Working … systems can scale effectively to meet changing customer demand. Drive automation initiatives that improve operational efficiency, reduce manual effort and enhance service reliability. Promote SRE best practices throughout the organisation, influencing teams through data-driven recommendations and continuous improvement initiatives. Who we are looking for Experience working with large-scale ...

Software Engineer Lead - Site Reliability

Location
Birmingham, England, United Kingdom
Description If you’re a proactive, self‐starting engineer who enjoys getting things done and improving the reliability of live digital services, Standard Life could be the place for you. We’re looking for a Lead DevOps Engineer to join our Digital Engineering team. This role … GitHub/GitHub Actions, Azure DevOps, Terraform and automated testing. Improve deployment safety, release readiness and operational readiness for customer‐facing digital services. Apply SRE principles pragmatically to improve availability, recoverability, monitoring and incident learning. Strengthen monitoring, logging, tracing, alerting and service‐health dashboards across digitally connected workloads. Reduce single ...

Director of Site Reliability Engineering

Location
Greater London, England, United Kingdom
will influence engineering standards, enhance operational frameworks, and foster a culture of continuous improvement across mission‐critical environments. Responsibilities Lead and scale a global SRE organization, focusing on engineering excellence and team empowerment Collaborate with product, platform, operations, and security teams to embed reliability within SDLC practices Define … deliver systemic improvements across production environments Establish observability strategies with standardized tooling for metrics, logs, and tracing to support distributed systems Adopt and enforce SRE practices, including SLIs, SLOs, SLAs, and error budgets across services Drive resilience strategies with highly available architectures and disaster recovery readiness Champion an automation‐first ...

Remote Senior Site Reliability Engineer Manager (Remote)

Location
Cambourne, England, United Kingdom
presence in London, Hong Kong, Amsterdam, and as well in Mumbai and now in New York in 2001. About the role : As the SRE Manager, you will play a critical role in ensuring the reliability, scalability, and performance of our infrastructure and services through both direct technical contribution along … streamline operational workflows and improve efficiency. Develop and maintain tools, scripts, and dashboards to monitor system health, performance, and reliability. Build a first class SRE team. Through a combination of leading by example, coaching and mentoring, mould the team would want to have around you. Provide leadership and guidance ...

Entry Level - Site Reliability Engineer - (Remote - United Kingdom)

Location
Leeds, England, United Kingdom
growing as engineers, and having fun in a collaborative environment. Do you want to build and manage scaleable, self-healing, globally-distributed systems? Our Site Reliability engineers keep Yelp fast, available, and growing, connecting users to great local businesses. No matter how many times we get searched, scraped … solutions don’t work at our scale and contribute upstream to open source projects. Participate in light on-call rotations - we have geographically distributed SRE teams for follow-the-sun support, which means nobody needs to be on-call 24h a day! What It Takes To Succeed Familiarity with Linux ...

Entry Level - Site Reliability Engineer - (Remote - United Kingdom)

Location
Birmingham, England, United Kingdom
growing as engineers, and having fun in a collaborative environment. Do you want to build and manage scaleable, self-healing, globally-distributed systems? Our Site Reliability engineers keep Yelp fast, available, and growing, connecting users to great local businesses. No matter how many times we get searched, scraped … solutions don’t work at our scale and contribute upstream to open source projects. Participate in light on-call rotations - we have geographically distributed SRE teams for follow-the-sun support, which means nobody needs to be on-call 24h a day! What It Takes To Succeed Familiarity with Linux ...

Entry Level - Site Reliability Engineer - (Remote - United Kingdom)

Location
City of Edinburgh, Scotland, United Kingdom
growing as engineers, and having fun in a collaborative environment. Do you want to build and manage scaleable, self-healing, globally-distributed systems? Our Site Reliability engineers keep Yelp fast, available, and growing, connecting users to great local businesses. No matter how many times we get searched, scraped … solutions don’t work at our scale and contribute upstream to open source projects. Participate in light on-call rotations - we have geographically distributed SRE teams for follow-the-sun support, which means nobody needs to be on-call 24h a day! What It Takes To Succeed Familiarity with Linux ...

Entry Level - Site Reliability Engineer - (Remote - United Kingdom)

Location
Belfast City District, Northern Ireland, United Kingdom
growing as engineers, and having fun in a collaborative environment. Do you want to build and manage scaleable, self-healing, globally-distributed systems? Our Site Reliability engineers keep Yelp fast, available, and growing, connecting users to great local businesses. No matter how many times we get searched, scraped … solutions don’t work at our scale and contribute upstream to open source projects. Participate in light on-call rotations - we have geographically distributed SRE teams for follow-the-sun support, which means nobody needs to be on-call 24h a day! What It Takes To Succeed Familiarity with Linux ...

Senior Site Reliability Engineer

Location
Greater London, England, United Kingdom
## Senior Site Reliability EngineerApplyremote type: Hybridlocations: London, United Kingdomtime type: Full timeposted on: Posted Todayjob requisition id: 2020341## **Meet the Team**Cisco's Webex Engineering Group is redefining the future of collaboration. We're building a world where people connect effortlessly to enjoy modern, uncompromised collaboration across … Adaptable & Problem-Solver**: Address complex challenges across configuration, policy, observability, and data services. Apply a data-driven approach using Prometheus and Grafana to improve reliability and performance.* **Ownership & Quality**: Own end-to-end configuration quality, enforcing governance with Open Policy Agent. Ensure secure, compliant deployments and full reproducibility, including ...

Site Reliability Engineer, Infra Platforms — Remote

Location
United Kingdom
GitLab is hiring Site Reliability Engineers to keep our production services running at scale, with a strong emphasis on automation, reliability, and security. You will work across Infrastructure Platforms, contributing to the observability stack and CI/CD workflows in a remote-first environment. You’ll join ...

Staff Software Engineer - Databases SRE | UK | Remote

Hiring Organisation
Grafana Labs
Location
United Kingdom
Salary
£ 70 K
looking for candidates from the UK, Sweden, Spain or Germany.About the role:We are looking for a Staff Software Engineer - SRE to help us support our highest value Grafana Cloud customers by increasing the reliability of our Cloud databases that are based on Mimir, Loki, Tempo, and Pyroscope. … provide these databases as a SaaS product from AWS, GCP, and Azure across all regions.The SRE team is embedded within the Mimir, Loki, and Tempo squads and focuses on ensuring that Grafana Cloud’s database products deliver exceptional reliability for our highest-SLA customers. In this role, you will ...

Senior AI-Driven SRE & Software Engineer (Hybrid)

Location
Manchester, England, United Kingdom
bet365 is seeking a Site Reliability Engineer to enhance reliability, observability, and performance of our critical systems. You will implement instrumentation, improve logging, and drive autonomous operations with AI‐driven telemetry. Collaboration across teams and mentoring colleagues will be key as we embed reliability throughout … software lifecycle. This role aligns with our hybrid working policy, offering flexibility to work from home while maintaining strong on-site presence where #J-18808-Ljbffr ...

Platform Engineering Manager (SRE)

Location
Bracknell, England, United Kingdom
cloud products fast, reliable, and trusted by some of the world's largest SAP-run businesses. We are hiring a Platform Engineering Manager (SRE) to own reliability and platform engineering for Klario and our Private Cloud estate. This is a hands-on leadership role and a genuine build. … management, observability, and developer tooling. Drive standardisation and reduce engineering friction through automation and self-service. Site Reliability Engineering Introduce and embed SRE practices across Engineering. Improve reliability, resilience, recoverability, and operational readiness. Stand up monitoring, alerting, logging, and service health capabilities, including public-facing uptime ...

Senior Site Reliability Engineer

Hiring Organisation
NICE Systems
Location
United Kingdom
Salary
£ 60 K
production environment by monitoring availability and taking a holistic view of system healthBuild software and systems to manage platform infrastructure and applicationsImprove reliability, quality, and time-to-market of our suite of software solutionsMeasure and optimize system performance, with an eye toward pushing our capabilities forward, getting ahead … release proceduresParticipate in system design consulting, platform management, and capacity planningCreate sustainable systems and services through automation and upliftsBalance feature development speed and reliability with well-defined service level objectivesHave you got what it takes 3-6 years of working experience in a similar role, with a focus ...

Senior AWS Site Reliability Engineer

Hiring Organisation
Spectrum IT Recruitment Limited
Location
City of London, London, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£70,000
Datadog, PagerDuty, or Rundeck Experience using configuration management platforms like Ansible, Puppet, or Chef Professional certifications in cloud DevOps, such as AWS Certified DevOps Engineer or Google Cloud Professional DevOps Engineer, or similar credentials Do You Have What It Takes? 3-6 years of hands-on experience … similar role, with a strong emphasis on systems engineering, automation, and service reliability Proficient in at least one programming language such as Python, Go, Java, or C#, along with scripting skills in Bash or PowerShell Solid grasp of cloud platforms like AWS, including an understanding of how core services ...

BI Engineer, SRE (Remote, International)

Hiring Organisation
PulsePoint
Location
United Kingdom
Salary
£ 70 K
DescriptionPosition TitleJob Title: BI Engineer, SREPosition OverviewOur BI team runs a set of GCP-based APIs and data services that a lot of internal products depend on. As we've grown, keeping things … running has increasingly been a side responsibility for engineers who are primarily building features — and that's not sustainable. We're looking for an SRE to own that space: service health, incident response, infrastructure monitoring, and making sure we're not blindly burning cloud budget.The BI Engineer, SRE will ...

Site Reliability Engineer- Spacetime UK

Location
Greater London, England, United Kingdom
Role Overview This isn't a "keep the lights on" SRE role. This is a strategic, high-impact opportunity to build the nervous system for a platform that transforms how networks of satellites, ground stations, and fleets are interconnected and orchestrated. You will be building the core observability stack that … cloud-native tools to a robust, scalable, and insightful platform built on best-in-class technologies (Prometheus, OpenTelemetry, etc.). If you are an SRE who thrives on platform-building challenges and wants to be relied upon to build a production-grade observability stack from the ground up, this role ...

Site Reliability Engineer II — Commercial Cloud

Location
Greater London, England, United Kingdom
CrowdStrike is seeking an Engineer II for the TechOps SRE team focused on our Commercial Cloud. You will be a deeply technical, hands-on engineer building automation and tooling to ensure mission-critical services run reliably across thousands of servers. You will work with Linux engineering, on-call ...

Senior or Staff Software Engineer, SRE/ Platform Team

Location
Greater London, England, United Kingdom
Senior or Staff Software Engineer, SRE/Platform Team OneSignal is a leading omnichannel customer engagement solution, powering personalized customer journeys across mobile and web push notifications, in-app messaging, SMS, and email. On a mission to democratize customer engagement, we enable businesses to keep their 1.5B monthly active … Go. This potent combination of high performance with efficient resource utilization has given us an incredible competitive edge. We are seeking a Platform Engineer to join our team and help us scale by managing and developing the next generation of our infrastructure. While we currently maintain a 99.95 % uptime ...

Site Reliability Engineer, K8s (Remote International)

Hiring Organisation
PulsePoint
Location
United Kingdom
Salary
£ 60 K
environment supports large-scale Kubernetes workloads, multi-petabyte data systems and business-critical services used across multiple engineering organizations.We're looking for an experienced engineer to help shape its next stage of growth.This is a hands-on role focused on architecture, reliability, automation and operational excellence. You will … infrastructureTechnologyYou'll work in an environment that includes:Kubernetes and platform servicesMulti-petabyte data infrastructureBare-metal and cloud environmentsGitOps and infrastructure automationModern observability and reliability engineering practicesTechnologies commonly used across the environment include Kubernetes, ArgoCD, Puppet, Terraform, OpenTelemetry, Prometheus, Alertmanager, Kafka, Redis and Ceph.Experience with every technology ...

Senior Site Reliability Engineer

Hiring Organisation
CISCO Systems
Location
London, United Kingdom
Salary
£ 70 K
risk.Adaptable & Problem-Solver: Address complex challenges across configuration, policy, observability, and data services. Apply a data-driven approach using Prometheus and Grafana to improve reliability and performance.Ownership & Quality: Own end-to-end configuration quality, enforcing governance with Open Policy Agent. Ensure secure, compliant deployments and full reproducibility, including backup ...

Site Reliability Engineer

Location
City Of London, England, United Kingdom
Adaptable & Problem-Solver : Address complex challenges across configuration, policy, observability, and data services. Apply a data-driven approach using Prometheus and Grafana to improve reliability and performance. Ownership & Quality : Own end-to-end configuration quality, enforcing governance with Open Policy Agent. Ensure secure, compliant deployments and full reproducibility, including ...

Staff Site Reliability Engineer

Hiring Organisation
Genomics
Location
Oxford, Oxfordshire, UK
Employment Type
Full-time
that usable, and MystraAI is the agentic layer we are building on top of it. This is a Staff-level role that owns the reliability, performance, security and integrity of that infrastructure end-to-end — and sets the technical direction that other teams build on. You will lead … source level rather than as a black box — and ideally have contributed code upstream. Reliability engineering for data platforms. You bring true SRE discipline — SLOs, observability, capacity planning and incident response — to analytical data systems and pipelines. Data-as-a-Service productisation. You think in terms of data ...

Staff Site Reliability Engineer

Location
Greater London, England, United Kingdom
that usable, and MystraAI is the agentic layer we are building on top of it. This is a Staff-level role that owns the reliability, performance, security and integrity of that infrastructure end-to-end — and sets the technical direction that other teams build on. You will lead … source level rather than as a black box — and ideally have contributed code upstream. Reliability engineering for data platforms. You bring true SRE discipline — SLOs, observability, capacity planning and incident response — to analytical data systems and pipelines. Data-as-a-Service productisation. You think in terms of data ...

Site Reliability Engineer, Big Data (Remote, International)

Hiring Organisation
PulsePoint
Location
United Kingdom
Salary
£ 70 K
ticket-driven operational role.You'll help define platform architecture, influence engineering standards and work on infrastructure that supports multiple engineering organizations.The engineer joining this role is expected to become a key technical contributor shaping the future of the platform.You can work fully remotely and get the opportunity to work ...