126 to 150 of 175 Permanent Site Reliability Engineer Jobs

Site Reliability Engineer, Infrastructure - ThousandEyes

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
deeply integrated across the Cisco technology portfolio, delivering AI‐powered assurance insights within Cisco’s Networking, Security, Collaboration, and Observability portfolios. Our distributed Site Reliability Engineering team of approximately nine engineers owns the availability, latency, performance, efficiency, monitoring, emergency response, and capacity planning of the platform while partnering … call rotation. Hands‐on experience with infrastructure‐as‐code tooling and codebases, preferably Terraform. Hands‐on experience leveraging AI as a force multiplier of SRE activities, such as automating toil away and improving operational efficiency. Professional experience administering and troubleshooting GNU/Linux systems, including system libraries, file systems, networking ...

Site Reliability Engineer, Infrastructure - ThousandEyes

Hiring Organisation
Jobleads-UK
Location
City Of London, England, United Kingdom
deeply integrated across the Cisco technology portfolio, delivering AI-powered assurance insights within Cisco’s Networking, Security, Collaboration, and Observability portfolios. Our distributed Site Reliability Engineering team of approximately nine engineers owns the availability, latency, performance, efficiency, monitoring, emergency response, and capacity planning of the platform while partnering … operational on-call rotation. Hands-on experience with infrastructure-as-code tooling and codebases, preferably Terraform. Hands-on experienceleveraging AIas a force multiplier of SRE activities, such as automati ng toil away and improving operational efficiency. Professional experience administering and troubleshooting GNU/Linux systems, including system libraries, file systems ...

SRE Engineer: Cloud-Native, Automation & Resilience

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Talent is seeking a Site Reliability Engineer to join their team in Central London. This permanent role offers a competitive salary up to £300k, depending on skills and experience. The candidate will contribute to the technology underpinning the business and improve system reliability while working ...

Elite FinTech SRE Engineer — Flexible Hours

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Hunter Bond is looking for a Site Reliability Engineer to join an elite FinTech firm in London. This role involves working with an extraordinary Linux team and offers opportunities to architect resilient, large-scale storage solutions. The ideal candidate will have a strong passion for Linux ...

Technical Lead - Site Reliability Engineering

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Reliability Engineering capabilities to strengthen reliability, observability, security, and operational excellence across our Markets and Risk Intelligence division.As a **Technical Lead SRE**, you will be a senior hands‐on technical person help shape the foundations of reliability across both new and existing platforms. You will collaborate … person who is passionate about reliability engineering and who bring a continuous improvement approach to everything they do!Lead the establishment of SRE foundations for new projects building environments, monitoring, alerting, and ensuring operational readiness from day one.Collaborate with Architecture and Engineering teams to embed reliability, scalability, security ...

SRE Fleet Engineer: Global Infra Automation & Reliability

Hiring Organisation
Jobleads-UK
Location
City Of London, England, United Kingdom
Cisco Systems, Inc. is seeking an experienced Site Reliability Engineer to maintain and expand automation across a global infrastructure. The role focuses on reliability, scalability, and operational excellence for a platform spanning thousands of devices and clouds. You will help design deployment pipelines, testing frameworks ...

Site Reliability Engineer

Hiring Organisation
Evantis Consulting
Location
Greater London, England, United Kingdom
SRE-Hyper-V-Infrastructure Engineer Contract-Inside IR35 London, UK-5 Days onsite a week Budget: £430/Day-Inside IR35 Skill Set :Skills: Hyper‐V, Windows infrastructure, PowerShell, production support/SR EFocus areas: TLM and OS upgrades, configuration drift remediation, operational stabilit ...

Principal Platform Engineer (SRE/Cloud)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
honesty ensuring our workforce is able to bring their full selves to work. ABOUT THE ROLE Principal Platform Engineers at Beamery solve the toughest reliability, scalability and infrastructure problems with the highest impact. Together they collaborate to set the standards for how Engineering will build, run and operate services … whole engineering organisation WHO ARE WE LOOKING FOR? We are seeking a hands-on technical leader with deep Site Reliability Engineering (SRE) and Cloud expertise who can set direction across the engineering organisation. Key skills/experience: A proven track record of designing and delivering scalable, reliable cloud ...

Senior AWS Site Reliability Engineer

Hiring Organisation
Spectrum IT Recruitment
Location
London, United Kingdom
Employment Type
Permanent
Salary
£60000 - £70000/annum Bonus, Medical Care
Datadog, PagerDuty, or Rundeck Experience using configuration management platforms like Ansible, Puppet, or Chef Professional certifications in cloud DevOps, such as AWS Certified DevOps Engineer or Google Cloud Professional DevOps Engineer, or similar credentials Do You Have What It Takes? 3-6 years of hands-on experience … similar role, with a strong emphasis on systems engineering, automation, and service reliability Proficient in at least one programming language such as Python, Go, Java, or C#, along with scripting skills in Bash or PowerShell Solid grasp of cloud platforms like AWS, including an understanding of how core services ...

Senior Site Reliability Engineer — Cloud & Automation

Hiring Organisation
Jobleads-UK
Location
Belfast City District, Northern Ireland, United Kingdom
Lucera is seeking an experienced SRE/DevOps Engineer to join our engineering team in Belfast. You will maintain reliable, scalable production infrastructure, implement IaC, and enhance CI/CD pipelines across our global trading platforms. You will work on container orchestration, observability, and automation, collaborating with developers ...

Site Reliability Engineer- Spacetime UK

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Role Overview This isn't a "keep the lights on" SRE role. This is a strategic, high-impact opportunity to build the nervous system for a platform that transforms how networks of satellites, ground stations, and fleets are interconnected and orchestrated. You will be building the core observability stack that … cloud-native tools to a robust, scalable, and insightful platform built on best-in-class technologies (Prometheus, OpenTelemetry, etc.). If you are an SRE who thrives on platform-building challenges and wants to be relied upon to build a production-grade observability stack from the ground up, this role ...

AI Engineer (Infrastructure SRE & Automation)

Hiring Organisation
Sky
Location
Livingston, West Lothian, Scotland, United Kingdom
Employment Type
Permanent, Work From Home
Salary
GBP per hour
Role/Team overview We are seeking an AI Specialist to design, build, and operationalise intelligent systems that enhance Site Reliability Engineering (SRE) platforms and automate large-scale infrastructure operations. This role has a strong emphasis on reliability, scalability, and operational efficiency. You will work closely with … SRE, platform, and cloud engineering teams , within a global technology support organisation to embed AI-driven decision-making, predictive analytics, and autonomous remediation into infrastructure platforms . What youll do Infrastructure Automation at Scal e Design and deploy AI-powered automation frameworks for incident response and remediation (self-healing systems ...

Senior or Staff Software Engineer, SRE/ Platform Team

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Senior or Staff Software Engineer, SRE/Platform Team OneSignal is a leading omnichannel customer engagement solution, powering personalized customer journeys across mobile and web push notifications, in-app messaging, SMS, and email. On a mission to democratize customer engagement, we enable businesses to keep their 1.5B monthly active … Go. This potent combination of high performance with efficient resource utilization has given us an incredible competitive edge. We are seeking a Platform Engineer to join our team and help us scale by managing and developing the next generation of our infrastructure. While we currently maintain a 99.95 % uptime ...

AI Engineer (Infrastructure SRE & Automation)

Hiring Organisation
Jobleads-UK
Location
Livingston, Scotland, United Kingdom
Role/Team overview We are seeking an AI Specialist to design, build, and operationalise intelligent systems that enhance Site Reliability Engineering (SRE) platforms and automate large-scale infrastructure operations. This role has a strong emphasis on reliability, scalability, and operational efficiency. You will work closely with … SRE, platform, and cloud engineering teams, within a global technology support organisation to embed AI-driven decision-making, predictive analytics, and autonomous remediation into infrastructure platforms. Whatyou’lldo Infrastructure Automation at Scal e Design and deploy AI-powered automation frameworks for incident response and remediation (self-healing systems). Automate ...

Senior Site Reliability Engineer

Hiring Organisation
Staffworx Limited
Location
London, United Kingdom
Employment Type
Permanent, Work From Home
team is expanding and hiring engineers now The role: Build, operate and maintain high-performance, scalable, reliable services across UK Government deployments Own production reliability: monitoring, alerting, configuration management and upgrades Lead automation to reduce manual operations, using modern platforms including LLM/AI tooling Deploy new products into … need: UK SC clearance, or eligibility to obtain it (active SC strongly preferred; no visa sponsorship) 1-5 years in infrastructure engineering or SRE, building and deploying production systems - not just debugging Hands-on Kubernetes/Docker in production; Terraform, Ansible or CI/CD pipeline experience Proficiency in Java ...

Site Reliability Engineer - Comcast Technology Solutions

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
help support you physically, financially and emotionally through the big milestones and in your everyday life. Please visit the benefits summary on our careers site for more details. Comcast is an equal opportunity workplace. We will consider all qualified applicants for employment without regard to race, color, religion ...

Senior Site Reliability Engineering Manager

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Role Overview Sr. Manager, Site Reliability Engineering (London) is an experienced leader responsible for overseeing a globally distributed team of SRE technologists with diverse skills in software development, systems, network, application, and/or database management. This role ensures seamless, continuous coverage of Cboe's real‐time ...

Site Reliability Engineer

Hiring Organisation
Jobleads-UK
Location
Belfast City District, Northern Ireland, United Kingdom
architecting everything around AI-native workflows. The Cloud Platform team is right at the center of that. We're looking for an SRE who doesn't just want to run reliable systems — but wants to use AI to make them more reliable. If you're already experimenting with … incident response, anomaly detection, or toil reduction, we want to talk. This isn't a "keep the lights on" SRE job. We're building the infrastructure that powers AI agents handling sensitive workflows for major financial and legal firms — and we want SREs who use AI as a force multiplier ...

Senior Site Reliability Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Adaptable & Problem-Solver : Address complex challenges across configuration, policy, observability, and data services. Apply a data-driven approach using Prometheus and Grafana to improve reliability and performance. Ownership & Quality : Own end-to-end configuration quality, enforcing governance with Open Policy Agent. Ensure secure, compliant deployments and full reproducibility, including ...

Site Reliability Engineering Manager

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
About the Role Nottingham Trent House (95002), United Kingdom, Nottingham, Nottinghamshire Site Reliability Engineering Manager This role requires a proven leader to develop technical staff, drive service excellence, and implement significant reliability improvements within complex, large‐scale, highly regulated systems. What You’ll Do Lead a cross … functional group of software engineers focused on managing and optimizing applications to maintain and improve reliability for our customers. Coach and nurture engineers to attain their technical, business, and personal goals. Collaborate with Senior Software Engineering managers to deliver improvements aligned with the technical roadmap and customer satisfaction. Ensure ...

Site Reliability Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
open and inclusive workplace. Our inclusive environment welcomes skills and experiences from diverse backgrounds, and defines who we are. We're hiring an SRE to help us run and evolve the infrastructure behind Signal AI's decision intelligence platform. You'd be joining a small, collaborative Infrastructure team … working on next AI-augmented operations : Claude Enterprise is deployed across Signal. We want this team to help define what good looks like for SRE: incident triage, runbook generation, capacity planning, cost analysis. This is a strategic investment, not a side project: and we'd love someone genuinely curious about ...

Database Site Reliability Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Responsibilities Design, configure, and operate highly available database and data storage systems Drive reliability, performance, scalability, and operational excellence across database environments Lead platform improvement initiatives Monitor, optimise, and automate database operations Identify and address technical debt and performance bottlenecks Support reliability of data pipelines and services Participate … ability to own services end-to-end Strong technical expertise Excellent problem-solving and collaboration skills Solid understanding of database design, performance optimisation, and reliability engineering principles Experience with monitoring, observability, alerting, and incident management practices Understanding of data platforms and data pipelines #J-18808-Ljbffr ...

Staff Site Reliability Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
that usable, and MystraAI is the agentic layer we are building on top of it. This is a Staff-level role that owns the reliability, performance, security and integrity of that infrastructure end-to-end — and sets the technical direction that other teams build on. You will lead … source level rather than as a black box — and ideally have contributed code upstream. Reliability engineering for data platforms. You bring true SRE discipline — SLOs, observability, capacity planning and incident response — to analytical data systems and pipelines. Data-as-a-Service productisation. You think in terms of data ...

Site Reliability Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Responsibilities Help run and evolve the infrastructure behind Signal AI's decision intelligence platform Join a collaborative Infrastructure team Shape the direction of the team Drive a multi-quarter workstream with clear direction Contribute insights ...

SRE Engineering Manager: Traffic Steering & DNS

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
United States Digital Space LLC is seeking a Site Reliability Engineer to lead a seasoned team ensuring the availability and performance of mission-critical services. You will automate responses, oversee incident management, and drive reliability across globally distributed systems. Join a culture of curiosity and collaboration ...