126 to 150 of 189 Site Reliability Engineer Jobs

Senior AI Platform SRE Engineer (Kubernetes & CI/CD)

Hiring Organisation
Jobleads-UK
Location
Leeds, England, United Kingdom
CreateFuture is seeking an experienced Site Reliability Engineer to bring SRE discipline to AI platform operations. You’ll design, build, and operate our Kubernetes-based infrastructure for AI workloads, and help run CI/CD pipelines tailored for AI and agentic systems. The role emphasizes reliable, scalable ...

Senior AI Platform SRE Engineer (Kubernetes & CI/CD)

Hiring Organisation
Jobleads-UK
Location
Manchester, England, United Kingdom
CreateFuture is seeking an experienced Site Reliability Engineer to bring SRE discipline to AI platform operations. You’ll design, build, and operate our Kubernetes-based infrastructure for AI workloads, and help run CI/CD pipelines tailored for AI and agentic systems. The role emphasizes reliable, scalable ...

Senior AI Platform SRE Engineer (Kubernetes & CI/CD)

Hiring Organisation
Jobleads-UK
Location
City of Edinburgh, Scotland, United Kingdom
CreateFuture is seeking an experienced Site Reliability Engineer to bring SRE discipline to AI platform operations. You’ll design, build, and operate our Kubernetes-based infrastructure for AI workloads, and help run CI/CD pipelines tailored for AI and agentic systems. The role emphasizes reliable, scalable ...

Senior AI Platform SRE Engineer (Kubernetes & CI/CD)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
CreateFuture is seeking an experienced Site Reliability Engineer to bring SRE discipline to AI platform operations. You’ll design, build, and operate our Kubernetes-based infrastructure for AI workloads, and help run CI/CD pipelines tailored for AI and agentic systems. The role emphasizes reliable, scalable ...

SRE - Site Reliability Engineer - Observability & Performance

Hiring Organisation
Sanderson Recruitment
Location
Bristol, Somerset, United Kingdom
Employment Type
Contract
Contract Rate
GBP 550 - 600 Daily
SRE - Observability and Performance Up to £600 per day outside IR35 6 month initial contract Bristol - Largely remote I'm currently working with a client who is looking for an SRE to implement and enhance observability across Java applications, middleware and Linux infrastructure using Grafana click apply for full ...

site reliability engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
digital engineering, cloud, and AI-enabled transformation services, focusing on complex software product development and digital platform engineering. Задачи Lead and scale a global SRE organization, focusing on engineering excellence and team empowerment Collaborate with product, platform, operations, and security teams to embed reliability within SDLC practices Define … deliver systemic improvements across production environments Establish observability strategies with standardized tooling for metrics, logs, and tracing to support distributed systems Adopt and enforce SRE practices, including SLIs, SLOs, SLAs, and error budgets across services Drive resilience strategies with highly available architectures and disaster recovery readiness Champion an automation-first ...

Software Engineer III, Site Reliability Engineering, Traffic Network Load Balancing

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Computer Science or Engineering. 2 years of experience designing, analyzing, and troubleshooting large-scale distributed systems. About the job Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed, fault‐tolerant systems. SRE ensures that the company Cloud's services—both … internally critical and our externally‐visible systems—have reliability, uptime appropriate to customer's needs and a fast rate of improvement. Additionally SRE’s will keep an ever‐watchful eye on our systems capacity and performance. Much of our software development focuses on optimizing existing systems, building infrastructure ...

SRE Platform Engineer Lead — High-Availability Cloud

Hiring Organisation
Jobleads-UK
Location
City Of London, England, United Kingdom
Goldman Sachs Group, Inc. seeks a Lead Site Reliability Platform Engineer (SRE) in London. This role emphasizes system reliability, performance, and collaboration across teams. Your mission is to design secure, scalable systems on AWS, streamline software delivery, and mentor technical staff. The ideal candidate has over … years of experience in SRE or DevOps and is skilled in AWS and infrastructure as code. You'll lead initiatives to enhance operational efficiency and contribute significantly to security and system integration. #J-18808-Ljbffr ...

Hybrid SRE Engineer — Observability & Cloud (London)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Reward Gateway’s London team is hiring a Site Reliability Engineer to help transform current workloads toward an SRE model while working in a hybrid setup, visiting the London office twice weekly. The role focuses on observability, high availability and incident management, with collaboration across Product Engineering ...

Site Reliability Engineer, Infrastructure - ThousandEyes

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
deeply integrated across the Cisco technology portfolio, delivering AI‐powered assurance insights within Cisco’s Networking, Security, Collaboration, and Observability portfolios. Our distributed Site Reliability Engineering team of approximately nine engineers owns the availability, latency, performance, efficiency, monitoring, emergency response, and capacity planning of the platform while partnering … call rotation. Hands‐on experience with infrastructure‐as‐code tooling and codebases, preferably Terraform. Hands‐on experience leveraging AI as a force multiplier of SRE activities, such as automating toil away and improving operational efficiency. Professional experience administering and troubleshooting GNU/Linux systems, including system libraries, file systems, networking ...

Site Reliability Engineer, Infrastructure - ThousandEyes

Hiring Organisation
Jobleads-UK
Location
City Of London, England, United Kingdom
deeply integrated across the Cisco technology portfolio, delivering AI-powered assurance insights within Cisco’s Networking, Security, Collaboration, and Observability portfolios. Our distributed Site Reliability Engineering team of approximately nine engineers owns the availability, latency, performance, efficiency, monitoring, emergency response, and capacity planning of the platform while partnering … operational on-call rotation. Hands-on experience with infrastructure-as-code tooling and codebases, preferably Terraform. Hands-on experienceleveraging AIas a force multiplier of SRE activities, such as automati ng toil away and improving operational efficiency. Professional experience administering and troubleshooting GNU/Linux systems, including system libraries, file systems ...

SRE Engineer: Cloud-Native, Automation & Resilience

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Talent is seeking a Site Reliability Engineer to join their team in Central London. This permanent role offers a competitive salary up to £300k, depending on skills and experience. The candidate will contribute to the technology underpinning the business and improve system reliability while working ...

Elite FinTech SRE Engineer — Flexible Hours

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Hunter Bond is looking for a Site Reliability Engineer to join an elite FinTech firm in London. This role involves working with an extraordinary Linux team and offers opportunities to architect resilient, large-scale storage solutions. The ideal candidate will have a strong passion for Linux ...

Technical Lead - Site Reliability Engineering

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Reliability Engineering capabilities to strengthen reliability, observability, security, and operational excellence across our Markets and Risk Intelligence division.As a **Technical Lead SRE**, you will be a senior hands‐on technical person help shape the foundations of reliability across both new and existing platforms. You will collaborate … person who is passionate about reliability engineering and who bring a continuous improvement approach to everything they do!Lead the establishment of SRE foundations for new projects building environments, monitoring, alerting, and ensuring operational readiness from day one.Collaborate with Architecture and Engineering teams to embed reliability, scalability, security ...

SRE Fleet Engineer: Global Infra Automation & Reliability

Hiring Organisation
Jobleads-UK
Location
City Of London, England, United Kingdom
Cisco Systems, Inc. is seeking an experienced Site Reliability Engineer to maintain and expand automation across a global infrastructure. The role focuses on reliability, scalability, and operational excellence for a platform spanning thousands of devices and clouds. You will help design deployment pipelines, testing frameworks ...

Site Reliability Engineer

Hiring Organisation
Evantis Consulting
Location
Greater London, England, United Kingdom
SRE-Hyper-V-Infrastructure Engineer Contract-Inside IR35 London, UK-5 Days onsite a week Budget: £430/Day-Inside IR35 Skill Set :Skills: Hyper‐V, Windows infrastructure, PowerShell, production support/SR EFocus areas: TLM and OS upgrades, configuration drift remediation, operational stabilit ...

Principal Platform Engineer (SRE/Cloud)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
honesty ensuring our workforce is able to bring their full selves to work. ABOUT THE ROLE Principal Platform Engineers at Beamery solve the toughest reliability, scalability and infrastructure problems with the highest impact. Together they collaborate to set the standards for how Engineering will build, run and operate services … whole engineering organisation WHO ARE WE LOOKING FOR? We are seeking a hands-on technical leader with deep Site Reliability Engineering (SRE) and Cloud expertise who can set direction across the engineering organisation. Key skills/experience: A proven track record of designing and delivering scalable, reliable cloud ...

SRE Security Engineer

Hiring Organisation
Opus Recruitment Solutions
Location
London, United Kingdom
Employment Type
Contract
SRE Security Engineer (SC Cleared) Location: Remote Clearance: Active SC Required We're seeking an experienced SRE Security Engineer to support a major government programme, helping to improve the security, reliability, and resilience of critical platforms and services. Key Skills: Splunk Security Engineering Platform & Infrastructure Security Vulnerability … security best practices. Contribute to monitoring, detection, and continuous improvement initiatives. Ideal for candidates with a background in Security Engineering, DevSecOps, Platform Security, or SRE within secure or regulated environments. ...

Senior AWS Site Reliability Engineer

Hiring Organisation
Spectrum IT Recruitment
Location
London, United Kingdom
Employment Type
Permanent
Salary
£60000 - £70000/annum Bonus, Medical Care
Datadog, PagerDuty, or Rundeck Experience using configuration management platforms like Ansible, Puppet, or Chef Professional certifications in cloud DevOps, such as AWS Certified DevOps Engineer or Google Cloud Professional DevOps Engineer, or similar credentials Do You Have What It Takes? 3-6 years of hands-on experience … similar role, with a strong emphasis on systems engineering, automation, and service reliability Proficient in at least one programming language such as Python, Go, Java, or C#, along with scripting skills in Bash or PowerShell Solid grasp of cloud platforms like AWS, including an understanding of how core services ...

Senior Site Reliability Engineer — Cloud & Automation

Hiring Organisation
Jobleads-UK
Location
Belfast City District, Northern Ireland, United Kingdom
Lucera is seeking an experienced SRE/DevOps Engineer to join our engineering team in Belfast. You will maintain reliable, scalable production infrastructure, implement IaC, and enhance CI/CD pipelines across our global trading platforms. You will work on container orchestration, observability, and automation, collaborating with developers ...

Site Reliability Engineer- Spacetime UK

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Role Overview This isn't a "keep the lights on" SRE role. This is a strategic, high-impact opportunity to build the nervous system for a platform that transforms how networks of satellites, ground stations, and fleets are interconnected and orchestrated. You will be building the core observability stack that … cloud-native tools to a robust, scalable, and insightful platform built on best-in-class technologies (Prometheus, OpenTelemetry, etc.). If you are an SRE who thrives on platform-building challenges and wants to be relied upon to build a production-grade observability stack from the ground up, this role ...

AI Engineer (Infrastructure SRE & Automation)

Hiring Organisation
Sky
Location
Livingston, West Lothian, Scotland, United Kingdom
Employment Type
Permanent, Work From Home
Salary
GBP per hour
Role/Team overview We are seeking an AI Specialist to design, build, and operationalise intelligent systems that enhance Site Reliability Engineering (SRE) platforms and automate large-scale infrastructure operations. This role has a strong emphasis on reliability, scalability, and operational efficiency. You will work closely with … SRE, platform, and cloud engineering teams , within a global technology support organisation to embed AI-driven decision-making, predictive analytics, and autonomous remediation into infrastructure platforms . What youll do Infrastructure Automation at Scal e Design and deploy AI-powered automation frameworks for incident response and remediation (self-healing systems ...

Principal SRE Engineer / Grafana Specialist - (Outside IR35)

Hiring Organisation
Sanderson Recruitment
Location
Bristol, Avon, South West, United Kingdom
Employment Type
Contract, Work From Home
Contract Rate
£650 - £750 per day + Outside IR35
Lead SRE/Observability Engineering Lead - (Outside IR35 Contract/Remote) Location: Bristol/London HQ - Largely Remote (Occasional Travel) Day Rate: Outside IR35 - £700 p/d Duration: 3-6 Months Initial - with intention to extend Payment Terms: Monthly Our client is a FTSE100 Wealth/Asset Management firm … seeking to engage a Lead SRE Engineer (Observability SME) to support the implementation and instrumentation of their new Observability solution. This role will be critical in delivering against our Digital OKRs by embedding observability best practices, frameworks, and tooling across digital platforms and engineering teams. Key Responsibilities: Strategy & Roadmap ...

Senior or Staff Software Engineer, SRE/ Platform Team

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Senior or Staff Software Engineer, SRE/Platform Team OneSignal is a leading omnichannel customer engagement solution, powering personalized customer journeys across mobile and web push notifications, in-app messaging, SMS, and email. On a mission to democratize customer engagement, we enable businesses to keep their 1.5B monthly active … Go. This potent combination of high performance with efficient resource utilization has given us an incredible competitive edge. We are seeking a Platform Engineer to join our team and help us scale by managing and developing the next generation of our infrastructure. While we currently maintain a 99.95 % uptime ...

AI Engineer (Infrastructure SRE & Automation)

Hiring Organisation
Jobleads-UK
Location
Livingston, Scotland, United Kingdom
Role/Team overview We are seeking an AI Specialist to design, build, and operationalise intelligent systems that enhance Site Reliability Engineering (SRE) platforms and automate large-scale infrastructure operations. This role has a strong emphasis on reliability, scalability, and operational efficiency. You will work closely with … SRE, platform, and cloud engineering teams, within a global technology support organisation to embed AI-driven decision-making, predictive analytics, and autonomous remediation into infrastructure platforms. Whatyou’lldo Infrastructure Automation at Scal e Design and deploy AI-powered automation frameworks for incident response and remediation (self-healing systems). Automate ...