1 to 25 of 28 Permanent Site Reliability Engineering Jobs in the Thames Valley

Platform Engineering Manager (SRE)

Location
Bracknell, England, United Kingdom
cloud products fast, reliable, and trusted by some of the world's largest SAP-run businesses. We are hiring a Platform Engineering Manager (SRE) to own reliability and platform engineering for Klario and our Private Cloud estate. This is a hands-on leadership role and a genuine … developer tooling. Drive standardisation and reduce engineering friction through automation and self-service. Site Reliability Engineering Introduce and embed SRE practices across Engineering. Improve reliability, resilience, recoverability, and operational readiness. Stand up monitoring, alerting, logging, and service health capabilities, including public-facing uptime and status ...

Senior Site Reliability Engineer

Location
Milton Keynes, England, United Kingdom
Senior Site Reliability Engineer Up to £75,000 plus bonus and on call allowance Milton Keynes (2 days on site a week) VIQU have partnered with a well-established B2B SaaS company who are going through a significant platform transformation. and so are hiring for a Senior … Site Reliability Engineer to build stability, respond to live incidents, and assist with system upkeep. The role will also play a key part in on implementing and adopting new tooling and processes surrounding the wider transformation. This is a genuine opportunity to own and operate how the cloud ...

Senior Site Reliability Engineer

Hiring Organisation
VIQU IT Recruitment
Location
Milton Keynes, Buckinghamshire, United Kingdom
Employment Type
Full-Time
Salary
£65,000 - £75,000 per annum
Senior Site Reliability Engineer Up to £75,000 plus bonus and on call allowance Milton Keynes (2 days on site a week) VIQU have partnered with a well-established B2B SaaS company who are going through a significant platform transformation. and so are hiring for a Senior … Site Reliability Engineer to build stability, respond to live incidents, and assist with system upkeep. The role will also play a key part in on implementing and adopting new tooling and processes surrounding the wider transformation. This is a genuine opportunity to own and operate how the cloud ...

Monitoring & Observability Engineer

Location
Reading, England, United Kingdom
monitor the health and performance of our live cryogenic systems, respond to operational incidents and continuously improve our observability capabilities. You'll collaborate across engineering, operations, software and reliability teams to develop dashboards, refine alerting strategies and automate operational responses that improve reliability and reduce downtime. What … Experience monitoring high-availability technical systems. Understanding of incident management, root cause analysis, reliability engineering or Site Reliability Engineering (SRE) principles. Experience with HTTP/REST APIs, Git, containerisation, Kubernetes or Infrastructure as Code. Continuous improvement, ownership or leadership experience. Why Join OQC You will ...

Site Reliability Engineer

Location
Slough, England, United Kingdom
Site Reliability Engineer (SRE) DevSecOps | Cloud Engineering | Observability | Production Environments | London SR2 is supporting a major 3-year programme and looking for an experienced Site Reliability Engineer (SRE) to join the Production Engineering team. This function underpins the reliability, security, and performance … likely) IR35: Inside Location: London twice a week (hybrid model) Clearance: SC level may be required depending on deployment If you’re an experienced SRE who thrives on building reliable, secure, and cost-efficient production systems. #J-18808-Ljbffr ...

Site Reliability Engineer

Location
Milton Keynes, England, United Kingdom
Description We are seeking an experienced Site Reliability Engineer (SRE) to join our Group Technology Team in Milton Keynes. ConnellsX is the company Technology's internal developer platform, built on Microsoft Azure. It simplifies cloud hosting, embeds security and compliance by default, and enables a frictionless developer experience. … operating this platform, you will play a hands‐on role in ensuring it is reliable, scalable, and observable. You will help establish and mature SRE practices, focusing on: Monitoring and observability Incident response Post‐incident review Reliability testing and capacity planning Toil reduction Enabling development velocity We offer ...

Senior Site Reliability Engineer

Location
Reading, England, United Kingdom
principles, operational knowledge, security, and automation to work towards platform/service production excellence from an angle of infrastructure, reliability, and security. The SRE team owns the foundation of AI Platform’s Core platform - the services and infrastructure that let us deploy to a multitude of public cloud providers … close partnership with our lead/backend/staff engineers. Who you are (must-haves) 5+ years in infrastructure engineering, DevOps, or SRE, operating large-scale, high-availability production systems using Kubernetes Production Operational experience - a live cluster under real load, not a lab. Fluent with Helm, and Terraform ...

Senior Site Reliability Engineer — Reliability Lead

Location
Milton Keynes, England, United Kingdom
VIQU IT Recruitment is partnering with a well-established B2B SaaS company to hire a Senior Site Reliability Engineer in Milton Keynes (2 days on-site per week). You will build stability, respond to live incidents, and help with the platform transformation, owning how the cloud ...

SRE: Cloud Reliability, DevSecOps & Observability — Hybrid London

Location
Slough, England, United Kingdom
seeking an experienced Site Reliability Engineer (SRE) to join the Production Engineering team in London on a 6-month contract. You will apply software engineering principles to automate, scale, and secure cloud-native environments. Responsibilities include building and maintaining production and demo environments, implementing observability with ...

Platform Engineer

Location
Milton Keynes, England, United Kingdom
CreatePay is building its Payment Facilitator (PayFac)/Acquiring platform on AWS and is building a dedicated platform engineering capability from the ground up. We are looking for a hands-on Platform Engineer to build and own the infrastructure the platform runs on, day to day. This … across these areas will be prioritised regardless of formal qualifications. Foundational platform experience: 2–5 years hands-on in a platform engineering, DevOps, SRE, or infrastructure role, with cloud experience (AWS preferred) Infrastructure as code: practical experience with Terraform, CloudFormation, or CDK CI/CD pipelines: built and maintained ...

Senior AI Platform SRE: Reliability & Automation

Location
Reading, England, United Kingdom
CloudFactory is seeking a Site Reliability Engineer to keep production systems reliable, scalable, and secure. You will work with engineers and operators to fuse engineering, operation, and security for platform and service excellence. The role emphasizes building golden paths, developer tooling, and end-to-end software delivery ...

Senior DevOps Engineer

Hiring Organisation
MarkIT Placements
Location
Didcot, Oxfordshire, South East, United Kingdom
Employment Type
Permanent
container image scanning. Implement and monitor infrastructure and application security controls. Support the organisation's ongoing compliance and certification requirements. Reliability & SRE Establish and maintain observability across distributed systems. Develop proactive monitoring, alerting and performance-tuning strategies. Help maintain service-level objectives and platform availability. Investigate and resolve infrastructure … with development teams to embed DevSecOps practices throughout the software lifecycle. What We're Looking For We're looking for a seasoned DevOps or SRE professional who combines strong hands-on technical skills with a pragmatic, collaborative approach . You'll ideally have: Proven professional experience in DevOps, SRE ...

Senior Platform Engineer

Location
Abingdon, England, United Kingdom
deliver sustainable fusion energy and maximise its scientific and economic impact. The Computing Division supports this through digital capabilities spanning research, simulation, data, engineering, business systems and communications. The Software Operations Group (SORG) provides the infrastructure and expertise needed for secure software development and operations, promoting DevSecOps and MLOps … improving processes and supporting the growth of technical capabilities. Drive continuous improvement and operational excellence across UKAEA, managing stakeholders, promoting secure‐by‐design and SRE practices, and maintaining high standards of safety and quality. What You’ll Bring Technical qualification or equivalent experience in a scientific, engineering or technical ...

Platform SRE Lead (Remote/Hybrid) – Build Reliable Cloud

Location
Bracknell, England, United Kingdom
Basis Technologies, Bracknell, is hiring a Platform Engineering Manager (SRE) to own reliability and platform engineering for Klario and Private Cloud. This hands-on leadership role focuses on building a proactive, scalable platform with a permanent team and strong ownership of uptime. You will define SRE practices ...

Senior Forward Deployment Engineer

Location
Slough, England, United Kingdom
Project Description: The project is focused on improving the security, release reliability, maintainability, and operational resilience of business-critical banking applications. Forward Deployed Engineers will conduct targeted, hands-on engagements with development teams that are facing challenges with … vulnerability remediation, automated testing, legacy technology, production stability, or achieving a safe and rapid release cadence. The role combines senior software engineering, DevOps, SRE, test automation, application security, and legacy modernization. Approximately 70% of the work involves hands-on coding and configuration. The technical scope includes building effective automated ...

Senior Azure Cloud Engineer

Hiring Organisation
hireful
Location
Milton Keynes, Buckinghamshire, UK
Employment Type
Full-time
help lead large-scale cloud transformation and elevate a multi-tenant SaaS infrastructure to new levels. You will be joining an R&D Platform Engineering team who are evolving the Azure environment that supports 100+ production customers globally. Keen to architect, build and operate scalable, secure, and automated cloud … infrastructure using cutting-edge Azure technologies? Role: Azure Cloud Engineer, Cloud Engineer, Azure Platform Engineer, Cloud Systems Engineer, Cloud Platform Engineer, DevOps Engineer, Site Reliability EngineerSalary: £50k - £55k base + bonus and great benefits! Benefits: 5% matched pension, 25 days holiday, 2 wellbeing days, Christmas shutdown, private healthcare ...

Logs Specialist

Location
Maidenhead, England, United Kingdom
Query Language) proficiency, and dashboarding.## **What will help you succeed****Qualifications & Requirements*** Experience: 5+ years in a domain, specialist, pre-sales, professional services, or SRE role, with at least 3 years specifically focused on Log Management or Big Data analytics.* Technical Depth: Advanced proficiency in Query Languages (e.g., Splunk … translate "bits and bytes" into "dollars and cents"—explaining how log management impacts MTTR and operational overhead.* Education: Bachelor's degree in - Computer science, Engineering, or equivalent practical experience.* Hands‐on exposure to modern observability pipelines, including Cribl solutions or OpenTelemetry logging specifications.* Familiarity with the Cribl ecosystem ...

Senior Infrastructure Operations Engineer

Location
Bracknell, England, United Kingdom
finance, emergency response, and government sectors. We’re looking for a technically skilled and solutions-driven Senior Infrastructure Operations Engineer to join our expanding Engineering team in Braacknell, Berkshire. If you have a passion for high-availability infrastructure, automation, and delivering mission-critical services at scale, we’d love … continuous improvement. About you... Bachelor’s degree in Computer Science, Engineering, or a related STEM field. 5+ years of experience in Infrastructure, SRE, DevOps or similar roles (or 3+ years plus 2+ years in customer support). Strong scripting and automation skills (PowerShell, Python, or similar). Experience with ...

Infrastructure Operations Engineer

Location
Bracknell, England, United Kingdom
Infrastructure & Automation Design and implement scalable, resilient infrastructure solutions using automation and infrastructure-as-code. Continuously improve platform reliability and streamline operational workflows through automation and tooling. Security & Observability Develop and maintain proactive monitoring, logging, and alerting solutions. Contribute to both offensive and defensive security strategies, including threat simulations … world's most vital organisations., Bachelor's degree in Computer Science, Engineering, or a related STEM field. 5+ years of experience in Infrastructure, SRE, DevOps or similar roles (or 3+ years plus 2+ years in customer support). Strong scripting and automation skills (PowerShell, Python, or similar). Experience ...

Senior Platform Engineer Enterprise Operations Oxford, England, United Kingdom

Location
Oxford, England, United Kingdom
meet internal audit standards Establish the foundational monitoring, alerting, and telemetry frameworkrequiredfor robust operations, defining clear SLOs, and setting the course for future SRE work Partner with Research and Data teams to build self-service capabilities that efficiently support diverse workloads, from Python notebooks to distributed clusters Essential Skills, Qualifications … Experience: Proven experience platform engineering, with a demonstrabletrack recordof architecting and automating operational processes A highly proactive attitude and a passion for introducing and automating operational structure Expertisewith at least one major cloud provider (OCI, AWS, GCP, or Azure) Proficiencywith Terraform for declarative, large-scale infrastructure provisioning Comfortable with ...

Operations Team Lead (Production & Reliability)

Location
High Wycombe, England, United Kingdom
improving. This is a hands‐on role. You’ll shape process, lead incidents, build the team, and move us from reactive firefighting to proactive reliability engineering. What You’ll Own Production Stability and availability of all live systems Operational readiness for new releases Safe production access and change coordination … Raise the bar on operational discipline You’re responsible for both system performance and team performance. What We’re Looking For Strong experience in SRE, DevOps, Infrastructure, or Production Engineering Prior experience leading technical teams Deep hands‐on incident management experience Strong observability and reliability mindset Calm under ...

Operations Team Lead (Production & Reliability)

Location
Slough, England, United Kingdom
improving. This is a hands‐on role. You’ll shape process, lead incidents, build the team, and move us from reactive firefighting to proactive reliability engineering. What You’ll Own Production Stability and availability of all live systems Operational readiness for new releases Safe production access and change coordination … Raise the bar on operational discipline You’re responsible for both system performance and team performance. What We’re Looking For Strong experience in SRE, DevOps, Infrastructure, or Production Engineering Prior experience leading technical teams Deep hands‐on incident management experience Strong observability and reliability mindset Calm under ...

Senior Infrastructure & Automation Engineer | SRE/DevOps

Location
Bracknell, England, United Kingdom
will shape scalable, secure infrastructure, drive automation, and ensure resilient services powering critical CX platforms. Ideal candidates bring 5+ years in Infrastructure/SRE/DevOps, strong scripting, Kubernetes, and cloud security knowledge. This role requires occasional night/weekend coverage and a proactive drive for continuous improvement. #J ...

Senior Connectivity Engineer

Hiring Organisation
Hackajob Ltd
Location
Wallingford, Oxfordshire, South East, United Kingdom
Employment Type
Permanent, Work From Home
repeatable and observable. Offices are based in Wallingford, Oxford, with this role offered on a flexible hybrid basis (approximately once every 4 weeks on site). Experience in robotics, telematics, connected vehicles, industrial IoT, or start-up/scale-up environments is desirable … along with experience operating fleets of devices across multiple geographies and carriers. We're open to candidates from a range of backgrounds - networking, platform, SRE, DevOps, embedded Linux, telematics - what matters is strong networking fundamentals paired with a tooling-first mindset. Please apply for more information. Responsibilities/Skills ...

Delivery Manager

Location
Wokingham, England, United Kingdom
manage complex dependencies across multiple teams and suppliers Excellent understanding of business and operational objectives, and stakeholder landscape, with the ability to challenge engineering, product or commercial decisions using this insight Proven delivery governance capability: RAID management, milestone planning, reporting, roadmap alignment and dependency management, co-ordinating with wider … proactive and collaborative manner to deliver results Demonstrated leadership behaviours including ownership, accountability, calm decision making, and structured problem solving Experience with CNI operations, SRE practices, DevOps models, or real-time operational systems Experience enabling product centric delivery and shaping cross-functional ways of working Strong communication (verbal and written ...