1 to 25 of 621 Site Reliability Engineering Jobs in the UK

AI Platform & Site Reliability Engineering Managing Consultant

Location
United Kingdom
build and scale secure, reliable and operationally effective AI platforms. You will combine expertise in platform engineering, Site Reliability Engineering (SRE), observability and intelligent operations to help organisations move from isolated AI experimentation to production-grade, enterprise-scale AI services. You will work with technology, engineering … observability, platform automation and operational guardrails. Enable reliable and repeatable delivery of AI services from experimentation through to production. Reliability Engineering & SRE: Establish SRE practices including SLIs, SLOs, error budgets, capacity planning, resilience engineering and reliability governance. Help clients shift from reactive operations to data-driven ...

Systems Engineering Manager, Site Reliability Engineering, ML Compute

Location
Greater London, England, United Kingdom
. Track record of mentoring technical leads. Proven success leading and influencing multiple technical teams. About the job Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed, fault-tolerant systems. SRE ensures that Google's services—both … internally critical and our externally-visible systems—have reliability, uptime appropriate to users' needs and a fast rate of improvement. Additionally SRE’s will keep an ever-watchful eye on our systems capacity and performance. Much of our software development focuses on optimizing existing systems, building infrastructure and eliminating ...

Cloud Operating Model - Managing Consultant

Location
Greater London, England, United Kingdom
build and scale secure, reliable and operationally effective AI platforms. You will combine expertise in platform engineering, Site Reliability Engineering (SRE), observability and intelligent operations to help organisations move from isolated AI experimentation to production-grade, enterprise-scale AI services.You will work with technology, engineering … observability, platform automation and operational guardrails. Enable reliable and repeatable delivery of AI services from experimentation through to production.• Reliability Engineering & SRE: Establish SRE practices including SLIs, SLOs, error budgets, capacity planning, resilience engineering and reliability governance. Help clients shift from reactive operations to data-driven ...

Head Of Infrastructure and Cloud - Internal Applicants Only

Location
Greater London, England, United Kingdom
transition from traditional infrastructure management to a platform-centric, product-led operating model, integrating platform engineering, DevOps, Site Reliability Engineering (SRE), and Network Operations (NOC) to enable scalable, automated, and resilient technology services. To place the interests of customers at the centre of all activities … YBIYRI) model with shared accountability for service delivery and operational outcomes. Establish and integrate Site Reliability Engineering (SRE) practices, defining and managing service‐level objectives (SLOs), error budgets, and proactive reliability engineering across critical services. Ensure end‐to‐end service reliability and resilience, including ...

Lead Cloud Site Reliability Engineer

Location
Manchester, England, United Kingdom
Summary End Date Monday 21 September 2026 Salary Range £92,701 - £109,060 Flexible Working Options Hybrid Working, Job Share Lead Site Reliability Engineer – Public Cloud Platform Location: Manchester or Bristol Salary: £92,701- £109,043 Working Pattern: Hybrid (2 days in office per week) About this opportunity … Skills and Experience Designing, building or operating large-scale cloud platforms within Azure, GCP or comparable cloud environments. Site Reliability Engineering (SRE), Platform Engineering, Infrastructure Engineering, Cloud Engineering or Production Operations. Observability and monitoring practices, including metrics, logging and distributed tracing. Incident management, problem ...

Site Reliability Engineer

Location
United Kingdom
TITLE: Site Reliability Engineer LOCATION(S): Edinburgh, Manchester or Leeds HOURS: Full-time WORKING PATTERN: Our work style is hybrid, which involves spending at least two days per week, or 40% of your time, at one of our above office sites About this opportunity We're transforming … DevOps and Site Reliability Engineering practices, including continuous improvement, operational excellence and collaborative delivery. The SBO Skills Library identifies DevOps and SRE & Service Engineering as core engineering skill areas. Familiarity with Infrastructure as Code concepts and tools such as Terraform, alongside source control and software ...

Senior Site Reliability Engineer

Hiring Organisation
VIQU IT
Location
United Kingdom, Morley, West Yorkshire
Employment Type
Permanent
Salary
£60000 - £66000/annum 15% Bonus
Senior Site Reliability Engineer (AWS CDK) Remote UK | Permanent, Full Time £60,000 - £66,000 + 15% bonus VIQU has partnered with a leading UK technology organisation investing significantly in its cloud platform and engineering capability. They are looking for a Senior Site Reliability Engineer … incident management. Improve cloud performance, resilience and operational efficiency. Work closely with engineering, architecture and delivery teams. Mentor and support engineers, promoting strong SRE and DevOps practices. Contribute to technical direction, architecture discussions and continuous improvement. Key Requirements of the Senior Site Reliability Engineer Strong commercial ...

Site Reliability Engineer - SC Cleared

Hiring Organisation
Searchability NS&D
Location
Gloucestershire, United Kingdom
Employment Type
Full-Time
Salary
£40,000 - £65,000 per annum, Negotiable
SITE RELIABILITY ENGINEER- SC CLEARED SITE RELIABILITY ENGINEER - Permanent opportunity for a Site Reliability Engineer with SC Clearance. - Salary up to £65,000 DOE - Hybrid opportunity with Gloucester based offices - To apply, please call Laura Jackson on , or email with an up-to-date … CV. WHO ARE WE? We're hiring for Site Reliability Engineers to join a top consultancy delivering cutting-edge solutions for industry-leading Defence and National Security clients. You'll have the opportunity to work across multiple high-impact, innovative and mission-critical projects, shaping solutions that make ...

Vice President - Site Reliability Engineering (SRE) - The Core Engineering - Birmingham Birmingham · United Kingdom · Vice President

Location
Birmingham, England, United Kingdom
Vice President - Site Reliability Engineering (SRE) - The Core Engineering - Birmingham location_on Birmingham, West Midlands, England, United Kingdom WHAT WE DO Site Reliability Engineering at Goldman Sachs sits at the intersection of software engineering, systems design, and production excellence. In this … This role is for software engineers who enjoy solving complex distributed system problems, building tools and platforms that make teams more effective, and championing SRE principles (such as SLOs, error budgets, and blameless post-mortems) across a large engineering organization. Key Responsibilities Partner with engineering leadership to establish ...

Site Reliability Engineer

Location
Swindon, England, United Kingdom
opportunities regardless of their gender and gender expression, disability, origin, religious belief and sexual orientation or any other criteria.**Site Reliability Engineer (SRE)****Location: Swindon (Hybrid)**Join Edenred and help shape resilient, future-ready infrastructure that supports millions of users worldwide.At Edenred, we’re looking for a Site Reliability Engineer (SRE) to join our Infrastructure Engineering team. This is an exciting opportunity for someone who is passionate about reliability, automation, operational excellence, and continuous improvement.**About the role**As a Site Reliability Engineer, you will play a key role in ensuring ...

Senior SRE Engineer

Hiring Organisation
SF Partners
Location
Birmingham, West Midlands (County), United Kingdom
Employment Type
Permanent
Salary
£100000 - £110000/annum great training and progression opp
some of the UK's largest and most complex enterprise platforms. Working within a highly skilled, multi-disciplinary engineering team, this Senior SRE Engineer will take technical ownership of observability and reliability across large-scale AWS environments, helping engineering teams improve platform performance, resilience and availability. … mindset, using automation, Infrastructure as Code and cloud-native tooling to improve reliability and reduce operational overhead. This role would suit a Senior SRE DevOps or Cloud Engineer who combines hands-on AWS expertise with genuine Dynatrace implementation experience, rather than simply having used Dynatrace as an end user. ...

Vice President, Site Reliability Engineering

Hiring Organisation
Hackajob Ltd
Location
South West London, London, United Kingdom
Employment Type
Permanent
hackajob is partnering directly with BNY Mellon to hire for this role. Were seeking a future team member for the role of Vice President - Site Reliability Engineer to join our team. This role is located in London. Role Summary BNY is seeking a Vice President - Site Reliability … driven automation, and modern software delivery practices. Experience supporting distributed systems, cloud-native platforms, or container-based architectures. Knowledge of Agile, DevOps, and SRE operating models, including continuous improvement and blameless post-incident practices. Ability to influence engineering standards and drive adoption of common tooling and automation patterns across ...

Director of Site Reliability Engineering

Hiring Organisation
Hackajob Ltd
Location
Glasgow, Lanarkshire, Scotland, United Kingdom
Employment Type
Permanent
your profession to new heights by contributing to revolutionary projects. You've discovered the perfect environment to have a major impact. As a Principal Site Reliability Engineer at JPMorgan Chase within the Corporate Technology and Enterprise Technology Team, you draw upon your advanced knowledge to identify new opportunities … lifecycle of software development for the firm. You will have the opportunity to manage, design, and implement infrastructure components to improve reliability and ensure operational efficiency. Job Responsibilities Identifies and solves problems of high complexity and drives improvements as outcomes Uses enterprise-authorized AI capabilities within the work environment ...

SRE Ops Lead

Hiring Organisation
Capgemini
Location
City of Bristol, United Kingdom
Employment Type
Full Time
contact the recruiter directly. About the job you’re considering We have an exciting opportunity for an experienced Site Reliability Engineering (SRE) Operational Lead to join Capgemini's Cloud & Infrastructure Services (CIS) managed services business. The role is responsible for leading operational service delivery across the SRE … managed service, ensuring reliability, availability, governance, and continual service improvement for our clients. Operating within a Cloud Pod structure, the SRE Operational Lead acts as the operational authority for service performance, driving operational excellence across Incident, Problem, Change, Availability, and Service Level Management. You will work closely with SRE ...

Site Reliability Engineer

Hiring Organisation
Hackajob Ltd
Location
Swindon, Wiltshire, South West, United Kingdom
Employment Type
Permanent
Salary
£85,000
opportunities regardless of their gender and gender expression, disability, origin, religious belief and sexual orientation or any other criteria. Site Reliability Engineer (SRE) Location: Swindon (Hybrid) Join Edenred and help shape resilient, future-ready infrastructure that supports millions of users worldwide. At Edenred, we're looking … Site Reliability Engineer (SRE) to join our Infrastructure Engineering team. This is an exciting opportunity for someone who is passionate about reliability, automation, operational excellence, and continuous improvement. About the role As a Site Reliability Engineer, you will play a key role in ensuring ...

Site Reliability Engineer

Hiring Organisation
Twinstream Limited
Location
Cheltenham, Montpellier, Gloucestershire, United Kingdom
Employment Type
Permanent
Salary
£75000 - £95000/annum
Site Reliability Engineer | Up to £95,000 DOE | Fully Remote Initially | Cheltenham | Future Hybrid Working Build resilient systems. Solve complex challenges. Make a real impact. Are you a Site Reliability Engineer or DevOps professional who enjoys getting stuck into complex infrastructure challenges? Do you want … Government residency and right-to-work requirements associated with the required level of clearance. Ready for your next challenge? If you’re an SRE or DevOps professional who enjoys solving difficult problems, improving infrastructure and working on projects that genuinely matter, we’d love to hear from you. Apply today ...
Hybrid / Remote Options View Job ❯

Custody Support Applications Support - Assistant Vice President

Location
Belfast City District, Northern Ireland, United Kingdom
modern observability platforms supporting securities processing and settlement functions. This position combines traditional application support responsibilities with Site Reliability Engineering (SRE) principles, automation, resiliency engineering, and operational risk management. The ideal candidate will be passionate about improving platform reliability, reducing operational toil, and driving continuous … enterprise-scale distributed applications within a financial services or mission-critical technology environment. Proven experience in Production Support, Site Reliability Engineering (SRE), Platform Support, or similar technical operations roles. Demonstrated experience managing or coordinating major incident responses. Strong analytical and problem-solving capabilities with a structured approach ...

Lead SRE - AWS Platform

Hiring Organisation
Hackajob Ltd
Location
Glasgow, Lanarkshire, Scotland, United Kingdom
Employment Type
Permanent
role in defining the future of a globally recognized firm and make a direct, meaningful impact in a space built for top achievers in site reliability engineering. As a Lead Site Reliability Engineer at JPMorganChase within Infrastructure Platforms, you bring deep technical expertise across multiple domains … play a key role in driving reliability outcomes for your team. You will conduct resiliency design reviews, break complex problems into digestible work, act as a technical authority for medium to large-sized products, and share your knowledge and experience with peers to raise the bar across the team. ...

Lead SRE- Azure & GCP

Hiring Organisation
Hackajob Ltd
Location
Glasgow, Lanarkshire, Scotland, United Kingdom
Employment Type
Permanent
hackajob is partnering directly with JPMorganChase to hire for this role. JOB DESCRIPTION We have a Lead Site Reliability Engineer (SRE) opportunity within our Google Cloud Site Reliability Engineering team. As a Lead Site Reliability Engineer at JPMorgan Chase within the Infrastructure Platform … Cloud Foundational Services SRE organization, you will join our Google Cloud Site Reliability Engineering team operating within a global follow-the-sun support model. Job Responsibilities: Lead and Implement SRE frameworks to support global google cloud environments and ensure the highest level of SLOs through operational excellence ...

Platform Engineering Manager (SRE)

Location
Bracknell, England, United Kingdom
cloud products fast, reliable, and trusted by some of the world's largest SAP-run businesses. We are hiring a Platform Engineering Manager (SRE) to own reliability and platform engineering for Klario and our Private Cloud estate. This is a hands-on leadership role and a genuine … developer tooling. Drive standardisation and reduce engineering friction through automation and self-service. Site Reliability Engineering Introduce and embed SRE practices across Engineering. Improve reliability, resilience, recoverability, and operational readiness. Stand up monitoring, alerting, logging, and service health capabilities, including public-facing uptime and status ...

Senior Site Reliability Engineer

Location
Knutsford, England, United Kingdom
experienced Senior Site Reliability Engineer to drive reliability, scalability and performance across critical banking systems. This role combines hands‐on SRE engineering with technical leadership, with a strong focus on observability, automation, continuous improvement and optimisation. Responsibilities Build and maintain reliable, scalable and secure infrastructure platforms … solutions. Apply SRE and software engineering practices to improve reliability, availability and performance. Monitor systems, manage incidents and lead complex troubleshooting and root cause analysis. Develop automation using programming and scripting to reduce manual intervention and improve efficiency. Develop and improve observability, monitoring, instrumentation and performance capabilities. ...

Senior Manager of SRE

Hiring Organisation
Hackajob Ltd
Location
Glasgow, Lanarkshire, Scotland, United Kingdom
Employment Type
Permanent
harnessing artificial intelligence and machine learning technologies to develop new products, improve productivity, and enhance risk management effectively and responsibly. As a Senior Manager - SRE at JPMorgan Chase within the AIML Data Platforms and Chief Data and Analytics Team , you will lead a team of 3-7 SRE engineers … teams. Job Responsibilities: Leads a team of Site Reliability Engineers and support critical application 24x7 Implements Site Reliability Engineering (SRE) best practices to ensure reliability, scalability, and performance of data platforms. Identifies opportunities to eliminate or automate remediation of recurring issues to improve overall ...

Senior Manager of SRE

Hiring Organisation
Hackajob Ltd
Location
Paisley, Scotland, United Kingdom
harnessing artificial intelligence and machine learning technologies to develop new products, improve productivity, and enhance risk management effectively and responsibly. As a Senior Manager - SRE at JPMorgan Chase within the AIML Data Platforms and Chief Data and Analytics Team , you will lead a team of 3-7 SRE engineers … teams. Job Responsibilities: Leads a team of Site Reliability Engineers and support critical application 24x7 Implements Site Reliability Engineering (SRE) best practices to ensure reliability, scalability, and performance of data platforms. Identifies opportunities to eliminate or automate remediation of recurring issues to improve overall ...

Senior Manager of SRE

Hiring Organisation
Hackajob Ltd
Location
Milton Keynes, England, United Kingdom
harnessing artificial intelligence and machine learning technologies to develop new products, improve productivity, and enhance risk management effectively and responsibly. As a Senior Manager - SRE at JPMorgan Chase within the AIML Data Platforms and Chief Data and Analytics Team , you will lead a team of 3-7 SRE engineers … teams. Job Responsibilities: Leads a team of Site Reliability Engineers and support critical application 24x7 Implements Site Reliability Engineering (SRE) best practices to ensure reliability, scalability, and performance of data platforms. Identifies opportunities to eliminate or automate remediation of recurring issues to improve overall ...

Senior Site Reliability Engineer

Hiring Organisation
GCS
Location
Glasgow, City of Glasgow, United Kingdom
Employment Type
Permanent
Salary
£75000 - £95000/annum Bonus
experienced Senior Site Reliability Engineer to drive reliability, scalability and performance across critical banking systems. This role combines hands-on SRE engineering with technical leadership, with a strong focus on observability, automation, continuous improvement and optimisation. Responsibilities: * Build and maintain reliable, scalable and secure infrastructure platforms … solutions. * Apply SRE and software engineering practices to improve reliability, availability and performance. * Monitor systems, manage incidents and lead complex troubleshooting and root cause analysis. * Develop automation using programming and scripting to reduce manual intervention and improve efficiency. * Develop and improve observability, monitoring, instrumentation and performance capabilities. ...