101 to 125 of 975 Site Reliability Engineering Jobs in the UK

Site Reliability Engineer – Fintech / Linux

Hiring Organisation
Quant Capital
Location
London, United Kingdom
Salary
£ 80 K
Site Reliability Engineer – Fintech/Linux Site Reliability Engineer – Fintech/Linux85,000 Plus BonusQuant Capital is urgently looking for a Site Reliability Engineer to join our high profile client.Our client is a major global financial exchange, driven by technology. They … latency trading environment for their clients. They have grown massively and recently were voted in the top 50 fintech firms globally.Day to Day the Site Reliability Engineer will:Analyzing and optimizing trading platform Monitoring development activities, change management tickets Monitoring U.S. production, disaster recovery, and certification systems ...

Senior PHP Engineer - Operational Support - Permanent - London/Hybrid - £70,000 - 85,000

Hiring Organisation
Robson Bale Ltd
Location
London, United Kingdom
Employment Type
Permanent
Salary
GBP 70,000 - 85,000 Annual
efficiency and reduce MTTR. Use Datadog, Kibana and Heap to investigate issues, understand customer impact and improve service health. Partner with Software Engineering, SRE and Platform teams to improve observability, reliability and operational readiness. Contribute to service onboarding, post-incident reviews and continuous improvement initiatives. What … Looking For Experience in Application Operations, Production Engineering, SRE or a similar operational engineering environment. Experience supporting cloud-hosted applications, ideally PHP services running on Kubernetes (EKS) with MySQL databases. Strong troubleshooting and diagnostic skills, with the ability to understand systems end-to-end. Hands-on experience with ...

Software Engineer III, Site Reliability Engineering, Traffic Network Load Balancing

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Science or Engineering. 2 years of experience designing, analyzing, and troubleshooting large-scale distributed systems. About the job Site Reliability Engineering (SRE) combines software and systems engineering to build and run large-scale, massively distributed, fault‐tolerant systems. SRE ensures that the company Cloud's services … internally critical and our externally‐visible systems—have reliability, uptime appropriate to customer's needs and a fast rate of improvement. Additionally SRE’s will keep an ever‐watchful eye on our systems capacity and performance. Much of our software development focuses on optimizing existing systems, building infrastructure ...

Site Reliability Engineer

Hiring Organisation
Jobleads-UK
Location
Ipswich, England, United Kingdom
This is a blended Network Operations and Site Reliability Engineering role within Professional Services, combining hands‐on network engineering with SRE principles to ensure the reliability of BT's fixed network infrastructure. You will implement flawless network change, resolve network and platform issues, and drive … automation to improve reliability and efficiency — developing your skills across network operations and SRE disciplines to deliver brilliant customer experience. You will build effective working relationships internally and externally and contribute as a technically talented network and reliability expert, challenging those around you to perform at their best. ...

Site Reliability Engineer (DV Security Clearance)

Hiring Organisation
CGI
Location
Gloucestershire, United Kingdom
Employment Type
Full Time
inclusive employer and a member of myGwork – the largest global platform for the LGBTQ+ business community. Please do not contact the recruiter directly. Site Reliability Engineer (DV Security Clearance) Position Description At CGI, we deliver mission-critical technology solutions that help protect the UK's national interests … support some of the country's most important programmes. As a Site Reliability Engineer, you will play a key role in ensuring the reliability, scalability and performance of critical services, helping teams deliver resilient platforms that operate at scale. Working within a collaborative and innovative environment ...

Site Reliability Engineer- London

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
extend and will be a hybrid role that will be based in London. Our client is seeking an experienced Site Reliability Engineer (SRE) with a strong focus on Observability and Monitoring Platforms. The successful candidate will play a key role in enhancing the organisation's monitoring, alerting … streamline operational processes and improve reliability. Collaborate with engineering, infrastructure, and support teams to improve system resilience and operational performance. Define and implement SRE best practices, including monitoring standards, alert management, incident response, and operational readiness. Perform troubleshooting and root cause analysis of platform and application issues. Support capacity ...

Software Engineer, GPU Infrastructure- ChatGPT Engineering

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
About the Team ChatGPT Engineering builds and operates the compute platform powering one of the world's largest AI products. Every ChatGPT conversation relies on massive GPU clusters serving inference workloads with high reliability, efficiency, and performance. As our GPU fleet continues to grow, we're investing … production infrastructure, preferably GPU clusters or other compute-intensive distributed systems. Have a background in Production Engineering, Site Reliability Engineering (SRE), Infrastructure Engineering, or Platform Engineering. Have built software that automates operational workflows rather than relying on manual processes. Have experience with Kubernetes, Linux systems ...

MongoDB-Site Reliability

Hiring Organisation
Barclays
Location
Knutsford, Cheshire, United Kingdom
Salary
£ 80 K
expected to demonstrate the Barclays Mindset – to Empower, Challenge and Drive – the operating manual for how we behave. Join our team as a MongoDB Site Reliability Engineer, where you'll be at the forefront of designing and maintaining robust, high-performance systems that power critical financial services. … solving, multi-layered problems and building systems that perform reliably amid shifting priorities, we encourage you to apply.To be successful as a MongoDB Site Reliability Engineer, you should have experience with:Working in Site Reliability Engineering, DevOps, and MongoDB administration in financial services.Using MongoDB features ...

Software & Data Engineers

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
native, low‐latency system that operates at global scale and underpins critical investment products used worldwide. You will work on challenging problems across software engineering where performance, data quality, and reliability are non‐negotiable, leveraging modern cloud (AWS) and AI‐assisted development tooling to accelerate delivery without compromising … encouraged to apply. Areas We Value Experience In Software Engineering Data Engineering Cloud & Platform Engineering Site Reliability Engineering (SRE) AI & Machine Learning DevOps & Automation Architecture & Distributed Systems Analytics & Data Platforms Benefits LSEG offers a range of tailored benefits and support, including healthcare, retirement planning ...

Lead Product Manager AIOPs

Hiring Organisation
S&P Global
Location
London, United Kingdom
Salary
£ 80 K
responsible for S&P Global's enterprise AIOps platform and strategy, driving the modernization of IT Operations and Site Reliability Engineering (SRE) through intelligent observability, event intelligence, automation, and AI-driven insights.DTS Platform & Tools – Service Enablement: We serve as thought leaders in AIOps, partnering across IT Operations … SRE, engineering, infrastructure, service management, and application teams to solve enterprise operational challenges. Our mission is to improve reliability, reduce operational complexity, optimize technology investments, and enable more proactive and resilient technology operations by applying AI.Responsibilities and Impact:Own and execute the AIOps product roadmap, aligning priorities with ...

Site Reliability Engineer

Hiring Organisation
Trimble Navigation
Location
Belfast, Down, United Kingdom
Salary
£ 60 K
runbooks and procedures for incident response and operational tasks.Collaborate with cross-functional teams to review and provide feedback on technical designs, ensuring alignment with SRE principles.Participate in on-call rotations and handle critical incidents with confidence and expertise.Continuously improve documentation for systems and services, contributing to a knowledge-sharing culture … tools and incident management processes like Prometheus, Grafana, New Relic, DataDog, Splunk, Cloudwatch, Sumologic etc.Extensive understanding of networking and security concepts.Bonus Points For:Specialized SRE observability experience with New Relic or DataDog.Familiarity with OpenTelemetry, AIOps, MLOps, or SecOps.Logistics:Location: Newcastle, UK - In-Office (at least 4 days per week ...

Lead Product Manager AIOPs

Hiring Organisation
S&P Global
Location
Greater London, United Kingdom
Employment Type
Full Time
responsible for S&P Global's enterprise AIOps platform and strategy, driving the modernization of IT Operations and Site Reliability Engineering (SRE) through intelligent observability, event intelligence, automation, and AI-driven insights. DTS Platform & Tools - Service Enablement: We serve as thought leaders in AIOps, partnering across … Operations, SRE, engineering, infrastructure, service management, and application teams to solve enterprise operational challenges. Our mission is to improve reliability, reduce operational complexity, optimize technology investments, and enable more proactive and resilient technology operations by applying AI. Responsibilities and Impact: Own and execute the AIOps product roadmap, aligning ...

Site Reliability Engineering (SRE) / Observability Technical Lead

Hiring Organisation
NTT DATA
Location
London, United Kingdom
Salary
£ 80 K
team you'll be working with:We are seeking an experienced Site Reliability Engineer (SRE)/Observability Technical Lead to join our team and drive the strategy and execution of observability and reliability projects across our clients. The ideal candidate will have deep expertise in Application Performance … will guide the design, implementation, and continuous improvement of observability solutions, ensuring system reliability, performance, and scalability while fostering best practices in SRE and DevOps.What you'll be doing:Lead the strategic development and management of observability and reliability frameworks across the organization, ensuring alignment with business goals ...

Platform Engineer (Security)

Hiring Organisation
Jobleads-UK
Location
United Kingdom
manages all applications and next steps. Our partner is looking for a Platform Engineer (Security) based in United Kingdom. Join a cloud-native engineering environment where security is embedded into every stage of platform development and operations. In this role, you will strengthen the security posture of modern infrastructure … integrate security into platform operations and development workflows. Requirements 3–6 years of experience in Platform Engineering, Site Reliability Engineering (SRE), DevOps, Cloud Engineering, Security Engineering, or a related field. Hands-on experience securing production cloud environments on AWS, Microsoft Azure, Google Cloud Platform ...

SRE Technical Lead

Hiring Organisation
Adecco
Location
Reading, Berkshire, England, United Kingdom
Employment Type
Full-Time
Salary
£70,000 - £90,000 per annum
SRE Technical Lead Reading/Hybrid (UK-based - mix of home, office, and client site) Must be eligible for SC Clearance We are seeking an experienced SRE Technical Lead to act as the technical authority for Site Reliability Engineering across complex, large-scale platforms. This … , availability, and operational excellence across multi-team and multi-vendor environments. You will combine hands-on engineering expertise with strategic leadership, ensuring SRE practices are embedded across the full service lifecycle-from design through to production operations. As the SRE Technical Lead, you will: Define and implement SRE ...

Engineering Manager, DevOps

Hiring Organisation
Jobleads-UK
Location
United Kingdom
About the Engineering Organization The Engineering Team at Loop is a balance of agility, consistency, and performance. These are the pillars that allow the team to constantly and consistently deliver value that matters to customers. That customer intimacy is what allows our engineering teams … joining high-severity incident calls when needed. Your experience: 2+ years of DevOps leadership managing DevOps or Site Reliability Engineering (SRE) teams. 7+ years in hands-on platform or infrastructure roles. You have a strong self-service track record, having delivered internal platforms or portals that empowered ...

Site Reliability Engineer, Intermediate to Senior Staff — Infrastructure Platforms

Hiring Organisation
GitLab
Location
United Kingdom
Salary
£ 60 K
mindset, and the ability to learn quickly. We'll support you in becoming successful with GitLab's tools, systems, and ways of working.How our SRE hiring worksBecause this is a single application for SRE roles across Infrastructure Platforms, our process is built to evaluate you once and match you well … looking for, and the level and teams that fit, so we can point your process in the right direction.Core Technical: The shared assessment every SRE candidate takes, regardless of eventual team. A low-stress, collaborative discussion covering source code, system architecture, and incident review.Peer Technical: Team-specific depth ...

SRE Architect (68019) (DEAI DS) Cloud & Data Engineering United Kingdom

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
balance reliability with feature velocity Conduct chaos engineering exercises and game days to validate resiliency and uncover hidden failure modes Mentor 2 SRE Engineers, establish engineering standards, and build a culture of reliability and continuous improvement Collaborate with Platform Engineering and Cloud teams to embed … Strong analytical and problem-solving mindset with attention to detail Ability to manage competing priorities across multiple workstreams simultaneously QUALIFICATIONS & EXPERIENCE 7+ years in SRE, DevOps, or production engineering with 3+ years in a senior or lead capacity Proven track record of improving availability, reducing MTTR, and implementing self ...

Senior Site Reliability Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Senior Site Reliability Engineer (SRE) - GCP/Kubernetes We are seeking an experienced and highly motivated Senior Site Reliability Engineer (SRE) to join our small, agile engineering team. This role offers the unique opportunity to drive the reliability, scalability, and performance of our core … Kubernetes application deployment. Monitoring & Observability: Implement and manage robust monitoring, alerting, and logging solutions to ensure clear system visibility and proactive issue identification. Reliability & Performance: Define, measure, and enforce Service Level Objectives (SLOs) and Service Level Indicators (SLIs). Participate in on-call rotation (if applicable) and lead post ...

Senior Site Reliability Engineer

Hiring Organisation
Brevan Howard
Location
London, United Kingdom
Salary
£ 80 K
Senior Site Reliability Engineer (SRE) - GCP/KubernetesAbout the RoleWe are seeking an experienced and highly motivated Senior Site Reliability Engineer (SRE) to join our small, agile engineering team. This role offers the unique opportunity to drive the reliability, scalability, and performance ...

Senior Site Reliability Engineering Manager

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Role Overview Sr. Manager, Site Reliability Engineering (London) is an experienced leader responsible for overseeing a globally distributed team of SRE technologists with diverse skills in software development, systems, network, application, and/or database management. This role ensures seamless, continuous coverage of Cboe's real‐time … features; monitor development activities, change‐management tickets, evaluate impact; approve and execute daily change tickets; organize testing prior to deployment; work with software engineering to resolve systemic issues; ensure compliance obligations are met. Incident Response & Escalation Management: Serve as senior escalation point for production incidents across European ...

Oracle Site Reliability Engineer

Hiring Organisation
Barclays
Location
Knutsford, Cheshire, United Kingdom
Salary
£ 60 K
DescriptionPurpose of the roleTo apply software engineering techniques, automation, and best practices in incident response, to ensure the reliability, availability, and scalability of the systems, platforms, and technology through them. AccountabilitiesAvailability, performance, and scalability of systems and services through proactive monitoring, maintenance, and capacity planning.Resolution, analysis and response … incident management, and production troubleshooting.Knowledge of Oracle RAC, Data Guard, GoldenGate, ASM, and disaster recovery architectures.Exposure to cloud platforms (OCI, AWS, Azure) and modern SRE practices such as reliability engineering, capacity planning, and service resilience.You may be assessed on the key critical skills relevant for success in role ...

Site Reliability Engineer

Hiring Organisation
Randstad
Location
London, United Kingdom
Salary
£ 55 K
Site Reliability Engineer (SRE) - 100% RemoteLocation: Fully Remote Duration: PermanentAre you passionate about building unbreakable systems and automating away the noise We are looking for a dedicated Site Reliability Engineer (SRE) to join our remote team. Your primary mission will be to design, implement, and maintain … tackling complex challenges in the Azure ecosystem and sharing your knowledge with others, we want you on our team!What You Will DoAs an SRE, you will be accountable for the delivery and support of production and non-production systems within the Azure ecosystem. Your day-to-day responsibilities will ...

Site Reliability Engineer

Hiring Organisation
Randstad Digital
Location
London, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£60,000
Site Reliability Engineer (SRE) - 100% Remote Location: Fully Remote Duration: Permanent Are you passionate about building unbreakable systems and automating away the noise? We are looking for a dedicated Site Reliability Engineer (SRE) to join our remote team. Your primary mission will be to design, implement … complex challenges in the Azure ecosystem and sharing your knowledge with others, we want you on our team! What You Will Do As an SRE, you will be accountable for the delivery and support of production and non-production systems within the Azure ecosystem. Your day-to-day responsibilities will ...

Site Reliability Engineer (SRE)

Hiring Organisation
Reward Gateway
Location
London, United Kingdom
Salary
£ 60 K
become available for a Site Reliability Engineer to join our team to help us transform our existing operational workloads to an SRE approach.What’s In It For Me A chance to be part of an extremely well established, stable and high growth ‘Unicorn’ SaaS company with over … teams work from our Dean Street office two days per week.What You’ll be Doing:Integrating tightly with our Product Engineering teamsFollowing SRE practices and maintaining high standards of complianceImplementing a new standard of observability utilising SLI/SLO/Error BudgetsContinually evolving our observability platforms for greater coverageUsing ...