276 to 300 of 357 Site Reliability Engineering Jobs in London

Head of DevOps and Platform Automation

Hiring Organisation
VIQU IT
Location
London, Cornhill, United Kingdom
Employment Type
Permanent
platform services Lead and develop a distributed DevOps and tooling team across UK and offshore locations, driving a culture of continuous improvement and engineering excellence Oversee end-to-end platform lifecycle management ensuring availability, resilience, security and compliance across all tooling environments Drive standardisation and optimisation of CI/… Automation: Extensive experience (10+ years) within DevOps, infrastructure, platform engineering or tooling/automation leadership roles Strong background leading enterprise-scale DevOps or SRE teams within complex, global or regulated environments Deep expertise in CI/CD tooling such as Jenkins, GitLab CI and strong understanding of Infrastructure ...

DevOps Engineer

Hiring Organisation
BAE Systems
Location
Greater London, United Kingdom
Employment Type
Full Time
working in a multiple disciplined team, and require a broad range of technical and soft skills to enable the team to implement sound DevOps engineering practices and deliver value quickly and continuously. These skill are categorised into the following domains. Automation skills : Automation is a key skill domain … work within a team using Agile methodology Scrum – DevOps engineers should be an active member of the scrum team and contribute to sprint ceremonies SRE – Should understand SRE principles and apply these to constantly improve the reliability and minimise the support burden within the team Security Clearance ...

Management Consultant - Cloud DevOps

Hiring Organisation
Capgemini
Location
Greater London, United Kingdom
Employment Type
Full Time
Salary
500000 GBP Annually
implement modern technologies, processes, and operating models, delivering sustainable transformations. Cross-Functional Leadership: Demonstrated leadership in cross-functional teams, fostering collaboration between product, engineering, security, and operations to deliver cohesive platform solutions. Innovation & Emerging Tech: Awareness of trends in platform engineering, such as platform orchestration, internal developer portals … improve speed, productivity, and quality, and implement product-centric operating models. Solid understanding of hybrid/multi-cloud environments, DevOps, CI/CD, SRE, DevSecOps models, DevX, build and deployment pipelines, observability, and ITIL.Optional: Required certifications, licenses or languages Understanding of observability and monitoring platform; Experience of having collaborated with ...

Applied AI ML - Senior Associate - Machine Learning Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
through careful design of libraries and services to be leveraged across the team. This role offers a unique blend of scientific research and software engineering, requiring a deep understanding of both mindsets. Job responsibilities Build robust Data Science capabilities which can be scaled across multiple business use cases Collaborate … both technical and non-technical audiences Document approaches taken, techniques used and processes followed to comply with industry regulation Collaborate closely with cloud and SRE teams while taking a leading role in the design and delivery of the production architectures for our solutions. Act as an individual contributor, though there ...

Applied AI ML - Senior Associate - Machine Learning Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
through careful design of libraries and services to be leveraged across the team. This role offers a unique blend of scientific research and software engineering, requiring a deep understanding of both mindsets. We recognize that our people are our strength and the diverse talents they bring to our global … both technical and non‐technical audiences Document approaches taken, techniques used and processes followed to comply with industry regulation Collaborate closely with cloud and SRE teams while taking a leading role in the design and delivery of the production architectures for our solutions. Act as an individual contributor, though there ...

Director, Cloud Infrastructure

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
able to spar with strong infrastructure engineers, make hard architectural calls, and build the operating model around them: what Platform owns centrally, what SRE enables through teams, and how product teams deploy and run their services with confidence. What you will do: Set the infrastructure strategy for Sanity's next … routing, gateways, CI/CD, infrastructure as code, and observability. Draw a clean line between Platform and SRE. Platform should own the shared foundations. SRE should help product teams run services well, with strong tooling, standards, and incident support. Raise the reliability bar across Sanity's production systems, including ...

Senior SRE: Cloud Reliability (Azure/AWS, Terraform, K8s)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
United States Digital Space LLC in London is seeking a Principal SRE to help scale automated, zero-downtime infrastructure as we move to public cloud (Azure). You will partner with global development teams to define patterns, standardise practices, and drive reliable, scalable systems. The role emphasizes hands‐on design ...

Senior Vice President, Senior Software Engineering, Developer Experience

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Overview We’re seeking a future team member for the role of Senior Vice President, Senior Software Engineering, Developer Experience to join our Developer Experience (DevEx) team. This role is located in London. Responsibilities Lead, mentor, and manage a high-performing Build Engineering team, setting clear technical direction … bottlenecks and failures. Champion secure software supply chain practices, artifact management standards, and consistent build strategies across the enterprise. Partner closely with Product Management, SRE, Security, and application engineering teams to ensure alignment with developer needs and organizational priorities. Foster innovation by evaluating and adopting emerging technologies, including ...

Platform Operations Director

Hiring Organisation
ClearCourse
Location
City of London, London, United Kingdom
estate - from end-user devices across 40+ portfolio companies through to cloud production environments. Reporting to the CTTO, the role owns infrastructure, internal IT, SRE, security operations implementation, and vendor cost management. The role operates at the intersection of operational excellence and commercial discipline - owning cloud FinOps outcomes, leading … across the group. Internal IT & Systems Manages internal IT and business systems administration (M365, NetSuite, SuccessFactors, SharePoint) -infrastructure, integrations, and IAM. Ensures observability and SRE capability is fit for purpose across cloud, hosted, and end-user environments. Vendor & Cost Management Drives cloud and vendor cost discipline - manages the Vendor & FinOps ...

Senior Kubernetes & Cloud Platform Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
, and partner with multiple engineering teams to scale infrastructure safely. The role combines hands-on engineering with strategic platform improvements, emphasizing SRE practices, incident response, and mentoring colleagues across the org. #J-18808-Ljbffr ...

Senior Software Engineer, Substrate

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
will also be responsible for ensuring scale, stability and security across a matrix of compliance regimes and hosting infrastructure types. Your team culture emphasizes engineering rigor and operational excellence at scale. This means issues in production should be pre-empted and deeply root-caused, and investments in automation … Deep familiarity with containers (Docker) and orchestration (Kubernetes) at scale Experience working with a cloud provider (AWS/Azure/GCE), or sysadmin/SRE experience in data centers Experience designing, building, and operating high-scale observability or infrastructure systems Working knowledge of networking fundamentals, experience with CNIs or cloud ...

Observability Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
best ideas take time to evolve. Together we’re building a world‐class platform to amplify our teams’ most powerful ideas. Role The Observability Engineering Team manages the doors – both entry and exit – to the telemetry backends at G‐Research, ensuring our engineers can effectively produce and consume telemetry … DevOps tooling (Terraform, ArgoCD, Helm, Jenkins) Experience with metrics, logs and tracing backends Coding in Go, Python, or similar Industry background in Observability or SRE Desirable Experience Profiling (eBPF, Pixie, Parca) Synthetic monitoring, AI‐observability tools and Kafka Benefits Highly competitive compensation plus annual discretionary bonus Lunch provided (via Just ...

SRE, London, UK

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
BIG. Operating at our scale, across multiple geographically dispersed data centers and servicing hundreds of millions of users presents unique challenges. As an SRE at the company, you'll need to solve these problems using data, teamwork, and your own expertise. SREs at the company own the full infrastructure stack … close partnership with our development teams and aim to design & build new services together. We're passionate about software and automation in SRE and develop a variety of tooling and infrastructure. Our services run on mixed & hybrid platforms. Responsibilities Create outstanding customer experience, and help developers write better code faster ...

Cloud SRE

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
looking to hire a Cloud SRE to help with our client's Azure infrastructure estate by contributing to discovery, planning, and supporting activities, including virtual machines, monitoring tools, networking, storage and databases. Day to day, you will monitor service health, resolve incidents, execute controlled changes, maintain patching and hardening standards … improve reliability through automation, observability, and clear operational documentation. This is a 6-month contract, to work remotely (outside IR35) Key Responsibilities Infrastructure discovery and assessment - Build and maintain an inventory of servers, applications, and configurations. Change execution - Plan and execute changes, including applications deployments, configurations changes and validation. ...

Senior Software Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
coordinated remediation alongside the incident commander. Use structured diagnostics before escalating — attach clear evidence, reproducibility steps, and impact assessments to every L3/SRE handoff. Feed operational findings into Problem Management and contribute to post-incident reviews; capture learning in improved runbooks, alerts, and automation. Help define, measure, and report … accurate, accessible, and kept up to date. Work closely with the Director of Application Operations, Problem Manager, and PETO peers (Platform, Infrastructure, Data, SRE) to ensure a coherent, joined-up operational approach. Partner with product-aligned engineering teams to understand application architecture, service dependencies, and failure modes; encode this ...

Senior Software Engineer, Event Streaming Systems

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Systems Join us in our mission to transform the way people shop and eat, where impact, innovation and growth drive everything we do. Our Engineering teams tackle complex technical challenges across a global, three-sided marketplace, building and scaling systems that serve millions of customers, riders and partners every … operational toil and improve platform resilience. Contribute to incident response, operational readiness and continuous improvement for Tier-0 infrastructure. Partner closely with EventBus, Storage, SRE, Data Platform and Core Infrastructure teams across US and EU regions. Help shape the long-term architecture and evolution of Deliveroo’s event streaming platform. ...

Cloud Platform Engineer – AWS & GCP | IaC/SRE

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
contract (London/Belfast) to accelerate its cloud transformation. You design, build, and run scalable cloud platforms on AWS and Google Cloud, focusing on reliability and security. … will implement IaC with Terraform, CloudFormation, and Deployment Manager, automate provisioning, and integrate with CI/CD pipelines while mentoring teammates and upholding SRE practices. #J-18808-Ljbffr ...

Senior DevOps Engineer: Platform & Reliability Lead

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Auriga is seeking a Senior DevOps Engineer to architect, lead and scale the platform engineering and operational backbone of our solutions. You will define and maintain roadmaps for infrastructure, CI/… observability, reliability and security, across cloud, on-premises, hybrid and customer-site environments. The role requires 5+ years in DevOps/SRE, strong AWS experience, and a background in robotics or mission-critical systems. #J-18808-Ljbffr ...

Head of Technology Resilience and Product Operations

Hiring Organisation
Deerfoot Recruitment Solutions Limited
Location
City, London, United Kingdom
Employment Type
Permanent
Salary
GBP 120,000 - 140,000 Annual
ability to lead confidently and calmly under pressure Desirable: ITIL 4, ITSM tooling (ServiceNow, Jira Service Management), ISO22301/ISO20000, CBCP/MBCI certification, SRE familiarity, or experience with tools such as Splunk, CyberArk PAM, or GenAI-driven service management Ready to take on a role where your leadership genuinely … Problem Management, Head of Change and Release Management, Business Continuity Director, Head of Operational Resilience, IT Service Continuity Manager, Head of DevOps/SRE Operations, ServiceNow, ITIL 4, ISO22301, DORA compliance, Disaster Recovery, AWS/Azure/Oracle Cloud, CBCP, MBCI. Deerfoot Recruitment Solutions Ltd is a leading independent tech ...

Head of Technology Resilience and Product Operations

Hiring Organisation
Deerfoot Recruitment Solutions Limited
Location
London, Coleman Street, United Kingdom
Employment Type
Permanent
Salary
£120000 - £140000/annum Benefits + Bonus + Hybrid Working
ability to lead confidently and calmly under pressure Desirable: ITIL 4, ITSM tooling (ServiceNow, Jira Service Management), ISO22301/ISO20000, CBCP/MBCI certification, SRE familiarity, or experience with tools such as Splunk, CyberArk PAM, or GenAI-driven service management Ready to take on a role where your leadership genuinely … Problem Management, Head of Change and Release Management, Business Continuity Director, Head of Operational Resilience, IT Service Continuity Manager, Head of DevOps/SRE Operations, ServiceNow, ITIL 4, ISO22301, DORA compliance, Disaster Recovery, AWS/Azure/Oracle Cloud, CBCP, MBCI. Deerfoot Recruitment Solutions Ltd is a leading independent tech ...

Data Scientist

Hiring Organisation
Jobleads-UK
Location
City Of London, England, United Kingdom
emerging AI technologies and recommend practical applications across the organisation. Data Science & Analytics Perform advanced statistical analysis and modelling across large-scale operational and engineering datasets. Develop feature engineering, experimentation and model evaluation frameworks. Create insights and recommendations that drive decision making for senior stakeholders. Build and maintain … tools. Desirable Skills Experience within financial services, regulated environments or market infrastructure. Familiarity with Azure AI services and Microsoft AI ecosystem. Experience with observability, SRE or operational analytics use cases. Experience developing AI-powered automation solutions. Knowledge of Responsible AI and model governance frameworks. Behaviours & Competencies Strong problem-solving capability ...

Data Scientist

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
emerging AI technologies and recommend practical applications across the organisation.### Data Science & Analytics* Perform advanced statistical analysis and modelling across large-scale operational and engineering datasets.* Develop feature engineering, experimentation and model evaluation frameworks.* Create insights and recommendations that drive decision making for senior stakeholders.* Build and maintain … tools.### Desirable Skills* Experience within financial services, regulated environments or market infrastructure.* Familiarity with Azure AI services and Microsoft AI ecosystem.* Experience with observability, SRE or operational analytics use cases.* Experience developing AI-powered automation solutions.* Knowledge of Responsible AI and model governance frameworks.## Behaviours & CompetenciesThe successful candidate will demonstrate ...

Global SRE Fleet Engineer — Automation & Reliability

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Cisco is seeking an experienced SRE/DevOps professional to join the SRE Fleet team. You will develop and maintain automation solutions that improve reliability, scalability, and efficiency of infrastructure spanning thousands of machines across global cloud environments. You will design deployment pipelines, testing frameworks, and tooling to support ...

Security Architect - Anti-Piracy

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
security assessments across IPTV ecosystems, applications, set‐top boxes (STBs), Android devices, cloud backends, and associated technologies. Partner with engineering, platform, data, and SRE teams to improve security visibility, reduce risk exposure, and enhance the overall security posture of Anti‐Piracy systems. Define, document, and promote security standards, architectural … both technical and non‐technical stakeholders. Strong collaboration and stakeholder engagement skills, with the ability to work effectively across engineering, platform, data, SRE, and business teams, and to create and maintain high‐quality architecture, security, and technical documentation. A proactive and curious mindset, demonstrating a willingness to learn, challenge ...

Engineering Manager, Agent Experience

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
stakeholder discussions, and align with product, design and other engineering teams on priorities. Ensure the team maintains high operational excellence – managed technical debt, SRE and healthy on‐call. Hire and grow the team, and stay technically engaged through architecture and design discussions. What You’ll Need to Thrive … balance servant leadership with making the tough calls. Nice to Have Experience with infrastructure and cloud tooling (e.g. Kubernetes, AWS, Datadog). An SRE background – on‐call experience, SLOs and alerts management. Experience building internal tooling, operational platforms or other complex, data‐dense product UIs. Why Join Us? Solve meaningful ...