176 to 200 of 359 Site Reliability Engineering Jobs in London

Senior DevOps Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
energized to learn and apply new technologies and skills.* An inquisitive nature, excellent analytical, communication and problem-solving skills* Working with a DevOps and SRE mindset* Excellent scripting skills with knowledge of scripting best practices* Understanding of CI/CD systems* Collaborative, team player with a positive attitude* Able ...

OpenSearch & Observability SRE – Hybrid (London)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
seeking a Site Reliability Engineer to join a London-based project in the finance sector. The role is hybrid and initially a 6-month contract with strong prospects to extend. You will focus on OpenSearch … deployment, monitoring, and observability across critical systems. You will design and operate OpenSearch environments, develop dashboards in Grafana, and support Geneos monitoring, automation, and SRE practices in collaboration with engineering teams. #J-18808-Ljbffr ...

Hybrid SRE Engineer — Observability & Cloud (London)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Reward Gateway’s London team is hiring a Site Reliability Engineer to help transform current workloads toward an SRE model while working in a hybrid setup, visiting the London office twice weekly. The role focuses on observability, high availability and incident management, with collaboration across Product Engineering ...

Director of Software Engineering - Executive Director

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
DESCRIPTION If you are a software engineering leader ready to take the reins and drive impact, we've got an opportunity just for you. As a Director of Software Engineering at JPMorgan Chase within the Engineer's Observability Platforms team, you lead a technical area and drive impact … solution delivery. Job responsibilities Leads technology and process implementations to achieve functional technology objectives in the Observability Platforms space, providing essential services for Site Reliability Engineers, Operations and Engineers across the whole firm Delivers technical solutions that can be leveraged across multiple businesses and domains, this will include ...

Lead DevOps Engineer (AI Early Stage Startup)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
audit-grade controls customers will hold you to Observability: telemetry, metrics, logging, tracing, dashboards, and alerting that keep a complex distributed system legible Reliability, incident response, and the operational runbooks that protect customer trust CI/CD and release automation that keep the shipping cadence high Developer experience: making … standard, and knows what world class infrastructure looks like because they've built it before. At least 4 years in DevOps, platform engineering, SRE, or infrastructure You’ve owned production systems rather than only contributing to them, including multi-tenant or customer-specific environments, and you understand the complexity ...

Platform Engineering Manager (Cloud Foundations) London, UK

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Platform Engineering Manager (Cloud Foundations) London, UK Our vision is to give everyone the belief they can make their move. We aim to make moving simpler, by giving everyone the best place to turn to and return to for access to the tools, expertise, trust, and belief to make … delivery in line with expectations. Align cloud platform strategy and delivery plans with business goals, partnering with technical product manager, DX, DBA, security, SRE, and data teams. Coach and mentor engineers to improve skills, confidence, and impact. Sets clear, achievable development goals and provides actionable feedback. Aligns individual growth plans ...

AI Platform Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
About Us the company is an engineering consultancy bridging Quality Engineering, Cloud Platforms and Developer Experience. We help enterprises reliably bring high-impact digital products to market faster, cheaper, and safer, working with technology leaders facing complex business challenges. We take as much pride in our people, culture … build, test, security scanning, artefact management, and deployment. Implement secrets management and certificate lifecycle automation using HashiCorp Vault or equivalent. Reliability & Security Embed SRE practices: SLOs, error budgets, runbooks, on‐call design, and blameless post‐mortems. Integrate security tooling (SAST, DAST, dependency scanning, policy‐as‐code) into delivery pipelines. ...

Site Reliability Engineer: Kubernetes & GitOps Expert

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Cisco's Webex Engineering Group in London is expanding its DevOps and platform engineering capabilities. The role focuses on Kubernetes-based microservices configurations, GitOps with Argo CD, and secure secret management using HashiCorp Vault. You will work in a hybrid London office, collaborating across teams to deliver scalable ...

AI Ops Platform Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
solutions are production‐ready, auditable, compliant, and scalable across merchant payment use cases. You will also be accountable for the end‐to‐end engineering of GenAI and ML platforms, embedding governance, observability and operational resilience by design, hile enabling teams to deploy and run AI solutions with clarity, assurance … cost optimisation, embedding governance by design through policy as code, alignment to model risk framework expectations, lifecycle traceability and audit‐ready evidence, supported by SRE‐grade monitoring and ongoing optimisation of token usage and compute cost across AI workloads. Some Other Highly Valued Skills May Include Retrieval Augmented Generation ...

SRE Fleet Engineer: Global Infra Automation & Reliability

Hiring Organisation
Jobleads-UK
Location
City Of London, England, United Kingdom
Cisco Systems, Inc. is seeking an experienced Site Reliability Engineer to maintain and expand automation across a global infrastructure. The role focuses on reliability, scalability, and operational excellence for a platform spanning thousands of devices and clouds. You will help design deployment pipelines, testing frameworks, and tooling ...

Senior SRE for Scaled SaaS — Reliability & Performance

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Java Script Works is looking for a Senior Site Reliability Engineer to maintain and monitor production environments in the Greater London area. The ideal candidate will have extensive experience with SaaS production infrastructure, focusing on security, stability, and uptime. Responsibilities include reacting to operational issues, participating ...

Kafka Admin Lead-6months-London

Hiring Organisation
Kirtana Consulting
Location
London, United Kingdom
Employment Type
Contract
Contract Rate
GBP Annual
/VPCs/VNETs; manage DNS, TLS SANs, LB configurations, advertised.listeners, and inter cluster networking. Required Skills & Experience 5-10 years in platform/SRE/DevOps; 3+ years dedicated to Kafka/Confluent admin in production. Strong hands on with Apache Kafka internals: brokers, controllers (KRaft), partitions/replication ...

Site Reliability Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
open and inclusive workplace. Our inclusive environment welcomes skills and experiences from diverse backgrounds, and defines who we are. We're hiring an SRE to help us run and evolve the infrastructure behind Signal AI's decision intelligence platform. You'd be joining a small, collaborative Infrastructure team … working on next AI-augmented operations : Claude Enterprise is deployed across Signal. We want this team to help define what good looks like for SRE: incident triage, runbook generation, capacity planning, cost analysis. This is a strategic investment, not a side project: and we'd love someone genuinely curious about ...

DevSecOps Engineer

Hiring Organisation
CGI
Location
Greater London, United Kingdom
Employment Type
Full Time
every stage of the lifecycle. We combine deep engineering expertise with a culture of collaboration and accountability, enabling you to shape DevOps and SRE best practice while driving measurable impact for our clients. Here, your ideas are valued, your ownership makes a difference, and your growth is supported … code and container platforms, embedding security and resilience from the outset. Working closely with architects, developers and operations teams, you will champion DevSecOps and SRE principles, improving reliability, performance and delivery speed. You will also play a key role in shaping engineering standards, mentoring colleagues and driving continuous ...

Hybrid Cloud SRE Engineer – AWS, Kubernetes & Terraform

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
SLR Consulting in London, UK is seeking a Cloud Infrastructure Engineer to design and implement scalable cloud platforms. You will build resilient systems leveraging AWS, Kubernetes, Terraform, and Helm while driving automation with GitHub Actions ...

Founding Cloud SRE — AI/ML Platform & GPU Compute

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Icehouseventures is seeking a Staff Cloud Site Reliability Engineer to shape the reliability of large-scale AI systems and GPU compute infrastructure. This founding role involves building and scaling reliability foundations for the AI cloud platform and ensuring cloud infrastructure resilience. Responsibilities include operationalizing SLOs, improving ...

Principal Cloud Engineer (Multiple)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
AMIs, Docker, or Serverless Define and lead technical strategy with a strong view on IaC modularisation Work directly with customers, driving cloud transformations Apply site reliability best practices and help shape mature security models What we’re looking for Deep AWS expertise, especially landing zones Strong hands … prevent, detect, and remediate them Proficient in Infrastructure as Code (IaC) using Terraform (or CloudFormation) Confident advising on IaC structure and modularisation Awareness of SRE principles and operational priorities Experience with CI/CD pipelines Strong system design skills Why join us? Industry‐leading AWS expert team ...

Cloud Solutions Architect

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
patterns* Drive adoption and migration from traditional technology platforms to public cloud services.* Recommend tools and methodologies to support Cloud Platform operations and implement SRE practices and team as part of the Digital transformation process.* Partner with Development teams, Cyber Security, Support and Enterprise Architecture teams to deliver and maintain ...

Network Engineer

Hiring Organisation
HCLTech
Location
City of London, London, United Kingdom
first operating model. The role focuses on Network Infrastructure as Code (NetIaC), CI/CD pipelines, AI-driven operations (AIOps), observability integration, and SRE-led reliability engineering. Key Responsibilities Develop and manage Network Infrastructure as Code (NetIaC) using Python, Ansible, and Terraform for provisioning and lifecycle management. Design … Model Context Protocol) and AI agent integrations. Ensure AI-driven compliance, security monitoring, cost optimization, and continuous network stability improvements. Collaborate with Cloud, IAM, SRE, and platform teams to deliver unified control plane operations across Azure, GCP, and Edge. Participate in design reviews focusing on resilience, failure domains, scalability ...

Platform Engineering Team Lead

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Platform Engineering Team Lead (London, UK)Job detailsFunctional Structure/Tech Ops - Platform Engineering (PE)CellPoint Digital LimitedFull-time**Join CellPoint Digital: Shape the Future of Payments with Us!**At CellPoint Digital, we’re revolutionizing the way businesses in the air, travel, and hospitality sectors manage their payments.With … Cloud SQL, Spanner), and robust telemetry pipelines. You will also participate in an on-call rotation, acting as a crucial escalation point for our SRE team during complex incidents.**Technical Expertise:*** **Cloud Native Mastery:** Deep expertise in Google Cloud Platform (GCP), specifically designing and managing high-availability, multi-regional infrastructure. ...

Founding Cloud SRE — AI Platform & GPU Clusters (Hybrid)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Wayve is seeking a Cloud Site Reliability Engineer to build and scale the reliability foundations … their AI cloud platform. This founding role involves defining frameworks, automation, and operational standards, ensuring the infrastructure operates efficiently. Candidates should have experience in SRE roles with strong skills in Kubernetes and cloud systems. The role is full-time with a hybrid working policy, allowing flexibility between office and remote ...

DevOps Engineer

Hiring Organisation
Anson Mccade
Location
South West London, London, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£55,000
government environments. Theyre looking for DevOps Engineers to build, automate and scale critical infrastructure supporting high-impact programmes. The Role Youll be responsible for engineering and automating cloud infrastructure, enabling fast, secure and reliable delivery across complex environments. Key responsibilities of the DevOps engineer: Design, build and manage cloud …/CD pipelines for continuous delivery Support containerised environments (Docker, Kubernetes) Develop automation scripts (Python, Bash, Shell) Work within cross-functional DevOps/SRE teams to improve system reliability and scalability Apply DevOps best practices across monitoring, logging and deployment The DevOps Engineer must have: 2+ years experience ...

DevOps Engineer

Hiring Organisation
Anson Mccade
Location
London, United Kingdom
Employment Type
Permanent, Work From Home
They're looking for DevOps Engineers to build, automate and scale critical infrastructure supporting high-impact programmes. The Role You'll be responsible for engineering and automating cloud infrastructure, enabling fast, secure and reliable delivery across complex environments. Key responsibilities of the DevOps engineer: Design, build and manage cloud …/CD pipelines for continuous delivery Support containerised environments (Docker, Kubernetes) Develop automation scripts (Python, Bash, Shell) Work within cross-functional DevOps/SRE teams to improve system reliability and scalability Apply DevOps best practices across monitoring, logging and deployment The DevOps Engineer must have: 2+ years experience ...

DevOps Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
tooling and documentation that lets researchers self‐serve infrastructure without waiting on you. Skills & Qualifications 4+ years in a DevOps, Platform Engineering, or SRE role. Strong proficiency with at least one major cloud provider and its core services (compute, storage, networking, IAM). Hands‐on experience with infrastructure ...

AWS SRE - Datadog

Hiring Organisation
Vallum Associates
Location
London, United Kingdom
Employment Type
Contract
Contract Rate
£475 - £500/day
Role: AWS SRE - Datadog Location: London, UK Position Type: Contract Inside IR35 Job Description: We are looking for someone with a strong background in Datadog, Geneos, Gitlab and Infra-as-Code (Terraform). You will play a crucial role in our technology team, contributing to the development, deployment, and maintenance … software within agreed upon timelines. Optimize alerting strategies to reduce noise and improve actionable insights. Mentor junior engineers and contribute to the evolution of SRE best practices. Oversee and continuously optimize cloud cost management strategies for our Observability infrastructure in line with Cloud FinOps principles. Collaborate with executive leadership ...