401 to 425 of 587 Site Reliability Engineering Jobs in London

Senior DevOps Engineer

Location
Greater London, England, United Kingdom
technical aspects of our platform. What you'll do Help develop and maintain a robust monitoring and observability framework to ensure system performance, reliability, and early issue detection. Leverage your experience in DevOps to design, maintain, and improve cloud infrastructure. Own and evolve incident management processes, including on-call … defining the role with us. You have a strong understanding of monitoring and observability tools (e.g., Datadog or Grafana). You have experience applying SRE principles such as SLOs, and error budgets. You have worked with incident management processes, including conducting postmortems and driving improvements. You have a solid understanding ...

AWS Engineering Lead

Hiring Organisation
Interact Consulting Limited
Location
South West London, London, United Kingdom
Employment Type
Permanent
Engineering Lead Remote-first | London office 1 day per month | £90,000-£100,000 base + 15%+ pension Lead the infrastructure and operations that power engineering at scale for this leader in the fintech space. You'll shape our AWS platform, architecture, security, reliability and developer … sustainable engineering. Mentor engineers and drive continuous improvement. Partner with engineering, security and business stakeholders. What you'll bring Leadership experience in Operations, SRE, DevSecOps, Cloud or Platform Engineering. Strong AWS, CI/CD, security and production operations experience. Expertise in Terraform or OpenTofu. Knowledge of GitHub Actions, Docker ...

Fastly: Senior SRE – Networks

Location
Greater London, England, United Kingdom
world’s most prominent companies, including GitHub, Yelp, Paramount, and JetBlue. We’re building a more trustworthy Internet. Come join us. Senior SRE – Networks Fastly’s Technical Operations (TechOps) team is responsible for building & operating the infrastructure that powers the Fastly Edge Cloud Platform. We oversee an extensive global network … media streaming services to e-commerce, to open source projects, nonprofits and more. We take this responsibility seriously! We’re looking for an experienced SRE with strong focus on Networks, SRE principals and knowledge of the internet who can bring engineering best practices, operational rigour, and automation skills ...

Trading Systems Engineer - Prop Trading

Hiring Organisation
Quant Capital
Location
London, UK
Employment Type
Full-time
applications at the bleeding edge. Required Skills & Experience: Experience supporting mission critical systems and high performance applicationsMinimum of 5 years working in trade support, site reliability engineeringBachelor's degree in STEM or related fieldFamiliarity with trading platforms and financial marketsThrives in high-pressure situations while working alongside traders … developers and other engineering teamsStrong problem-solving skills and the ability to troubleshoot technical issues under pressureKnowledge of Linux/Unix environmentsExperience with scripting languages such as Python and Bash for automation tasksAbility to devise complex SQL database queries and updatesBasic networking knowledge, including multicast, TCP/ ...

Senior Platform Engineer

Location
Greater London, England, United Kingdom
your institutional memory. The team you’ll be joining and impact you’ll have A small Platform team with a very wide remit: cloud, SRE, DevOps, developer experience, compliance, and the infrastructure behind our AI products. We run AWS first across multiple regions, with GCP and Azure alongside. Our model … control. What you’ll be doing from day one: Owning our infrastructure as code in Terraform, plus alerting, observability (Prometheus, Grafana) and reliability, including load testing and disaster recovery exercises. Making CI/CD faster (GitHub Actions, ArgoCD) and taking obstacles out of engineers’ way, from build times ...

Senior Cloud Infrastructure Engineer (cloud sandboxes)

Location
Greater London, England, United Kingdom
almost doubled in size and continue to grow across all areas in 2026. LocalStack is headquartered in Zurich/Switzerland, with a small engineering office in Vienna/Austria and remote team members from 25 countries including the US, FR, UK, CA, ES, and many more! Check our Notion … Chef cookbooks, automation scripts and tools) Experience we expect you to bring to the role 6+ years of experience as a Cloud Engineer/SRE/DevOps/DevSecOps/Infrastructure Engineer or similar role. 5+ years of expertise in using AWS, such as EC2, S3, IAM, Lambda ...

DevOps Engineer

Location
Greater London, England, United Kingdom
developer tooling that enable product teams to move fast without breaking things. Our work is often behind the scenes, but its impact is everywhere: reliability, scalability, security, and developer experience. We solve hard, systemic problems, make deliberate trade-offs, and hold ourselves accountable for the long-term health … cloud infrastructure in AWS Enable development teams to be owners of their systems and to deliver high impact products to our users Partner with SRE and security teams to help developers build resilient and secure applications Participate in our on-call and incident management process as part ...

Azure DevOps Engineer - SC Cleared

Hiring Organisation
CBSbutler Holdings Limited trading as CBSbutler
Location
London, United Kingdom
Employment Type
Contract
Contract Rate
£625 - £675/day
PowerShell Experience producing technical designs and documentation Excellent communication and problem-solving skills Passion for automation, innovation and emerging technologies Desirable Skills Understanding of SRE principles Experience with PKI and TLS Experience working within secure or government environments Security Requirements - Essential Active and transferable UK SC clearance Clearance generally less … national identity documentation Client and project details are confidential. Skills: DevOps, CI/CD, AWS, Azure, GCP, Kubernetes, Docker, Jenkins, Git, Python, Bash, PowerShell, SRE ...

Head of Platform Engineering & Cloud

Location
Greater London, England, United Kingdom
resilient platform supporting critical banking services while transitioning to a platform‐centred, product‐led operating model. The successful candidate will drive platform engineering, SRE, and NOC capabilities, embed DevOps and IaC practices, manage large budgets and vendors, and ensure #J-18808-Ljbffr ...

CD&A Data Engineer

Location
Greater London, England, United Kingdom
operations across BigQuery, Airbyte, Prefect, dbt, and related analytics systems. Contribute to program governance, modernization activities, observability improvements, and platform hardening. Assist with monitoring, SRE practices, L3 incident support, and uptime/data quality tracking. Provide reporting and analytics support, including validations, schedules, permissions, and catalogue maintenance. This … technical problems. Experience working with AI-assisted software engineering and rapid prototyping techniques. What will help you on the job Understanding of SRE concepts (SLO/SLI’s, uptime, alerts, MTTR). Experience with GCP DevOps, GitHub, YAML, AI assisted coding, and basic repo/pull request workflows. Exposure ...

Senior SRE - AI Automation & Self-Healing (Remote UK)

Location
Greater London, England, United Kingdom
Principle HR is partnering with a global technology leader to hire a Senior Site Reliability Engineer for AI-driven automation. This remote UK role focuses on building self-healing systems that reduce firefighting and keep a large VR product healthy. You will own backend and cloud services, deploy ...

Junior Endpoint SRE – Hybrid London | Equity

Location
Greater London, England, United Kingdom
MLabs is seeking a skilled DevOps Engineer in London to manage corporate endpoint hardware, ensuring optimal onboarding processes and ongoing support. This hybrid role offers competitive salary between £50K and £60K, with opportunities for professional ...

Principal Platform Engineer

Hiring Organisation
Trayport
Location
London, UK
Employment Type
Full-time
hiring a Principal Platform Engineer to be a senior technical anchor in our Platform/Operations function. This is a 70% hands-on engineering role: you'll design, build, and operate the infrastructure that keeps a global trading platform running, while acting as a technical mentor and design authority … oneContribute to the technical roadmap for the platform function, balancing reliability investment against deliveryWhat we're looking forMust haveA background in Operations or SRE running highly available, redundant production platforms — you understand failure domains, graceful degradation, and what "five nines" costsDeep hands-on experience with AWS (VPC design, networking ...

Middle/Senior DevOps Engineer

Location
Greater London, England, United Kingdom
professionals who want ownership, commercial impact, and the opportunity to build systems that scale in B2B iGaming. Head of Sales Head of SEO Junior+ Site Reliability Engineer Middle/Senior DevOps Engineer For high-level projects and exploring the roles, visit https://lnkd.in/dPzYzZUT #J ...

Lead AI Infrastructure & Distributed Systems Engineer

Hiring Organisation
LinuxRecruit
Location
London, UK
Employment Type
Full-time
training bottlenecks, and drastically accelerate research cycles. This is a deeply technical, hands on position. Rather than managing people or executing traditional DevOps/SRE maintenance, the successful candidate will be coding daily, designing scalable distributed systems from scratch, and squeezing maximum throughput out of large scale GPU compute clusters. ...

Edge SRE: Scale Global Reliability & Automation

Location
Greater London, England, United Kingdom
Cloudflare is seeking a Senior Systems Reliability Engineer (SRE) for our Edge platform. You will build and operate a globally distributed infrastructure, focusing on automation, reliability, and scale across more than 320 cities in over 120 countries. The role emphasizes improving availability and performance while supporting a follow ...

Cloud Infrastructure Engineer - Digital Assets

Location
Greater London, England, United Kingdom
operational ownership through incident response, post-incident improvement, and continuous hardening. Who you are Demonstrated expertise in owning production infrastructure as an engineer (Platform, SRE, or Infrastructure engineering background). Deep cloud infrastructure knowledge, especially around VPC design, compute optimization, and performance networking features. Proven ability to write ...

Senior Platform Engineer

Hiring Organisation
9fin
Location
London, UK
Employment Type
Full-time
ensure that engineers can release as early and often as possible. Designing and implementing a developer portal, to provide a service catalog to the engineering team, and also author many other useful DevOps plugins. Contributing to observability best practices and providing key SLI/SLO metric reporting, so that … CloudFront, S3, RDSGood understanding of monitoring and logging solutions. We use OpenTelemetry, AWS Cloudwatch and SigNoz so experience with them is a bonus. Basic SRE knowledge, and experience in alerting and incident management platforms (eg. incident.io, Pagerduty)Proven ability to provide and support strong and scalable CI/CD pipelinesLinux ...

Senior PM, Storage & Networking for AI GPU Platform

Location
Greater London, England, United Kingdom
networking strategy underpinning our GPU platform. You’ll influence bare‐metal clusters, Kubernetes, high‐performance storage, and data‐centre networking, partnering with engineering, SRE, and operations to deliver scalable, reliable infrastructure for AI workloads. You will drive prioritisation, define success metrics, and guide capacity planning while balancing performance, reliability ...

DevOps & Environment Lead (AWS / Cloud Platforms)

Location
Greater London, England, United Kingdom
deployment validation and health checks into CI/CD. Build monitoring and alerting with minimal false positives. Manage and mentor DevOps engineers. Partner with Engineering, Product, and Operations teams. Ensure operational excellence and secure, scalable cloud environments. Core Responsibilities Deliver end-to-end environment lifecycle automation. Create templates … Value Strong experience managing and automating cloud environments on AWS. Expertise in CI/CD and DevOps practices. Experience with monitoring, alerting, and SRE methodologies. Financial industry experience is a plus. How We Work: Engineering at FTSE Russell Our engineers operate with a clear 'why' — staying aligned to strategic ...

Platform Engineer

Location
Greater London, England, United Kingdom
feature roadmaps. Our squads include Product Managers and Product Designers and part of our DNA as a company is achieving amazing things through collaboration. Engineering squads follow DevOps principles, with teams running what they build – and the platform team sits alongside them, making it easy for them … plans for our business; Receive mentoring, development, and training to ensure your skills are kept up to date; Have an opportunity to contribute to SRE and incident response. Who will you work with Lead Platform Engineer, Engineering Team etc. About you Hands‐on experience with running services ...

Engineering Lead AWS Platform, Infrastructure & Operations

Hiring Organisation
Interact Consulting Limited
Location
East London, London, United Kingdom
Employment Type
Permanent, Work From Home
long standing and exciting client of ours in the Fintech space are once again expanding and looking to bring in an Engineering Lead whose focus will be on transforming financial services through technology and helping more people feel comfortable with their finances. This is a remote role with once … deliver scalable, secure solutions. Provide technical leadership while empowering your team to own outcomes. You must have: Proven experience leading engineering, infrastructure, DevOps, SRE or Operations teams of 3years+ Strong AWS and cloud infrastructure expertise - 10years+ Experience with CI/CD, Infrastructure as Code and observability. Knowledge of Docker ...

Mid-Level DevOps Engineer

Location
Greater London, England, United Kingdom
share of the client ticket queue without escalating routine work. Comfortable on the on-call rota and working through the incidents we see (scaling, site-down responses, health issues from bad traffic), with a senior to escalare to for deeper investigation. Hands-on with our main IaC stack (Terraform … . Knowing the platform well enough to notice when something’s off. What we need from you Two to five years in a DevOps, SRE, platform, or sysadmin role, with at least two of those on AWS in production. Comfortable with infrastructure-as-code. We use Terraform and Chef. Strong ...

Senior Application Security Engineer (AI)

Location
Greater London, England, United Kingdom
challenge the status-quo and help it reach new horizons. Where you come in? As our Senior Application Security Engineer, you’ll empower our engineering and product teams to safely adopt AI and embed robust technical security controls directly into our services and workflows. By designing cloud-native security … lead technical security discussions. Strong ability to write and review code (preferably in Python, TypeScript, Kotlin and Terraform), with a background in software development, SRE, DevOps, or practical security engineering. In-depth understanding of securing AI/LLM applications, agentic tooling, and developer AI tools, including prompt injection mitigation, model ...

Principal, Technology Business Partner

Hiring Organisation
Pearson
Location
London, UK
Employment Type
Full-time
understand objectives, challenges, market context, and technology needs. Translate business priorities into technology strategies and roadmaps that align OCTO initiatives with business unit goals. Engineering Delivery Leadership: Work closely with engineering, product, architecture, and platform teams to execute complex technology initiatives with strong governance, operational discipline, modern engineering … system integration, cybersecurity, and digital product delivery, with the credibility to guide architecture and engineering decisions. Delivery Excellence: Demonstrated success implementing Agile, DevOps, SRE, product-centric operating models, and engineering best practices that improve delivery quality, speed, and business value. Executive Influence: Experience partnering with C-suite executives ...