276 to 300 of 587 Site Reliability Engineering Jobs in London

Senior SRE: Front‐Office Trading Reliability

Location
Greater London, England, United Kingdom
Engineer within its Trading Technology group in London. You will embed with the software engineering team that builds front‐office trading platforms, shaping SRE patterns, observability, and resilience across globally distributed systems and AI‐accelerated incident response. You will partner with traders and senior stakeholders, lead incident responses, implement … reliable code, and drive modern SRE practices across CI/CD, telemetry, and automated #J-18808-Ljbffr ...

DevOps / SRE Engineer (London)

Location
Greater London, England, United Kingdom
Bumper continues to scale across the UK and Europe — and we’re excited to welcome a passionate and experienced DevOps/SRE Engineer , based in our London office, to help us take our platform reliability and engineering excellence to the next level. Our Head Office is in Sheffield … with the vision of being the leading automotive payment and insights platform! A bit about the role... We’re looking for a DevOps/SRE Engineer to join our growing Infrastructure team. In this role, you’ll help shape and strengthen the reliability, scalability and observability of our cloud ...

Senior DevOps Analyst

Hiring Organisation
NTT DATA
Location
London, UK
Employment Type
Full-time
Azure/GCP.Implement Infrastructure as Code (IaC) using Terraform, CloudFormation, or ARM templates. Build and manage containerized applications using Docker and Kubernetes. Ensure system reliability, scalability, security, and performance. Implement monitoring, logging, and alerting solutions. Collaborate with development, QA, and security teams. Troubleshoot production issues and lead incident response. … stack, Splunk. Version Control: Git (GitHub/GitLab/Bitbucket).Security: IAM, secrets management, vulnerability scanning. Experience & Qualifications8+ years of experience in DevOps/Site Reliability/Infrastructure Engineering. Strong understanding of DevOps and CI/CD best practices. Experience supporting high-availability, production systems. Experience in Agile ...

Golang Lead Engineer

Hiring Organisation
Lloyds Banking Group
Location
London, UK
Employment Type
Full-time
lead the technical development and delivery of backend services that are reliable, secure, scalable and easy to operate! This role sits within Software Engineering and offers the opportunity to shape the Group's Mobile Security Core and wider backend capabilities. You'll combine hands-on engineering with technical … skills, with the ability to communicate technical trade-offs clearly and influence outcomes. The Skills Taxonomy identifies relevant skills including Software Engineering, DevOps, SRE & Service Engineering, Technology Leadership, Critical Thinking, Decision Making, Coaching & Feedback and Stakeholder Management. And any experience of this would be really usefulExperience modernising legacy ...

Platform Engineer

Location
Greater London, England, United Kingdom
Daintta are a rapidly growing, values-driven consultancy delivering mission-critical data, technology, and engineering solutions across the UK public sector. We partner with clients in Defence, National Security, Law Enforcement, and wider Government to strengthen the UK’s resilience and operational advantage through innovative, data-driven insights. … teams to design, develop, and implement platform and infrastructure solutions that meet our clients’ requirements. Key Responsibilities Collaborate with clients to understand their platform engineering needs, business objectives, and operational constraints, acting as a trusted technical advisor Design, develop, and implement secure, reliable, and scalable platform solutions across cloud ...

Staff Platform Engineer - AI Native SaaS Platform

Location
City Of London, England, United Kingdom
2025. Their platform is redefining the sector, and with revenues nearly 10x since the start of last year, they're continuing to expand their Engineering team to match the ambition of their product and customers. The product is real-time, data-rich and AI-native, creating complex engineering … infrastructure as code Strong understanding of networking, IAM, managed services and secure cloud architecture Experience owning observability, logging, APM or distributed tracing tooling An SRE mindset across SLOs, reliability, incident response and toil reduction A track record of technical leadership, mentoring and delivering complex projects through others Strong systems ...

Founding Platform Engineer

Location
Greater London, England, United Kingdom
implementation quarters. Make data residency, access controls, observability and recovery part of the product from the start. Shape Hive's security posture alongside the engineering team. Keep the platform dependable, understandable and calm under pressure. What you bring 5+ years of experience in cloud infrastructure, platform engineering, DevOps … site reliability engineering. Strong experience running production infrastructure with AWS, Azure or GCP, Kubernetes, infrastructure as code and CI/CD. A degree in Computer Science, Engineering or a related technical field. Relevant cloud and security certifications are a strong plus. Experience working with security, data-residency ...

Managed Service Operations - Head of Practice

Location
Greater London, England, United Kingdom
service operations including incident, problem, change, event, monitoring, resilience, continuity, capacity, and on‐call models. Strong understanding of ITIL practices blended with modern DevOps, SRE, Agile and platform‐engineering approaches. Broad technical awareness across cloud platforms, application architectures, data platforms, networks, observability tooling, security‐by‐design, and automation. Ability … engineering, service readiness, and live‐service best practices. Key experiences Running and growing operational or engineering teams in a Managed Service, SRE, DevOps, or live‐service environment—with responsibility for hiring, coaching, development and performance. Leading high‐pressure operational functions including incident management, problem resolution, major incident ...

Senior SRE, Observability & Cloud Reliability

Location
Greater London, England, United Kingdom
company is seeking a Principal Site Reliability Engineer, Infrastructure Observability to guide a team of SREs focused on observability, reliability, and scalable cloud/on‐prem solutions. … role requires hands‐on expertise and collaboration with diverse partners to drive measurable improvements. The ideal candidate has extensive cloud experience, DevOps/SRE leadership, and strong automation skills, with a track record in designing resilient systems and implementing effective monitoring and #J-18808-Ljbffr ...

Senior SRE Technical Lead — Reliability & Observability

Location
Greater London, England, United Kingdom
leading global financial markets infrastructure provider is seeking a Technical Lead SRE in Greater London. In this role, you will enhance the reliability engineering capabilities, collaborating with various teams to establish observability standards and ensure operational excellence. The ideal candidate will have over 10 years of experience … SRE or related fields, strong AWS and Kubernetes skills, and a proven track record in building resilient platforms. Join us to make a significant impact in financial markets infrastructure. #J-18808-Ljbffr ...

SRE | Permanent | London, Hybrid, AWS

Hiring Organisation
Source Group International
Location
London, UK
Employment Type
Full-time
Reference: 56925About the job Role: Site Reliability EngineerType: Full-time permanent roleLocation: Hybrid, London City - 3 days per week on-siteSalary: 90,000 per annumIndustry: Technology - Gaming Platforms Our Client is a premier provider of high-volume software solutions for the global iGaming and predictive analytics sector. With … analysis, and continual improvement actions. Collaborate cross-functionally to raise standards for stability, security, performance, and compliance. Required skills & experience 3+ years' experience in SRE, Platform, or DevOps roles within production environments. Strong Kubernetes operational experience (on-prem and AWS EKS).Hands-on experience defining and operating SLOs/SLIs ...

SRE for AI-Driven Financial Infrastructure

Location
Greater London, England, United Kingdom
company Group is seeking a Site Reliability Engineer (SRE) to improve, manage, and monitor production-critical infrastructure and data pipelines. You will work on production systems, sometimes embedded with software teams, to increase reliability and reduce operational risk. We value a growth mindset, strong problem-solving ...

Head of Engineering, Supply Chain

Location
Greater London, England, United Kingdom
most senior leaders sit, which makes it a natural base for a Head of Engineering. So, now we are looking for our Head of Engineering, Supply Chain to join us in the UK. About The Role POS Hardware and Device Platforms owns the software behind the physical side … strategy and architecture for inventory and logistics tools and repair management software. Leading an engineering organisation of approximately 25 engineers, including managers, QA, SRE and Data Engineering. Drivingplatformisation of inventory and supply chain systems so operational data is visible, trustworthy and consistent across millions of terminals. Bridging ...

Principal Engineer (AWS & Java)

Hiring Organisation
London Stock Exchange Group
Location
London, UK
Employment Type
Full-time
most senior hands‐on technical expert for our leading Risk Screening platform. This role partners closely with architecture, service, product, and engineering teams to ensure that agreed designs are implemented to the highest technical and operational standards. A core focus of the role is ownership of the system … Java development to tackle the hardest technical challenges, proactively identify operational risks, and provide deep technical guidance that sets the highest standard across all engineering teams working on the product. This is the most senior individual contributor role with broad influence, focused on technical excellence, delivery confidence, and long ...

Senior Systems Reliability Engineer (SRE), Edge

Location
Greater London, England, United Kingdom
Senior Systems Reliability Engineer (SRE), Edge Cloudflare ·United States, London, United Kingdom Job Description About Us At Cloudflare, we are on a mission to help build a better Internet. Today the company runs one of the world’s largest networks that powers millions of websites and other Internet properties … capabilities. We own a wide portfolio of applications and services, running a tight feedback loop of developer and operator patterns. The ideal SRE candidate has a passionate curiosity about how the Internet fundamentally works and has a strong knowledge of networking, Linux and TLS along with coding ability ...

Site Reliability Engineer

Hiring Organisation
Wheely
Location
London, UK
Employment Type
Full-time
About WheelyWheely is redefining premium transportation across major cities in Europe, the US, and the Middle East. We blend cutting-edge technology with the craft of five-star chauffeuring to deliver an experience trusted by ...

Senior Site Reliability Engineer

Hiring Organisation
CISCO Systems
Location
London, UK
Employment Type
Full-time
Meet the TeamCisco's Webex Engineering Group is redefining the future of collaboration. We're building a world where people connect effortlessly to enjoy modern, uncompromised collaboration across every room, desk, pocket, and application. Our technology powers industry-leading products like Webex Meetings, Webex Calling, and Webex Contact Center. … Adaptable & Problem-Solver: Address complex challenges across configuration, policy, observability, and data services. Apply a data-driven approach using Prometheus and Grafana to improve reliability and performance. Ownership & Quality: Own end-to-end configuration quality, enforcing governance with Open Policy Agent. Ensure secure, compliant deployments and full reproducibility, including ...

Lead DevSecOps Engineer

Location
Greater London, England, United Kingdom
Legal Entity: AECOM Engineering Technology Services Ltd Business Line: Corporate Work Location Model: On-Site Operating Group: Corporate Company Description Work with Us. Change the World. At AECOM, we're delivering a better world. Whether improving your commute, keeping the lights on, providing access to clean water … make clear trade-offs across developer experience, reliability, performance, cost, and security Qualifications Must-Have Qualifications Demonstrated hands-on experience across DevOps, SRE, platform engineering, DevSecOps, or related roles, with experience guiding technical decisions or mentoring colleagues Experience building and maintaining CI/CD pipelines and release automation ...

Senior Software Engineer (GO)

Hiring Organisation
Source Group International
Location
London, UK
Employment Type
Full-time
scale cloud-native backend platforms that power enterprise AI and GenAI initiatives across a global banking environment. This role will focus on developing reusable engineering blueprints, establishing LLMOps best practices, and delivering secure, scalable microservices architectures that operate across hybrid cloud environments spanning public cloud providers and private data … event processing, and backend performance for large-scale distributed systems. Observability & Reliability Implement comprehensive observability solutions using: PrometheusOpenTelemetryGrafanaEstablish monitoring, tracing, logging, alerting, and SRE best practices. Drive operational excellence through performance tuning, reliability engineering, and proactive incident prevention. ...

Senior Lead SRE - Part-Time Job-Share (Reliability)

Location
Greater London, England, United Kingdom
London seeks a Lead Site Reliability Engineer for a jobshare opportunity. You will set the vision and operating model for an SRE transformation, enabling business-aligned teams to achieve higher reliability and better end-user experience. The role emphasizes AI-enabled reliability and scalable practices across ...

Senior/Principal Product Manager - Infrastructure

Location
Greater London, England, United Kingdom
prioritisation, and delivery of the core infrastructure that underpins Radiant’s platform. You’ll work closely with platform engineering, software engineering, networking, SRE, security, and infrastructure teams across Kubernetes, networking, storage, and shared platform services. Key Responsibilities: Own the roadmap, backlog, prioritisation, and delivery for Radiant’s core … platform infrastructure. Work with engineering teams to understand requirements across Kubernetes, networking, storage, compute, and shared infrastructure services. Translate technical needs into clear initiatives, stories, acceptance criteria, dependencies, and priorities. Manage work across new platform capabilities, upgrades, configuration changes, reliability improvements, technical debt, and security requirements. Coordinate infrastructure ...

Site Reliability Engineer

Location
Greater London, England, United Kingdom
Meet the Team Cisco's Webex Engineering Group is redefining the future of collaboration. We're building a world where people connect effortlessly to enjoy modern, uncompromised collaboration across every room, desk, pocket, and application. Our technology powers industry-leading products like Webex Meetings, Webex Calling, and Webex Contact … Adaptable & Problem-Solver : Address complex challenges across configuration, policy, observability, and data services. Apply a data-driven approach using Prometheus and Grafana to improve reliability and performance. Ownership & Quality : Own end-to-end configuration quality, enforcing governance with Open Policy Agent. Ensure secure, compliant deployments and full reproducibility, including ...

Observability SRE — Reliability & Telemetry Engineer, London

Location
Greater London, England, United Kingdom
HCLTech in London is seeking an Observability SRE to join the Group Platform Services & Engineering division. The role focuses on administering the production environment, building scalable monitoring, and embedding reliability in products and services. You will work with a global, agile team to enhance telemetry, observability and incident ...

SRE Lead, Athena Core — AI-Driven Reliability

Location
Greater London, England, United Kingdom
JPMorgan Chase & Co. is seeking a Lead Site Reliability Engineer to shape the reliability strategy for the Athena Core team within Markets Technology. You will lead a team, drive resiliency reviews, and mentor engineers across multiple domains. The role focuses on incident leadership, service-level governance … enhanced reliability workflows to prevent outages and enable rapid recovery. Excellent technical breadth and communication are essential. #J-18808-Ljbffr ...

Release Engineer

Location
Greater London, England, United Kingdom
zero-trust integrity, and reproducible build standards. You will build and maintain the secure build systems, artifacts, release tooling, and deployment automation that allow engineering teams to ship high-velocity code without compromising security, compliance, or operational resiliency. What you’ll do: Own the Release Lifecycle: Build, maintain … government and commercial compliance standards. What we are looking for: 4+ years of professional experience in Release Engineering, DevOps, Platform Engineering, or SRE, with a strong focus on secure software delivery pipelines. Deep Cloud-Native & Kubernetes Expertise: Strong hands-on experience with Kubernetes orchestration, Helm, container registries ...