201 to 225 of 518 Site Reliability Engineering Jobs in London

Trading Systems Engineer, Trading Platform

Location
Greater London, England, United Kingdom
communicate directly with exchanges, traders and developers to iron out any problems that arise. Qualifications & Skills: Minimum of 5 years working in trade support, site reliability engineering or related fields Bachelor’s degree in STEM or related field Familiarity with trading platforms and financial markets Thrives … high-pressure situations while working alongside traders, developers and other engineering teams Strong problem-solving skills and the ability to troubleshoot technical issues under pressure Excellent communication skills, both written and verbal Knowledge of Linux/Unix environments Experience with scripting languages such as Python and Bash for automation ...

Site Reliability Engineer / Production Support

Hiring Organisation
Hackajob Ltd
Location
London, United Kingdom
Employment Type
Permanent
Salary
GBP Annual
hackajob is partnering directly with Monument to hire for this role. Site Reliability Engineer/Production Support Location London (Oxford Circus) Hybrid: 2 days per week Reports to Head of Cloud Operations ABOUT MONUMENT We're building something genuinely rare: a financial brand designed for the mass affluent ...

Site Reliability Engineer

Location
Greater London, England, United Kingdom
iGaming company based in London. Estimated Benefits Health Insurance Pension Stock Options Benefits estimated based on industry standards We’re hiring a Site Reliability Engineer to join our London team This is a fantastic opportunity for someone passionate about reliability, scalability and automation. You’ll be pivotal ...

Junior Trading Support Engineer - Prop Trading

Hiring Organisation
Quant Capital
Location
London, UK
Employment Type
Full-time
incidents. Required Skills & Experience: Experience supporting mission critical systems and high performance applications Minimum of 1 years working in trade support, app support or site reliability engineering Bachelor's degree in STEM or related field Have exposure to VCS, particularly Git/Github Have demonstrated scripting abilities … Comfortable with Linux and the command-line Financial Services experience Someone who thrives in high-pressure situations while working alongside traders, developers and other engineering teams Experience developing proprietary process automation and monitoring tools to streamline software configuration and rollout procedures The environment is that of Facebook or Google ...

Principal Platform Engineer

Hiring Organisation
Sanderson Recruitment
Location
City of London, London, United Kingdom
Employment Type
Permanent
large-scale distributed systems and database platforms? We're looking for a hands-on technical leader to help shape the future of our platform engineering capability. This is an opportunity to lead complex engineering initiatives, define technical strategy, and act as a subject matter expert across AWS infrastructure … technical authority for distributed database and persistence technologies Required Experience 8+ years' experience in Platform Engineering, Infrastructure Engineering, DevOps, SRE or Software Engineering Expert-level AWS infrastructure experience Strong Infrastructure as Code expertise with Terraform Strong Linux systems administration and networking knowledge Experience designing and operating distributed ...

Principal Platform Engineer

Hiring Organisation
Sanderson
Location
London, South East England, United Kingdom
Employment Type
Full-Time
Salary
£100,000 - £150,000 per annum, Inc benefits
large-scale distributed systems and database platforms? We're looking for a hands-on technical leader to help shape the future of our platform engineering capability. This is an opportunity to lead complex engineering initiatives, define technical strategy, and act as a subject matter expert across AWS infrastructure … technical authority for distributed database and persistence technologies Required Experience 8+ years' experience in Platform Engineering, Infrastructure Engineering, DevOps, SRE or Software Engineering Expert-level AWS infrastructure experience Strong Infrastructure as Code expertise with Terraform Strong Linux systems administration and networking knowledge Experience designing and operating distributed ...

Senior Platform Engineer

Location
Greater London, England, United Kingdom
diversified range of trading strategies. We employ over 130 colleagues in Jersey, Geneva, London, Singapore, New York and Shanghai. This is a hands‐on engineering role within the Platform Engineering team, part of Technology Operations. Platform Engineering is responsible for building and operating the infrastructure, platforms … comply with all organisational, statutory and regulatory policies and procedures. Experience, Knowledge & Skills Five or more years of experience in platform engineering, DevOps, SRE, infrastructure engineering or a closely related role. Strong experience operating production or production‐like infrastructure, ideally across hybrid cloud and on‐premises environments. Hands ...

DevOps Engineer – Security & Intelligence

Location
Greater London, England, United Kingdom
secure platforms that underpin mission-critical digital services. You'll operate within multi-disciplinary Agile teams, collaborating closely with software engineers, test engineers, architects, SRE specialists and mission stakeholders to ensure robust, production-grade delivery. This role is suited to someone who thrives in complex, secure environments and enjoys working … automated platform capabilities Supporting AWS-based environments, including Kubernetes, OpenShift, EKS and ECS Implementing observability, monitoring, logging and alerting for live services Supporting SRE practices, cloud migration activities and production platform operations Job Responsibilities Design, implement and maintain secure CI/CD pipelines to support efficient, automated software delivery Develop ...

Senior SRE: Scale, Reliability & Automation Leader

Location
Greater London, England, United Kingdom
Jobtailor is seeking a senior Site Reliability Engineer to tackle reliability, scalability, and efficiency challenges across SRE and development teams. You will build and run large-scale, distributed fault-tolerant systems that power the Genesis platform, optimize existing infrastructure, and cut toil through automation to improve uptime. ...

Application Support Engineer

Hiring Organisation
London Stock Exchange Group
Location
London, UK
Employment Type
Full-time
hours on‐call rota, ensuring continuity and resilience of critical clearing services. Support extended clearing operations as part of a late‐shift Site Reliability Engineering rota (up to 10:30 pm London time) to accommodate increased US trading hours, with scope evolving as the global support model … stability and operational efficiency. Contribute to IT Change and Governance forums to prioritise enhancements supported by clear casesCollaborate with development teams to improve release reliability and automation, using the approved LSEG DevOps toolsetLead capacity and performance management, improving monitoring capabilities to ensure platforms scale in line with business growth ...

Platform Engineer Graduate Considered

Hiring Organisation
RedTech Recruitment Ltd
Location
East London, London, United Kingdom
Employment Type
Permanent
Salary
£55,000
company working on complex, large-scale software systems. You will join a specialist platform team responsible for the infrastructure, tooling and automation that enables engineering teams to develop, deploy and operate software effectively. This is a hands-on role with Kubernetes at its core, offering the opportunity to work … . Keywords: Platform Engineer/Kubernetes Engineer/DevOps Engineer/Cloud Platform Engineer/Infrastructure Engineer/Site Reliability Engineer/SRE/Cloud Engineer/Kubernetes Platform Engineer/Systems Engineer/Production Engineer/Cloud Infrastructure Engineer/Kubernetes/Terraform/CloudFormation/Infrastructure ...

Senior Network SRE: Cloud Reliability & IaC

Location
Greater London, England, United Kingdom
Miro is seeking a Senior Network Site Reliability Engineer to help strengthen reliability, availability, and scalability of our production environment. You will focus on cloud automation, IaC, and governance across our AWS infra, contributing to highly available services for millions of users. You will own automation, observability ...

AI-Powered Observability Tech Lead

Location
Greater London, England, United Kingdom
Collaboration Technology Group in London seeks a Technical Leader to drive architectural vision and implement an AI-powered Production Intelligence platform. You will blend Site Reliability Engineering with agentic AI to improve monitoring, incident response, and auto-remediation across global SaaS infrastructure. You will mentor engineers, shape ...

Senior Product Manager for AI Observability

Hiring Organisation
London Stock Exchange Group
Location
London, UK
Employment Type
Full-time
with AI Evaluation PM (previous role), Model Risk, GSSR, Legal and Compliance to align telemetry with governance frameworks. Work closely with Engineering and SRE teams to drive observability improvements and reliability engineering for AI systems. Optimisation & Insights Identify cost inefficiencies across model and MCP usage, and drive … support auditability, compliance and explainability requirements. Skills & Competencies Required Experience in product management with a strong foundation in observability, telemetry, data platforms, monitoring, or SRE/DevOps-driven products. Understanding of LLMs, embeddings, vector search, MCP tools, and AI inference workflows. Deep familiarity with logging, tracing, metrics, and event-based ...

Solutions Architect, Studios — Pre-Sales & Architecture

Location
Greater London, England, United Kingdom
solutions across client contracts, new business opportunities, and procurement activity. The role emphasizes pre-sales engagement, proofs of concept, and collaboration with DevOps and Site Reliability Engineering teams to ensure automatable, observable, and secure solutions. The position reports into IMG Studios and involves working across commercial, technical ...

Manager, DE , TC, FS

Location
Greater London, England, United Kingdom
go. Join EY and help to build a better working world. Experience: 8-12 years | Rank: Manager | Team: Financial Services Organisation, Technology Consulting, Digital Engineering Role Summary: Hands-on Digital Engineering Manager responsible for leading software engineering delivery across EY FSO client engagements. The role combines engineering … such as Kafka, API management, data quality and operational data analysis Production support Monitoring, logging, incident management, RCA, performance tuning, runbooks, ServiceNow/Jira, SRE‐aligned operational readiness Experience delivering technology change for banking, insurance, wealth, asset management, risk, finance, payments or market infrastructure clients Consulting capabilities Stakeholder management, workshops ...

Senior SRE Engineer: Reliability, Cloud & Automation

Location
Greater London, England, United Kingdom
London Stock Exchange Group is looking for a Senior Engineer in Site Reliability who will join a driven team focused on system availability, performance, and scalability. Responsibilities include maintaining service level objectives, writing automation for system resilience, and partnering with development teams. Required qualifications include a Bachelor … computer science, experience in Object Oriented programming and cloud systems, and DevOps familiarity. The role is pivotal in ensuring 24/7 system reliability and promoting engineering best practices. #J-18808-Ljbffr ...

Research Engineer, Safety Oversight, DeepMind

Location
Greater London, England, United Kingdom
technical products. Experience in the domain area of generative AI and Large Language Models (LLMs). Preferred qualifications: Master’s degree or PhD in Engineering, Computer Science, or a related technical field. 3 years of experience developing code, running experiments and analyses collaboratively with coding agents. Experience building large … Software Engineer, Full Stack, Google AdsGoogle-2w agoLondon, UKFull-time14DetailsG### Software Engineer III, Full Stack, Publisher InventoryGoogle-2w agoLondon, UKFull-time15DetailsG### Software Engineer III, Site Reliability Engineering, Traffic Network Load BalancingGoogle-2w agoLondon, UKFull-time15Details## Explore related hubsCountry hubUnited Kingdom JobsCompany pageGoogle JobsSalary pageSoftware Engineer SalaryVisa pageSkilled ...

Staff Software Engineer - Physical AI

Location
Greater London, England, United Kingdom
services. Implement reliable workflow orchestration patterns. Own CI/CD pipelines, automated testing, and deployment strategy for platform services. Drive reliability, observability, and SRE practices (monitoring, alerting, incident response, performance and scaling). Team Mentorship & Collaboration Mentor senior and mid‐level engineers, elevating the technical capabilities of the entire … experience, including time in a staff‐level engineering role. Broad software engineering depth spanning deployment and CI/CD, automated testing, and SRE/production reliability (observability, incident response). Deep expertise in distributed systems architecture, including building and operating large‐scale distributed backend systems ...

AI-Powered Production Intelligence Tech Lead

Location
Greater London, England, United Kingdom
Cisco is seeking a Technical Leader to drive architectural vision for an AI-powered Production Intelligence platform. You will blend Site Reliability Engineering with agentic AI to improve monitoring, diagnosis, and auto-remediation across global SaaS infrastructure. You will lead architecture, mentor engineers, and partner across teams ...

Principal Platform Engineer

Location
Greater London, England, United Kingdom
impact***** ****Build and maintain infrastructure-as-code (Terraform or similar), CI/CD pipelines, and Kubernetes platforms as products consumed by engineering teams********SRE & reliability***** ****Define and drive SLOs, error budgets, and observability standards (metrics, logging, tracing) across the platform***** ****Take part in post-incident reviews and help … roadmap for the platform function, balancing reliability investment against delivery****## ****What we're looking for********Must have***** ****A background in Operations or SRE running highly available, redundant production platforms — you understand failure domains, graceful degradation, and what "five nines" costs***** ****Deep hands-on experience with AWS (VPC design ...

Senior SRE: GCP, Kubernetes & Automation Leader

Location
Greater London, England, United Kingdom
Brevan Howard CFD LTD is seeking a Senior Site Reliability Engineer (SRE) to enhance the reliability, scalability, and performance of its core platform. The successful candidate will provide operational support and lead infrastructure projects. This role requires hands-on experience with Google Cloud Platform (GCP) and Kubernetes ...

Site Reliability Engineer- Spacetime UK

Location
Greater London, England, United Kingdom
Role Overview This isn't a "keep the lights on" SRE role. This is a strategic, high-impact opportunity to build the nervous system for a platform that transforms how networks of satellites, ground stations, and fleets are interconnected and orchestrated. You will be building the core observability stack that … cloud-native tools to a robust, scalable, and insightful platform built on best-in-class technologies (Prometheus, OpenTelemetry, etc.). If you are an SRE who thrives on platform-building challenges and wants to be relied upon to build a production-grade observability stack from the ground up, this role ...

Infrastructure Engineer

Location
Greater London, England, United Kingdom
Build resilient, scalable, fault-tolerant infrastructure for WRITER's high-traffic enterprise generative AI platform Move between SRE, DevOps, Infrastructure, and Platform initiatives as priorities shift Automate operational tasks and infrastructure management with Python or Go Design and operate infrastructure across AWS, GCP, and Azure Work with Kubernetes, Helm, Terraform … requests Encode recurring infrastructure tasks as reusable internal skills for human and agent teammates Lead incident response, post-mortems, and root-cause analyses Own reliability, performance, and efficiency of core services end-to-end Define and uphold SLOs and error budgets and carry the on-call pager Balance immediate ...

Senior Cloud Engineer

Hiring Organisation
Pacific Life
Location
London, UK
operational excellence, security, reliability, and lifecycle management of cloud applications across multiple regions. The role sits at the intersection of cloud architecture, SRE, security, and developer experience, helping to raise the overall standard of cloud usage across PL Re. While experience in platform provisioning is required, this role does … hours on‐call rotas as required. Support the full lifecycle of application workloads, including maintenance, patching, backups, upgrades, and controlled decommissioning. Reliability, SRE & Observability Drive improvements in reliability, resilience, performance, and efficiency through strong SRE practices. Own and enhance observability and monitoring, ensuring meaningful alerting, clear operational dashboards ...