126 to 150 of 158 Site Reliability Engineer Jobs in London

SRE Engineer: Build Reliable, Scalable Systems

Location
Greater London, England, United Kingdom
leading recruitment agency is seeking an SRE Engineer to combine software and IT engineering principles for building reliable systems. Responsibilities include scoping projects, designing software, and automating infrastructure management. Ideal candidates will have proficiency in programming languages like Python or Go, experience with Terraform, and familiarity with CI/ ...

Software Engineering or SRE, PhD Intern, 2027

Hiring Organisation
Hackajob Ltd
Location
South West London, London, United Kingdom
Employment Type
Permanent
hackajob is partnering directly with Google to hire for this role. We offer a range of internships in either Software Engineering or Site-Reliability Engineering across EMEA. Durations and start dates will vary according to project and location. Our recruitment team will determine where you fit best based … continue to push technology forward. You will design, test, deploy and maintain software solutions as you grow and evolve during your internship. Site Reliability Intern: Our engineers create, fix, extend and scale the code to keep it working and to harden it against all the bad actors ...

Senior AWS Site Reliability Engineer

Hiring Organisation
Spectrum IT Recruitment Limited
Location
City, London, United Kingdom
Employment Type
Permanent
Salary
GBP 70,000 Annual
The company deliver cutting-edge enterprise software solutions across both cloud and on-premises environments, empowering organisations to enhance customer experiences, maintain regulatory compliance, and proactively fight fraud. The company are trusted by businesses worldwide ...

Engineer II, Site Reliability (Hybrid, London)

Location
Greater London, England, United Kingdom
commitment to our customers, our community and each other. About the Role CrowdStrike is looking to hire an Engineer II to the TechOps SRE team that will have a focus on our Commercial Cloud. We’re looking for a deeply-technical, hands-on engineer, who loves to develop … with well in a diverse, team-focused environment with other SREs and Engineers Ability to broadly communicate and present recommended conventions defined by the reliability team broadly Proven experience utilizing AI technologies to enhance decision-making, streamline workflows and processes, improve efficiency and drive business outcomes. #LI-GO1 #LI ...

Engineer II, Site Reliability (Hybrid, London)

Hiring Organisation
CrowdStrike
Location
London, United Kingdom
Salary
£ 70 K
mission that matters? The future of cybersecurity starts with you.About the Role:CrowdStrike is looking to hire an Engineer II to the TechOps SRE team that will have a focus on our Commercial Cloud. We’re looking for a deeply-technical, hands-on engineer, who loves to develop … work with well in a diverse, team-focused environment with other SREs and EngineersAbility to broadly communicate and present recommended conventions defined by the reliability team broadlyProven experience utilizing AI technologies to enhance decision-making, streamline workflows and processes, improve efficiency and drive business outcomes.#LI-GO1 #LI-HybridBenefits ...

Software Engineering Tech Lead (SRE + AI)

Location
Greater London, England, United Kingdom
years, or PhD + 3 years in Computer Science, Software Engineering, or a related technical field. Proven record as a Technical Lead or Lead SRE/Software Engineer delivering distributed, high-availability SaaS platforms at scale. Strong proficiency in Python, Go, Java, or C++ with experience designing microservices, APIs … production automation. Deep experience with Kubernetes, Docker, and container orchestration in large-scale multi-cluster environments. Proven background in SRE practices: SLI/SLO design, observability platforms (metrics/logs/traces), incident management, and automated RCA. Preferred Qualifications AI & Agentic Systems: Hands‐on experience building LLM pipelines, AI Agents ...

Software Engineering Tech Lead (SRE + AI)

Hiring Organisation
CISCO Systems
Location
London, United Kingdom
Salary
£ 80 K
years, or PhD + 3 years in Computer Science, Software Engineering, or a related technical field.Proven record as a Technical Lead or Lead SRE/Software Engineer delivering distributed, high-availability SaaS platforms at scale.Strong proficiency in Python, Go, Java, or C++ with experience designing microservices, APIs, and production … automation.Deep experience with Kubernetes, Docker, and container orchestration in large-scale multi-cluster environments.Proven background in SRE practices: SLI/SLO design, observability platforms (metrics/logs/traces), incident management, and automated RCA.Preferred QualificationsAI & Agentic Systems: Hands-on experience building LLM pipelines, AI Agents, Model Context Protocol (MCP) servers ...

Observability SRE — Reliability & Telemetry Engineer, London

Location
Greater London, England, United Kingdom
HCLTech in London is seeking an Observability SRE to join the Group Platform Services & Engineering division. The role focuses on administering the production environment, building scalable monitoring, and embedding reliability in products and services. You will work with a global, agile team to enhance telemetry, observability and incident response ...

Senior Cloud Engineer, AI Platform SRE

Location
Greater London, England, United Kingdom
hurry, no SLOs, no runbooks, and a cost line nobody can explain. We're engaging a Senior Cloud Engineer to bring proper SRE discipline to a major AI platform programme with a client in a high-traffic, heavily regulated consumer sector. You'll own how the platform runs. That … role for someone who finds that work satisfying rather than thankless, and who wants to do it on a platform where the SRE patterns are still being written. What you'll be doing Technical Delivery & Implementation Kubernetes for AI workloads: Design, build and operate the Kubernetes infrastructure behind model serving ...

Senior Platform Engineer — SRE & Cloud

Location
Greater London, England, United Kingdom
Beamery in London seeks a hands-on Principal Platform Engineer to lead reliability and scalability across our cloud-based platform. You will shape architecture, drive incident response, and mentor engineers while partnering with product and leadership to deliver scalable services. You will own platform roadmap, advance Kubernetes, Terraform ...

Senior Platform Engineer – Remote UK (Cloud/SRE)

Location
Greater London, England, United Kingdom
Hudl in London, United Kingdom, is seeking a Senior Engineer to join our Platform Engineering team. You’ll work on site reliability, cloud infrastructure, observability and production operations to keep Hudl’s platform highly available, scalable and secure. You’ll lead with technical excellence, mentor engineers ...

Software Engineer – SRE

Location
Greater London, England, United Kingdom
Build and improve product capabilities for enterprise customers using Geordie’s AI agent visibility, governance, and risk platform Own service reliability by defining service health, instrumenting services, and acting on operational data Improve observability, alerting, and incident response practices Strengthen infrastructure, delivery pipelines, and automation, including infrastructure as code … automation using modern AI tooling Core Competencies Demonstrates expertise in building and improving product capabilities using AI technologies, with a strong focus on service reliability, observability, and infrastructure automation. Proficient in leading software projects and implementing best practices in secure software development and operational efficiency. Highest-signal resume keywords ...

Cloud-Native DevOps Engineer | SRE & CI/CD Pro

Location
Greater London, England, United Kingdom
OpenTalent is seeking a skilled DevOps/SRE for delivery in Greater London. The role requires hands-on experience with Linux, cloud platforms, containerisation, and modern CI/CD pipelines. You will design, implement, and maintain scalable infra, automate deployments, and contribute to reliability initiatives. The candidate will work ...

Storage Platform Engineer (SRE)

Location
Greater London, England, United Kingdom
Selection changes the language of the page/content London, England, United Kingdom Software and Services Apple Cloud infrastructure is BIG. The storage SRE teams of Apple Cloud are building and runningthe next generation distributed storage systems to support Apple’s most critical services. Operating atour scale, across multiple geographically … dispersed data centres, and servicing users with vast dataneed presents unique challenges. As a member of Storage SRE at Apple, you'll need to solve theseproblems using your deep understanding of infrastructure, storage, data analysis, programming,teamwork, and expertise in Linux system internals. Description We are looking for seasoned software ...

Senior SRE & Platform Engineer - High-Impact Infra

Location
Greater London, England, United Kingdom
Paragon Alpha in London seeks a Senior SRE/Platform Engineer to scale the infrastructure powering its elite quant hedge fund. You’ll own reliability, latency, and platform performance across Kubernetes, observability, and internal tooling. You’ll collaborate with researchers and traders, develop Python-based services with ...

Production Support Engineer: SRE, DevOps & Backend Code

Location
Greater London, England, United Kingdom
intro is seeking a Product Support Engineer in London to tackle production issues across applications, APIs and cloud infrastructure, with … hands-on work in backend code and automation. You will own L2/L3 incidents, drive RCA, improve monitoring and observability, and collaborate with SRE, DevOps and Software Engineering to raise platform reliability. London four-days onsite with a £55,000-£85,000 package. #J-18808-Ljbffr ...

Senior Platform Engineer: SRE & Infra (Kubernetes)

Location
Greater London, England, United Kingdom
OneSignal, Inc. is looking for a Senior or Staff Software Engineer to join their Platform Team in Greater London. The role focuses on managing and developing infrastructure, ensuring uptime and performance. The ideal candidate will have extensive experience in platform engineering and be knowledgeable in Kubernetes and PostgreSQL. ...

Senior Storage Platform Engineer (SRE) - Scalable Cloud

Location
Greater London, England, United Kingdom
Apple is seeking a seasoned Storage SRE to join the Object Storage team in London. You will design, build, and operate scalable storage services used by hundreds of millions of users, ensuring high availability and reliability. The role requires ownership of features from design through delivery, collaboration with cross-functional ...

Hybrid Cloud Platform Engineer — SRE & Internal Tools

Location
Greater London, England, United Kingdom
Onyx-Conseil is looking for a Platform Engineer to enhance our cloud banking platform. In this role … will design and scale cloud infrastructure using Java and Golang, contributing to our continuous deployment systems and internal tooling. You will embrace a strong SRE mindset and participate in a hybrid working approach, collaborating closely with cross-functional teams. If you are passionate about cloud technologies and automation, this position ...

Senior SRE Platform Engineer Capital Markets

Hiring Organisation
Huxley Associates
Location
London, United Kingdom
Salary
£ 120 K
Tier Investment bank requires a SRE rockstar, build and defend the next gen reliability frameworks, observability platforms, and performance engineering systemes that traders rely on daily.Lead incident response like a battlefied commanderDive deep into Java Kotlin, Python, Microservices, Kafka, Grafana, Splunk, Dynatrace and full modern SRE arsenalWork shoulder … relaibility means in the world of high stakes financeThis isn't just a job, it's where elite engineering meets adrenaline.If youi're an SRE who gets turned on ny relaibility at scale, performace under presssure and build systems that never break!To find out more about Huxley, please visit ...

Cloud Operations Engineer (SRE) – Market Data

Location
Greater London, England, United Kingdom
Morningstar UK Ltd. in London is seeking an experienced IT Operations/SRE professional to monitor and support the Market Data stack. You will collaborate with the London and Mumbai squads, deploy updates, and help shape tooling and automation during cloud migration initiatives. You will join a hybrid work model … rotating shifts that include late finishes and weekends, with 40 hours per week after training. This role emphasizes reliability, fast resolution, and cross-functional teamwork. #J-18808-Ljbffr ...

SRE Systems Engineering Manager, ML Compute

Location
Greater London, England, United Kingdom
Google is hiring a Systems Engineering Manager for Site Reliability Engineering focusing on ML Compute in London. You will lead a team responsible for uptime, availability and performance of core services and drive automation across large-scale infrastructure. Own end-to-end reliability, mentor engineers, manage ...

Software Engineering or SRE, PhD Intern, 2027

Hiring Organisation
Hackajob Ltd
Location
London, United Kingdom
Employment Type
Permanent
Salary
GBP Annual
hackajob is partnering directly with Google to hire for this role. We offer a range of internships in either Software Engineering or Site-Reliability Engineering across EMEA. Durations and start dates will vary according to project and location. Our recruitment team will determine where you fit best based ...

AI-Driven Principal Platform Engineer - Multi-Cloud SRE

Location
Greater London, England, United Kingdom
Trayport seeks a Principal Platform Engineer to anchor our Platform/Operations in London. The role is 70% hands-on—designing, building and operating infrastructure for a global trading platform, and mentoring the broader team. You will drive … cloud engineering across AWS and Azure, own networking design, and lead IaC, CI/CD, and Kubernetes platforms as products. A strong focus on SRE, AI-enabled workflows, and reliability is required. #J-18808-Ljbffr ...