151 to 175 of 592 Site Reliability Engineering Jobs in London

Lead Site Reliability Engineer (Kubernetes Required) - Hybrid

Location
Greater London, England, United Kingdom
anticipating our clients’ needs and exceeding their expectations. About the Role We are looking for a skilled and motivated Lead Site Reliability Engineer to join our team. In this role, you will be responsible for ensuring the reliability, scalability, and performance of our systems and services. … pressure, particularly during incident response Commitment to a blameless culture and continuous learning Nice to Have Experience contributing to open-source projects Familiarity with SRE principles as defined by the Google SRE handbook Previous experience in a DevOps or Platform Engineering role Company Overview FactSet (NYSE:FDS | NASDAQ ...

Senior Site Reliability Engineer - Reliability, Automation & Observability

Location
Greater London, England, United Kingdom
JPMorgan Chase is seeking a Site Reliability Engineer to join the International Consumer Bank. You will enhance reliability and operability of customer-facing digital banking services, building automation, and implementing AI-assisted engineering practices. The role emphasizes reducing toil, designing for scale, and collaborating with cross ...

Senior Site Reliability Engineer (SRE)

Hiring Organisation
fortice
Location
London, UK
Employment Type
Full-time
Hybrid | London80,000 – 110,000/annum plus benefitsRole: As a Senior Site Reliability Engineer for a global consultancy, you'll initially be aligned to a Defence-related project, where you'll lead a team in helping to relocate data to a new cloud platform. Working hybrid … would expect to be on site 2 days/month in Central London. This will require you hold an active UK Government Security Clearance, which you would be sponsored through, if not currently held. This role will see you be 50% operations-focused, 50% automation-focused – from systems builds ...

Lead Site Reliability Engineer (Dynatrace)

Location
London, United Kingdom
looking for Strong hands-on Dynatrace implementation and administration experience Experience designing and implementing observability/monitoring solutions end-to-end Strong SRE and production engineering background Experience configuring instrumentation, metrics, alerting and monitoring Understanding of technologies such as OneAgent, ActiveGate, distributed tracing and application/infrastructure monitoring Experience … mentoring other engineers The opportunity You'll join a sizeable engineering capability working across complex, large-scale environments, taking a leading role in SRE and observability engineering. There is flexibility around some of the wider cloud/platform technology stack for candidates with genuinely strong Dynatrace and SRE expertise. ...

Lead Site Reliability Engineer (Dynatrace)

Hiring Organisation
SF Partners Admin
Location
London, UK
looking for Strong hands-on Dynatrace implementation and administration experience Experience designing and implementing observability/monitoring solutions end-to-end Strong SRE and production engineering background Experience configuring instrumentation, metrics, alerting and monitoring Understanding of technologies such as OneAgent, ActiveGate, distributed tracing and application/infrastructure monitoring Experience … mentoring other engineers The opportunity You'll join a sizeable engineering capability working across complex, large-scale environments, taking a leading role in SRE and observability engineering. There is flexibility around some of the wider cloud/platform technology stack for candidates with genuinely strong Dynatrace and SRE expertise. ...

Site Reliability Engineer

Hiring Organisation
JAM Recruitment Ltd
Location
London, United Kingdom
Employment Type
Permanent
Salary
£700 - £750 per day
Site Reliability Engineer Contract £700-750 per day (inside IR35) London 3-4 days per week on site 'Developed Vetting (DV) clearance is required for this role.' Our client is at the forefront of technology, innovation and national security. Bringing together talented people, cutting-edge digital capabilities ...

Site Reliability Engineer

Hiring Organisation
JAM Recruitment Ltd
Location
London, United Kingdom
Employment Type
Contract
Contract Rate
GBP 700 - 750 Daily
Site Reliability Engineer Contract (Apply online only) per day (inside IR35) London 3-4 days per week on site "Developed Vetting (DV) clearance is required for this role." Our client is at the forefront of technology, innovation and national security. Bringing together talented people, cutting-edge digital ...

Frontend Engineering Associate Manager (GenAI experience)

Location
Greater London, England, United Kingdom
Role- Frontend Engineering Associate Manager Location: London Level- Associate Manager Accenture Song accelerates growth and value for our clients by combining creativity, technology, and data-driven intelligence. Within Song, our Agentic Commerce practice helps organisations design, build, and scale modern digital commerce ecosystems - embedding Generative AI and Agentic … APIs) Experience in agile environments for digital product creation Broad experience in modern engineering practices including: Systems Architecture Front End Engineering Platform & SRE Engineering Quality Engineering CI/CD Set yourself apart: Presented your work at an event or conference Experience developing short & long-term plans ...

Site Reliability Engineer

Hiring Organisation
Capital On Tap
Location
London, UK
Employment Type
Full-time
just getting started! ðLondon, Old Street | ð 2 Days in OfficeSRE at Capital On Tap ðAt Capital On Tap, we run a hybrid embedded SRE model. We aim to work closely with the teams within Capital On Tap to provide them the best support. Our main objective currently … much visibility into our platform's health while offering scalable solutions. What You'll be doing: As a Site Reliability Engineer (SRE) you will help ensure our platforms are fast, reliable, and scalable. You'll design, build, and monitor systems, prevent issues before they happen. Using SLAs, SLIs ...

Remote-First Site Reliability Engineer - Azure & Kubernetes

Location
Greater London, England, United Kingdom
Vertus Partners is seeking an experienced Site Reliability Engineer to join a growing function within a leading financial services organisation. The role offers ownership of projects and a pathway to shape SRE/DevOps across a large enterprise. The successful candidate will work across Azure, Kubernetes, automation and production reliability, collaborating with engineering and delivery ...

Platform Engineer

Location
Greater London, England, United Kingdom
operate AI workload infrastructure, including model gateways, retrieval services, orchestration components, and supporting cloud or Kubernetes resources. Observability, Monitoring & Site Reliability (SRE) Instrument services and implement monitoring, logging, and alerting as code using standard tooling (Prometheus, Grafana, OpenTelemetry). Participate in the on‐call rotation, responding to incidents … least one major cloud platform (AWS or Azure) and Kubernetes/Docker. Familiarity with observability tooling (Grafana, Datadog, Splunk, ELK, OpenTelemetry) and basic SRE practices. Exposure to test automation, policy-as-code, and platform security practices. Familiarity with ITIL best practices (incident, change, and problem management) preferred. Experience with Lean ...

Principal SRE (AWS, Azure, Terraforms, Kubernetes)

Hiring Organisation
Fourth
Location
London, UK
Employment Type
Full-time
Bulgaria, China, Australia, and UAE. Interested in joining our smart, fun, and talented team? Position OverviewFourth is actively seeking an experienced and pragmatic Principal SRE to join our worldwide team. We are progressing rapidly in developing automated, highly reliable, and zero-downtime infrastructure pipelines that are becoming the standard across … valuable and achievable chunks. You have excellent written and verbal communication skills, allowing you to work effectively with our worldwide development teams and SRE community to select the right patterns and practices. You understand the importance of standardisation of technology and practices and have experience of implementing these ...

Site Reliability Engineer - Core

Hiring Organisation
Blockchain
Location
London, UK
Employment Type
Full-time
distributed financial platform tackles some of the most interesting problems in the crypto for millions of our customers and continues to grow rapidly. The SRE team at blockchain combines software and systems engineering to provide a platform that abstracts complexity for increased security, reliability and rapid product delivery. … SRE organization at Blockchain is a work in progress - our focus is always on how to make our existing systems better. We pride ourselves on having created an environment where individuals have a high degree of freedom in proposing, discussing, designing and implementing changes. We are a team that places ...

Trainee DevOps Engineer | No experience needed (Ref: 7501)

Hiring Organisation
Qualify Nation Recruitment
Location
London, South East England, United Kingdom
Employment Type
Full-Time
Salary
£28,000 - £38,000 per annum
Platforms (AWS, Microsoft Azure and Google Cloud) Configuration Management Monitoring and Logging Security Best Practices (DevSecOps) Networking Fundamentals Automation and Scripting Incident Management and Reliability Engineering Practical Experience You will work on realistic DevOps projects that may include: Building CI/CD pipelines Deploying applications to cloud environments … completion, learners may pursue roles such as: Junior DevOps Engineer DevOps Engineer Cloud Support Engineer Platform Engineer Infrastructure Engineer Site Reliability Engineer (SRE) Build and Release Engineer Cloud Operations Engineer Systems Administrator Cloud Infrastructure Engineer Apply Today If you are looking to start a career in DevOps ...

Site Reliability Engineer

Location
Greater London, England, United Kingdom
benefits packages, technology talks by our experts, a beautiful modern office, daily catered lunches, and more. As a Site Reliability Engineer (SRE), you will work at the intersection of production operations and software development as you improve, manage, and monitor production-critical infrastructure and data pipelines. … make a real difference: your contributions will make our critical systems more reliable, lower operational risk, and increase the efficiency of our engineering effort. Responsibilities Improve fault-tolerance and maintainability of code in proprietary data pipelines and trading systems Diagnose and fix bugs in code Lead complex deployments Automate ...

Front End Engineering Manager

Hiring Organisation
Accenture
Location
London, UK
Employment Type
Full-time
culture that encourages you to push boundaries and think outside the box. We're especially looking for individuals who are experts in front-end engineering for mobile or web platforms, with deep experience in modern frameworks and performance optimization. You'll bring hands-on experience with generative AI technologies … agile environments for digital product creation Broad experience in modern engineering practices including: o Systems Architectureo Front End Engineering o Platform & SRE Engineering o Quality Engineering o CI/CDComfortable in articulating trends in technology and engineering practices Experience improving speed and quality outcomes using ...

SRE Director — AI-Driven Reliability & Scale

Location
Greater London, England, United Kingdom
EPAM Systems in London, United Kingdom, is seeking a Director of Site Reliability Engineering to lead a global SRE organization in a hybrid work setting. The role focuses on reliability, operational excellence, and governance across mission‐critical platforms, with an emphasis on AI‐enabled automation … improving engineering standards. The successful candidate will drive resilience, define KPIs, and collaborate across product, platform, operations, and security teams to embed reliability #J-18808-Ljbffr ...

Mid-Level DevOps Engineer

Location
Greater London, England, United Kingdom
engineers and technical leadership to ensure our production systems remain reliable, secure, and performant as the company grows. In short, you will help the engineering team keep our production systems running smoothly while contributing to infrastructure improvements and automation initiatives. What You Bring 3–5 years of DevOps, Site Reliability Engineering, or Infrastructure Engineering experience Experience working with AWS and cloud-based infrastructure Hands-on experience with Kubernetes and Docker in production environments Familiarity with networking concepts including VPCs, VPNs, and secure infrastructure design Experience building or maintaining CI/CD pipelines using CircleCI, Jenkins ...

Senior Site Reliability Engineer - AI Automation - £115K

Location
Greater London, England, United Kingdom
systems that detect problems, work out the fix and land it, so recurring operational work drops by 80-90%. This role suits an SRE or platform engineer who has run backend services at scale, has already put AI to work in their tooling, and would rather own a problem … based in the UK and hold full UK right to work, as sponsorship is not available. What you'll bring 8+ years in software, SRE or platform engineering, ideally within a large multinational tech company Hands-on experience building AI-driven, self-healing systems (LLMs, agents, auto-remediation) that ...

SRE Lead: Cloud Reliability & Platform Excellence

Location
City Of London, England, United Kingdom
LexisNexis Risk Solutions is seeking an experienced Site Reliability Engineering Lead to provide technical leadership across multiple product portfolios. You will drive reliability, scalability, security, and operational excellence in mission-critical platforms while mentoring engineers and guiding cloud modernization. You will collaborate with engineering, architecture … security, and operations to implement SLOs/SLIs, IaC, and automated resilience measures, shaping enterprise-wide engineering #J-18808-Ljbffr ...

IDP Technical Delivery Lead

Hiring Organisation
NTT DATA
Location
London, UK
Employment Type
Full-time
working with: You will lead the enterprise delivery and rollout of an Internal Developer Platform (IDP), working with client technology leaders, engineering teams, architects and platform engineers to establish the platform as a strategic capability across the organisation. The role will lead the programme from roadmap and initial platform … into ongoing operation and continuous improvement. What experience you'll bring: Technical depth: Strong previous experience within Platform Engineering, DevOps, Cloud Engineering, SRE or a closely related engineering discipline. Candidates who have only managed technical programmes without working closely with the underlying technologies are unlikely to have ...

Senior Network SRE: Automation, Reliability & Observability

Location
Greater London, England, United Kingdom
leading IT solutions provider in London is seeking a Senior Network Site Reliability Engineer (SRE) with extensive experience in network engineering and automation. The ideal candidate will design and maintain high-availability network solutions while employing SRE principles for improved reliability and performance. A strong background ...

Site Reliability Engineer

Location
Greater London, England, United Kingdom
passionate about building reliable, scalable cloud platforms that empower engineering teams to deliver at speed? Do you enjoy solving complex infrastructure challenges, driving automation, and improving operational excellence across a modern cloud environment? About the Business LexisNexis® Risk Solutions provides customers with solutions and decision tools that combine public … feature rich Kubernetes clusters that are ready to operate workloads in a safe and consistent way. About the Role We're looking for a Site Reliability Engineer to design, build and operate cloud infrastructure across AWS and Azure. You'll play a key role in enabling engineering ...

Senior Data & MLOps Engineer

Location
Greater London, England, United Kingdom
proud to be a Living Wage accredited Employer. What You’ll Do The Data Science team is focused on developing an advanced reliability platform. This system covers various aspects of data processing and analysis, including data intake, deriving meaningful metrics, identifying unusual patterns, predicting potential issues, finding slow processes … level performance analysis systems. Experience developing agentic or LLM‐powered reasoning systems for diagnostics or operational intelligence. Background in reliability engineering or SRE practices. Wondering if you’re a good fit? We believe in investing in our people, and value candidates who can bring their own diversified experiences ...

Director of Platform Engineering

Location
Greater London, England, United Kingdom
incident triage, anomaly detection, runbook automation, knowledge search, root cause support or service desk workflows. Knowledge of observability and reliability practices, including SRE principles, service‐level objectives, alert tuning, capacity planning and production readiness reviews. Experience operating SaaS products for financial services, enterprise technology or other regulated customers with … requirements. Experience with container security, policy‐as‐code, image scanning, secrets management, role‐based access control and Kubernetes security hardening. A background in DevOps, SRE, Cloud Operations or Platform Engineering, with a track record of improving automation, reliability and operational maturity. Health Insurance and Dental Health Cover ...