201 to 225 of 792 Site Reliability Engineering Jobs in England

VP, Site Reliability Engineering: Scale & Resilience

Location
Birmingham, England, United Kingdom
Goldman Sachs is seeking a VP-level Site Reliability Engineer to architect and operate highly reliable platforms that support critical … services at scale. You will collaborate across engineering teams to improve production systems and enable rapid delivery of new services. The role emphasizes SRE principles such as SLOs, error budgets, and blameless post-mortems, with leadership opportunities in incident response and on-call design within a financial services context. ...

Core AI Engineer

Hiring Organisation
London Stock Exchange Group
Location
London, UK
Employment Type
Full-time
Artificial Intelligence, Automation and Intelligent Engineering. We are building enterprise-scale AI capabilities that improve service resilience, automate operational workflows, accelerate engineering productivity and enhance customer outcomes. As a Core AI Engineer, you will play a leading technical role in the design, development and deployment of AI solutions across … systems design. AI observability, evaluation and governance frameworks. Desirable ExperienceExperience within Financial Services or highly regulated environments. Knowledge of Service Reliability Engineering (SRE) principles. Experience developing AI-powered operational tooling. Experience building internal AI platforms or developer enablement capabilities. Familiarity with Microsoft AI ecosystem, Copilot technologies and Azure ...

Lead Data Platform Engineer - DataOps

Location
Nottingham, England, United Kingdom
seeking a Lead Data Platform Engineer to own the resilience, scalability, and security of our core financial data ecosystem. This senior, hands‐on engineering role requires a technical expert who thrives on designing, building, and running high‐throughput, mission‐critical batch and real‐time data platforms in a regulated … standards across a distributed team while contributing to long‐term technical strategy and cloud modernization efforts. What We’re Looking For A Platform or SRE Background: Solid experience in platform engineering, site reliability (SRE), or data operations, ideally within a large‐scale or regulated environment. AWS Experience ...

Monitoring & Observability Engineer (Dynatrace)

Hiring Organisation
Computacenter
Location
London, United Kingdom
Salary
£ 80 K
some of the world’s most well-known organisations. You’ll play a key role in helping our customers achieve greater visibility, performance, and reliability across their IT estates—contributing to their operational success through proactive insight and incident prevention.What you'll doDesign, implement, and manage observability solutions using … with a passion for continuous improvement and knowledge sharingCertificationsDynatrace Associate & ProSplunk Core Certified Power User Desirable ExperienceDevOps or Site Reliability Engineering (SRE) experienceAutomation with Terraform or similar toolsBuilding CI/CD pipelinesExperience with Docker and Kubernetes for packaging and deploymentAbility to adapt to new technologies in fast ...

Monitoring & Observability Engineer (Dynatrace)

Location
Greater London, England, United Kingdom
some of the world’s most well-known organisations. You’ll play a key role in helping our customers achieve greater visibility, performance, and reliability across their IT estates—contributing to their operational success through proactive insight and incident prevention. What you'll do Design, implement, and manage observability … passion for continuous improvement and knowledge sharing Certifications Dynatrace Associate & Pro Splunk Core Certified Power User DevOps or Site Reliability Engineering (SRE) experience Automation with Terraform or similar tools Experience with Docker and Kubernetes for packaging and deployment Ability to adapt to new technologies in fast-paced ...

Production Engineer

Hiring Organisation
Liquidnet
Location
London, United Kingdom
Salary
£ 80 K
issues across our trading platform. You will leverage deep expertise in FIX, Linux, Windows Server, DevOps, databases, networking, and cloud technologies to ensure platform reliability and performance.This is a hands-on leadership role involving complex troubleshooting across cross-platform market-leading technologies, driving automation and tooling improvements, and acting … working hoursPositive approach to the day-to-day, with the resilience to handle high-pressure production incidentsDesiredExperience with Site Reliability Engineering (SRE) practices, including monitoring, incident response, and post-mortem analysisProven experience applying AI or machine-learning models to optimise workflows, identify patterns, and drive intelligent automation ...

Production Engineer

Location
Greater London, England, United Kingdom
issues across our trading platform. You will leverage deep expertise in FIX, Linux, Windows Server, DevOps, databases, networking, and cloud technologies to ensure platform reliability and performance.This is a hands-on leadership role involving complex troubleshooting across cross-platform market-leading technologies, driving automation and tooling improvements, and acting … Positive approach to the day-to-day, with the resilience to handle high-pressure production incidentsDesired* Experience with Site Reliability Engineering (SRE) practices, including monitoring, incident response, and post-mortem analysis* Proven experience applying AI or machine-learning models to optimise workflows, identify patterns, and drive intelligent ...

Production Engineer

Location
City Of London, England, United Kingdom
issues across our trading platform. You will leverage deep expertise in FIX, Linux, Windows Server, DevOps, databases, networking, and cloud technologies to ensure platform reliability and performance. This is a hands-on leadership role involving complex troubleshooting across cross-platform market-leading technologies, driving automation and tooling improvements … approach to the day-to-day, with the resilience to handle high-pressures production incidents Desired Experience with Site Reliability Engineering (SRE) practices, including monitoring, incident response, and post-mortem analysis Proven experience applying AI or machine-learning models to optimise workflows, identify patterns, and drive intelligent ...

Lead Data Platform Engineer - DataOps

Location
Nottingham, England, United Kingdom
Lead Data Platform Engineer to own the resilience, scalability, and security of our core financial data ecosystem. This is a senior, hands‐on engineering role for a technical expert who thrives on the unique challenge of designing, building, and running high‐throughput, mission‐critical batch and real‐time data … across a distributed team while contributing to the long‐term technical strategy and cloud modernization efforts. What we’re looking for A Platform or SRE Background: You have a solid background in platform engineering, site reliability (SRE), or data operations, ideally within a large‐scale or regulated ...

Site Reliability Engineer, Infrastructure - ThousandEyes

Location
Greater London, England, United Kingdom
deeply integrated across the Cisco technology portfolio, delivering AI‐powered assurance insights within Cisco’s Networking, Security, Collaboration, and Observability portfolios. Our distributed Site Reliability Engineering team of approximately nine engineers owns the availability, latency, performance, efficiency, monitoring, emergency response, and capacity planning of the platform while … call rotation. Hands‐on experience with infrastructure‐as‐code tooling and codebases, preferably Terraform. Hands‐on experience leveraging AI as a force multiplier of SRE activities, such as automating toil away and improving operational efficiency. Professional experience administering and troubleshooting GNU/Linux systems, including system libraries, file systems, networking ...

Site Reliability Engineer, Infrastructure - ThousandEyes

Location
City Of London, England, United Kingdom
deeply integrated across the Cisco technology portfolio, delivering AI-powered assurance insights within Cisco’s Networking, Security, Collaboration, and Observability portfolios. Our distributed Site Reliability Engineering team of approximately nine engineers owns the availability, latency, performance, efficiency, monitoring, emergency response, and capacity planning of the platform while … operational on-call rotation. Hands-on experience with infrastructure-as-code tooling and codebases, preferably Terraform. Hands-on experienceleveraging AIas a force multiplier of SRE activities, such as automati ng toil away and improving operational efficiency. Professional experience administering and troubleshooting GNU/Linux systems, including system libraries, file systems ...

Cloud Operations Engineer (remote – London)

Hiring Organisation
Quant Capital
Location
London, United Kingdom
Salary
£ 80 K
remote) Cloud Operations Engineer/Site Reliability Engineer – Fintech80,000 Plus Bonus + 10% non-cont pension + 10-15k bonus and sharesQuant Capital is urgently looking for a Site Reliability Engineer to join or well-known Fintech50 client who produces software disrupting the wealth … TechniquesSolid understanding of the OSI ModelExperience in database technology and basic query writing MSSQL, Postgres.This role suits a senior Engineer from a DevOps or SRE background who is a real technologist and cloud specialist interested in the latest tooling and technologies that support software development and infrastructure. The firm ...

VodafoneThree - SRE III

Hiring Organisation
VodafoneThree
Location
Newbury, Berkshire, United Kingdom
Employment Type
Permanent
Salary
GBP Annual
data. What you'll do We're looking for someone who thrives on building resilient platforms, reducing operational complexity and helping engineering teams move faster with confidence. If you enjoy solving complex reliability challenges, automating away toil, and influencing how modern cloud platforms are built and operated … automation to reduce operational burden and improve service resilience. Coach and mentor engineers across the organisation, sharing knowledge and helping others develop modern SRE practices. Who you are You are an experienced Site Reliability Engineer with strong cloud and software engineering skills, comfortable owning critical services running ...

AMBG - Cloud Security & Exposure Management Architect

Location
Greater London, England, United Kingdom
recovery sequencing. Identify risks, vulnerabilities, and single points of failure across workloads and operational processes. Recommend improvements aligned with Azure Well-Architected Framework, SRE principles, and ITIL practices. Engage customer stakeholders to understand RTO/RPO objectives and recovery workflows. Produce professional documentation outlining findings, risks, and recommended improvements. About … Azure architecture including availability zones, backup, recovery, and monitoring services. Familiarity with cloud-native resiliency patterns and site reliability engineering (SRE) practices. Experience designing and assessing Major Incident Response Plans (MIRPs). Experience in business continuity planning and operational resilience. Strong communication and documentation skills across technical ...

Lead Site Reliability Engineer

Hiring Organisation
Inspire People
Location
Birmingham, United Kingdom
Employment Type
Permanent
Salary
GBP 80,000 Annual
support economic growth across the UK. The Department for Business, Innovation, Science and Trade (BIST), in partnership with Inspire People, is seeking a Senior SRE Squad Lead with experience leading and developing engineers, strong DevOps and Site Reliability Engineering expertise, cl click apply for full job details ...

Lead Site Reliability Engineer

Hiring Organisation
Inspire People
Location
Salford, Manchester, United Kingdom
Employment Type
Permanent
Salary
GBP 80,000 Annual
support economic growth across the UK. The Department for Business, Innovation, Science and Trade (BIST), in partnership with Inspire People, is seeking a Senior SRE Squad Lead with experience leading and developing engineers, strong DevOps and Site Reliability Engineering expertise, cl click apply for full job details ...

Lead Site Reliability Engineer

Hiring Organisation
Inspire People
Location
Darlington, County Durham, United Kingdom
Employment Type
Permanent
Salary
GBP 80,000 Annual
support economic growth across the UK. The Department for Business, Innovation, Science and Trade (BIST), in partnership with Inspire People, is seeking a Senior SRE Squad Lead with experience leading and developing engineers, strong DevOps and Site Reliability Engineering expertise, cl click apply for full job details ...

Senior DevOps / Platform Engineer (Google Cloud)

Hiring Organisation
Datatonic
Location
London, United Kingdom
Salary
£ 80 K
Cloud's premier partner in AI, driving transformation for world-class businesses. We push the boundaries of technology with expertise in machine learning, data engineering, and analytics on Google Cloud Platform. By partnering with us, clients future-proof their operations, unlock actionable insights, and stay ahead of the curve … scale-up environmentContainerisation/Virtualisation Expertise: Proficiency with technologies such as Terraform and KubernetesSRE Principles: Experience in implementing Site Reliability Engineering (SRE) principlesCloud Native Architecture: Hands-on experience with cloud-native architectures, ideally on Google CloudClient-Facing Role: Prior experience in a client-facing positionSDN Knowledge: Understanding ...

Data Platform Engineer

Hiring Organisation
MONY Group
Location
London, United Kingdom
Salary
£ 80 K
personalised customer experiences. We work closely with teams across the business to make data clean, reliable, secure and accessible for decision-making. Data & AI Engineering is a cross-functional team of engineers and scientists. We integrate with the group's operational data stores, maintain shared data models, build … ability to apply automation responsibly to real delivery and operational problems. You might come from data engineering, platform engineering, software engineering, SRE, analytics engineering, MLOps, or cloud infrastructure. What matters most is that you enjoy reducing toil, improving developer experience, and building secure, observable systems that ...

Network Site Reliability Engineer

Hiring Organisation
Quant Capital
Location
London, UK
Employment Type
Full-time
Network SRE – 250,000-350,000 total compensation – 4 days in officeQuant Capital is urgently looking Network SRE for our high profile client. Our client is a leading quantitative trading company and liquidity provider. Their focus on technology has allowed them to deeply penetrate the market and gain market share. … Shared Engineering team that focuses on designing, developing, and maintaining infrastructure and tools. The team requires a Network Site Reliability Engineer (SRE) with strong network fundamentals, problem-solving skills, and a keen interest in diverse tools and techniques. The role involves collaborative work across various teams, exploring ...

VodafoneThree - SRE III

Location
Newbury, England, United Kingdom
easy for teams to consume. Drive a culture of 'build once, automate forever' by identifying repetitive operational work and replacing it with sustainable engineering solutions. Enable engineering teams to leverage advanced metrics and tracing using Datadog to reduce SLA. Partner with engineers, architects, security teams and product teams … automation to reduce operational burden and improve service resilience. Coach and mentor engineers across the organisation, sharing knowledge and helping others develop modern SRE practices. Who you are You are an experienced Site Reliability Engineer with strong cloud and software engineering skills, comfortable owning critical services running ...

Site Reliability Engineer

Hiring Organisation
UK Tote Group
Location
Wigan, Greater Manchester, United Kingdom
Salary
£ 55 K
mission to deliver a seamless and reliable digital experience for racing fans across the UK and beyond. As a Site Reliability Engineer (SRE), you’ll play a critical role in keeping our online platforms and infrastructure fast, stable, and scalable — especially during the most exciting moments … performance and stability. You’ll analyse telemetry data, identify bottlenecks, and drive improvements across our infrastructure and applications.You’ll lead the development of our SRE strategy, defining standards, best practices, and ways of working that embed reliability into everything we build. Working closely with engineering, operations, and product ...

Site Reliability Engineer - Service Assurance Systems

Location
Greater London, England, United Kingdom
enrich raw data, making it readily consumable by our stakeholders across operations, engineering and management. As a Site Reliability Engineer (SRE), you will play a key role in bridging the gap between software development and operational reliability. You will be responsible for supporting the applications built … deployments and operational improvements to the SAS group and wider stakeholders. Support on-call and out-of-hours incident response as required. Liaise with engineering and infrastructure teams to ensure system changes are communicated and operationally risk-assessed before deployment. What you'll need Familiarity with both Windows ...

Site Reliability Engineer - Banking & Finance

Location
Greater London, England, United Kingdom
Ready to take the next step in your career? Join a leading technology-driven trading firm where engineering, automation, and high-performance infrastructure are central to supporting global trading operations. The organisation invests heavily in modern platform engineering practices, enabling teams to build reliable, scalable, and highly automated … Have: Strong experience programming with Python, Go and/or C++ Strong Linux knowledge and understanding of distributed systems. Experience with monitoring, observability or SRE practices. Experience with CI/CD pipelines, Git and infrastructure automation. Familiarity with Kubernetes and containerised workloads. Strong analytical and troubleshooting skills. Benefits: Build ...

Senior DevOps Engineer

Hiring Organisation
Anson Mccade
Location
Manchester, North West, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£75,000
Engineer Location: Manchester (Client Travel Required) Why Join This Senior DevOps Engineer Opportunity? Join a leading technology consultancy delivering large-scale cloud and platform engineering solutions for some of the world's most recognisable organisations. You'll work with modern cloud technologies, influence engineering best practices and play … Engineer Design and support cloud-native platforms across AWS and Azure Implement DevSecOps and Infrastructure as Code best practices Drive platform reliability using SRE principles and observability tools Support incident management and continuous improvement initiatives Implement Terraform-based infrastructure solutions Leverage automation and AI-assisted engineering tools ...