76 to 100 of 357 Site Reliability Engineering Jobs in London

Senior Azure DevOps Engineer

Hiring Organisation
REVYBE IT RECRUITMENT LIMITED
Location
City of London, London, United Kingdom
Employment Type
Permanent, Work From Home
continue to invest in their cloud and DevOps capabilities. They're looking for a Senior Azure DevOps Engineer who wants more than just another engineering role. This is an opportunity to influence technical decisions, build highly automated cloud infrastructure, and play a key role in scaling a modern Azure … platform. You'll join a collaborative engineering team where your ideas are encouraged, ownership is expected, and you'll have the opportunity to help shape the future direction of the DevOps function. If you're passionate about Azure, Kubernetes, automation, and Infrastructure as Code, we'd love to hear ...

DevOps Engineer (Security Cleared)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Solirius Reply, part of the Reply Group, is a technology consultancy and digital transformation partner that helps organisations solve complex challenges through strategy, design, engineering, and delivery. We work closely with our clients to deliver secure, accessible, user-focused services that evolve with their needs. By combining deep technical … Ministry of Housing, Communities and Local Government, UEFA, International Olympic Committee, and Mercedes-Benz. Our services span the full digital delivery lifecycle, including architecture, engineering, delivery management, user-centred design, business analysis, data, DevOps, and AI. We operate as a collaborative and inclusive organisation that empowers our people ...

Azure Devops Engineer

Hiring Organisation
VIQU IT
Location
London, Candlewick, United Kingdom
Employment Type
Contract
Contract Rate
£500 - £600/day Inside IR35
Detection and Response (EDR) solution across Azure cloud environments. Deploy and support EDR across virtual machines, Kubernetes clusters and containerised workloads. Collaborate with DevOps, Engineering and Platform teams to ensure secure and reliable deployments. Build and maintain Infrastructure as Code (IaC) using Terraform. Develop and manage CI/… environments Monitoring and observability tools such as Prometheus, Grafana, Dynatrace, AppDynamics, Splunk or similar DevOps and/or Site Reliability Engineering (SRE) environments Networking fundamentals including DNS, VPNs, load balancing, firewalls and network protocols Working with cross-functional teams and managing technical stakeholders Excellent written and verbal ...

Senior DevOps / Platform Engineer (Google Cloud)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Cloud's premier partner in AI, driving transformation for world-class businesses. We push the boundaries of technology with expertise in machine learning, data engineering, and analytics on Google Cloud Platform. By partnering with us, clients future-proof their operations, unlock actionable insights, and stay ahead of the curve … Experience: Previous experience working in a start-up or scale-up environment Containerisation/Virtualisation Expertise: Proficiency with technologies such as Terraform and Kubernetes SRE Principles: Experience in implementing Site Reliability Engineering (SRE) principles Cloud Native Architecture: Hands-on experience with cloud-native architectures, ideally on Google ...

DevOps Engineer (AWS)

Hiring Organisation
IT Graduate Recruitment
Location
London, South East, England, United Kingdom
Employment Type
Full-Time
Salary
£30,000 per annum
enjoys automation, solving infrastructure challenges and working with modern cloud-native technologies. About the Role You'll work across cloud infrastructure, automation and platform engineering, building Infrastructure as Code, improving CI/CD pipelines and supporting containerised workloads. You'll collaborate with engineering teams to deliver secure, reliable … Personal cloud projects, home labs or GitHub repositories demonstrating practical experience. Key Words: Cloud Engineer, DevOps Engineer, Platform Engineer, Site Reliability Engineer, SRE, Cloud Infrastructure, Infrastructure Engineering, Infrastructure as Code, IaC, Terraform, OpenTofu, Terraform Modules, AWS, Amazon Web Services, EC2, VPC, S3, RDS, IAM, Route 53, Cloud ...

Technology Business Partner

Hiring Organisation
INFUSED SOLUTIONS LIMITED
Location
London, United Kingdom
Employment Type
Permanent, Work From Home
advisor to senior business leaders, translating strategic priorities into technology roadmaps that drive innovation, operational excellence and measurable business outcomes. You'll work across engineering, architecture, product and operations teams, ensuring technology investments deliver real value while championing modern engineering practices, AI-enabled transformation and continuous improvement. … product delivery. AI, machine learning, automation, cybersecurity, data platforms and systems integration. Driving delivery excellence through Agile, DevOps, Site Reliability Engineering (SRE) and product-led operating models. Influencing C-suite and executive stakeholders, translating technical capability into commercial value. Building and leading high-performing multidisciplinary technology teams. ...

Senior Platform Engineer

Hiring Organisation
REVYBE IT RECRUITMENT LIMITED
Location
City of London, London, United Kingdom
Employment Type
Permanent, Work From Home
Platform We're partnering with one of London's most exciting and rapidly growing fintech companies as they continue to invest in their Platform Engineering function. They're looking for a Senior Platform Engineer who wants more than just another engineering role. This is an opportunity to help … shape the platform, influence technical decisions, and build the cloud infrastructure that powers a rapidly scaling business. You'll join a collaborative engineering team where your ideas are encouraged, ownership is expected, and you'll play a key role in driving automation, reliability, and platform maturity. ...

Staff Implementation Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
solutions for enterprise customers. You will have an opportunity to work with Harness Engineering and various customer functions, such as DevOps, SRE, Cloud, Finance and Engineering Analytics teams. You will develop best practices and automations to streamline Harness platform deployments in the most efficient, scalable, repeatable and reliable … using the Harness platform. You will collaborate with Harness Engineering, Sales Engineering, and Customer Success, as well as customer teams including DevOps, SRE, Platform Engineering, Cloud Infrastructure, Finance, and Engineering Leadership. Your mission is to help customers modernize and scale their CI/CD practices ...

Software Engineering III - AI/ML Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Description As a Site Reliability Engineer for AI/ML Data Platforms, you will be instrumental in building scalable, resilient and market‐leading data solutions. You will engage in root cause analysis, production changes, budgetary considerations, and staffing challenges. Your experience will be vital in managing and mentoring … company‐wide standards. Ability to work collaboratively in teams and build meaningful relationships to achieve common goals. Preferred Qualifications 4+ years in an SRE or production support role with AWS Cloud, Databricks, Snowflake or similar Technologies. AWS and Databricks certifications. We do not discriminate on the basis of any protected ...

Site Reliability Engineer (SRE) - Data Platform in London - Apple Inc.

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Site Reliability Engineer (SRE) – Data Platform London, United Kingdom Full time, visa sponsorship not provided. Tech Stack Python Go S3 Reliability Responsibilities Design, author, and release code in Go or Python. Manage and scale distributed systems in public, private, or hybrid cloud environments. Preferred Qualifications Experience with ...

Specialized Cloud Network Engineer - Multi-Cloud

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
services. Develop scripts and tools (e.g.,Python, Go, Bash) to streamline network operations, ensure consistency,and improve efficiency. Site Reliability Engineering (SRE) for Networks: Embrace a "you build it, you run it" mindset fornetwork services. Take ownership of the reliability, performance, andavailability of the cloud network … Code (IaC) Hands-on Expertise: Demonstrable experience with the following: Programming Languages: Python and Go. Agile and DevOps Mindset: Familiaritywith Agile Development, DevOps, and SRE practices. Adaptability: A demonstratedability to quickly learn new technologies and adapt to changing projectrequirements. Strategic Thinking: Experienceevaluating complex requirements and rationalizing them into a consistentservice ...

Monitoring & Observability Engineer (Dynatrace)

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
some of the world’s most well-known organisations. You’ll play a key role in helping our customers achieve greater visibility, performance, and reliability across their IT estates—contributing to their operational success through proactive insight and incident prevention. What you'll do Design, implement, and manage observability … passion for continuous improvement and knowledge sharing Certifications Dynatrace Associate & Pro Splunk Core Certified Power User DevOps or Site Reliability Engineering (SRE) experience Automation with Terraform or similar tools Experience with Docker and Kubernetes for packaging and deployment Ability to adapt to new technologies in fast-paced ...

Production Engineer

Hiring Organisation
Jobleads-UK
Location
City Of London, England, United Kingdom
issues across our trading platform. You will leverage deep expertise in FIX, Linux, Windows Server, DevOps, databases, networking, and cloud technologies to ensure platform reliability and performance.This is a hands-on leadership role involving complex troubleshooting across cross-platform market-leading technologies, driving automation and tooling improvements, and acting … working hoursPositive approach to the day-to-day, with the resilience to handle high-pressure production incidentsDesiredExperience with Site Reliability Engineering (SRE) practices, including monitoring, incident response, and post-mortem analysisProven experience applying AI or machine-learning models to optimise workflows, identify patterns, and drive intelligent automation ...

Production Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
issues across our trading platform. You will leverage deep expertise in FIX, Linux, Windows Server, DevOps, databases, networking, and cloud technologies to ensure platform reliability and performance.This is a hands-on leadership role involving complex troubleshooting across cross-platform market-leading technologies, driving automation and tooling improvements, and acting … Positive approach to the day-to-day, with the resilience to handle high-pressure production incidentsDesired* Experience with Site Reliability Engineering (SRE) practices, including monitoring, incident response, and post-mortem analysis* Proven experience applying AI or machine-learning models to optimise workflows, identify patterns, and drive intelligent ...

Production Engineer

Hiring Organisation
Jobleads-UK
Location
City Of London, England, United Kingdom
issues across our trading platform. You will leverage deep expertise in FIX, Linux, Windows Server, DevOps, databases, networking, and cloud technologies to ensure platform reliability and performance. This is a hands-on leadership role involving complex troubleshooting across cross-platform market-leading technologies, driving automation and tooling improvements … approach to the day-to-day, with the resilience to handle high-pressures production incidents Desired Experience with Site Reliability Engineering (SRE) practices, including monitoring, incident response, and post-mortem analysis Proven experience applying AI or machine-learning models to optimise workflows, identify patterns, and drive intelligent ...

Principal Solutions Engineer - Observe by Snowflake

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
sales cycle, from initial contact to final contract. Feedback and Product Collaboration: Serve as the voice of the customer to the Product Management and Engineering teams, providing crucial feedback on product capabilities, market needs, and competitive landscape. Content and Training: Develop and maintain technical sales collateral, including demo environments … articulate complex technical concepts to both technical and non-technical audiences. Preferred Experience selling to Developer, DevOps, and Site Reliability Engineering (SRE) personas. Prior experience in a fast-paced, high-growth Observability or Application Performance Monitoring (APM) company. Snowflake is growing fast, and we’re scaling ...

Site Reliability Engineer, Infrastructure - ThousandEyes

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
deeply integrated across the Cisco technology portfolio, delivering AI‐powered assurance insights within Cisco’s Networking, Security, Collaboration, and Observability portfolios. Our distributed Site Reliability Engineering team of approximately nine engineers owns the availability, latency, performance, efficiency, monitoring, emergency response, and capacity planning of the platform while … call rotation. Hands‐on experience with infrastructure‐as‐code tooling and codebases, preferably Terraform. Hands‐on experience leveraging AI as a force multiplier of SRE activities, such as automating toil away and improving operational efficiency. Professional experience administering and troubleshooting GNU/Linux systems, including system libraries, file systems, networking ...

Site Reliability Engineer, Infrastructure - ThousandEyes

Hiring Organisation
Jobleads-UK
Location
City Of London, England, United Kingdom
deeply integrated across the Cisco technology portfolio, delivering AI-powered assurance insights within Cisco’s Networking, Security, Collaboration, and Observability portfolios. Our distributed Site Reliability Engineering team of approximately nine engineers owns the availability, latency, performance, efficiency, monitoring, emergency response, and capacity planning of the platform while … operational on-call rotation. Hands-on experience with infrastructure-as-code tooling and codebases, preferably Terraform. Hands-on experienceleveraging AIas a force multiplier of SRE activities, such as automati ng toil away and improving operational efficiency. Professional experience administering and troubleshooting GNU/Linux systems, including system libraries, file systems ...

Site Reliability Engineer – NS London

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Location(s): [[mfield3]] The role of a Site Reliability Engineer (SRE) at BAE Systems Digital Intelligence involves combining operations and software engineering to automate system support and enhance reliability for a key national security customer. The SRE team works on continuous improvement of system health … deploy monitoring products, creating custom tools as needed to provide comprehensive, intelligent observations that demonstrate daily improvements. Participate in the wider DevOps/SRE community within the organization. Qualifications Experience in web development and object‐oriented programming. Knowledge of database technologies such as Oracle SQL, MongoDB, and PostgreSQL. Proficiency with ...

AMBG - Cloud Security & Exposure Management Architect

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
recovery sequencing. Identify risks, vulnerabilities, and single points of failure across workloads and operational processes. Recommend improvements aligned with Azure Well-Architected Framework, SRE principles, and ITIL practices. Engage customer stakeholders to understand RTO/RPO objectives and recovery workflows. Produce professional documentation outlining findings, risks, and recommended improvements. About … Azure architecture including availability zones, backup, recovery, and monitoring services. Familiarity with cloud-native resiliency patterns and site reliability engineering (SRE) practices. Experience designing and assessing Major Incident Response Plans (MIRPs). Experience in business continuity planning and operational resilience. Strong communication and documentation skills across technical ...

Technical Account Manager

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Verda. The role works in two directions: outward, as the customer's trusted technical advisor, and inward, as their advocate inside Verda's engineering, infrastructure, and operations teams. Outward (customer-facing): Serve as the primary technical point of contact for assigned customers, building trusted relationships with their ML, research … facing role such as pre‐sales solutions architecture, post‐sales technical account management, solution architecture, or customer‐oriented site reliability engineering (SRE). Practical understanding of the HPC/AI stack: GPU compute, job schedulers (e.g., Slurm, Kubernetes), high‐performance networking (InfiniBand/RDMA), and parallel ...

Hybrid SRE Manager: Scale Reliability & Platforms

Hiring Organisation
Jobleads-UK
Location
City of Westminster, England, United Kingdom
Holland & Barrett is seeking a Site Reliability Engineering Manager to lead a high … performing team and drive reliability, scalability and security across cloud platforms and digital services. You will champion modern engineering practices, shape the SRE roadmap, and partner with Engineering, Security, Data and Product teams to embed operational excellence from design through production. This role combines technical leadership with ...

Systems Operations Lead

Hiring Organisation
Hays Technology
Location
City of London, London, United Kingdom
Employment Type
Contract
Contract Rate
£750 - £800/day Up to £800pd inside ir35 via umbrella
team of technical SMEs, ensuring workloads are prioritised and delivered effectively. Act as an escalation point for operational and infrastructure-related issues. Drive service reliability, operational excellence and continuous improvement across the environment. Required Experience Strong infrastructure background with … experience across Linux and Windows server environments. Good understanding of storage, backup and wider infrastructure technologies. Experience in Site Reliability Engineering (SRE), Infrastructure Operations, or Production Support environments. Proven experience leading and developing technical teams. Comfortable remaining hands-on and involved in technical delivery on a daily ...

Systems Operations Lead

Hiring Organisation
Hays Specialist Recruitment Limited
Location
London, South East, England, United Kingdom
Employment Type
Contractor
Contract Rate
£750 - £800 per day
team of technical SMEs, ensuring workloads are prioritised and delivered effectively. Act as an escalation point for operational and infrastructure-related issues. Drive service reliability, operational excellence and continuous improvement across the environment. Required Experience Strong infrastructure background with … experience across Linux and Windows server environments. Good understanding of storage, backup and wider infrastructure technologies. Experience in Site Reliability Engineering (SRE), Infrastructure Operations, or Production Support environments. Proven experience leading and developing technical teams. Comfortable remaining hands-on and involved in technical delivery on a daily ...

SRE Managing Consultant - Cloud Operating Model

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
# SRE Managing Consultant - Cloud Operating ModelManchester, LondonApply for this job* Permanent* Experienced Professionals* Strategy & Transformation* ID 428177-en\_GB## **Capgemini Invent**At Capgemini Invent, we believe difference drives change. As inventive transformation consultants, we blend our strategic, creative and scientific capabilities, collaborating closely with clients to deliver cutting-edge … tomorrow. Informed and validated by science and data. Superpowered by creativity and design. All underpinned by technology created with purpose.## **Your Role**As an SRE Consultant (Manager) at Capgemini Invent you will be part of our Cloud Advisory capability within the wider Business Technology capability unit. Our cloud advisory capability ...