126 to 150 of 538 Site Reliability Engineering Jobs in London

Site Reliability Engineer III: Scale & Automation

Location
Greater London, England, United Kingdom
Google London, UK, is seeking a Mid-level Software Engineer III in Site Reliability Engineering for the GCE AI team. You will help build and run large‐scale, fault‐tolerant systems, with emphasis on reliability, uptime, and performance. Expect collaboration across software and systems teams, mentoring ...

Site Reliability Engineer (London) - Banking & Finance

Location
Greater London, England, United Kingdom
collaboration and technical excellence, the organisation continues to push the boundaries of low-latency infrastructure and reliable system design. The team is hiring a Site Reliability Engineer (London) to build, monitor, and optimise mission-critical trading systems. The role will focus on automation, system scalability, and incident response … improving the infrastructure. Drive automation and operational excellence by leveraging your Linux expertise, Kubernetes, and Python scripting skills. Monitor and ensure high availability and reliability of trading applications while being on top of system alerts and incidents. Key Requirements: 1-5 years working experience The right candidate will come ...

Scala Engineer

Location
City Of London, England, United Kingdom
Experience working within Continuous Integration environments Strong understanding of Agile methodologies Experience with testing and automation Awareness of Site Reliability Engineering (SRE) principles and support Experience troubleshooting incidents and restoring services following outages Experience working in a you build it, you run it environment Strong collaborative ...

Consultant - Service Management Transformation

Location
Greater London, England, United Kingdom
Microsoft AI Fundamentals (AI-900), AWS Cloud Practitioner, Google Cloud Digital, equivalent AI credentials.Knowledge of DevOps, Site Reliability Engineering (SRE) and modern operational practices would be beneficial.Strong analytical and problem-solving skills, with the ability to gather information, analyse data and contribute to practical recommendations.Excellent communication … address the entire breadth of their business needs. It delivers end-to-end services and solutions leveraging strengths from strategy and design to engineering, all fueled by its market leading capabilities in AI, cloud and data, combined with its deep industry expertise and partner ecosystem. The Group reported ...

Senior Network Site Reliability Engineer

Location
Greater London, England, United Kingdom
About the Team Miro is a fast-growing engineering organization building a business-critical collaboration platform used by companies around the world. As our product and infrastructure scale, we are looking for a Senior Network Site Reliability Engineer to help strengthen the reliability, availability, and scalability … What you’ll need 8+ years of professional experience in infrastructure, reliability, networking, or software engineering 6+ years of experience as an SRE, DevOps Engineer, Network Engineer, Software Engineer, or similar Hands-on experience with AWS infrastructure, including EC2, VPC, ALB, S3, Route 53, and CloudFront Confident networking ...

Data Platform Engineer

Hiring Organisation
MONY Group
Location
London, UK
Employment Type
Full-time
personalised customer experiences. We work closely with teams across the business to make data clean, reliable, secure and accessible for decision-making. Data & AI Engineering is a cross-functional team of engineers and scientists. We integrate with the group's operational data stores, maintain shared data models, build … ability to apply automation responsibly to real delivery and operational problems. You might come from data engineering, platform engineering, software engineering, SRE, analytics engineering, MLOps, or cloud infrastructure. What matters most is that you enjoy reducing toil, improving developer experience, and building secure, observable systems that ...

Scala Engineer

Location
City Of London, England, United Kingdom
services. Support incremental re-architecting initiatives to reduce technical complexity and improve maintainability. Develop clean, testable and maintainable code using Scala and modern engineering practices. Design, build and maintain secure APIs, databases and applications. Collaborate with Product Owners, Business Analysts, Data Engineers and wider technical teams to deliver effective … design and development experience. Experience working with databases and SQL. Hands-on AWS cloud experience. Understanding of Site Reliability Engineering (SRE) principles. Experience supporting and restoring production services during incidents. Strong appreciation of testing, automation and software quality practices. Experience working within Agile environments. Experience with Continuous ...

Site Reliability Engineer - Service Assurance Systems

Location
Greater London, England, United Kingdom
enrich raw data, making it readily consumable by our stakeholders across operations, engineering and management. As a Site Reliability Engineer (SRE), you will play a key role in bridging the gap between software development and operational reliability. You will be responsible for supporting the applications built … deployments and operational improvements to the SAS group and wider stakeholders. Support on-call and out-of-hours incident response as required. Liaise with engineering and infrastructure teams to ensure system changes are communicated and operationally risk-assessed before deployment. What you'll need Familiarity with both Windows ...

Site Reliability Engineer - Banking & Finance

Location
Greater London, England, United Kingdom
Ready to take the next step in your career? Join a leading technology-driven trading firm where engineering, automation, and high-performance infrastructure are central to supporting global trading operations. The organisation invests heavily in modern platform engineering practices, enabling teams to build reliable, scalable, and highly automated … Have: Strong experience programming with Python, Go and/or C++ Strong Linux knowledge and understanding of distributed systems. Experience with monitoring, observability or SRE practices. Experience with CI/CD pipelines, Git and infrastructure automation. Familiarity with Kubernetes and containerised workloads. Strong analytical and troubleshooting skills. Benefits: Build ...

Network Site Reliability Engineer - Trading

Hiring Organisation
Quant Capital
Location
London, UK
Employment Type
Full-time
Network SRE – 250,000-300,000 total compensation – 4 days in officeQuant Capital is urgently looking Network SRE for our high profile client. Our client is a leading quantitative trading company and liquidity provider. Their focus on technology has allowed them to deeply penetrate the market and gain market share. … Shared Engineering team that focuses on designing, developing, and maintaining infrastructure and tools. The team requires a Network Site Reliability Engineer (SRE) with strong network fundamentals, problem-solving skills, and a keen interest in diverse tools and techniques. The role involves collaborative work across various teams, exploring ...

SRE Systems Engineering Manager, ML Compute

Location
Greater London, England, United Kingdom
Google is hiring a Systems Engineering Manager for Site Reliability Engineering focusing on ML Compute in London. You will lead a team responsible for uptime, availability and performance of core services and drive automation across large-scale infrastructure. Own end-to-end reliability, mentor engineers ...

SVP, Global Head of Platform Operations

Location
Greater London, England, United Kingdom
Operations leads how TT delivers, runs and continuously improves its production platform. This is a newly consolidated executive role that brings Systems Engineering, SRE/DevOps, Deployment, Platform Engineering, Service Improvement and Release Management into one organization with a single mandate: make TT's delivery faster, safer … shape TT's global operating footprint, and act as the operational owner for integrating acquired products onto the TT platform. Direct reports : Head of SRE; Head of Systems Engineering; team leads for Deployment, Platform Engineering, Service Improvement and Release Management. Key partners : CTPO, Engineering leadership, SVP Infrastructure ...

Nework Site Reliability Engineer - Algo trading

Hiring Organisation
Quant Capital
Location
London, UK
Employment Type
Full-time
Network SRE – 250,000-300,000 total compensation – 4 days in officeQuant Capital is urgently looking Network SRE for our high profile client. Our client is a leading quantitative trading company and liquidity provider. Their focus on technology has allowed them to deeply penetrate the market and gain market share. … Shared Engineering team that focuses on designing, developing, and maintaining infrastructure and tools. The team requires a Network Site Reliability Engineer (SRE) with strong network fundamentals, problem-solving skills, and a keen interest in diverse tools and techniques. The role involves collaborative work across various teams, exploring ...

Head of Cloud Platform Engineering

Location
Greater London, England, United Kingdom
exciting point in our journey. As we continue to evolve our platform and expand our SaaS offering, we're investing in the engineering foundations that will support the next phase of our growth. About the role This is a rare opportunity to shape the future of Totara's cloud … world. We're clear on where we're heading, but how we get there is still being built. As Head of Cloud Platform Engineering, you'll play a central role in defining that journey. Today our infrastructure capability spans multiple teams, regions, and platforms, including environments inherited through acquisition. ...

Sr. Network Site Reliability Engineer (SREs)

Location
Greater London, England, United Kingdom
/ML Technologies and Professional services in the UK and EU market. Job Description Overview We are seeking a highly experienced Senior Network SRE with deep expertise across multi-vendor network infrastructure, automation, and reliability engineering. The ideal candidate will possess strong technical leadership, hands‐on engineering capabilities … resilient, scalable, and observable network environments. Key Responsibilities Design, implement, and maintain highly available network solutions across routing, switching, firewalling, and wireless technologies. Apply SRE principles to improve network reliability, scalability, and performance. Develop and maintain automation workflows using Ansible, Salt, and related frameworks to reduce operational toil. Build ...

Software Engineering or SRE, PhD Intern, 2027

Hiring Organisation
Hackajob Ltd
Location
London, United Kingdom
Employment Type
Permanent
Salary
GBP Annual
hackajob is partnering directly with Google to hire for this role. We offer a range of internships in either Software Engineering or Site-Reliability Engineering across EMEA. Durations and start dates will vary according to project and location. Our recruitment team will determine where ...

Senior Network Site Reliability Engineer

Hiring Organisation
Miro
Location
London, UK
Employment Type
Full-time
About the TeamMiro is a fast-growing engineering organization building a business-critical collaboration platform used by companies around the world. As our product and infrastructure scale, we are looking for a Senior Network Site Reliability Engineer to help strengthen the reliability, availability, and scalability … develop our CloudWan infrastructureWhat you'll need8+ years of professional experience in infrastructure, reliability, networking, or software engineering6+ years of experience as an SRE, DevOps Engineer, Network Engineer, Software Engineer, or similarHands-on experience with AWS infrastructure, including EC2, VPC, ALB, S3, Route 53, and CloudFrontConfident networking knowledge, including ...

Senior Site Reliability Engineer (SRE)

Hiring Organisation
fortice
Location
London, UK
Employment Type
Full-time
Hybrid | London80,000 – 110,000/annum plus benefitsRole: As a Senior Site Reliability Engineer for a global consultancy, you'll initially be aligned to a Defence-related project, where you'll lead a team in helping to relocate data to a new cloud platform. Working hybrid … would expect to be on site 2 days/month in Central London. This will require you hold an active UK Government Security Clearance, which you would be sponsored through, if not currently held. This role will see you be 50% operations-focused, 50% automation-focused – from systems builds ...

Site Reliability Engineer — AI-Driven Finance Production

Location
Greater London, England, United Kingdom
Voleon Group is seeking a Site Reliability Engineer (SRE) to join a production-focused team responsible for reliability and efficiency of data pipelines and trading systems. You will work across production operations and software development to improve, manage, and monitor critical infrastructure. The role emphasizes fault-tolerance ...

Core AI Engineer

Location
Greater London, England, United Kingdom
Artificial Intelligence, Automation and Intelligent Engineering. We are building enterprise-scale AI capabilities that improve service resilience, automate operational workflows, accelerate engineering productivity and enhance customer outcomes.As a Core AI Engineer, you will play a leading technical role in the design, development and deployment of AI solutions across … systems design.* AI observability, evaluation and governance frameworks.Desirable Experience* Experience within Financial Services or highly regulated environments.* Knowledge of Service Reliability Engineering (SRE) principles.* Experience developing AI-powered operational tooling.* Experience building internal AI platforms or developer enablement capabilities.* Familiarity with Microsoft AI ecosystem, Copilot technologies and Azure ...

Principal SRE (AWS, Azure, Terraforms, Kubernetes)

Location
Greater London, England, United Kingdom
China, Australia, and UAE. Interested in joining our smart, fun, and talented team? Position Overview Fourth is actively seeking an experienced and pragmatic Principal SRE to join our worldwide team. We are progressing rapidly in developing automated, highly reliable, and zero-downtime infrastructure pipelines that are becoming the standard across … valuable and achievable chunks. You have excellent written and verbal communication skills, allowing you to work effectively with our worldwide development teams and SRE community to select the right patterns and practices. You understand the importance of standardisation of technology and practices and have experience of implementing these ...

Entry Level - Site Reliability Engineer - (Remote - United Kingdom)

Location
Greater London, England, United Kingdom
Summary Yelp engineering culture is driven by our values: we’re a cooperative team that values individual authenticity and encourages creative solutions to problems. All new engineers deploy working code their first week, and we strive to broaden individual impact with support from managers, mentors, and teams. … solutions don’t work at our scale and contribute upstream to open source projects. Participate in light on-call rotations - we have geographically distributed SRE teams for follow-the-sun support, which means nobody needs to be on-call 24h a day! What It Takes To Succeed Familiarity with Linux ...

Platform Engineer

Location
Greater London, England, United Kingdom
operate AI workload infrastructure, including model gateways, retrieval services, orchestration components, and supporting cloud or Kubernetes resources. Observability, Monitoring & Site Reliability (SRE) Instrument services and implement monitoring, logging, and alerting as code using standard tooling (Prometheus, Grafana, OpenTelemetry). Participate in the on‐call rotation, responding to incidents … least one major cloud platform (AWS or Azure) and Kubernetes/Docker. Familiarity with observability tooling (Grafana, Datadog, Splunk, ELK, OpenTelemetry) and basic SRE practices. Exposure to test automation, policy-as-code, and platform security practices. Familiarity with ITIL best practices (incident, change, and problem management) preferred. Experience with Lean ...

Trainee DevOps Engineer | No experience needed (Ref: 7501)

Hiring Organisation
Qualify Nation Recruitment
Location
London, South East England, United Kingdom
Employment Type
Full-Time
Salary
£28,000 - £38,000 per annum
Platforms (AWS, Microsoft Azure and Google Cloud) Configuration Management Monitoring and Logging Security Best Practices (DevSecOps) Networking Fundamentals Automation and Scripting Incident Management and Reliability Engineering Practical Experience You will work on realistic DevOps projects that may include: Building CI/CD pipelines Deploying applications to cloud environments … completion, learners may pursue roles such as: Junior DevOps Engineer DevOps Engineer Cloud Support Engineer Platform Engineer Infrastructure Engineer Site Reliability Engineer (SRE) Build and Release Engineer Cloud Operations Engineer Systems Administrator Cloud Infrastructure Engineer Apply Today If you are looking to start a career in DevOps ...

Senior Site Reliability Engineer

Location
Greater London, England, United Kingdom
more about life at DeepL on LinkedIn, Instagram, and our Blog. Meet the team behind this journey We currently have two teams in the SRE track, and some SRE peers in other parts of the business. The in-track teams work closely, often collaborating. The first team SRE: Excellence focussed … making it easier to use our systems and provide tools and services to help run and monitor our products. The second team SRE: Accelerate works closely with our product development teams to use these as effectively as possible and embed better practice in teams. Both teams help ensure the services ...