326 to 350 of 426 Remote/Hybrid Site Reliability Engineering Jobs

Backend Software Engineer

Location
City Of London, England, United Kingdom
build elegant, performant, maintainable, operable and secure Golang microservices that power innovative financial products and services. You’ll collaborate closely with product, engineering and business stakeholders to deliver customer-focused solutions, contribute to system design and architecture, champion engineering best practice, and help continuously improve development processes. … resolve issues across the application stack, including network, server and database layers Experience building concurrent, distributed applications at scale Experience working with DevOps or SRE teams to support production environments Strong communication and collaboration skills A passion for learning, continuous improvement and delivering high-quality software And it would ...

Senior Backend Engineer - Databases - Loki Query | UK | Remote

Hiring Organisation
Grafana Labs
Location
United Kingdom
Salary
£ 70 K
Grafana Labs is the company behind Grafana Cloud, the fully managed observability platform trusted by more than 10,000 organizations to ensure reliability, resolve incidents faster, and optimize telemetry at scale. Built on open source and open standards and designed for interoperability across any stack, Grafana Cloud brings … hobby/homelab projects)Exposure to microservices architecture and distributed systems, and a desire to learnFamiliarity with being on-call and performing operations/SRE tasks or with the concept of infrastructure as codeCompensation & Rewards:In the UK, the Base compensation range for this role is 91,000 - 114,000. ...

Senior DevOps Engineer: Cloud Platform & SRE (Hybrid)

Location
Reigate, England, United Kingdom
DevOps Engineer to join our global SaaS platform team. You will help evolve Radar Live SaaS, applying IaC, CI/CD, cloud automation, and SRE practices to ensure reliability and security at scale. You’ll work across Dev, Ops, Security and Architecture in an Agile environment, building automated deployments ...

AI Platform Engineer 3194

Location
Guildford, England, United Kingdom
processes. Partner with AI Gateway, governance owners, hyperscalers, AI frontier labs for regulated industry features, terms of service and roadmap influence. Optimize platform scalability, reliability, and cost across regions and environments; drive capacity planning, release automation, and risk recovery. What You Bring Proven track record building & operating … identity/risk management. Strong software engineering skills (e.g., Python/TypeScript/Go), infrastructure as code (Terraform), CI/CD, and SRE practices for reliability, performance, and cost efficiency. Experience implementing governance, compliance, safety guardrails, and AI FinOps (quotas, budgeting, and policy enforcement) at the platform layer. ...

Senior DevOps/SRE - Fintech London

Hiring Organisation
Adaptive Financial Consulting Ltd
Location
Greater London, United Kingdom
Employment Type
Full Time
proven track record of delivering powerful, elegant and intuitive trading technology solutions with a global reach. We are now looking for a Senior/SRE engineer to join our team! YOU ARE: A great team player who loves sharing knowledge and learning from others A great communicator Excited about learning … easy to deploy and manage, through automation & standardisation. Participating in best practice definition for project installations. WHY US: To be immersed in high-standard engineering culture. Our fantastic team takes pride in crafting elegant solutions to complex technical problems but also loves sharing their knowledge and helping you grow ...

Lead DevOps Engineer

Hiring Organisation
Shortlist Recruitment
Location
Chester, Cheshire, United Kingdom
Salary
£ 80 K
alongside supporting the infrastructure behind innovative AI products.Responsibilities:Provide technical leadership, mentoring and day-to-day direction to a small, distributed DevOps teamEstablish consistent engineering standards, tooling and best practices across the teamAct as a senior technical point of contact for stakeholders, development and infrastructure teamsLead the design …/CD automation and pipeline toolingExperience designing and operating resilient, cloud-native platformsExperience mentoring engineers or providing technical leadership within a DevOps, Platform or SRE environmentExcellent communication and stakeholder management skills, with the confidence to represent a technical team across the wider businessThe Lead DevOps Engineer role is remote-first ...

Managing Consultant / Senior Manager - Cloud Programme Manager

Location
Greater London, England, United Kingdom
client-facing environment.Delivering cloud migration and modernisation with at least one of AWS, Azure or GCP; familiarity with DevOps, FinOps, platform engineering and SRE is desirable.Strong programme governance capability – business case, roadmap, RAID, dependency management, executive reporting, and scope/commercial management.Stakeholder management across senior executives, finance, procurement, engineering … address the entire breadth of their business needs. It delivers end-to-end services and solutions leveraging strengths from strategy and design to engineering, all fueled by its market leading capabilities in AI, cloud and data, combined with its deep industry expertise and partner ecosystem. The Group reported ...

Managing Consultant / Senior Manager - Cloud Programme Manager

Location
Manchester, England, United Kingdom
facing environment. Delivering cloud migration and modernisation with at least one of AWS, Azure or GCP; familiarity with DevOps, FinOps, platform engineering and SRE is desirable. Strong programme governance capability – business case, roadmap, RAID, dependency management, executive reporting, and scope/commercial management. Stakeholder management across senior executives, finance … procurement, engineering, security and operations, with the ability to enlist support and commitment from peers in a matrixed organisation. Experience of programmes involving AI or data workloads, and/or using AI tools to accelerate programme delivery, is highly desirable. Certifications in one or more of MSP, PMP, SAFe ...

Managing Consultant / Senior Manager - Cloud Programme Manager

Location
United Kingdom
facing environment. Delivering cloud migration and modernisation with at least one of AWS, Azure or GCP; familiarity with DevOps, FinOps, platform engineering and SRE is desirable. Strong programme governance capability – business case, roadmap, RAID, dependency management, executive reporting, and scope/commercial management. Stakeholder management across senior executives, finance … procurement, engineering, security and operations, with the ability to enlist support and commitment from peers in a matrixed organisation. Experience of programmes involving AI or data workloads, and/or using AI tools to accelerate programme delivery, is highly desirable. Certifications in one or more of MSP, PMP, SAFe ...

BI Engineer, SRE (Remote, International)

Hiring Organisation
PulsePoint
Location
United Kingdom
Salary
£ 70 K
running has increasingly been a side responsibility for engineers who are primarily building features — and that's not sustainable. We're looking for an SRE to own that space: service health, incident response, infrastructure monitoring, and making sure we're not blindly burning cloud budget.The BI Engineer, SRE will ensure … Business Intelligence team's GCP-hosted APIs and data infrastructure. This role is responsible for proactive monitoring, incident response, and continuous improvement of platform reliability across a cloud-native stack. The engineer will work closely with backend and data engineers to maintain service health and drive operational excellence. This ...

Senior DevOps/SRE: Cloud, CI/CD & Automation Lead

Location
Greater London, England, United Kingdom
Adaptive is seeking a Senior/SRE engineer to join our London-based team. You will guide CI/CD, maintain infrastructure, and ensure high availability of customer-critical systems in a hybrid work setup. Ideal candidates have strong AWS, Linux, Terraform and Docker experience, plus Python/Bash scripting ...

Senior Principal Platform Lead || AI & Agentic Systems

Hiring Organisation
IFS
Location
London, United Kingdom
Salary
£ 80 K
world's largest organisations manage assets, operations and critical services. This is an opportunity to work at the forefront of modern AI engineering, building intelligent products that combine Large Language Models (LLMs), agentic AI and cloud-native technologies to solve complex, real-world business challenges at enterprise scale. … product teams and IFS R&D domain teams to understand where they’re struggling and prioritise what to fixPartner with IAM, infrastructure, and SRE teams to ensure the developer-facing surface of the platform is coherent, reliable, and secureEstablish feedback loops, developer surveys, usage analytics, office hours ...

DevOps Engineer

Hiring Organisation
Hackajob Ltd
Location
South West London, London, United Kingdom
Employment Type
Permanent, Part Time, Work From Home
working in a multiple disciplined team, and require a broad range of technical and soft skills to enable the team to implement sound DevOps engineering practices and deliver value quickly and continuously. These skill are categorised into the following domains. Automation skills : Automation is a key skill domain … work within a team using Agile methodology Scrum DevOps engineers should be an active member of the scrum team and contribute to sprint ceremonies SRE Should understand SRE principles and apply these to constantly improve the reliability and minimise the support burden within the team Security Clearance ...

Principal Security Engineer, Product & Infrastructure

Location
City Of London, England, United Kingdom
reproduce and triage vulnerabilities, dig through infrastructure configuration and build the automation that closes gaps. You’ll also set direction and bring product and engineering with you, but that influence comes from technical credibility and not process: the engineers you work with will take you seriously because … triage them, design or validate the mitigation, and confirm it actually worked. Improve the KPIs that tell you whether the process is holding. Detection Engineering - Build and improve our detection capability alongside the infrastructure and engineering teams: identify the signals worth collecting, write rules that catch real attacks ...

Senior SRE & DevTools Engineer (CI/CD & Observability)

Location
Ham, England, United Kingdom
Visa is seeking a Software Engineer + SRE hybrid to join its UK Cloud platform team. You will safeguard reliability, automate resolution of recurring issues, and work with developers to optimize CI/CD pipelines. The role combines hands-on SRE with software engineering, requiring experience with GitHub ...

Senior Software Engineer, Substrate

Hiring Organisation
Palantir Technologies
Location
London, United Kingdom
Salary
£ 80 K
will also be responsible for ensuring scale, stability and security across a matrix of compliance regimes and hosting infrastructure types. Your team culture emphasizes engineering rigor and operational excellence at scale. This means issues in production should be pre-empted and deeply root-caused, and investments in automation … boltsDeep familiarity with containers (Docker) and orchestration (Kubernetes) at scaleExperience working with a cloud provider (AWS/Azure/GCE), or sysadmin/SRE experience in data centersExperience designing, building, and operating high-scale observability or infrastructure systemsWorking knowledge of networking fundamentals, experience with CNIs or cloud networking infrastructure preferredWhat ...

DevOps Engineer - Dexory

Location
Wallingford, England, United Kingdom
DevOpsEngineer/SRE/Cloud/Docker/Kubernetes Location:Wallingford-Oxfordshire(Hybrid)onsitecirca1dayperweek Permanent AtDexory,wearetransformingthewarehousingandlogisticsindustrythroughdataintelligence. We’re currently recruiting for a devops engineer to define, build and operate the infrastructure and tooling necessary for our rapidly growing business. You will be responsible for supporting high-velocity deployment … user-facing applications, internal services, and infrastructure at scale, while ensuring stability, reliability and automation. OfficesbasedinWallingford,Oxfordbutflexiblehybridbasis(circaonedayeveryweekonsite) Anyexperiencewithinroboticsorstartup/scaleupenvironmentsdesirablealongwithexperienceworkingonmultiplecloudplatformsvsbeingtiedtojustvendor. Opentothecareerbackgroundcandidatesarecomingfrom(infrastructure,DevOps,sitereliability,platformetc) Responsibilities/Skills DevelopandMaintainCI/CDpipelines anddevelopertoolingthatsupportsmultipleengineeringteamswithdifferenttechnicalrequirements–enablingself-service,repeatabledeployments,andhigh-throughputdelivery. Managecloudinfrastructure acrossawiderangeofcloudproviders,usingKubernetes,forbothuser-facingapplicationsandinternalservices–provision,monitor,optimisecost,performanceandreliability. Buildinfrastructure ...

Principal Engineer, CSRE Provisioning (Remote, United Kingdom)

Location
Greater London, England, United Kingdom
events they love. It truly is a unique and rewarding environment. You will be part of the Provisioning team within CSRE (Customer/Client Site Resilience Engineering). Provisioning reduces operational complexity across the Ticketmaster hosted CSRE platform estate — we own the lifecycle of the platforms behind Resilience … decommissioning. Deep expertise designing and troubleshooting distributed systems, with a track record of resolving the most complex cross-service failure modes. Expert understanding of SRE principles — SLIs, SLOs, error budgets — and how to apply them consistently across a large, heterogeneous estate. Extensive experience with on-premises data centre infrastructure ...

Senior Software Engineer (Application Operations)

Location
Greater London, England, United Kingdom
coordinated remediation alongside the incident commander. Use structured diagnostics before escalating — attach clear evidence, reproducibility steps, and impact assessments to every L3/SRE handoff. Feed operational findings into Problem Management and contribute to post‐incident reviews; capture learning in improved runbooks, alerts, and automation. Quality, Process, and Continuous Improvement … accessible, and kept up to date. Stakeholder Collaboration Work closely with the Director of Application Operations, Problem Manager, and PETO peers (Platform, Infrastructure, Data, SRE) to ensure a coherent, joined‐up operational approach. Partner with product‐aligned engineering teams to understand application architecture, service dependencies, and failure modes; encode ...

Remote UK SRE: Observability & Reliability Lead

Location
United Kingdom
Orex Nova, Inc. is seeking an experienced SRE/Platform Engineer to help keep our systems fast, reliable, and observable. This fully remote role covers the UK and requires strong incident response experience. You will own blameless post-mortems, drive reliability improvements, and collaborate with a team focused ...

Sr Director Analyst - IT Service Management

Hiring Organisation
Gartner
Location
United Kingdom
Salary
£ 80 K
such as dependency mapping/CMDB optimization—and delivering high-impact content that shapes industry directionWorking experience with modern operations practices such as DevOps, SRE, platform engineering or agile product teamsDemonstrate executive presence; can immediately establish credibility with executives and additional stakeholdersStrong organizational skills; ability to work under tight ...

Senior Architect Private Cloud

Hiring Organisation
Randstad Technologies Recruitment
Location
Sheffield, South Yorkshire, United Kingdom
Employment Type
Contract
Contract Rate
£500 - £550/day Inside IR35 via Umbrella
Role: Senior Architect - Private Cloud Location: Sheffield, UK (Hybrid: 2 days/week on-site) Duration: 6 Months (Inside IR35) Experience Level: 10+ Years Role Overview We are seeking a Senior Private Cloud Architect to define, design, and drive the evolution of our enterprise private cloud infrastructure. You will … policy-as-code, vulnerability management, and robust RBAC. Ensure production readiness across multi-region, high-availability environments. Stakeholder Engagement: Collaborate across Product, Cyber Security, SRE, and CTO domains to influence technical decision-making and mentor engineering teams. Key Requirements Core Expertise: 10+ years in platform architecture/engineering ...

AWS Cloud Engineer

Hiring Organisation
Anson McCade
Location
City of London, London, United Kingdom
Security I'm working with a leading National Security and Defence organisation to find an experienced AWS/DevOps Engineer to join a growing engineering team in Manchester or London. This is a genuinely hands-on cloud engineering role, working on complex systems supporting UK National Security customers. … Strong engineering community and long-term development opportunities What We're Looking For We're interested in experienced AWS Cloud, DevOps, Platform and SRE Engineers with strong hands-on experience across AWS, Infrastructure as Code and modern DevOps practices. Security clearance is required, with candidates holding ...

Java/Kotlin Software Engineer – AVP - (Developer Enablement)

Hiring Organisation
Citigroup
Location
Belfast, Down, United Kingdom
Salary
£ 60 K
takes it further. This blue-sky project looks at ways we can greatly improve the SDLC process at Citi, taking ideas like grading engineering teams and rewarding those who follow best practices, perhaps moving them from manual approval process to an automatic one. The use of Artificial Intelligence … requirements of the firm are upheld, whilst driving industry best practices (DORA) Why you'll love working here: You get to work in the engineering focused part of the bank, the Chief Technology Office, building tools for other engineersYou’ll work and lead small, agile team, in an organisation ...

Senior Backend Engineer - Databases Pyroscope | UK | Remote

Hiring Organisation
Grafana Labs
Location
United Kingdom
Salary
£ 70 K
development to reducing friction across the full experience: better onboarding, deeper integration with the rest of Grafana, and profiles that work for operational and SRE workflows, not just performance specialists.Over the next year, you will help us:Ship Adaptive Profiles as the default ingestion strategy, so customers automatically collect … robust, performant software that others can maintain, and you know when to optimize versus when to ship.Operational mindset. You have carried a pager, done SRE-style work or infrastructure as code, and treat reliability as a feature.Pragmatism. You break complex problems into short feedback loops: analyze, design, deliver ...