226 to 250 of 645 Permanent High Availability Jobs

Lead Site Reliability Engineer

Hiring Organisation
Jobleads-UK
Location
Glasgow, Scotland, United Kingdom
with team members to identify comprehensive service level indicators and stakeholders to establish reasonable service level objectives and error budgets with customers Demonstrates a high level of technical expertise within one or more technical domains and proactively identifies and solves technology-related bottlenecks in your areas of expertise Acts … logging best practices. Certification in AWS, Kubernetes, or relevant technologies. Proven track record in system health monitoring, capacity management, and blameless postmortems for high-availability services Deep understanding of distributed system design principles, networking (TCP/IP, DNS, load balancing), Linux internals. Contributions to open-source observability ...

Senior DevSecOps Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
Senior DevSecOps Engineer, you will be a key driver in embedding security into every phase of the software development lifecycle. You’ll join a high-impact team responsible for securing a highly available, multi-tenant platform built on cloud-native infrastructure (GCP and Kubernetes), taking a proactive and automated … expertise to the security incident response function, helping manage and resolve security events swiftly. Your Profile Essential: Deep, practical experience designing, managing, and securing high-availability infrastructure within GCP. Proficiency in API security: reviewing, defining patterns for, and upskilling engineers on secure API design. Expert knowledge of hardening ...

Network Infrastructure Consultant

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
strategic goals and technology initiatives. Qualification Key Responsibilities Design, Implement, and Manage Network Infrastructure Lead the design and deployment of robust, scalable, and high-performing network architectures, including LAN, WAN, WiFi, and cloud networking solutions. Expertise in routing, switching, and network segmentation to support business operations across multiple sites … environment. Cloud Networking Expertise Architect and manage cloud-based network solutions (e.g., AWS, Azure, GCP) to support a range of applications and workloads. Ensure high availability and performance of network services in cloud environments, leveraging virtual network functions, VPNs, and direct connections. Design cloud-to-cloud and hybrid ...

Site Reliability Engineer

Hiring Organisation
Jobleads-UK
Location
Watford, England, United Kingdom
will work closely with Senior SREs and engineering teams to ensure services remain stable, scalable, and well-instrumented across both steady-state and high-demand events. What you’ll be doing Objectives of the role Maintain reliable production services across digital platforms Improve monitoring, alerting, and observability coverage Reduce … deployment processes and CI/CD improvements Platform & cloud Work with AWS services (primarily ECS, with exposure to EKS/Kubernetes) Support scalability and availability improvements Assist in performance tuning and capacity planning Collaboration Work closely with engineers to: + Improve service reliability + Support releases and production readiness ...

DevOps Engineer

Hiring Organisation
ISR Recruitment
Location
United Kingdom
excellent opportunity to work within a modern AWS cloud environment, supporting the deployment, monitoring and continuous improvement of live services used across a high-profile government programme. The successful candidate will be a hands-on DevOps Engineer with strong AWS expertise, excellent troubleshooting skills and experience supporting production environments. … other highly regulated environments Role and Responsibilities: Support the deployment, operation and continuous improvement of cloud-based services hosted within AWS. Monitor the health, availability and performance of live production environments. Investigate, diagnose and resolve infrastructure, application and performance issues. Collaborate with development teams to identify root causes ...

Infrastructure Lead - FinTech, Azure, Security

Hiring Organisation
Quant Capital
Location
London, UK
support of IT infrastructure services. • Design, implement, and maintain secure, scalable, and resilient infrastructure solutions. • Supporting our primarily Microsoft-based infrastructure, ensuring its high availability, resilience, and performance. Both hybrid cloud and on-premise environments. • Oversee server, storage, network, backup, cloud and virtualisation platforms. • Ensure infrastructure availability … required. • Leadership and Team Management • Lead and develop the small Infra team, strong leadership and people management skills. What you’ll need • A high level of experience with Microsoft technologies including Windows Server andDesktop, Active Directory, Entra, Hybrid domains, Group Policies, 365, Intune, Teams,SharePoint, NTFS etc. • Networking ...

DevOps Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
continuous improvement of our deployment pipelines, cloud and bare metal environments, and Kubernetes‐based orchestration layer — ensuring our platform scales reliably to support high‐performance AI and GPU workloads. You'll set the standard for engineering excellence, partnering with Platform, SRE and Product Engineering teams to align on strategy … Ansible, ensuring infrastructure is version‐controlled, peer‐reviewed, and consistently provisioned across dev, staging, and production Manage and maintain hosted application environments, ensuring high availability, scalability, and smooth deployment of services into production Partner closely with Go engineers to deeply understand their development workflows and release processes, identifying ...

Senior Cloud Engineer

Hiring Organisation
Robinhood Financial
Location
London, UK
part of the Robinhood family, the Exchange Platform team owns the full service lifecycle. We are the architects of a modern, high-velocity ecosystem that enables our global expansion, ensuring the world’s longest-running exchange remains unshakeable.The RoleAs a Senior Backend Engineer integrated into the Exchange Platform team … essential part of building the new generation infrastructure for low-latency services in the cloud, adopting cutting-edge technologies to drive our high-velocity ecosystem.This role is based in our London office(s), with in-person attendance expected at least 3 days per week. At Robinhood, we believe ...

Principal Network Engineer - remote across UK or Ireland

Hiring Organisation
17918
Location
Belfast, County Antrim, United Kingdom
world-class engineering team responsible for designing, securing, and optimising a large-scale global network. You will play a key role in evolving a high-performance infrastructure that supports complex routing, automation, and next-generation security. This is a deeply technical, hands-on position, ideal for engineers who want … disaster recovery planning, testing, and monitoring/forensics capability across the network estate Work closely with global peers to ensure scalability, security, and high availability Act as a technical mentor and thought leader within the network engineering function Contribute to continuous improvement and adoption of new technologies Required ...

Principal Platform Security Engineer

Hiring Organisation
Jobleads-UK
Location
York and North Yorkshire, England, United Kingdom
DevOps/Platform Engineering experience delivering solutions in Azure and/or GCP. Full‐stack application and infrastructure solution design with robust security controls, high availability, and operational resilience. Working knowledge of vulnerability and compliance management (scanning to remediation), patch management, endpoint protection/anti‐malware, and access … delivery focus, capable of prioritizing effectively and delivering outcomes in a fast‐paced environment with shifting demands. Ability to operate effectively in a small, high‐impact team while collaborating across a wider product/engineering organisation. Excellent communication and stakeholder‐management skills, able to influence at all levels ...

Site Reliability Engineer

Hiring Organisation
ISR Recruitment
Location
United Kingdom
excellent opportunity to work within a modern AWS cloud environment, supporting the deployment, monitoring and continuous improvement of live services used across a high-profile government programme. The successful candidate will be a hands-on engineer with strong AWS expertise, excellent troubleshooting skills and experience supporting production environments. … call back in the strictest confidence. If you're an experienced Site Reliability Engineer (DevOps) with strong AWS expertise and a passion for supporting high-availability cloud services; please contact Edward Laing here at ISR to learn more about our client and how they are leading ...

Principal Platform Security Engineer

Hiring Organisation
Hiscox AG
Location
London, UK
years’ DevOps/Platform Engineering experience delivering solutions in Azure and/or GCP.Full‐stack application and infrastructure solution design, ensuring robust security controls, high availability, and operational resilience throughout.Working knowledge of vulnerability and compliance management (scanning through to remediation), patch management, endpoint protection/anti-malware … value.Strong delivery focus, able to prioritise effectively and deliver outcomes in a fast-paced environment with shifting demands.Able to operate effectively in a small, high-impact team while collaborating across a wider product/engineering organisation.Excellent communication and stakeholder-management skills, able to influence at all levels and present ...

Principal Platform Security Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
DevOps/Platform Engineering experience delivering solutions in Azure and/or GCP.* Full‐stack application and infrastructure solution design, ensuring robust security controls, high availability, and operational resilience throughout.* Working knowledge of vulnerability and compliance management (scanning through to remediation), patch management, endpoint protection/anti-malware … delivery focus, able to prioritise effectively and deliver outcomes in a fast-paced environment with shifting demands.* Able to operate effectively in a small, high-impact team while collaborating across a wider product/engineering organisation.* Excellent communication and stakeholder-management skills, able to influence at all levels ...

Senior Platform Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
commercial and residential installation companies. Access to our unique on-line portal and design tool, plus our emphasis on product quality, consistency and availability sets us apart in the market. Segen is a fast-moving business that responds quickly to any market changes supported by its bespoke ERP system … guidance. Promote cloud security best practices and ensure compliance within infrastructure design. Contribute to the evolution of our monitoring and alerting systems to maintain high availability and performance. What We’re Looking For Deep experience in Microsoft Azure, including Hub and Spoke networking. Strong proficiency with Azure DevOps ...

Lead Network Operations Engineer

Hiring Organisation
G Research
Location
London, UK
world-class platform to amplify our teams’ most powerful ideas.As part of our engineering team, you’ll shape the platforms and tools that drive high-impact research - designing systems that scale, accelerate discovery and support innovation across the firm.Take the next step in your career.The roleThis is an exciting … security infrastructure across datacentre and office environments.This is a hands-on technical leadership role focused on operational excellence, automation, observability, and incident response — ensuring high availability, resilience, and a strong security posture.Key responsibilities:Own the day-to-day performance, stability, availability, and security of network platforms across ...

Senior Cloud Engineer

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
part of the Robinhood family, the Exchange Platform team owns the full service lifecycle. We are the architects of a modern, high-velocity ecosystem that enables our global expansion, ensuring the world’s longest-running exchange remains unshakeable. The Role As a Senior Backend Engineer integrated into the Exchange … essential part of building the new generation infrastructure for low-latency services in the cloud, adopting cutting-edge technologies to drive our high-velocity ecosystem. This role is based in our London office(s), with in-person attendance expected at least 3 days per week. At Robinhood, we believe ...

Senior Cloud Platform Engineer (GCP | Kubernetes | DevSecOps)

Hiring Organisation
Jobleads-UK
Location
Bolsterstone, England, United Kingdom
service mesh technologies and secure service-to-service communication. Manage ingress, API gateways and network routing components. Implement platform resilience, backup, disaster recovery and high availability capabilities. Support workload identity and platform authentication mechanisms. Ensure platform compliance with enterprise security, governance and regulatory requirements. Optimise platform performance, scalability ...

Senior IT Technician

Hiring Organisation
IT Talent Solutions
Location
Walsall, West Midlands (County), United Kingdom
Employment Type
Permanent
Salary
£40000 - £47000/annum + Bens
cloud services, vendor management, and continuous improvement. Key Responsibilities Own and manage the Group's IT infrastructure across on-premise and cloud environments. Ensure high availability, performance, resilience, and security of all IT systems. Lead cybersecurity initiatives including MFA, endpoint protection, patching, and vulnerability management. Manage Microsoft ...

Lead Site Reliability Engineer

Hiring Organisation
FNZ
Location
London, UK
This role focuses on deploying, integrating, and providing ongoing operational support for mission-critical systems, leveraging modern automation and cloud-native practices.Key Responsibilities· Maintain high availability and performance of FNZ platforms.· Implement monitoring, alerting, and observability solutions to proactively detect and resolve issues.· Collaborate with engineering teams ...

Senior Data Architect

Hiring Organisation
Stantec
Location
London, UK
sources, optimizing application performance and planning capacity for growth. Ensure that data storage, indexing, and retrieval mechanisms are efficient and that disaster recovery and high-availability solutions are in place.Security and Compliance: Ensure robust data security measures and compliance protocols are implemented across PM2PA, Oracle EBS and integrated ...

Frontend Architect - React

Hiring Organisation
HCLTech
Location
City of London, London, United Kingdom
Developer (React, TypeScript and Low-Latency) Role Overview- We are seeking a highly skilled UI Developer with strong expertise in building time-critical, high-availability, and high-performance applications. The ideal candidate will have hands-on experience across the full UI development lifecycle, including design, development, testing ...

Site Reliability Engineer, Infrastructure - ThousandEyes

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
powered assurance insights within Cisco’s Networking, Security, Collaboration, and Observability portfolios. Our distributed Site Reliability Engineering team of approximately nine engineers owns the availability, latency, performance, efficiency, monitoring, emergency response, and capacity planning of the platform while partnering closely with application development teams. We believe in operations, infrastructure … scale, highly available distributed systems that allow the ThousandEyes platform to process a constantly growing volume of telemetry data. Use AI Tooling to write high‐quality code and automated solutions that enable fast, reliable releases, reduce operational expense, and allow our infrastructure and platforms to scale thoughtfully across regions. ...

Lead Software Engineer - GLCM Postings

Hiring Organisation
JP Morgan Chase
Location
Bournemouth, Dorset, UK
goals and industry best practicesLead and mentor a team of software engineers, providing guidance, support, and professional development opportunitiesDrive the design and implementation of high-performance backend services using Java and cloud-native technologiesChampion the development of highly resilient, share-nothing, multi-region architectures to achieve fault tolerance, high availability, and disaster recoveryDrive the adoption and optimization of MongoDB (including change streams), Kubernetes, and other modern platformsSet and enforce standards for API design and development (REST/gRPC), including best practices for versioning, documentation, and error handling to ensure robust, secure, and scalable interfacesGuide the integration ...

Site Reliability Engineer, Infrastructure - ThousandEyes

Hiring Organisation
Jobleads-UK
Location
City Of London, England, United Kingdom
powered assurance insights within Cisco’s Networking, Security, Collaboration, and Observability portfolios. Our distributed Site Reliability Engineering team of approximately nine engineers owns the availability, latency, performance, efficiency, monitoring, emergency response, and capacity planning of the platform while partnering closely with application development teams. We believe in operations, infrastructure … scale, highly available distributed systems that allow the ThousandEyes platform to process a constantly growing volume of telemetry data.Use AI Tooling to w rite high-quality code and automated solutions that enable fast, reliable releases, reduce operational expense, and allow our infrastructure and platforms to scale thoughtfully across regions. ...

Senior Site Reliability Engineering Manager

Hiring Organisation
Jobleads-UK
Location
Greater London, England, United Kingdom
real‐time low‐latency trading platforms around the clock and provides direct platform support for Cboe's European operations. Responsibilities Technical Leadership & System Availability: Provide technical leadership, support and operational oversight to sustain resiliency and high availability of critical business operations across European and GTH market sessions ...