26 to 50 of 966 High Availability Jobs in the UK

Global Banking Markets - Software Engineer - Equities Data Platform Engineering - London - Vice President

Location
City Of London, England, United Kingdom
leads on the ground‐up rebuild of our Electronic Trading data infrastructure into a cloud‐native, AI‐first platform. Design, build, and maintain a high‐performance, highavailability, high‐capacity platform that processes and persists massive volumes of business‐critical data in near real time, distributing … raise the bar on engineering practices across the team. Required Qualifications & Skill‐Sets Around 5–10 years of professional experience with experience building high‐performance, low‐latency systems (sub‐second) – at Vice President level. Proven ability to drive complex, industrial‐scale development independently, while also leading and developing engineers ...

DevOps Engineer- Night Shift(10:00 PM - 6:00 AM)

Location
United Kingdom
production environments healthy and performant, while simultaneously designing and maintaining the CI/CD pipelines, infrastructure‐as‐code frameworks, and tooling that enable rapid, high‐quality software delivery. You are the connective tissue between engineering, platform, and operations — someone who is equally comfortable in an incident bridge call … production downtime, performance degradation, and security‐related incidents in a timely, structured manner. Perform end‐to‐end operational duties covering application server health, service availability, and platform integrity in accordance with documented processes and runbooks. Review and manage client service request tickets in adherence to defined SLAs, ensuring accountability ...

Infrastructure Engineer

Location
Greater London, England, United Kingdom
Build resilient, scalable, fault-tolerant infrastructure for WRITER's high-traffic enterprise generative AI platform Move between SRE, DevOps, Infrastructure, and Platform initiatives as priorities shift Automate operational tasks and infrastructure management with Python or Go Design and operate infrastructure across AWS, GCP, and Azure Work with Kubernetes, Helm … director of engineering Requirements 5+ years of experience in infrastructure engineering, DevOps, or a similar role focused on building and operating large-scale, high-availability production systems at a high-growth product company Experience running containerisation in production Experience with Helm and Terraform or Pulumi ...

Senior Backend Engineer (Ruby), AI Engineering: AI Coding

Hiring Organisation
GitLab
Location
United Kingdom, UK
Employment Type
Full-time
into their daily workflows to drive efficiency, innovation, and impact. GitLab is where careers accelerate, innovation flourishes, and every voice is valued. Our high-performance culture is driven by our values and continuous knowledge exchange, enabling our team members to reach their full potential while collaborating with industry leaders … broader event platform, the core infrastructure that many of GitLab's AI features rely on to know when and how to run. Architect for high availability and throughput, since this part of the system sits at the center of automated AI workflows across GitLab and needs to reliably ...

Systems Integration Advisor

Hiring Organisation
NTT DATA
Location
London, UK
Employment Type
Full-time
environments, from development to mission-critical production systems. Configure and maintain database servers and processes, including monitoring of system health and performance, to ensure high levels of performance, availability, and security. Database performance tuning. Planning, deploying, maintaining, troubleshooting of high-availability database environments (RAC. Dataguard etc. … monitoring and troubleshooting tools Experience with backups, restores and recovery models Experience with database administration Experience with incident and problem queue management. Knowledge of High Availability (HA) and Disaster Recovery (DR) options for Oracle Experience working with Linux, Windows & Unix servers Excellent written and verbal communication, problem solving ...

Systems Integration Advisor

Hiring Organisation
NTT DATA
Location
Dunstable, Bedfordshire, UK
Employment Type
Full-time
environments, from development to mission-critical production systems. Configure and maintain database servers and processes, including monitoring of system health and performance, to ensure high levels of performance, availability, and security. Database performance tuning. Planning, deploying, maintaining, troubleshooting of high-availability database environments (RAC. Dataguard etc. … monitoring and troubleshooting tools Experience with backups, restores and recovery models Experience with database administration Experience with incident and problem queue management. Knowledge of High Availability (HA) and Disaster Recovery (DR) options for Oracle Experience working with Linux, Windows & Unix servers Excellent written and verbal communication, problem solving ...

Senior Data Engineer

Location
West of England, England, United Kingdom
resilience, and availability.* Troubleshoot and resolve database-related issues in development and production environments.* Implement and maintain database security controls, backup strategies, replication, and high-availability solutions.* Support change management activities and continuous improvement of data platform services.* Develop tools for data-driven insights and task automation* Collaborate … Spark, Hadoop, NiFi* Exposure to infrastructure scaling e.g. Kubernetes* Exposure to Ansible and Terraform* Understanding of ML algorithms and deployment* Experience with SQL Server high-availability technologies, including replication, mirroring, or clustering.* Experience with Windows Server administration.* Exposure to PowerShell scripting and automation.* Experience working within Agile delivery ...

DevOps Engineer, Studios

Location
Greater London, England, United Kingdom
pipelines across our digital and broadcast platforms. The successful candidate will play a key role in enabling continuous delivery, improving system reliability, and supporting high-profile clients and live event services, ensuring optimal performance and resilience across all environments. Key Responsibilities and Accountabilities Design, build, and maintain scalable cloud … efficient, reliable software delivery across multiple teams. Automate infrastructure provisioning using Infrastructure as Code tools such as Terraform, CloudFormation, or similar. Monitor system performance, availability, and reliability using observability tools such as Prometheus, Grafana, and ELK stack. Ensure high availability and disaster recovery strategies are in place ...

Senior Fullstack Engineer (Python + React.js)

Location
Greater London, England, United Kingdom
maintain scalable, secure, and efficient server-side applications. The ideal candidate should have experience in microservices architecture, API development, database management, and frontend ensuring high availability and performance of backend and frontend services. In this role, you will collaborate closely with frontend engineers, product managers, and other stakeholders … Redux or React Query Collaborate with frontend developers to ensure efficient API integration and a seamless user experience. Troubleshoot and resolve production issues, ensuring high availability and minimal downtime. Write unit and integration tests to maintain code reliability and ensure high- quality releases. Continuously monitor and optimize ...

Software Engineering Manager

Location
Manchester, England, United Kingdom
across major global airport hubs. You will lead a team of senior engineers, partnering closely with Product Management, Data Science, and Design to scale high-availability, distributed microservices handling high-volume traffic. What You'll Do Partner with Product, Data, and Design to execute a multi-year … platform strategy. Manage, mentor, and coach an engineering team, fostering a high-trust, collaborative culture. Drive recruitment efforts to build out a strong team of senior individual contributors. Establish operational processes to guarantee high availability, security, and reliability across distributed systems. Leverage test automation, TDD/ ...

AWS DevOps Platform Engineer - eSC/eDV Clearance

Location
Leicester, England, United Kingdom
gain hands on experience with cutting edge technologies, and deliver solutions that create real business impact. From day one, you’ll work on meaningful, high profile programmes that stretch your skills and accelerate your growth. We invest heavily in you—supporting continuous learning, in demand skills development, and long … seamless integration with DevOps toolchains and CI/CD pipelines Apply knowledge of Storage, Compute, and Security Services to administer hybrid cloud environments Configure High Availability (HA) and Disaster Recovery (DR) solutions and manage Kubernetes clusters Join our team and contribute to the development of innovative AWS infrastructure ...

platform engineer in hybrid cloud

Location
Greater London, England, United Kingdom
deployment models Integrate applications with DevOps toolchains and CI/CD pipelines Administer hybrid cloud environments using Storage, Compute, and Security Services Configure High Availability and Disaster Recovery solutions Manage Kubernetes clusters Contribute to innovative AWS infrastructure solutions that drive business success. Требования: Hands‐on experience designing … cloud‐native architecture patterns Experience integrating CI/CD pipelines and DevOps toolchains such as GitLab CI, Jenkins, or GitHub Actions Experience implementing High Availability and Disaster Recovery strategies in AWS environments Strong Linux systems knowledge and troubleshooting capability in cloud and container environments Ability to work collaboratively ...

Database Administrator

Location
Greater London, England, United Kingdom
solutions under the guidance of the DBA Team Participate in team meetings and proactively contribute to database strategy discussions Other requirements raised by management High Availability & Scalability Support high-availability and disaster recovery (HA/DR) solutions, including replication and failover strategies Optimize databases for scalability ...

Site Reliability Engineer, Studios

Location
Uxbridge, England, United Kingdom
project stakeholders. Key Responsibilities And Accountabilities Design, build, and maintain reliable, scalable infrastructure and platform services across on‐premises and cloud environments. Improve service availability, latency, performance, and operational efficiency through engineering‐led reliability practices. Build and enhance observability across services and infrastructure, including monitoring, logging, alerting, dashboards … Drive root cause analysis and corrective actions following incidents, with a focus on prevention and continuous improvement. Support the design, testing, and documentation of high availability, backup, failover, and disaster recovery arrangements. Help enforce security, access control, patching, and operational best practices across infrastructure and services. Optimise system ...

Site Reliability Engineer, Studios

Location
Greater London, England, United Kingdom
project stakeholders. Key Responsibilities and Accountabilities Design, build, and maintain reliable, scalable infrastructure and platform services across on-premises and cloud environments. Improve service availability, latency, performance, and operational efficiency through engineering-led reliability practices. Build and enhance observability across services and infrastructure, including monitoring, logging, alerting, dashboards … Drive root cause analysis and corrective actions following incidents, with a focus on prevention and continuous improvement. Support the design, testing, and documentation of high availability, backup, failover, and disaster recovery arrangements. Help enforce security, access control, patching, and operational best practices across infrastructure and services. Optimise system ...

Infrastructure Specialist - London

Hiring Organisation
SystemRS
Location
London, United Kingdom
Employment Type
Permanent
Salary
GBP 80,000 - 100,000 Annual
Dell server hardware and enterprise storage platforms. Strong understanding of virtual infrastructure and underlying physical environments. Backup and recovery technologies including Veeam and Rubrik. High availability and disaster recovery design and implementation. Experience with Zerto replication and failover solutions. Microsoft Technologies Windows Server . Active Directory and Entra … Code and configuration management tools such as Ansible, DSC, Chef, or Puppet. Experience leveraging AI-assisted coding and automation tools. Additional Experience SQL Server high availability deployments. Citrix administration. SolarWinds monitoring and alerting. Server hardware troubleshooting and vendor management. Personal Attributes 10+ years' experience in enterprise infrastructure roles ...

Engineer - Site Reliability

Location
Greater London, England, United Kingdom
support model for its US Global Trading Hours (GTH) markets, providing critical overnight and early‐session coverage from London that ensures continuous, highavailability operations across Cboe's real‐time low‐latency trading platforms. The London‐based SRE provides technical support to Cboe Trade Desk and Operations Support … timely, precise communication to stakeholders during active incidents and contribute to post‐incident reviews and remediation tracking to drive long‐term platform stability. System Availability & Technical Support: Provide technical support and operational oversight to sustain resiliency and high availability of critical business operations. Monitor production, disaster recovery ...

Senior MS SQL Database Administrator

Hiring Organisation
KBR
Location
Guildford, Surrey, UK
Employment Type
Full-time
part of that new company. Our Mission Technology Solutions business partners with governments and defense, intelligence, space, aviation, and critical infrastructure customers to deliver high-end engineering, science, technology, and mission support solutions. From national security and readiness to advanced research, cyber, logistics, and life-cycle sustainment, our teams … leveraging technology to create lasting business value. About the RoleThe Senior MS SQL Database Administrator is responsible for the administration, maintenance, performance, security, and availability of enterprise Microsoft SQL Server environments supporting critical business applications worldwide. This role serves as a senior member of the Global Database Team, providing ...

Senior Network Engineer

Location
Greater London, England, United Kingdom
team responsible for network architecture, deployment, and operational readiness. You will play a key role in ensuring connectivity solutions align with performance, security, and availability requirements. The ideal candidate combines deep networking expertise with practical experience in cloud networking, compute platforms, and enterprise infrastructure. Success in this role requires … network technologies including routing, switching, wireless, firewalls, and secure remote access solutions. Lead network architecture decisions with a focus on scalability, resiliency, and security (high availability, redundancy, segmentation). Implement and enforce network security controls including segmentation, zero trust principles, and secure access patterns. Support hybrid connectivity models ...

Cloud Infrastructure Consultant

Location
York and North Yorkshire, England, United Kingdom
alongside them. Working knowledge of Microsoft 365 is desirable. You will work closely with pre-sales, Principal Consultants, Architects, and operational teams to deliver high quality, supportable solutions aligned to agreed standards. What You Will Do Provide hands-on design and delivery expertise across datacentre, hybrid, and Azure environments … Services, Group Policy, DNS, DHCP, certificate services, and hybrid identity through Entra Connect. Design and deploy resilient compute and storage platforms, including failover clustering, high availability, shared and software-defined storage, and the backup and replication that underpins them. Lead and support migration of customer server estates, including ...

Java Software Developer

Location
Greater London, England, United Kingdom
Fasanara Digital is a quantitative investment team applying a scientific, high frequency investment style in digital assets, seeking to achieve exceptional risk-adjusted returns for our investors. We were founded in 2018 and have grown to a 30-person strong team, managing over $500m USD in a basket … delta-neutral trading strategies. Our team members come from diverse backgrounds. We are fully dedicated to building out our globally deployed, 24/7 availability trading platform, which allows us to capture trading opportunities on >15 crypto liquidity venues, and maintain our position as one of the top trading ...

Senior DevOps Analyst

Hiring Organisation
NTT DATA
Location
London, UK
Employment Type
Full-time
/CD, cloud platforms, automation, containerization, and infrastructure management. The ideal candidate will drive DevOps best practices, enable scalable deployments, and ensure high availability and performance of applications. Key ResponsibilitiesDesign, implement, and maintain CI/CD pipelines for automated build, test, and deployment. Manage and optimize cloud infrastructure … Qualifications8+ years of experience in DevOps/Site Reliability/Infrastructure Engineering. Strong understanding of DevOps and CI/CD best practices. Experience supporting high-availability, production systems. Experience in Agile/Scrum environments. Bachelor's degree in Computer Science, Engineering, or equivalent. Good to HaveExperience with DevSecOps ...

Senior / Lead Site Reliability Engineer

Location
Watford, England, United Kingdom
role At Allwyn, the Senior/Lead Site Reliability Engineer is responsible for technical leadership of reliability engineering across the digital estate, ensuring high availability, performance, and resilience of customer-facing systems during both normal operation and peak lottery events. The role combines hands-on engineering, incident leadership … reporting, working across platform, product, and operational teams. Objectives of the role Own reliability outcomes across services using SLOs, SLIs, and error budgets Improve availability, latency, and scalability across Instant-Win and Draw-based platforms Lead incident response and operational readiness, including peak jackpot events Drive automation and platform ...

IT Support Engineer I, IT Services

Hiring Organisation
Amazon
Location
Manchester, UK
Employment Type
Full-time
passionate about solving technical challenges and helping people? Do you thrive in a dynamic, high-impact environment? Join Amazon's IT Services team as an ITS Support Engineer and be part of the engine that powers Amazon's seamless operations. We're seeking customer-focused, innovative problem-solvers … standards, systems and equipment deployed throughout Amazon. They can work independently or collaborate with partner teams and contractors managing projects while maintaining a high level of productivity to meet goals. Quickly adapting to new processes and procedures, they act as a mentor and main partner for escalations within ...

IT Support Engineer I, IT Services

Hiring Organisation
Amazon
Location
Edinburgh, UK
Employment Type
Full-time
passionate about solving technical challenges and helping people? Do you thrive in a dynamic, high-impact environment? Join Amazon's IT Services team as an ITS Support Engineer and be part of the engine that powers Amazon's seamless operations. We're seeking customer-focused, innovative problem-solvers … standards, systems and equipment deployed throughout Amazon. They can work independently or collaborate with partner teams and contractors managing projects while maintaining a high level of productivity to meet goals. Quickly adapting to new processes and procedures, they act as a mentor and main partner for escalations within ...