1 to 25 of 590 High Availability Jobs

OT Technical Support Manager

Hiring Organisation
Hays
Location
Perth, Perthshire, Scotland, United Kingdom
Employment Type
Contract
assignment within the utilities sector. Your new role As the OT Technical Support Manager, you will lead the operational support, maintenance and resilience of high-availability Operational Technology systems. You will oversee service delivery and incident response while ensuring that patching, vulnerability remediation and planned maintenance activities … role will also involve workforce planning, capability development and leadership of a technical support team. Key responsibilities Lead the operational support and maintenance of high-availability OT systems. Manage incident response, fault diagnosis, system recovery and operational resilience. Oversee service delivery across the supported OT environment. Coordinate ...

SQL Database Administrator

Hiring Organisation
Hays Specialist Recruitment Limited
Location
London, South East England, United Kingdom
Employment Type
Full-Time
Salary
Salary negotiable
performance optimization, and security of Microsoft SQL Server environments. The role requires experience supporting production databases, troubleshooting complex issues, implementing database changes, and ensuring high availability and data integrity.The ideal candidate combines strong technical expertise with a proactive approach to problem-solving and effective collaboration with application, infrastructure … support Microsoft SQL Server environments. Perform database installations, configuration, upgrades, patching, and migrations. Manage databases, instances, security settings, users, roles, and permissions. Ensure database availability, reliability, and operational stability. Performance & Optimization Monitor and optimize database performance. Investigate and resolve performance bottlenecks, locking issues, and query-related problems. Analyze execution ...

Lead Software Engineer - Java, Payments

Location
Bournemouth, England, United Kingdom
scalable, and resilient way. As a core technical contributor, you are responsible for designing and delivering critical technology solutions across multiple technical areas, supporting high-volume payment initiation platforms and complex distributed systems that are essential to the firm’s business objectives. This role requires deep engineering expertise … building and operating production systems at scale, with strong emphasis on distributed systems, low-latency processing, high availability, resiliency, recoverability, observability, and operational excellence . Job Responsibilities: Executes creative software solutions, design, development, and technical troubleshooting with the ability to think beyond routine or conventional approaches to build ...

Lead DevOps Engineer

Location
Greater London, England, United Kingdom
Automate infrastructure provisioning, deployment, configuration, database schema creation, and operational workflows across cloud and on-prem platforms.* Support production systems with a focus on high availability, disaster recovery, incident response, vulnerability management, patching, and change management.* Containerize applications and support deployments using container orchestration platforms.* Partner with development … management with Terraform, Ansible, Chef, or similar.* **Infrastructure & Operations** + Strong understanding of networking, routing, and traffic flow across distributed systems. + Knowledge of high availability, resiliency, capacity, and latency considerations. + Experience supporting customer-facing, low-latency, high-availability, or other business-critical platforms. + ...

Database Engineer (MongoDB, Postgres)

Hiring Organisation
CGG
Location
Asthall Leigh, Oxfordshire, UK
Employment Type
Full-time
transition and infrastructure challenges. Job DetailsViridien is seeking a Database Engineer to design, maintain, and optimise database platforms that support global production systems and high-performance computing workloads. This role focuses on managing MongoDB and relational database environments, ensuring they remain secure, scalable, and highly available. You will work … Viridien's global software and HPC environments. The team works closely with software developers, infrastructure engineers, and operations teams to deliver reliable, scalable, and high-performing database services that support data-intensive applications across the business. Key Responsibilities-Database AdministrationManage, maintain, and support MongoDB and relational database environments. Install ...

Global Banking & Markets - Software Engineer - Vice President - London

Hiring Organisation
Goldman Sachs
Location
London, UK
Employment Type
Full-time
production idioms, while meeting the diverse and often complex business requirements of a global 247 trading operation, while balancing stringent non-functional demands for availability, latency, and resilience. Critically, you will lean heavily into AI-driven development. Using Goldman Sachs' AI tooling and agentic coding assistants, you will govern … magnitude higher volumes at lower operational costs, directly contributing to the firm's competitive and commercial edge. Build for the future: Design and implement high-availability, multi-region, event-driven services on a modern cloud-native platform, setting the architectural standard for years to come. What You Will ...

IAM Secrets Management Engineering - SRE Platform Engineer - VP - London London · United Kingdo[...]

Location
Greater London, England, United Kingdom
management firm that provides a wide range of services worldwide to a substantial and diversified client base that includes corporations, financial institutions, governments and high net‐worth individuals. The Role We are seeking a skilled and experienced Lead Site Reliability Platform Engineer (SRE) to join our team. The ideal … candidate will be responsible for ensuring the reliability, performance, and scalability of mission‐critical, highavailability, high‐throughput systems and infrastructure. This role involves leading collaboration with cross‐functional teams and implementing best practices in SRE, DevOps, and cyber security to enhance our operational efficiency and security ...

OT Technical Support Manager

Hiring Organisation
Experis
Location
Perth, Perthshire, Scotland, United Kingdom
Employment Type
Contract
operational support, maintenance, and resilience of critical Operational Technology (OT) systems within a Utilities environment. This is a leadership role focused on ensuring the availability, performance, security, and continuous improvement of high-availability OT platforms. The successful candidate will be responsible for incident management, fault recovery, service … maintenance of critical OT platforms and services. Manage incident response activities, fault diagnosis, service restoration, and root cause analysis. Ensure operational resilience across high-availability environments, supporting business continuity and recovery activities. Oversee OT patching, maintenance schedules, and vulnerability remediation programmes. Drive service delivery excellence and continuous improvement ...

CICS Mainframe Systems Programmer

Hiring Organisation
NTT
Location
London, UK
Employment Type
Full-time
growing team in the UK and rest of Europe. About the roleThe CICS Mainframe Systems Programmer is responsible for ensuring the stability, availability, performance, and continuous improvement of CICS services within the IBM Mainframe environment. The role requires strong technical expertise, domain knowledge, and specialist skills to operate, maintain … solution design, deployment planning, configuration, testing, and ongoing operational support. The successful candidate will oversee the management and support of all CICS services, ensure high availability of business-critical applications, provide technical leadership during incident resolution, and collaborate with development and application support teams to deliver robust solutions. ...

CICS Mainframe Systems Programmer

Location
Greater London, England, United Kingdom
growing team in the UK and rest of Europe. About the role The CICS Mainframe Systems Programmer is responsible for ensuring the stability, availability, performance, and continuous improvement of CICS services within the IBM Mainframe environment. The role requires strong technical expertise, domain knowledge, and specialist skills to operate … solution design, deployment planning, configuration, testing, and ongoing operational support. The successful candidate will oversee the management and support of all CICS services, ensure high availability of business-critical applications, provide technical leadership during incident resolution, and collaborate with development and application support teams to deliver robust solutions. ...

Infrastructure Engineer-Hyper-V

Hiring Organisation
NEEV LIMITED
Location
London, United Kingdom
Employment Type
Contract
Contract Rate
From £400 to £450 per day 400 - 450 GBP/day InsideIR35
Required Technical Skills Site Reliability Engineering Strong understanding of Site Reliability Engineering principles and operational excellence. Experience with infrastructure reliability, service availability, resiliency, and performance optimization. Storage Space Direct a nd failover clustering technical expertise . (Storage Spaces Direct enables you to build highly available, software-defined storage … external SAN, S2D uses the servers' local storage to create a resilient shared storage pool) Experience managing production-critical infrastructure environments with high availability requirements. Experience with incident management, problem management, RCA, and continuous operational improvement. Knowledge of monitoring, observability, alerting, and performance management. Microsoft Hyper-V (Core ...

IT Infrastructure Engineer

Hiring Organisation
Paragon Banking Group
Location
Solihull, West Midlands, UK
Employment Type
Full-time
which will allow them to effectively triage tickets and queries. What you'll be doing Monitor, manage, and optimise enterprise network environments to ensure high availability, performance, scalability, and security across LAN, WAN, Network Security, Cloud and Hybrid infrastructures. Design and deliver complex end-to-end network solutions … preventative measures to minimise recurrence. Provide technical guidance, and mentorship to peers, promoting best practice, upskilling, and knowledge sharing to contribute to supporting a high-performing and resilient team. About YouWhat you'll bring to the team Expert-level knowledge of network and infrastructure technologies, including enterprise routing ...

Network and Voice Engineer

Hiring Organisation
Source Group International
Location
London, UK
Employment Type
Full-time
Network and Voice Engineer to design, support and enhance resilient enterprise networks and unified communications within a regulated financial services environment. You will ensure high availability, performance, security and service quality across LAN/WAN, data centre connectivity, cloud networking and voice platforms, working closely with infrastructure, security … service management teams. Key responsibilities Design, implement and maintain routing and switching, including VLANs, STP, QoS, OSPF/BGP and high-availability patterns. Support WAN and remote connectivity (MPLS/SD-WAN, VPN, IPSec/SSL), optimising performance and resilience. Administer network security controls such as firewalls ...

z/OS Mainframe Systems Programmer

Hiring Organisation
NTT
Location
London, UK
Employment Type
Full-time
environments. Plan and execute operating system upgrades and software maintenance activities using SMP/E.Administer and support large-scale Parallel Sysplex environments to ensure high availability and performance. Perform system performance monitoring, capacity planning, and resource optimization. Analyse and resolve complex system-level issues involving z/… security, and operations teams to ensure platform stability. Implement software maintenance, fixes, service packs, and vendor-recommended updates. Support disaster recovery, business continuity, and high-availability initiatives. Maintain system security controls and ensure compliance with organizational standards. Create and maintain technical documentation, procedures, and operational runbooks. Participate ...

z/OS Mainframe Systems Programmer

Location
Greater London, England, United Kingdom
Plan and execute operating system upgrades and software maintenance activities using SMP/E. Administer and support large-scale Parallel Sysplex environments to ensure high availability and performance. Perform system performance monitoring, capacity planning, and resource optimization. Analyse and resolve complex system-level issues involving z/… security, and operations teams to ensure platform stability. Implement software maintenance, fixes, service packs, and vendor-recommended updates. Support disaster recovery, business continuity, and high-availability initiatives. Maintain system security controls and ensure compliance with organizational standards. Create and maintain technical documentation, procedures, and operational runbooks. Participate ...

Software Engineer

Location
Milton Keynes, England, United Kingdom
build distributed systems on AWS using serverless-first principles Liaise with 3rd Party Engineering teams on all aspects of the engineering lifecycle Architect for high availability, fault tolerance, and resilience by default Validate 3rd Party deliverables, workshop solutions, troubleshoot issues Design scalable data approaches and reporting structures aligned … secure, PCI-aware designs aligned with payments industry expectations Data & Infrastructure Design and maintain data pipelines and reporting structures using AWS native tooling Apply high availability and fault tolerance patterns across all infrastructure Ensure codebases and infrastructure are maintainable and transferable across engineers Technology Environment AWS (primary platform ...

Network & Security Consultant

Hiring Organisation
Solutions Through Knowledge
Location
London, United Kingdom
Employment Type
Full-Time
Salary
£400.00 - £450.00 per day
with particular emphasis on Cisco Secure Firewall and advanced BGP. Key Responsibilities Design and deployment of a new Cisco Secure Firewall Management Center (FMC) High Availability pair Deployment and configuration of HA pair Cisco Secure Firewalls Migration of existing legacy firewall configurations to the new environment Review … Strong hands-on experience with Cisco Secure Firewall/FTD Strong experience with Cisco Firewall Management Center (FMC) Experience deploying FMC and firewalls in High Availability configurations Proven experience with firewall migrations and configuration optimisation Advanced BGP knowledge, including configuration and troubleshooting Strong understanding of routing protocols ...

Network & Security Consultant

Hiring Organisation
STK Recruitment
Location
City of London, London, United Kingdom
Employment Type
Contract
Contract Rate
£450 per day - Outside IR35 / Ltd Company
with particular emphasis on Cisco Secure Firewall and advanced BGP. Key Responsibilities Design and deployment of a new Cisco Secure Firewall Management Center (FMC) High Availability pair Deployment and configuration of HA pair Cisco Secure Firewalls Migration of existing legacy firewall configurations to the new environment Review … Strong hands-on experience with Cisco Secure Firewall/FTD Strong experience with Cisco Firewall Management Center (FMC) Experience deploying FMC and firewalls in High Availability configurations Proven experience with firewall migrations and configuration optimisation Advanced BGP knowledge, including configuration and troubleshooting Strong understanding of routing protocols ...

Lead Software Engineer - Java, Payments

Hiring Organisation
JP Morgan Chase
Location
Bournemouth, Dorset, UK
Employment Type
Full-time
scalable, and resilient way. As a core technical contributor, you are responsible for designing and delivering critical technology solutions across multiple technical areas, supporting high-volume payment initiation platforms and complex distributed systems that are essential to the firm's business objectives. This role requires deep engineering expertise … building and operating production systems at scale, with strong emphasis on distributed systems, low-latency processing, high availability, resiliency, recoverability, observability, and operational excellence. Job Responsibilities: Executes creative software solutions, design, development, and technical troubleshooting with the ability to think beyond routine or conventional approaches to build solutions ...

DevOps Engineer- Night Shift(10:00 PM - 6:00 AM)

Location
United Kingdom
production environments healthy and performant, while simultaneously designing and maintaining the CI/CD pipelines, infrastructure-as-code frameworks, and tooling that enable rapid, high-quality software delivery. You are the connective tissue between engineering, platform, and operations — someone who is equally comfortable in an incident bridge call … production downtime, performance degradation, and security-related incidents in a timely, structured manner. Perform end-to-end operational duties covering application server health, service availability, and platform integrity in accordance with documented processes and runbooks. Review and manage client service request tickets in adherence to defined SLAs, ensuring accountability ...

Service Delivery Manager

Hiring Organisation
Hays
Location
Southampton, UK
Employment Type
Full-time
cloud, and hybrid environments while translating business needs into clear technical roadmaps and delivery plans. The role also carries accountability for critical systems, ensuring high availability, performance, and recoverability in a demanding operational setting. In addition, you will lead and develop a team of senior engineers, providing direction … ITIL.A proactive approach to continuous improvement, resilience planning, and service delivery will be key, along with the ability to operate effectively in a high-availability, high-pressure environment. What you'll get in returnIn return, you will join a forward-thinking organisation where technology underpins critical public ...

Senior Backend Engineer (Ruby), AI Engineering: AI Coding

Hiring Organisation
GitLab
Location
United Kingdom, UK
Employment Type
Full-time
into their daily workflows to drive efficiency, innovation, and impact. GitLab is where careers accelerate, innovation flourishes, and every voice is valued. Our high-performance culture is driven by our values and continuous knowledge exchange, enabling our team members to reach their full potential while collaborating with industry leaders … broader event platform, the core infrastructure that many of GitLab's AI features rely on to know when and how to run. Architect for high availability and throughput, since this part of the system sits at the center of automated AI workflows across GitLab and needs to reliably ...

Systems Integration Advisor

Hiring Organisation
NTT DATA
Location
London, UK
Employment Type
Full-time
environments, from development to mission-critical production systems. Configure and maintain database servers and processes, including monitoring of system health and performance, to ensure high levels of performance, availability, and security. Database performance tuning. Planning, deploying, maintaining, troubleshooting of high-availability database environments (RAC. Dataguard etc. … monitoring and troubleshooting tools Experience with backups, restores and recovery models Experience with database administration Experience with incident and problem queue management. Knowledge of High Availability (HA) and Disaster Recovery (DR) options for Oracle Experience working with Linux, Windows & Unix servers Excellent written and verbal communication, problem solving ...

Systems Integration Advisor

Hiring Organisation
NTT DATA
Location
Dunstable, Bedfordshire, UK
Employment Type
Full-time
environments, from development to mission-critical production systems. Configure and maintain database servers and processes, including monitoring of system health and performance, to ensure high levels of performance, availability, and security. Database performance tuning. Planning, deploying, maintaining, troubleshooting of high-availability database environments (RAC. Dataguard etc. … monitoring and troubleshooting tools Experience with backups, restores and recovery models Experience with database administration Experience with incident and problem queue management. Knowledge of High Availability (HA) and Disaster Recovery (DR) options for Oracle Experience working with Linux, Windows & Unix servers Excellent written and verbal communication, problem solving ...

Senior Fullstack Engineer (Python + React.js)

Location
Greater London, England, United Kingdom
maintain scalable, secure, and efficient server-side applications. The ideal candidate should have experience in microservices architecture, API development, database management, and frontend ensuring high availability and performance of backend and frontend services. In this role, you will collaborate closely with frontend engineers, product managers, and other stakeholders … Redux or React Query Collaborate with frontend developers to ensure efficient API integration and a seamless user experience. Troubleshoot and resolve production issues, ensuring high availability and minimal downtime. Write unit and integration tests to maintain code reliability and ensure high- quality releases. Continuously monitor and optimize ...