351 to 375 of 898 High Availability Jobs

Senior Network Engineer

Location
Greater London, England, United Kingdom
across the US, EU, Singapore and Australia. Your work directly underpins the platform experience for millions of users running live cyber simulations — low-latency, high-availability networking isn't a nice-to-have, it's the product. Design and operate production-grade, multi-site network architectures using … multi‐region, multi‐continent data center environment Familiarity with Ceph or other distributed storage systems from a network‐performance perspective (NVMEoF, RDMA, high‐throughput east‐west traffic) Exposure to Kubernetes networking (CNI plugins, service meshes, load balancers) JNCIA/JNCIS/JNCIP or equivalent networking certifications Understanding of Golang ...

Senior DevOps/SRE: Cloud, CI/CD & Automation Lead

Location
Greater London, England, United Kingdom
Adaptive is seeking a Senior/SRE engineer to join our London-based team. You will guide CI/CD, maintain infrastructure, and ensure high availability of customer-critical systems in a hybrid work setup. Ideal candidates have strong AWS, Linux, Terraform and Docker experience, plus Python/ ...

Senior DevOps Engineer - Remote, AWS, IaC & Automation

Location
Dundee, Scotland, United Kingdom
role with an office in Scotland. You will own AWS infrastructure, automate with Terraform/Ansible, and drive CI/CD improvements while ensuring high availability and security of production systems. The role expects 5+ years in similar spheres, strong Linux and AWS experience, and a proactive approach ...

Senior Cloud Services Engineer

Location
Greater London, England, United Kingdom
ascending levels: Awareness, Working, Proficient, and Expert. Skills needed for this role level: Architecture Design & Strategy: (Level: Practitioner): Develop and implement robust, scalable, and high-availability architectural designs for AWS and Nutanix environments, focusing on services like Route 53, EC2, and integrating infrastructure automation through Terraform and Ansible. ...

Senior Infrastructure Engineer

Location
United Kingdom
complex enterprise environment? We're recruiting a Senior Infrastructure Engineer to support and enhance critical IT infrastructure across multiple UK & Ireland locations, ensuring high availability, resilience, security, and performance. Working closely with business stakeholders, global IT teams, and external partners, you'll play a key role in delivering ...

Solutions Architect

Location
Leeds, England, United Kingdom
cloud-native applications using Java, Spring Boot, and Microservices. Define architectural best practices, standards, and frameworks for GCP-based solutions. Ensure security, scalability, and high availability of applications. Design microservices architecture with RESTful APIs and event-driven patterns (Kafka/Pub-Sub). Good experience in Authentication ...

Senior Infrastructure Engineer (DV Cleared)

Location
Newbury, England, United Kingdom
Skills & Experience Extensive experience designing and deploying Microsoft Hyper-V clusters in enterprise environments. Strong expertise in Windows Server , Failover Clustering , Storage Replica , and High Availability solutions. Proven experience migrating physical and virtual server estates from VMware and legacy platforms. Deep understanding of Microsoft PKI , Certificate Services, certificate ...

Microsoft Azure, Messaging & Collaboration Lead

Location
City Of London, England, United Kingdom
support of Microsoft 365 services including Exchange Online, Teams, OneDrive, and SharePoint Online. Design, implement, and maintain messaging and collaboration solutions to ensure high availability and performance. Manage mail flow, email security, retention, archiving, and compliance policies. Endpoint & Mobility Management Manage and support Microsoft Intune and Mobile Device ...

Principal Engineer

Location
London, United Kingdom
engineering best practices Experience working with cloud platforms such as Azure or AWS Experience developing and deploying microservices-based applications Strong understanding of high-availability, scalability, and performance optimization concepts Experience working within Agile software development environments Nice to have Experience with Swift and the Vapor framework ...

Trade Floor Operations - FICC

Location
City Of London, England, United Kingdom
Responsibilities: Provide end-to-end production support for Fixed Income trading platforms (Rates, Credit, FX, Bonds, Derivatives). Monitor and maintain system health, ensuring high availability and low latency of critical trading applications. Support and enhance monitoring capabilities using ITRS Geneos, initially within FICC and expanding to other ...

Principal Engineer

Hiring Organisation
Hackajob Ltd
Location
South West London, London, United Kingdom
Employment Type
Permanent
engineering best practices Experience working with cloud platforms such as Azure or AWS Experience developing and deploying microservices-based applications Strong understanding of high-availability, scalability, and performance optimization concepts Experience working within Agile software development environments Nice to have Experience with Swift and the Vapor framework ...

Director of Site Reliability Engineering

Location
Greater London, England, United Kingdom
Terraform or CloudFormation), and CI/CD practices Deep understanding of incident management processes, ITSM standards, and ITIL principles Knowledge of resilience design patterns, high availability, and fault‐tolerant architectures Familiarity with AI/ML‐driven approaches for operational efficiency and system reliability Ability to lead transformation, influence ...

cloud engineer in fintech

Location
Greater London, England, United Kingdom
architects; Document significant technical decisions as ADRs; Act as the Subject Matter Expert for team-owned services; Provide advanced operational support and ensure high availability for global payment systems; Produce exemplary code and documentation; Spread engineering best practices and standards across the group; Manage technical dependencies inside ...

Engineering Manager - Payments Integrations

Location
Greater London, England, United Kingdom
documentation. Ensure the team follows secure software development practices, including OWASP standards, data privacy, and compliance requirements. Guide architectural decisions for low‐latency, high‐availability backend systems and microservices. Champion an AI‐first engineering culture, setting standards for AI‐assisted development, code generation, and automated testing, ensuring your ...

Python Developer

Hiring Organisation
Networking People (UK) Limited
Location
Glasgow, Lanarkshire, Scotland, United Kingdom
Employment Type
Contract
Contract Rate
£500 - £550 per day + Inside of IR35
supports a broad range of infrastructure technologies, including hardware platforms, hypervisors, Linux operating systems, system configuration tooling, critical infrastructure services, automation platforms, load balancing, high availability solutions, and virtualization management technologies. The team plays a key role in enabling DevOps practices by delivering real time APIs and automation ...

DevOps Engineer (f/m/d)

Location
Greater London, England, United Kingdom
changes through version control and code review. Experience with CI/CD pipelines, containerized applications and production deployments. Familiarity with backup and disaster recovery, high-availability configurations and safe database migrations. The ability to independently deliver well-defined infrastructure tasks, explain technical trade-offs and contribute to architectural ...

Consulting Principal Lead (Presales)

Location
London, England, United Kingdom
opportunities preferred. Experience working with a cloud service provider (AWS, Azure, or GCP), preference for cloud data skills. Experience in large-scale, secure, and high-availability solutions with multi–AZ Cloud Architecture. Significant experience with cloud modernization and migration solutioning: Discovery, Assessment, Roadmap, SOW Creation, Migration Planning ...

Infrastructure & Operations Engineer

Location
United Kingdom
PaloAlto, Cisco ASA, Watchguard or similar)•Network segmentation and security (IT/OT best practices)**Backup & Disaster Recovery:**•Veeam Backup & Replication•Disaster recovery planning & high availability solutions•**Database & Application Support:**•Working knowledge of Microsoft SQL Server (2000-2022) is advantageous•Oracle Database (various versions)•PostgreSQL/MySQL (desirable ...

Software Engineering Tech Lead (SRE + AI)

Location
Greater London, England, United Kingdom
Computer Science, Software Engineering, or a related technical field. Proven record as a Technical Lead or Lead SRE/Software Engineer delivering distributed, high-availability SaaS platforms at scale. Strong proficiency in Python, Go, Java, or C++ with experience designing microservices, APIs, and production automation. Deep experience with ...

IT Infrastructure & Security Lead

Location
Bradford, England, United Kingdom
premise environments, including servers, storage, networking, and virtualisation. Oversee Microsoft 365, Azure/AWS environments, Active Directory/Entra ID, and endpoint management. Ensure high availability, disaster recovery, backup, and business continuity capabilities. Monitor infrastructure performance, capacity, and reliability. Manage third‐party IT vendors, MSPs, and technology partners. ...

Infrastructure & Platform Specialist Solutions Architect (SSA)

Location
Greater London, England, United Kingdom
such as AWS PrivateLink/Azure Private Link/GCP Private Service Connect), network routing, performance optimisation, and large-scale deployment management Platform Administration: High availability, disaster recovery, cluster orchestration, observability and audit (e.g. Amazon CloudWatch/CloudTrail, Azure Monitor, Google Cloud Operations Suite), and cloud cost management ...

Tanium Engineer

Location
United Kingdom
visibility, patching and software deployment issues. Supporting the integration of endpoint security platforms with wider IT and security tooling. Monitoring platform performance and maintaining high availability. Supporting audits and compliance with internal and external security standards. Developing automation and reporting using PowerShell and/or Python . Providing technical ...

Tanium Engineer

Hiring Organisation
Circle Group
Location
Exeter, Devon, South West, United Kingdom
Employment Type
Contract
visibility, patching and software deployment issues. Supporting the integration of endpoint security platforms with wider IT and security tooling. Monitoring platform performance and maintaining high availability. Supporting audits and compliance with internal and external security standards. Developing automation and reporting using PowerShell and/or Python . Providing technical ...

Infrastructure Engineer

Location
Manchester, England, United Kingdom
user computing estate. Regular travel to customer sites across the UK is expected. Key Responsibilities Own the day-to-day management, monitoring, performance tuning, availability and security of the Omnissa Horizon VDI platform (Connection Servers, Unified Access Gateway, Instant Clones, desktop pools, App Volumes, Dynamic Environment Manager and related … timely patch management across Horizon components, thin clients and laptops, monitor for vulnerabilities, and drive rapid remediation. Perform capacity planning, proactive performance optimisation and high-availability improvements for the Horizon platform. Provide Tier 2/3 support and root-cause analysis for complex VDI, thin client and managed ...

MLOps Engineer

Hiring Organisation
DGH Recruitment
Location
City of London, London, United Kingdom
Employment Type
Permanent
MLOps Engineer (Platform/DevOps) - AI Platform - 3 days per week onsite We're partnering with a leading organisation on a high-profile AI initiative, seeking an experienced MLOps/Platform Engineer to build and scale infrastructure supporting advanced AI solutions. The Role You'll play a key role … platform for deploying AI workloads and agents. Working closely with Data Science and Engineering teams, you'll ensure robust infrastructure, seamless deployment pipelines, and high platform reliability. Key Responsibilities - Design, deploy, and manage AI platforms and agent infrastructure - Build and maintain CI/CD pipelines and DevOps workflows - Implement ...