251 to 275 of 860 Permanent High Availability Jobs

Senior Software Engineer, Machine Learning Services

Location
Greater London, England, United Kingdom
senior engineers building the core platform that powers UiPath's large‐scale AI and Document Understanding products. Our world is one of distributed systems, high‐throughput model serving, and complex asynchronous training workflows. We're looking for a systems‐level engineer who wants to work on the gnarly, foundational … distributed job queue that orchestrates it all. Solve hard concurrency, performance, and distributed systems problems to ensure our platform is bulletproof for high‐volume production workloads. Work directly with product and ML science teams to understand their needs and build the scalable infrastructure required to bring their models ...

Python Technical Lead FinTech

Hiring Organisation
Run-Time Group Ltd
Location
City of London, London, United Kingdom
Employment Type
Permanent
Technical Lead Salary: £90-120K work model: Hybrid Were looking for a Python Technical Lead to drive the architecture, development, and delivery of high-performance financial systems. Youll lead a team of engineers building scalable platforms that power real-time transactions, risk analytics, and next-generation digital financial … build robust, secure, and scalable microservices using Python (FastAPI, Django, Flask). System Architecture Lead the design of distributed systems, event-driven architectures, and high-availability platforms. Team Mentorship Coach engineers, conduct code reviews, and foster a culture of continuous improvement. FinTech Platform Development Build systems for payments ...

Lead Site Reliability / DevOps Engineer

Location
Auchentibber, Scotland, United Kingdom
with team members to identify comprehensive service level indicators and stakeholders to establish reasonable service level objectives and error budgets with customers Demonstrates a high level of technical expertise within one or more technical domains and proactively identifies and solves technology-related bottlenecks in your areas of expertise Acts … logging best practices. Certification in AWS, Kubernetes, or relevant technologies. Proven track record in system health monitoring, capacity management, and blameless postmortems for high-availability services Deep understanding of distributed system design principles, networking (TCP/IP, DNS, load balancing), Linux internals. Contributions to open-source observability ...

Infrastructure Analyst

Location
Liverpool, England, United Kingdom
resolving complex technical issues. Contributing to cloud, infrastructure and application projects. Working closely with service delivery, project and architecture teams. Supporting disaster recovery, high availability and business continuity initiatives. Driving automation and continuous improvement across the technology estate. Skills & Experience Strong VMware administration experience. Microsoft Azure and Entra ...

Senior .NET developer (with Snowflake)

Location
Greater London, England, United Kingdom
Enablemento Collaborate with data engineers to optimize deployment patterns for Snowflake workloads Monitoring & Reliabilityo Implement monitoring, logging, and alerting across the data platformo Ensure high availability and performance of pipelines and integrationso Troubleshoot production issues and drive root cause analysis Security & Complianceo Embed security best practices across infrastructure ...

Cloud Infrastructure Engineer

Location
Lichfield, England, United Kingdom
engineering, platform design, and continuous improvement , with opportunities to shape our cloud strategy. Key Responsibilities Design, deploy, and manage Azure services (IaaS & PaaS) Ensure high availability, scalability, and resilience across platforms Implement governance, security controls, and cost optimisation strategies Build and maintain IaC solutions (Terraform, Bicep, or similar ...

Senior CMMC Network Engineer

Hiring Organisation
Vailexa
Location
United States
Employment Type
Permanent
Salary
USD Annual
C3PAO assessment activities by providing network evidence, diagrams, and control narratives. Work with service providers for circuit turn-ups, troubleshooting, and escalations. Ensure high availability, redundancy, and disaster recovery readiness. Mentor junior engineers and provide technical guidance. Participate in on-call rotation and after-hours maintenance. Cloud Networking ...

Cloud Security Engineer

Hiring Organisation
Pinnacle Technical Resources
Location
Mc Lean, Virginia, United States
Employment Type
Permanent
Salary
USD 70 Hourly
Protection platform. 2+ years of experience with Linux file system concepts (Fuse, GPFS, NFS). 2+ years of experience with Linux storage subsystems, clustering, high availability, fault tolerance, and replication concepts. 3+ years of experience with DevOps principles and Agile development practices. About PTR Global: PTR Global ...

Deployed Architect, Professional Services (London)

Location
Greater London, England, United Kingdom
Terraform, Helm) and GitOps practices Knowledge of database systems (relational databases, in-memory data stores) including HA, replication, backup strategies, and sizing Experience designing high-availability and disaster recovery solutions Strong understanding of networking, security (SSO/RBAC, TLS, secrets management), and observability (Prometheus, Grafana, Datadog) Experience with ...

Senior Forward Deployment Engineer

Location
Slough, England, United Kingdom
OpenShift, showing understanding of container orchestration platforms. Experience designing hot‐hot, active‐active, multi‐region, or cross‐datacenter deployment architectures, denoting knowledge of high‐availability systems. Experience with resilience testing, disaster recovery, controlled failover, and production‐recovery exercises, highlighting skills in system reliability and continuity planning. Experience with ...

Platform Engineer

Location
Greater London, England, United Kingdom
hosted platform Competent in Git and the GitOps philosophy Familiarity with concepts for managing very large application load (e.g. CDNs, load balancers) and providing high availability services Keen interest/experience in leveraging AI and agentic approaches in an Ops or Developer Platform context Experience of working ...

DevSecOps

Location
Greater London, England, United Kingdom
tangible impact on our world. Whether you join our London HQ or the wider global organisation, you’ll be a part of collaborative, high-performing teams, creating cutting-edge software, platforms, and infrastructure. The Role Join us as a DevSecOps and help us build the future of data sovereignty … seeking an DevSecOps passionate about creating high-performance, secure, scalable, and reliable services for our production infrastructure and security. You'll have a direct impact, improving existing systems and developing innovative solutions to complex challenges. Our small, collaborative engineering teams own the full lifecycle of their services, from development ...

Senior ML Infrastructure Engineer Enterprise Operations Oxford, England, United Kingdom

Location
Oxford, England, United Kingdom
service guardrails that accelerate experimentation and turn ideas into results - faster, at scale, and with confidence. Your Responsibilities: Build,operate, and continuously optimise our high-performance GPU training and inference clusters, focusing on robust, high-availability scheduling, isolation, and automated lifecycle management. Drive systems design and implementation … high-throughput data paths, optimising I/O, caching, and data locality across compute and storage (including our current Lustre implementation). Proactively benchmark, profile, and resolve performance bottlenecks across the compute, network, and orchestration layers to maximise efficiency for distributed training and inference. Establish comprehensive observability, resilience ...

Lead AI Engineer

Location
Newcastle upon Tyne, England, United Kingdom
about user-centred design, agile delivery, and building digital services that make a real difference — and we're now scaling that expertise into large, high-stakes AI adoption programmes across the public sector. Overview We're looking for a Lead AI Engineer to be the hands‐on technical builder … into working, production‐grade systems — semantic search, RAG pipelines, and broader generative AI capability — integrated into complex legacy and multi-cloud environments handling high-volume, sensitive public sector data. You'll work inside a collaborative "Rainbow Team" alongside civil servants and the wider delivery team, staying close ...

DevOps Engineer

Hiring Organisation
Anson Mccade
Location
Manchester, North West, United Kingdom
Employment Type
Permanent
Salary
£80,000
experienced DevOps Engineer to join an Agile engineering team working on a large-scale programme, building a bespoke software solution for a high-profile customer. What youll be doing Designing, implementing and maintaining AWS cloud infrastructure Deploying and managing Kubernetes clusters using technologies such as Docker and Helm Building … automation for deployment, monitoring, testing and infrastructure management Optimising infrastructure for high availability, resilience and performance Troubleshooting and resolving infrastructure and application issues Supporting security, compliance and disaster recovery Mentoring junior DevOps engineers What were looking for Commercial experience in a DevOps/Cloud Engineering role Strong ...

DevOps Engineer

Hiring Organisation
Anson Mccade
Location
South West London, London, United Kingdom
Employment Type
Permanent
Salary
£80,000
experienced DevOps Engineer to join an Agile engineering team working on a large-scale programme, building a bespoke software solution for a high-profile customer. What youll be doing Designing, implementing and maintaining AWS cloud infrastructure Deploying and managing Kubernetes clusters using technologies such as Docker and Helm Building … automation for deployment, monitoring, testing and infrastructure management Optimising infrastructure for high availability, resilience and performance Troubleshooting and resolving infrastructure and application issues Supporting security, compliance and disaster recovery Mentoring junior DevOps engineers What were looking for Commercial experience in a DevOps/Cloud Engineering role Strong ...

DevOps Engineer

Hiring Organisation
Anson Mccade
Location
Bristol, Avon, South West, United Kingdom
Employment Type
Permanent
Salary
£60,000
experienced DevOps Engineer to join an Agile engineering team working on a large-scale programme, building a bespoke software solution for a high-profile customer. What youll be doing Designing, implementing and maintaining AWS cloud infrastructure Deploying and managing Kubernetes clusters using technologies such as Docker and Helm Building … automation for deployment, monitoring, testing and infrastructure management Optimising infrastructure for high availability, resilience and performance Troubleshooting and resolving infrastructure and application issues Supporting security, compliance and disaster recovery Mentoring junior DevOps engineers What were looking for Commercial experience in a DevOps/Cloud Engineering role Strong ...

Lead Ai Engineer (Contract)

Location
Newcastle upon Tyne, England, United Kingdom
about user-centred design, agile delivery, and building digital services that make a real difference - and we're now scaling that expertise into large, high-stakes AI adoption programmes across the public sector. Overview We're looking for a Lead AI Engineer to be the hands-on technical builder … into working, production-grade systems - semantic search, RAG pipelines, and broader generative AI capability - integrated into complex Legacy and multi-cloud environments handling high-volume, sensitive public sector data. You'll work inside a collaborative "Rainbow Team" alongside civil servants and the wider delivery team, staying close ...

DevOps Engineer

Location
Bristol, Gloucestershire, United Kingdom
experienced DevOps Engineer to join an Agile engineering team working on a large-scale programme, building a bespoke software solution for a high-profile customer. What you'll be doing Designing, implementing and maintaining AWS cloud infrastructure Deploying and managing Kubernetes clusters using technologies such as Docker and Helm … Building automation for deployment, monitoring, testing and infrastructure management Optimising infrastructure for high availability, resilience and performance Troubleshooting and resolving infrastructure and application issues Supporting security, compliance and disaster recovery Mentoring junior DevOps engineers What we're looking for Commercial experience in a DevOps/Cloud Engineering role ...

Sr. Software Engineer

Location
Cambridge, England, United Kingdom
complex projects end-to-end, from architecture through deployment, working from loose briefs Contribute to larger team-scoped projects through some combination of producing high-level designs, delivering high-quality code, large/complex test strategies, or debugging challenging field bugs in unfamiliar code Adopt AI-assisted tooling …/Qualifications 5+ years of experience as a software engineer, covering the full software development lifecycle, in telecoms or a similarly complex domain with high availability requirements Experience using frontier LLMs to improve outcomes, across the SDLC Advanced programming concepts such as low-level resource optimizations and high ...

Senior Network Security Specialist

Location
Greater London, England, United Kingdom
conducted on the LME’s three trading platforms totalling $21 trillion, 191 million lots and 4 billion tonnes notional with a market open interest high of 2.1 million lots in 2025.The metals community uses the LME, an HKEX Group company, as a venue to transfer or take on price … Overall Purpose of Role**As a Senior Network Security Specialist, you will design, implement and govern the network security controls that protect our modern, high‐performance enterprise network. You will take a hands‐on lead role in shaping the network security roadmap, defining policies and standards and driving ...

Senior Server, Storage & Network Engineer

Hiring Organisation
Jobot
Location
Greenville, South Carolina, United States
Employment Type
Permanent
Salary
USD 130,000 Annual
routers, firewalls, wireless networks, and VPN solutions. Serve as a senior escalation resource for complex customer infrastructure issues. Design solutions for multi-campus environments, high-density wireless access, and large student/staff device populations. Contribute technical documentation for education procurement and E-Rate projects, as applicable. What …/IP, DNS, DHCP, VLANs, routing, switching, wireless networking, firewalls, and VPNs. Experience with SAN/NAS storage, backup and recovery systems, and high-availability configurations. Willingness to travel throughout the Southeastern U.S. (75%), including occasional overnight travel. Availability for scheduled evening or weekend work tied ...

IT Support Manager

Location
Greater London, England, United Kingdom
hands‐on technical oversight. You will own day‐to‐day IT support operations, oversee maintenance and incident resolution for GPU clusters, and ensure high service levels across hardware, network, and field support activities. You will act as a key escalation point, drive operational improvements, and align support execution with … Lead and manage L1, L2, IT infrastructure and Field Network Engineers supporting data center IT infrastructure and GPU clusters and region IX colocations. Ensure high availability and reliability of GPU clusters and IT infrastructure Act as the primary escalation point for complex hardware, network, and operational incidents Monitor ...

Staff Network Engineer

Location
Greater London, England, United Kingdom
Nscale is the GPU cloud engineered for AI. We provide cost-effective, high-performance infrastructure for AI start-ups and large enterprise customers. Nscale enables AI-focused companies to achieve superior results by reducing the complexity of AI development. Our GPU cloud strengthens technical capabilities and directly supports strategic … design, validation, and ongoing operation of all networking services that underpin both the internal management platform and the customer-facing cloud infrastructure — including high-performance Ethernet fabrics, InfiniBand, WAN connectivity, and data center networking. The team also acts as a 3rd/4th line escalation point for the support ...

Senior Systems Engineer (Server and Cloud Technologies Specialist)

Hiring Organisation
Trusted Technology Partnership
Location
Ringwood, Hampshire, South East, United Kingdom
Employment Type
Permanent, Work From Home
Salary
£55,000
Intune. Strong experience in managing Windows Server environments (2016, 2019, 2022, 2025). Proficiency in automation and scripting using PowerShell. Experience with virtualisation and high availability solutions, including Hyper-V (clustered and standalone), Windows Admin Centre, and Failover Clustering. Working knowledge of SQL Server in both clustered … troubleshooting and problem-solving skills, with the ability to diagnose and resolve complex technical issues. Excellent communication skills, including the ability to produce clear, high-quality documentation such as processes, procedures, and technical designs. Positive, proactive attitude with a strong sense of ownership. Desirable: Experience designing and deploying highly ...