176 to 200 of 486 High Availability Jobs in London

Sr Lead Software Engineer - KDB+ / Q

Location
Greater London, England, United Kingdom
design, development, and technical troubleshooting with ability to think beyond routine or conventional approaches to build solutions or break down technical problems Develops secure high-quality production code, and reviews and debugs code written by others Identifies opportunities to eliminate or automate remediation of recurring issues to improve overall … experience developing/running large datasets and optimizing query performance. Practical experience scaling and load-balancing of KDB applications. Practical experience building resilient and high-availability KDB applications. Preferred qualifications, capabilities, and skills Experience with market data venue and vendor data platforms. AWS Experience. Experience in Terraform ...

Principal Platform Security Engineer

Hiring Organisation
Hiscox AG
Location
London, UK
Employment Type
Full-time
years' DevOps/Platform Engineering experience delivering solutions in Azure and/or GCP.Full‐stack application and infrastructure solution design, ensuring robust security controls, high availability, and operational resilience throughout. Working knowledge of vulnerability and compliance management (scanning through to remediation), patch management, endpoint protection/anti-malware … delivery focus, able to prioritise effectively and deliver outcomes in a fast-paced environment with shifting demands. Able to operate effectively in a small, high-impact team while collaborating across a wider product/engineering organisation. Excellent communication and stakeholder-management skills, able to influence at all levels ...

Cloud Engineering Team Lead

Hiring Organisation
Capital On Tap
Location
London, UK
Employment Type
Full-time
culture of continuous improvement, providing meaningful mentorship and constructive feedback that helps engineers grow. Driving team performance by resolving bottlenecks, fostering accountability, and sustaining high engagement to ensure consistent delivery & resource efficiency. Driving the adoption of AI within the team, ranging from day to day operations to large scale … writing, managing, and optimising infrastructure with tools such as Terraform. Experience using CI/CD tools, building pipelines, templates and troubleshooting. Experience with applying high availability across cloud architectures, ideally including multi-cloud. Experience with containerisation technologies such as Kubernetes and Docker. A customer-focused, commercially aware ...

Principal Platform Security Engineer

Location
Greater London, England, United Kingdom
DevOps/Platform Engineering experience delivering solutions in Azure and/or GCP.* Full‐stack application and infrastructure solution design, ensuring robust security controls, high availability, and operational resilience throughout.* Working knowledge of vulnerability and compliance management (scanning through to remediation), patch management, endpoint protection/anti-malware … delivery focus, able to prioritise effectively and deliver outcomes in a fast-paced environment with shifting demands.* Able to operate effectively in a small, high-impact team while collaborating across a wider product/engineering organisation.* Excellent communication and stakeholder-management skills, able to influence at all levels ...

Senior Network Engineer

Location
Greater London, England, United Kingdom
across the US, EU, Singapore and Australia. Your work directly underpins the platform experience for millions of users running live cyber simulations — low-latency, high-availability networking isn't a nice-to-have, it's the product. Design and operate production-grade, multi-site network architectures using … multi‐region, multi‐continent data center environment Familiarity with Ceph or other distributed storage systems from a network‐performance perspective (NVMEoF, RDMA, high‐throughput east‐west traffic) Exposure to Kubernetes networking (CNI plugins, service meshes, load balancers) JNCIA/JNCIS/JNCIP or equivalent networking certifications Understanding of Golang ...

Senior DevOps/SRE: Cloud, CI/CD & Automation Lead

Location
Greater London, England, United Kingdom
Adaptive is seeking a Senior/SRE engineer to join our London-based team. You will guide CI/CD, maintain infrastructure, and ensure high availability of customer-critical systems in a hybrid work setup. Ideal candidates have strong AWS, Linux, Terraform and Docker experience, plus Python/ ...

DevOps Platform Engineer

Hiring Organisation
Context Recruitment Limited
Location
London, South East England, United Kingdom
Employment Type
Full-Time
Salary
£90,000 - £100,000 per annum, Inc benefits
environments. Supporting and administering MySQL, PostgreSQL and MongoDB databases. Managing monitoring, alerting and logging platforms. Automating infrastructure through Infrastructure as Code and scripting. Maintaining high availability and resilience across critical systems. Managing access controls and permissions within cloud platforms. Troubleshooting, performance tuning and platform optimisation. Developing technical standards ...

Trade Flow Support - Trading Infrastructure Support

Hiring Organisation
SQUAREPOINT CAPITAL
Location
London, UK
Employment Type
Full-time
Position Overview: The Trading Infrastructure Support Engineer will be part of a highly talented team of support engineers responsible for ensuring high availability of Squarepoint's mission critical services such Infrastructure as a Service (IaaS), Build & Deploy, Research and Trading. The candidate will be responsible for providing ...

Microsoft Azure, Messaging & Collaboration Lead

Location
City Of London, England, United Kingdom
support of Microsoft 365 services including Exchange Online, Teams, OneDrive, and SharePoint Online. Design, implement, and maintain messaging and collaboration solutions to ensure high availability and performance. Manage mail flow, email security, retention, archiving, and compliance policies. Endpoint & Mobility Management Manage and support Microsoft Intune and Mobile Device ...

Citrix Server Remediation Engineer

Hiring Organisation
HCLTech
Location
Greater London, England, United Kingdom
Directory Patch Management Tools Preferred Skills Azure Virtual Desktop (AVD) VMware ESXi/Hyper-V Microsoft Azure Infrastructure Load Balancing Technologies Disaster Recovery and High Availability Soft Skills Strong troubleshooting abilities Change management experience Risk-based decision making Cross-functional collaboration ...

Trade Floor Operations - FICC

Location
City Of London, England, United Kingdom
Responsibilities: Provide end-to-end production support for Fixed Income trading platforms (Rates, Credit, FX, Bonds, Derivatives). Monitor and maintain system health, ensuring high availability and low latency of critical trading applications. Support and enhance monitoring capabilities using ITRS Geneos, initially within FICC and expanding to other ...

Principal Engineer

Location
London, United Kingdom
engineering best practices Experience working with cloud platforms such as Azure or AWS Experience developing and deploying microservices-based applications Strong understanding of high-availability, scalability, and performance optimization concepts Experience working within Agile software development environments Nice to have Experience with Swift and the Vapor framework ...

Principal Engineer

Hiring Organisation
Hackajob Ltd
Location
South West London, London, United Kingdom
Employment Type
Permanent
engineering best practices Experience working with cloud platforms such as Azure or AWS Experience developing and deploying microservices-based applications Strong understanding of high-availability, scalability, and performance optimization concepts Experience working within Agile software development environments Nice to have Experience with Swift and the Vapor framework ...

Director of Site Reliability Engineering

Location
Greater London, England, United Kingdom
Terraform or CloudFormation), and CI/CD practices Deep understanding of incident management processes, ITSM standards, and ITIL principles Knowledge of resilience design patterns, high availability, and fault‐tolerant architectures Familiarity with AI/ML‐driven approaches for operational efficiency and system reliability Ability to lead transformation, influence ...

cloud engineer in fintech

Location
Greater London, England, United Kingdom
architects; Document significant technical decisions as ADRs; Act as the Subject Matter Expert for team-owned services; Provide advanced operational support and ensure high availability for global payment systems; Produce exemplary code and documentation; Spread engineering best practices and standards across the group; Manage technical dependencies inside ...

Engineering Manager - Payments Integrations

Location
Greater London, England, United Kingdom
documentation. Ensure the team follows secure software development practices, including OWASP standards, data privacy, and compliance requirements. Guide architectural decisions for low‐latency, high‐availability backend systems and microservices. Champion an AI‐first engineering culture, setting standards for AI‐assisted development, code generation, and automated testing, ensuring your ...

DevOps Engineer (f/m/d)

Location
Greater London, England, United Kingdom
changes through version control and code review. Experience with CI/CD pipelines, containerized applications and production deployments. Familiarity with backup and disaster recovery, high-availability configurations and safe database migrations. The ability to independently deliver well-defined infrastructure tasks, explain technical trade-offs and contribute to architectural ...

Consulting Principal Lead (Presales)

Location
London, England, United Kingdom
opportunities preferred. Experience working with a cloud service provider (AWS, Azure, or GCP), preference for cloud data skills. Experience in large-scale, secure, and high-availability solutions with multi–AZ Cloud Architecture. Significant experience with cloud modernization and migration solutioning: Discovery, Assessment, Roadmap, SOW Creation, Migration Planning ...

Software Engineering Tech Lead (SRE + AI)

Location
Greater London, England, United Kingdom
Computer Science, Software Engineering, or a related technical field. Proven record as a Technical Lead or Lead SRE/Software Engineer delivering distributed, high-availability SaaS platforms at scale. Strong proficiency in Python, Go, Java, or C++ with experience designing microservices, APIs, and production automation. Deep experience with ...

Infrastructure & Platform Specialist Solutions Architect (SSA)

Hiring Organisation
DataBricks
Location
London, UK
Employment Type
Full-time
connectivity such as AWS PrivateLink/Azure Private Link/GCP Private Service Connect), network routing, performance optimisation, and large-scale deployment managementPlatform Administration: High availability, disaster recovery, cluster orchestration, observability and audit (e.g. Amazon CloudWatch/CloudTrail, Azure Monitor, Google Cloud Operations Suite), and cloud cost managementInfrastructure ...

Infrastructure & Platform Specialist Solutions Architect (SSA)

Location
Greater London, England, United Kingdom
such as AWS PrivateLink/Azure Private Link/GCP Private Service Connect), network routing, performance optimisation, and large-scale deployment management Platform Administration: High availability, disaster recovery, cluster orchestration, observability and audit (e.g. Amazon CloudWatch/CloudTrail, Azure Monitor, Google Cloud Operations Suite), and cloud cost management ...

MLOps Engineer

Hiring Organisation
DGH Recruitment
Location
City of London, London, United Kingdom
Employment Type
Permanent
MLOps Engineer (Platform/DevOps) - AI Platform - 3 days per week onsite We're partnering with a leading organisation on a high-profile AI initiative, seeking an experienced MLOps/Platform Engineer to build and scale infrastructure supporting advanced AI solutions. The Role You'll play a key role … platform for deploying AI workloads and agents. Working closely with Data Science and Engineering teams, you'll ensure robust infrastructure, seamless deployment pipelines, and high platform reliability. Key Responsibilities - Design, deploy, and manage AI platforms and agent infrastructure - Build and maintain CI/CD pipelines and DevOps workflows - Implement ...

Principal Architect

Location
Greater London, England, United Kingdom
strategy as we scale our global fintech platform. You will be the primary technical authority for our system design, responsible for ensuring that our high-throughput transaction engines and complex microservices ecosystem remain resilient, performant, and secure under massive load. You will partner with Engineering, Product, and Security leadership … Doing Own the Architectural Vision: Define and drive the long-term technical architecture for KAST, ensuring our platform evolves to meet the demands of high-throughput transactional volume. Design for Scale & Reliability: Lead the design and review of our distributed systems, ensuring our microservices architecture is built for fault ...

Lead Software Engineer

Location
Greater London, England, United Kingdom
ensure technical solutions align with business priorities. You'll also support the development of up to four engineers, helping them grow while maintaining high standards of engineering quality and delivery. "The best person for this role will be someone who loves writing software and solving problems but also enjoys … helping other engineers succeed. You'll lead through technical expertise, practical coaching and a commitment to delivering high‐quality solutions that make a real difference." Head of Software Engineering The team & culture Engineering at easyJet is organised around Missions and Squads, bringing technology teams closer to the parts ...

Platform Engineer @ Hnry

Location
Greater London, England, United Kingdom
needs and improve developer experience Design, build and maintain improvements to the infrastructure and platform of Hnry, currently hosted on AWS Improve the reliability, availability and performance of our platform Evolve and optimise our CI/CD pipelines to enable fast, safe and consistent deployments Implement and enhance observability … monitoring, alerting, logging) using tools like DataDog to ensure high-availability APM Apply best practices in security, scalability and cost optimisation across our infrastructure Support incident response and proactively identify and resolve system issues Produce work that meets the expected level of test coverage and improve test coverage ...