451 to 475 of 1,013 High Availability Jobs

MLOps Engineer

Location
City Of London, England, United Kingdom
MLOps Engineer (Platform/DevOps) - AI Platform - 3 days per week onsite We\'re partnering with a leading organisation on a high-profile AI initiative, seeking an experienced MLOps/Platform Engineer to build and scale infrastructure supporting advanced AI solutions. The Role You\'ll play a key role … platform for deploying AI workloads and agents. Working closely with Data Science and Engineering teams, you\'ll ensure robust infrastructure, seamless deployment pipelines, and high platform reliability. Key Responsibilities Design, deploy, and manage AI platforms and agent infrastructure Build and maintain CI/CD pipelines and DevOps workflows Implement ...

MLOps Engineer

Hiring Organisation
DGH Recruitment
Location
City of London, London, United Kingdom
Employment Type
Permanent
MLOps Engineer (Platform/DevOps) - AI Platform - 3 days per week onsite We're partnering with a leading organisation on a high-profile AI initiative, seeking an experienced MLOps/Platform Engineer to build and scale infrastructure supporting advanced AI solutions. The Role You'll play a key role … platform for deploying AI workloads and agents. Working closely with Data Science and Engineering teams, you'll ensure robust infrastructure, seamless deployment pipelines, and high platform reliability. Key Responsibilities - Design, deploy, and manage AI platforms and agent infrastructure - Build and maintain CI/CD pipelines and DevOps workflows - Implement ...

Principal Architect

Location
Greater London, England, United Kingdom
strategy as we scale our global fintech platform. You will be the primary technical authority for our system design, responsible for ensuring that our high-throughput transaction engines and complex microservices ecosystem remain resilient, performant, and secure under massive load. You will partner with Engineering, Product, and Security leadership … Doing Own the Architectural Vision: Define and drive the long-term technical architecture for KAST, ensuring our platform evolves to meet the demands of high-throughput transactional volume. Design for Scale & Reliability: Lead the design and review of our distributed systems, ensuring our microservices architecture is built for fault ...

Platform Engineer: Scalable AWS, CI/CD & Observability

Location
Greater London, England, United Kingdom
needs and improve developer experience Design, build and maintain improvements to the infrastructure and platform of Hnry, currently hosted on AWS Improve the reliability, availability and performance of our platform Evolve and optimise our CI/CD pipelines to enable fast, safe and consistent deployments Implement and enhance observability … monitoring, alerting, logging) using tools like DataDog to ensure high-availability APM Apply best practices in security, scalability and cost optimisation across our infrastructure Support incident response and proactively identify and resolve system issues Produce work that meets the expected level of test coverage and improve test coverage ...

Senior Multi-Cloud DevOps & Infra Engineer

Location
Greater London, England, United Kingdom
Senior DevOps & Cloud Infrastructure Engineer to design, automate, and scale multi-cloud infrastructure across AWS, Google Cloud, and Vercel, with a focus on high availability and secure deployments. You will build robust CI/CD pipelines, manage Kubernetes workloads, and collaborate with engineering teams to implement IaC (Terraform … OpenTofu) and optimize deployment pipelines for a high-traffic SaaS product. #J-18808-Ljbffr ...

AMBG - Cloud Security & Exposure Management Architect

Location
Greater London, England, United Kingdom
passionate about improving reliability, continuity, and disaster recovery capabilities of enterprise cloud environments. You will work closely with customers to provide high-impact assessments and practical recommendations to strengthen operational resilience. Core to the Role Assess the resiliency posture of customer environments from both technical and operational perspectives. Evaluate … proven ability to assess complex environments, identify risks, and deliver clear, actionable recommendations to both technical and executive stakeholders. Professional Skillset Deep experience with high availability and disaster recovery design in Microsoft Azure. Strong understanding of Azure architecture including availability zones, backup, recovery, and monitoring services. Familiarity ...

Senior Engineer (Fincrime)

Hiring Organisation
ebury
Location
London, UK
Employment Type
Full-time
team is responsible for Ebury's AML (Anti-Money Laundering) platform. We are seeking a Senior Software Engineer to help us build and maintain high-reliability, low-latency systems that directly impact our regulatory compliance. In this role, you will work on mission-critical features, including real-time sanctions … highest standards of safety and auditability in a regulated global environment. ResponsibilitiesContribute to the development of our AML platform's services, focusing on high reliability and scalability for real-time sanctions screening and transaction monitoring. You will be responsible for both delivering new features and improving/automating existing ...

Back-End Engineer, Trading

Location
Greater London, England, United Kingdom
trading workflows to client-facing APIs. Develop and improve services for order management, market data processing, and trade lifecycle handling in a real-time, high-volume environment. Build and maintain REST APIs and WebSocket streams that deliver real-time prices, order updates, and account data. Collaborate with cross-functional … teams — product, frontend, trading — to adapt the platform quickly to new features and market conditions. Ensure high availability, reliability, and scalability of trading services, with a sharp focus on latency, observability, and testing. Contribute to architecture discussions and technical decisions, writing clean, maintainable code and sharing knowledge across ...

Back-End Engineer, Trading

Hiring Organisation
Blockchain
Location
London, UK
Employment Type
Full-time
trading workflows to client-facing APIs. Develop and improve services for order management, market data processing, and trade lifecycle handling in a real-time, high-volume environment. Build and maintain REST APIs and WebSocket streams that deliver real-time prices, order updates, and account data. Collaborate with cross-functional … teams — product, frontend, trading — to adapt the platform quickly to new features and market conditions. Ensure high availability, reliability, and scalability of trading services, with a sharp focus on latency, observability, and testing. Contribute to architecture discussions and technical decisions, writing clean, maintainable code and sharing knowledge across ...

Senior Database Administrator

Hiring Organisation
N P Associates
Location
London, United Kingdom
Employment Type
Full-Time
Salary
£80,000 - £95,000 per annum
hands-on technical role covering day-to-day administration, reliability, design and support across a varied stack that includes MySQL, Redis, QuestDB (a high-performance time-series database used for market data and operational metrics), Dolt (a version-controlled, Git-for-data SQL database), and Elasticsearch. The successful candidate … databases to deliver and operate our exchange technology products. Responsibilities and Duties Primary resource responsible to administer, monitor, and maintain the database estate, ensuring high availability, performance, and reliability. Plan and execute database upgrades, patch management, and version migrations with minimal disruption to production systems. Design, implement ...

Senior Linux Infrastructure Engineer - Systems & Platform

Hiring Organisation
Quant Capital
Location
London, UK
Employment Type
Full-time
PlatformLondon – Hybrid200,000-350,000 totalQuant Capital is partnered with a global systematic trading firm seeking a Senior Linux Infrastructure Engineer to join a high-impact platform engineering function. The team owns core infrastructure across datacentres, cloud environments and global offices, supporting both investment activity and large-scale engineering … workloads. This role suits a deeply technical engineer who understands large distributed systems, low-level Linux internals, and the realities of running secure, high-availability infrastructure at scale. What you'll do Design, operate and improve global Linux-based infrastructure Engineer secure, reliable systems across datacentres, cloud ...

Devops Engineer

Location
United Kingdom
platform is moving towards a modern, event-driven architecture, leveraging APIs, cloud‐native services, and advanced DevOps tooling to deliver secure, scalable, and high‐performing solutions. We work closely with partners across the Group, including Data & Machine Learning, Consumer Lending, and Cloud Services, to deliver strategic account data products … Core Banking, you’ll play a pivotal role in designing, building, and supporting the tools and processes that enable our teams to deliver high‐quality software at pace. You’ll work collaboratively across engineering, product, and operations teams to ensure our platform is secure, resilient, and continuously improving. What ...

Technical Lead (Java/.Net)

Location
United Kingdom
Reduce production incidents and improve system observability through thoughtful architecture and proactive monitoring. Raise team performance standards through active mentoring, peer feedback, and maintaining high standards of accountability. What You'll Need to be Successful B.S. in Computer Science or Engineering with 12+ years of professional software engineering experience. … C#) and maintainable frontend applications with React (TypeScript). Experience with distributed systems, RESTful/GraphQL APIs, system performance trade-offs, load balancing, and high availability engineering. Experience with AWS (Azure or GCP welcome), Infrastructure as Code (Terraform), CI/CD pipelines (e.g., GitLab), Docker, and Kubernetes. Experience ...

Software Engineer

Location
Greater London, England, United Kingdom
with the Smarkets product roadmap. Smarkets is looking for talented and passionate engineers like you for an exciting opportunity to create a unified and high‐performing system that will not only optimize our services but also elevate Smarkets to new heights of success. We believe in using the best … critical path operations and a slower interpreted one (Python) for others. Our Kafka pub/sub MQ, the heart of our system, offers high availability, low latency and message persistence. We provide gRPC and HTTP APIs for various metadata, while PostgreSQL and ElasticSearch serve ...

Full Stack Engineer

Location
Greater London, England, United Kingdom
identity management, as well as SAML 2.0 and Single Sign‐On protocols. Must be capable of effectively implementing and troubleshooting SSO and SCIM integrations. High‐Quality Code: Proven track record of writing clean, testable, and maintainable code that meets high standards of software quality. A developer who consistently … levels‐up the code base. Problem‐Solving & Scalability: Strong problem‐solving skills, with the ability to develop scalable and durable features in highavailability environments. Adaptability & Communication: Ability to thrive in a fast‐paced, dynamic environment, with excellent communication skills to support both internal teams and external customers. ...

Database Reliability Engineer

Location
Manchester, England, United Kingdom
regressions and health issues before they impact our customers Support Replication & Mobility: Support data streaming and "Zero-Downtime" migration strategies, ensuring data consistency and availability Fortify Business Continuity (BCP): Design and implement rigorous Business Continuity Planning and Disaster Recovery strategies. You will build the automation that ensures data durability … regressions and health issues before they impact our customers Support Replication & Mobility: Support data streaming and "Zero-Downtime" migration strategies, ensuring data consistency and availability Fortify Business Continuity (BCP): Design and implement rigorous Business Continuity Planning and Disaster Recovery strategies. You will build the automation that ensures data durability ...

KDB Software Engineer III

Hiring Organisation
JP Morgan Chase
Location
London, UK
Employment Type
Full-time
solutions, design, development, and technical troubleshooting with ability to think beyond routine or conventional approaches to build solutions or break down technical problemsDevelops secure high-quality production code, and reviews and debugs code written by othersAdds to team culture of diversity, equity, inclusion, and respectDevelop core systems and frameworks … data organization, performance implications of different approaches. Practical experience developing/running large datasets and optimizing query performance. Practical experience building resilient and high-availability KDB applications. Preferred qualifications, capabilities, and skillsExperience in Terraform and Kubernetes from managing a Production Plant in Public Cloud. AWS Experience. Experience other ...

SAP BTP Architect

Location
Manchester, England, United Kingdom
across SAP S/4HANA, SAP ECC, SAP BTP, SAP PI/PO, SAP Fiori/UI5, and integrated enterprise systems. This role ensures high‐performance, scalable, secure, and future‐ready SAP solutions that align with business strategy and IT roadmaps. The architect collaborates closely with business leaders, solution … system landscape design, including DEV/QAS/PRD/DR environments. In‐depth knowledge of SAP Basis, system sizing, performance tuning, highavailability (HA) and disaster recovery (DR) configurations. Strong experience with SAP HANA database architecture, optimization, backup/restore strategies, and security. Expertise in SAP integration ...

SAP BTP Architect

Location
Greater London, England, United Kingdom
across SAP S/4HANA, SAP ECC, SAP BTP, SAP PI/PO, SAP Fiori/UI5, and integrated enterprise systems. This role ensures high‐performance, scalable, secure, and future‐ready SAP solutions that align with business strategy and IT roadmaps. The architect collaborates closely with business leaders, solution … system landscape design, including DEV/QAS/PRD/DR environments. In‐depth knowledge of SAP Basis, system sizing, performance tuning, highavailability (HA) and disaster recovery (DR) configurations. Strong experience with SAP HANA database architecture, optimization, backup/restore strategies, and security. Expertise in SAP integration ...

Software Engineer III - KDB

Location
Greater London, England, United Kingdom
design, development, and technical troubleshooting with ability to think beyond routine or conventional approaches to build solutions or break down technical problems Develops secure high-quality production code, and reviews and debugs code written by others Adds to team culture of diversity, equity, inclusion, and respect Develop core systems … data organization, performance implications of different approaches. Practical experience developing/running large datasets and optimizing query performance. Practical experience building resilient and high-availability KDB applications. Preferred qualifications, capabilities, and skills Experience in Terraform and Kubernetes from managing a Production Plant in Public Cloud. AWS Experience. Experience ...

KDB / Q - Lead Software Engineer - Vice President

Location
Greater London, England, United Kingdom
design, development, and technical troubleshooting with ability to think beyond routine or conventional approaches to build solutions or break down technical problems Develops secure high-quality production code, and reviews and debugs code written by others Identifies opportunities to eliminate or automate remediation of recurring issues to improve overall … experience developing/running large datasets and optimizing query performance. Practical experience scaling and load‐balancing of KDB applications. Practical experience building resilient and highavailability KDB applications. Preferred qualifications, capabilities, and skills Experience with market data venue and vendor data platforms. AWS Experience. Experience in Terraform ...

Infrastructure Engineer

Location
Holme, England, United Kingdom
Responsibilities Administer,maintainand enhance the organisation’s Nutanix virtualisation platform and supporting infrastructure services. Monitor, troubleshoot and resolve complex infrastructure issues, ensuringhigh levelsof service availability and performance. Manage backup,recoveryand disaster recovery solutions, regularlyvalidatingresilience and recovery capabilities. Support the lifecycle management of server,storageand infrastructure platforms, including upgrades,patchingand … Windows Server and Red Hat Linux environments. Experience administering enterprise virtualisation platforms, ideally including Nutanix & VMware Experience supporting server,storageand compute infrastructure in a high-availability environment. Experience managing enterprise backup and recovery solutions, ideally using technologies such as NetBackup, Veeam or Commvault. Strong experience administering Microsoft cloud ...

Principal Software Engineer

Location
Waterbeach, England, United Kingdom
technical leader, mentoring other engineers and driving best practices across the development lifecycle. Key Responsibilities Lead the architecture, design, and implementation of high-performance, resilient, and secure communication systems using C#/.NET. Develop robust, low‐latency applications that handle high-volume data traffic (e.g., messaging queues, real … highly scalable, distributed systems (e.g., microservices architecture). Experience with development of Web Applications. Proven experience with protocols and technologies common in communication or highavailability systems (e.g., TCP/IP, gRPC, messaging services like Kafka or RabbitMQ). Expertise in performance tuning, concurrency, and multithreading to achieve ...

Principal Software Engineer

Hiring Organisation
Sepura
Location
Cambridge, Cambridgeshire, UK
Employment Type
Full-time
progress your career within this innovative technology company, based in Waterbeach, Cambridge. Role: Specific responsibilities will include: Lead the architecture, design, and implementation of high-performance, resilient, and secure communication systems using C#/.NET.Develop robust, low-latency applications that handle high-volume data traffic (e.g., messaging queues … developing highly scalable, distributed systems (e.g., microservices architecture).Experience with development of Web Applications. Proven experience with protocols and technologies common in communication or high-availability systems (e.g., TCP/IP, gRPC, messaging services like Kafka or RabbitMQ).Expertise in performance tuning, concurrency, and multithreading to achieve ...

Identity Infrastructure Engineer

Hiring Organisation
Hays Specialist Recruitment Limited
Location
London, South East England, United Kingdom
Employment Type
Full-Time
Salary
£447.00 per day
regulated environment, supporting critical Identity and Access Management (IDAM) services. Working closely with security, engineering and IDAM teams, you will help maintain the resilience, availability and security of core identity infrastructure platforms. Your new role As a Senior Identity Infrastructure Engineer, you will be responsible for the administration, support … recovery and archival Experience supporting enterprise PKI environments. Infrastructure Platforms Windows Server and 2022 administration. PowerShell scripting and automation. Security hardening and privileged administration. High availability and disaster recovery planning. Infrastructure monitoring and operational support. Identity Technologies Working knowledge of: Microsoft Entra ID Microsoft Entra Connect/Cloud ...