151 to 175 of 1,025 High Availability Jobs

Senior Software Engineer Python

Hiring Organisation
MarkIT Placements
Location
Didcot, Oxfordshire, South East, United Kingdom
Employment Type
Permanent, Work From Home
usable in the real world. If you enjoy owning backend services end to end, care deeply about reliability and API quality in a high-impact environment, this role offers both autonomy and real impact. Requirements Backend Architecture & APIs (Primary Focus) Design and evolve scalable backend services in Python using … FastAPI for high-availability, high-throughput workloads. Build well-versioned RESTful APIs aligned to OpenAPI/Swagger, with strong conventions for consistency, idempotency, and backward compatibility. Implement authentication and authorization using OAuth2/OIDC, session management, and fine-grained permissions. Design and maintain event-driven architectures ...

Head of Architecture - Data Platform & Account & Identity (BBC)

Location
Lancashire, England, United Kingdom
audience insight through to regulatory compliance. You'll operate as one of the most senior technologists in the organisation, owning architectural direction across a high-volume, event-driven ecosystem with near-zero tolerance for failure. This role is hybrid, with flexibility to be based from your nearest … aligned to both technical strategy and wider organisational goals. Key Responsibilities Define and evolve platform capabilities supporting billions of audience events per day, ensuring high availability, scalability and security Lead the transition toward domain-oriented, platform-based architectures, embedding data product thinking and ownership models Own and govern ...

Lead Data Center Technician

Hiring Organisation
Visionaire Partners
Location
Tucker, Georgia, United States
Employment Type
Permanent
Salary
USD Annual
enterprise organization is seeking a Lead Data Center Technician to oversee the end-to-end deployment, migration, and optimization of critical infrastructure within a high-availability environment. This role focuses on scaling server, network, and SAN environments while elevating physical layer standards, optimizing power/airflow, and managing … inventory records. Maintenance: Safely execute installations, moves, and modifications while keeping a clean facility. Required : Experience: 5+ years of hands-on technical experience in high-availability data centers. Deployment: Lead rack/stack operations for servers, network, and SAN systems, owning layouts, elevations, and labeling. Infrastructure: Deep knowledge ...

Global Banking Markets - Low-Latency Java Software Engineer - Vice President - London

Location
Greater London, England, United Kingdom
outside the office. Goldman Sachs Electronic Trading (GSET) strives to be the top provider in electronic trading by building superior technology and delivering high quality products by investing in people, platforms, and products. Join the team and participate in the development and launch of best-in-class products … should be willing to participate in the full product lifecycle from requirements gathering, design, implementation, testing, support, and monitoring. RESPONSIBILITIES Design, build and maintain high-performance, high-availability, high-capacity, yet nimble and adaptive platforms satisfying a range of business needs. Work in partnership with ...

Global Banking Markets - Low-Latency Java/C++ Software Engineer - Vice President - London

Location
Greater London, England, United Kingdom
outside the office. Goldman Sachs Electronic Trading (GSET) strives to be the top provider in electronic trading by building superior technology and delivering high quality products by investing in people, platforms, and products. Join the team and participate in the development and launch of best-in-class products … should be willing to participate in the full product lifecycle from requirements gathering, design, implementation, testing, support, and monitoring. RESPONSIBILITIES Design, build and maintain high-performance, high-availability, high-capacity, yet nimble and adaptive platforms satisfying a range of business needs Work in partnership with ...

Global Banking Markets - Software Engineer - Equities Data Platform Engineering - London - Vice President

Hiring Organisation
Goldman Sachs
Location
London, UK
Employment Type
Full-time
leads on the ground-up rebuild of our Electronic Trading data infrastructure into a cloud-native, AI-first platform. Design, build, and maintain a high-performance, high-availability, high-capacity platform that processes and persists massive volumes of business-critical data in near real time, distributing … raise the bar on engineering practices across the team. Required Qualifications & Skill-Sets Around 5–10 years of professional experience with experience building high-performance, low-latency systems (sub-second) – at Vice President level. Proven ability to drive complex, industrial-scale development independently, while also leading and developing engineers ...

Global Banking Markets - Software Engineer - Equities Data Platform Engineering - London - Vice President

Location
Greater London, England, United Kingdom
leads on the ground-up rebuild of our Electronic Trading data infrastructure into a cloud-native, AI-first platform. Design, build, and maintain a high-performance, high-availability, high-capacity platform that processes and persists massive volumes of business-critical data in near real time, distributing … raise the bar on engineering practices across the team. Required Qualifications & Skill-Sets Around 5–10 years of professional experience with experience building high-performance, low-latency systems (sub-second) – at Vice President level. Proven ability to drive complex, industrial-scale development independently, while also leading and developing engineers ...

Global Banking Markets - Low-Latency Java/C++ Software Engineer - Vice President - London

Hiring Organisation
Goldman Sachs
Location
London, UK
Employment Type
Full-time
outside the office. Goldman Sachs Electronic Trading (GSET) strives to be the top provider in electronic trading by building superior technology and delivering high quality products by investing in people, platforms, and products. Join the team and participate in the development and launch of best-in-class products … should be willing to participate in the full product lifecycle from requirements gathering, design, implementation, testing, support, and monitoring. RESPONSIBILITIESDesign, build and maintain high-performance, high-availability, high-capacity, yet nimble and adaptive platforms satisfying a range of business needsWork in partnership with the wider engineering ...

Sr. Site Reliability Engineer/ SWE

Location
Ham, England, United Kingdom
instead of infrastructure. You’ll promote observability best practices and automate resolution of recurring issues, working closely with software engineering teams to support security, availability, and performance. Responsibilities include triaging issues, collaborating on infrastructure management, and setting up monitoring for full coverage. Hands‐on expertise is required, especially with … integrations for reliability and scalability. Collaborate with development teams to improve workflows and automation. Site Reliability Engineering Design, implement, and maintain systems for high availability, scalability, and performance. Monitor and improve application reliability through proactive measures and incident response. Develop and maintain observability solutions (metrics, logging, tracing). ...

Sr. Site Reliability Engineer/ SWE

Hiring Organisation
Visa
Location
Basingstoke, Hampshire, UK
Employment Type
Full-time
instead of infrastructure. You'll promote observability best practices and automate resolution of recurring issues, working closely with software engineering teams to support security, availability, and performance. Responsibilities include triaging issues, collaborating on infrastructure management, and setting up monitoring for full coverage. Hands-on expertise is required, especially with … integrations for reliability and scalability. Collaborate with development teams to improve workflows and automation. Site Reliability Engineering Design, implement, and maintain systems for high availability, scalability, and performance. Monitor and improve application reliability through proactive measures and incident response. Develop and maintain observability solutions (metrics, logging, tracing). ...

Software Engineer/ SRE (Linux)

Hiring Organisation
Visa
Location
Basingstoke, Hampshire, UK
Employment Type
Full-time
instead of infrastructure. You'll promote observability best practices and automate resolution of recurring issues, working closely with software engineering teams to support security, availability, and performance. Responsibilities include triaging issues, collaborating on infrastructure management, and setting up monitoring for full coverage. Hands-on expertise is required, especially with … integrations for reliability and scalability. Collaborate with development teams to improve workflows and automation. Site Reliability Engineering Design, implement, and maintain systems for high availability, scalability, and performance. Monitor and improve application reliability through proactive measures and incident response. Develop and maintain observability solutions (metrics, logging, tracing).Participate ...

Software Engineer/ SRE (Linux)

Hiring Organisation
Visa
Location
London, England, United Kingdom
instead of infrastructure. You'll promote observability best practices and automate resolution of recurring issues, working closely with software engineering teams to support security, availability, and performance. Responsibilities include triaging issues, collaborating on infrastructure management, and setting up monitoring for full coverage. Hands-on expertise is required, especially with … reliability and scalability. \n Collaborate with development teams to improve workflows and automation. \n Site Reliability Engineering Design, implement, and maintain systems for high availability, scalability, and performance. \n Monitor and improve application reliability through proactive measures and incident response. \n Develop and maintain observability solutions (metrics, logging ...

Senior Backend Engineer

Location
Greater London, England, United Kingdom
mission is to enable everyone to build wealth We reinvent how trading and investing work by creating exceptional products people love. That takes high standards, high velocity and engineers who own what they ship. Today, we serve over 6 million clients, with more than €38 billion in assets … under management. At that scale, performance, reliability and data integrity are part of the product. Our platform is powered by a modern, high-performance engineering stack designed for scale and reliability: Java 21 (Spring), Go, Node.js (TypeScript), PostgreSQL, MariaDB, Redis, Kafka, ClickHouse, Elasticsearch, Docker, Grafana, AWS, Kubernetes and Kibana. ...

Platform Chapter Lead - Engineering

Location
Greater London, England, United Kingdom
developer tooling – Backstage especially welcome. Familiarity with tools like Terraform, GitHub Actions and Octopus (or equivalents), and FinOps/cloud cost management. Regulated, highavailability or high‐transaction‐volume domains (travel, retail, payments). What success looks like in the first 12 months A clear platform roadmap ...

Senior DevOps & Systems Engineer

Location
United Kingdom
Provisioning and configuration of new virtual machines, servers, services and supporting infrastructure. - Deployment and configuration of applications and infrastructure components. - Monitoring the health, performance, availability and security of production and internal systems. - Investigating and resolving infrastructure incidents, performance issues and system failures. - Designing and implementing appropriate monitoring, alerting … performance management troubleshooting. - Experience provisioning and configuring servers and virtual machines from initial deployment through to production readiness. - Strong understanding of production infrastructure reliability, availability and operational support. - Experience with system patching, upgrades and lifecycle management. - Experience diagnosing complex infrastructure issues independently. Monitoring, Reliability & Security - Strong experience with system ...

Senior DevOps & Systems Engineer

Hiring Organisation
Centric Talent
Location
Blackburn, Lancashire, North West, United Kingdom
Employment Type
Permanent
Salary
£65,000
Provisioning and configuration of new virtual machines, servers, services and supporting infrastructure. - Deployment and configuration of applications and infrastructure components. - Monitoring the health, performance, availability and security of production and internal systems. - Investigating and resolving infrastructure incidents, performance issues and system failures. - Designing and implementing appropriate monitoring, alerting … performance management troubleshooting. - Experience provisioning and configuring servers and virtual machines from initial deployment through to production readiness. - Strong understanding of production infrastructure reliability, availability and operational support. - Experience with system patching, upgrades and lifecycle management. - Experience diagnosing complex infrastructure issues independently. Monitoring, Reliability & Security - Strong experience with system ...

Senior Software Development Engineer

Location
Greater London, England, United Kingdom
build for travelers everywhere. The team As a Fullstack engineer in the Product Experience team, you will be responsible for building and maintaining our high-availability, high-transactional Property Detail Pages on the app. You will be part of a multi-functional team of Product Managers, TPMs … Engineering Managers, and Software Engineers based in Gurgaon, Bangalore, London, and Madrid. We prioritise building high-quality software with a focus on availability, performance, scalability, and system resiliency. Our London team is looking for curious, empathetic, and creative problem solvers with a growth mindset. We value our team ...

Lead C++ / Java Developer

Hiring Organisation
London Stock Exchange Group
Location
London, UK
Employment Type
Full-time
responsible for designing, building, and operating mission‐critical FX market infrastructure and matching platforms using C++ and Java. The role focuses on low‐latency, high‐throughput distributed systems with strict availability, resiliency, and data integrity requirements. This role owns technical outcomes for complex components and services, influences architecture … trading workflows. Own technical delivery for key platform components, ensuring alignment with LSEG architectural principles, performance standards, and operational controlsBuild and evolve low‐latency, highavailability systems handling high message volumes and time‐critical processing. Ensure systems meet non‐functional requirements, including latency, throughput, resiliency, fault tolerance ...

Backend Technical Expert, Gaming

Location
Greater London, England, United Kingdom
stability. Online Service Development and Maintenance: Participate in building Online Services for games (such as login, matchmaking, real‐time battles, data synchronization). Ensure high availability and low latency to support stable global player access. Cross-Team Collaboration: Collaborate with top‐tier global game studio teams to drive … required). Proficient in at least one programming language (expertise in C++ or Golang is preferred), with strong experience in threads, coroutines, and building high‐performance, high‐concurrency, and highly available systems. Familiar with distributed systems and core backend technologies such as microservices architecture, message queues (such ...

Senior AWS Infrastructure Engineer - eSC or eDV Clearance Required

Location
Cheltenham, England, United Kingdom
gain hands on experience with cutting edge technologies, and deliver solutions that create real business impact. From day one, you’ll work on meaningful, high profile programmes that stretch your skills and accelerate your growth. We invest heavily in you—supporting continuous learning, in demand skills development, and long … Balancers, Auto Scaling, and backup/recovery solutions in line with public sector and defence security requirements Manage and optimise Linux server environments, ensuring high availability, hardening, patching, performance, and compliance with security standards and operational policies Support and maintain core infrastructure technologies including networking, middleware, DNS, VPNs ...

Principal Engineer - Control System Performance

Location
Stafford, England, United Kingdom
unprecedented growth in demand for and generation of electricity. To accommodate this growth, the capacity of transmission system is projected to grow three-fold. High voltage direct current (HVDC) transmission systems will play a key role in connecting remote and offshore renewables, interconnecting energy markets and reinforcing AC grids. … technical consultation on product problems throughout the business including supplier and field support and perform technical rescues when needed. Actively mentor and coach identified high potential Engineering talents within one’s business lines. Chair Design reviews for individual components, sub-assemblies and key engineering deliverables at tendering and contract ...

DevOps Engineer

Location
Greater London, England, United Kingdom
hosted in AWS. Replace repetitive manual processes with scripts and automation. Operate, maintain, and optimise Kubernetes clusters and containerised applications. Monitor platform health, performance, availability, and capacity. Investigate and resolve platform, infrastructure, and application‐related incidents. Work closely with Development, DevOps, and Application Support teams to improve system reliability … alerting, logging, and observability solutions. Participate in technical projects including platform upgrades, migrations, and cloud transformation initiatives. Contribute to disaster recovery, business continuity, and highavailability strategies. Create and maintain technical documentation, standards, and operational procedures. Participate in an Overnight on‐call rota. Mentoring other members ...

DevOps Engineer

Location
Greater London, England, United Kingdom
hosted in AWS. Replace repetitive manual processes with scripts and automation. Operate, maintain, and optimise Kubernetes clusters and containerised applications. Monitor platform health, performance, availability, and capacity. Investigate and resolve platform, infrastructure, and application‐related incidents. Work closely with Development, DevOps, and Application Support teams to improve system reliability … alerting, logging, and observability solutions. Participate in technical projects including platform upgrades, migrations, and cloud transformation initiatives. Contribute to disaster recovery, business continuity, and highavailability strategies. Create and maintain technical documentation, standards, and operational procedures. Participate in an Overnight on‐call rota. Mentoring other members ...

Lead Engineer

Location
Greater London, England, United Kingdom
reviews; partner with risk/compliance and audit functions on evidence and remediation.* Solve complex issues end-to-end, including proxy/connector components, high-availability clusters, and multi-region architectures.## Technical/job functional knowledgeConsistent record implementing and operating PAM at scale in complex enterprises (multi-site ...

IT Infrastructure Architect - contract

Location
City Of London, England, United Kingdom
Systems designs. Responsibilities To act as an Infrastructure Architect on a variety of internal and external IT delivery projects, key responsibilities include: Produce high quality High Level Design documentation describing technical systems solutions, providing overall design direction and recommendations for resolution of complex technical issues Provide technical leadership … communication skills Leadership, presentation and stakeholder management skills Ability to work on complex projects with globally distributed teams and tight timelines Ability to design high availability solutions, including consideration to fault tolerant analysis and resiliency Ability to work with and manage vendors/suppliers such as IBM Experience ...