301 to 325 of 905 High Availability Jobs

Senior Dev Ops Engineer

Location
Greater London, England, United Kingdom
solutions, developing and supporting automation frameworks, managing CI/CD pipelines, and providing technical support to development teams to enable the efficient delivery of high-quality services and applications. Key Deliverables Build and maintain Kubernetes clusters on-prem and in cloud Access and IAM management in Google Cloud, Looker … with on-prem sharded mongo cluster Extensive experience in building, managing, scaling Kubernetes clusters (GKE, manual Kubernetes, etc.) Extensive experience in designing and building high availability, resilient infrastructure Experience in setting up and maintaining monitoring, alerting, log collection (rsyslog, Grafana, Prometheus, elk) Automation and scripting experience (terraform, packer ...

Lead Site Reliability Engineer

Hiring Organisation
Hackajob Ltd
Location
Glasgow, Lanarkshire, Scotland, United Kingdom
Employment Type
Permanent
with team members to identify comprehensive service level indicators and stakeholders to establish reasonable service level objectives and error budgets with customers Demonstrates a high level of technical expertise within one or more technical domains and proactively identifies and solves technology-related bottlenecks in your areas of expertise Acts … logging best practices. Certification in AWS, Kubernetes, or relevant technologies. Proven track record in system health monitoring, capacity management, and blameless postmortems for high-availability services Deep understanding of distributed system design principles, networking (TCP/IP, DNS, load balancing), Linux internals. Contributions to open-source observability ...

Infrastructure Team Lead

Location
Greater London, England, United Kingdom
support of IT infrastructure services. Design, implement, and maintain secure, scalable, and resilient infrastructure solutions. Supporting our primarily Microsoft-based infrastructure, ensuring its high availability, resilience, and performance. Both hybrid cloud and on-premise environments. Oversee server, storage, network, backup, cloud and virtualisation platforms. Ensure infrastructure availability … required. Leadership and Team Management Lead and develop the small Infra team, strong leadership and people management skills. What you'll need A high level of experience with Microsoft technologies including Windows Server andDesktop, Active Directory, Entra, Hybrid domains, Group Policies, 365, Intune, Teams,SharePoint, NTFS etc. Networking ...

Database Administrator (DBA)

Hiring Organisation
NuAxis Innovations
Location
Washington, Washington DC, United States
Employment Type
Permanent
Salary
USD Annual
join our team NOW! We are currently seeking a talented and motivated Database Administrator (DBA) for a Full-Time position. Role Summary Owns database availability, performance, security, and migration readiness across a legacy IBM DB2 estate and modern cloud-managed relational databases. Keeps mission-critical data services healthy while … auditing, and privileged access management. • Maintain baseline configurations, configuration management records, and system component inventories on a defined cadence. • Manage backup, recovery, replication, and high-availability configurations; test and document recovery procedures. • Remediate database-layer vulnerability and STIG findings; keep platform versions within supported release windows. • Support developers ...

Director of Platform Engineering

Location
Greater London, England, United Kingdom
been excelling here for 10+ years. With headquarters in London and teams across the US, Europe, and Asia, ITRS combines the agility of a high-impact tech business with the stability of a private equity-backed global partner Scope ITRS is looking for an experienced and accomplished, or aspiring … vulnerability management and service continuity. You will have: Proven experience leading SaaS Hosting teams in a hands-on technical leadership role, ideally within a high-availability, enterprise or regulated environment. A track record of building, leading and improving operational teams, including setting direction, developing engineers, raising standards ...

eTrading Python Developer

Location
Greater London, England, United Kingdom
Location: London (Hybrid – 3 days in office) Type: Full‐time Are you ready to step into a high-performance environment at the heart of global markets? We're looking for a motivated and technically skilled eTrading Production Support Engineer to join our fast‐paced Application Support team based … similar Networking: solid understanding of firewall infrastructure and multicast messaging Familiarity with electronic trading systems (FX or other asset classes) Exposure to low‐latency, high‐availability trading environments Project delivery experience in a production or infrastructure setting We Need Someone Who Is: Strong communicator and natural collaborator Analytical ...

Infrastructure / DevOps Engineer

Location
Birmingham, England, United Kingdom
thinks in systems, automates everything, and takes platform reliability personally. What you'll do Design, build, and manage AWS infrastructure for a real-time, high-availability voice AI platform Own CI/CD pipelines — build, test, deploy, and rollback automation across all environments Define and implement infrastructure … rest/in transit, audit trails, access controls) Nice to have Site Reliability Engineering background or formal SRE experience Experience supporting real-time or high-throughput systems (voice, streaming, or similar) AWS certifications (Solutions Architect, DevOps Engineer, or SysOps) Experience with multi-region or active-active deployment architectures Database ...

Forward Deployed Engineer - Infrastructure

Location
Greater London, England, United Kingdom
design and prepare suitable cloud infrastructure to ensure Thought Machine Vault products can be tested and run successfully at scale. Includes planning for high availability, disaster recovery, backup, redundancy, capacity and security. Deploying and configuring Thought Machine Vault products on client, SaaS and internal cloud infrastructure Developing deep … Azure (ideally certified Solutions Architect Expert) Experience of enterprise secrets management systems, e.g. HashiCorp Vault, AWS secrets manager Experience in supporting production systems for high profile, mission critical systems, ideally for a tier 1 financial institution. Experience with hybrid cloud technologies including OpenShift, Google Anthos, AWS EKS Anywhere ...

5G Specialists

Hiring Organisation
LA International Computer Consultants Ltd
Location
London, United Kingdom
Employment Type
Contract
Contract Rate
GBP 460 Daily
Good understanding of DTLS/SRTP protocol. MANDATORY: Extensive experience in mobile telecommunications industry in the role of a performance SME with business-critical, high-availability/high-throughput systems and networks. Able to provide technical solutions, processes definition for network performance & optimisation of 4G based … recruitment process, please let us know and we will work with you to ensure a fair and accessible experience. Please Note: If a high volume of applications is received, only candidates shortlisted will be contacted. ...

L4 Network Operations Engineer

Location
Gloucester, England, United Kingdom
compliance with financial services regulations Oversee lifecycle management including firmware governance, vulnerability remediation, and audit readiness Lead, mentor and develop L1-L3 engineers, fostering high performance through structured coaching, regular feedback, and the consistent application of best practices and continuous improvement initiatives. Provide design assurance for new deployments … monitoring and telemetry enrichment for improved visibility and governance Key Skills Deep multi-vendor expertise (Cisco, Fortinet, F5, SASE platforms) Advanced routing, segmentation, and high-availability design (BGP, OSPF, ACLs, NAT, HA clusters) Network security best practices (least privilege, zero trust, vulnerability remediation) Experience with IPAM/ ...

Platform Engineer – Vault SME IRC298440

Location
Greater London, England, United Kingdom
technical and cross-functional stakeholders. Job responsibilities Deploy and Maintain: Manage the lifecycle of HashiCorp Vault clusters across multiple environments (Dev, Stage, Prod). High Availability: Ensure 99.9% uptime through proper configuration of storage backends (like Raft or Consul) and unseal workflows (Auto-unseal via KMS/Cloud … offer Our goal is to build an inclusive positive culture where everyone can feel comfortable being themselves, empowering people to create their own high standards and therefore more value. We work together to promote fairness while recognising, valuing and embracing differences - providing a transparent support structure and generous training ...

Senior Platform Engineer

Location
Greater London, England, United Kingdom
commercial and residential installation companies. Access to our unique on-line portal and design tool, plus our emphasis on product quality, consistency and availability sets us apart in the market. Segen is a fast-moving business that responds quickly to any market changes supported by its bespoke ERP system … guidance. Promote cloud security best practices and ensure compliance within infrastructure design. Contribute to the evolution of our monitoring and alerting systems to maintain high availability and performance. What We’re Looking For Deep experience in Microsoft Azure, including Hub and Spoke networking. Strong proficiency with Azure DevOps ...

Vice President, DevOps Production Services

Hiring Organisation
Hackajob Ltd
Location
Manchester, North West, United Kingdom
Employment Type
Permanent
based applications in a fast-paced production environment. The role requires hands-on expertise in monitoring, incident management, troubleshooting, release support, and ensuring high availability and stability of business-critical platforms. In this role, youll make an impact in the following ways: Provide L2/L3 production support … environments. Familiarity with scripting and troubleshooting middleware/interfaces. Strong knowledge of release support, service recovery, and operational governance. Ability to work in a high-pressure environment with strong ownership and accountability. Demonstrated ability to ramp up quickly on new applications, platforms, and support processes, with strong learning agility ...

Technical Delivery Manager

Location
United Kingdom
/CD pipeline milestones. Ensure pre-release performance and queue testing gates are scheduled to prevent image-loading bottlenecks or monitoring delays during high-volume event spikes. Coordinate with third-party Out-of-Hours (OOH) support providers to keep operational runbooks and escalation paths up to date for overnight … business intelligence, data reporting software solutions. Exposure to mobile application release schedules (Apple App Store/Google Play Store) or NativePHP. Experience working in high-availability, real-time, or safety-critical software environments (e.g., IoT, telematics, Alarm Receiving Centres, or security operations). Familiarity with cloud infrastructure, messaging ...

Senior Cloud Engineer

Location
Greater London, England, United Kingdom
part of the Robinhood family, the Exchange Platform team owns the full service lifecycle. We are the architects of a modern, high-velocity ecosystem that enables our global expansion, ensuring the world’s longest-running exchange remains unshakeable. The Role As a Senior Backend Engineer integrated into the Exchange … essential part of building the new generation infrastructure for low-latency services in the cloud, adopting cutting-edge technologies to drive our high-velocity ecosystem. This role is based in our London office(s), with in-person attendance expected at least 3 days per week. At Robinhood, we believe ...

Senior CI/CD Automation Engineer

Hiring Organisation
BC Forward
Location
Jersey City, New Jersey, United States
Employment Type
Permanent
Salary
USD 92 Hourly
Security, and Engineering teams. Build secure, highly automated deployment pipelines. Advance AI-assisted vulnerability response and AI-enabled SDLC practices. Design deployment architectures for high availability and resiliency. Required Skills & Qualifications: Strong CI/CD and release engineering experience. Jenkins, XLR, Datical. Cucumber and Playwright for test automation. ...

Technical DevOps Engineer

Location
Greater London, England, United Kingdom
knowledge of helm – Strong analytic and problem solving skills – Experience on the following is a plus – monitoring and toning the CI environment to ensure high availability, efficiency, scalability and traceability – expanding the adoption and integration of CI build chain – improving degree of automation for repetitive tasks and reporting ...

Sr. Oracle Engineer

Hiring Organisation
Jobot
Location
Cincinnati, Ohio, United States
Employment Type
Permanent
Salary
USD 145,000 Annual
RMAN Oracle Enterprise Manager PL/SQL Performance tuning and optimization Experience supporting mission-critical production environments. Strong knowledge of backup, recovery, replication, and high-availability architectures. Experience with Linux/Unix operating systems. Excellent troubleshooting and analytical skills. Preferred Qualifications Experience with Oracle Cloud Infrastructure (OCI). ...

Azure Platform Architect

Location
Slough, England, United Kingdom
subscription architecture Azure Migrate and associated discovery tooling Microsoft Defender for Cloud and Microsoft Sentinel Azure Monitor, Log Analytics and KQL Zero Trust architecture High availability and disaster recovery FinOps and cloud cost optimisation Use of GitHub Copilot to accelerate Infrastructure as Code development The Person ...

Platform Lead: Fintech Architecture & SRE

Location
United Kingdom
making process for platform technologies, balancing in-house development with best-in-class third-party solutions* Drive a Site Reliability Engineering (SRE) culture, ensuring high availability, low latency, and robust disaster recovery capabilities* Manage and optimize our cloud infrastructure, focusing on Infrastructure-as-Code (e.g., Terraform), containerization (e.g. ...

Platform Lead - UK

Location
United Kingdom
making process for platform technologies, balancing in-house development with best-in-class third-party solutions* Drive a Site Reliability Engineering (SRE) culture, ensuring high availability, low latency, and robust disaster recovery capabilities* Manage and optimize our cloud infrastructure, focusing on Infrastructure-as-Code (e.g., Terraform), containerization (e.g. ...

OT Infrastructure Engineer / Consultant (PAYE Contract)

Location
Glasgow, Scotland, United Kingdom
Enterprise Linux (RHEL) Experience supporting application hosting and associated infrastructure Strong understanding of Active Directory, Group Policy, DNS and DHCP Experience with high availability, clustering and infrastructure resilience Experience with enterprise server hardware, including Dell, HPE or equivalent Knowledge of SAN, NAS and enterprise storage technologies Experience with ...

DevSecOps Engineer

Location
Greater London, England, United Kingdom
best practices for Azure, including networking, RBAC, managed identities, logging, and monitoring. Kubernetes security: network policies, PodSecurityStandards, RBAC, image security, secrets, ingress security, and high‐availability configurations. Terraform and Terraform‐security concepts (policies, modules, scanning tools). Azure DevOps CI/CD, including YAML pipelines, secure build pipelines ...

Infrastructure & Platform Senior Specialist Solutions Engineer

Location
Greater London, England, United Kingdom
such as AWS PrivateLink/Azure Private Link/GCP Private Service Connect), network routing, performance optimisation, and large‐scale deployment management Platform Administration: High availability, disaster recovery, cluster orchestration, observability and audit (e.g. Amazon CloudWatch/CloudTrail, Azure Monitor, Google Cloud Operations Suite), and cloud cost management ...

Solutions Engineer

Location
Manchester, England, United Kingdom
Product, Engineering, Customer Success, and Sales teams to deliver successful customer outcomes. Strong understanding of cloud architecture principles, including multi-tenant SaaS environments, scalability, high availability, disaster recovery, and global data residency requirements. Knowledge of AWS and/or Azure cloud technologies, networking concepts, infrastructure design, monitoring, security ...