226 to 250 of 905 High Availability Jobs

Cloud and Infrastructure Engineer

Location
Greater London, England, United Kingdom
scalable IaaS and PaaS solutions. Hands-on experience managing VMware environments and Hyper-Converged Infrastructure (e.g. HPE SimpliVity). Proven experience designing and managing high availability, disaster recovery and backup solutions. Strong knowledge of network infrastructure, including routing, firewalls, VPNs, SD-WAN and load balancing. Experience implementing infrastructure ...

DevOps Engineer

Location
Greater London, England, United Kingdom
cloud, enterprise networking, or multi-account AWS environments. Experience with Ansible or similar configuration management tools. Experience in financial services, trading, fintech, or another high-availability environment. Relevant AWS certification. Compensation Jain Global offers a total compensation package which includes a base salary, discretionary bonus, and comprehensive benefits. ...

Senior DevOps Engineer

Location
Greater London, England, United Kingdom
experience could also be relevant where you can demonstrate the ability to transfer that knowledge into an Azure environment. Experience with high‐availability infrastructure, PostgreSQL or MongoDB, AI/ML infrastructure, ETL/data pipelines, FinOps or SOC 2/ISO 27001 environments would be a bonus. Above ...

Azure DevOps Architect

Location
Penarth, Wales, United Kingdom
optimize containerized platforms leveraging Kubernetes, OpenShift, Azure Kubernetes Service (AKS), or Azure Red Hat OpenShift (ARO). Develop resilient cloud solutions that support high availability, disaster recovery, security, and operational excellence. Establish and promote DevSecOps practices, including automated security testing, secrets management, compliance controls, and policy enforcement. Drive ...

Principal Engineer - Kubernetes

Location
Greater London, England, United Kingdom
standards and modern SDLC practises are maintained. Is required to be a technical authority for Kubernetes. Understand security, risk and compliance frameworks, disaster recovery, high availability architectures, hardware, operating systems, and networking connectivity Collaborates with internal teams (including Product Owners). Provide solutions/guidance on Cloud budget ...

Lead Site Reliability Engineer

Hiring Organisation
Boeing
Location
Berkeley, Missouri, United States
Employment Type
Permanent
Salary
USD 267,950 Annual
Atlassian products in an enterprise environment Deep experience with PostgreSQL architecture and operations, including backup and recovery, replication, performance tuning, storage planning, maintenance, and high-availability patterns Experience with AWS, Microsoft Azure, Infrastructure as Code, Ansible, configuration management, containers, Docker, Kubernetes, virtualization, artifact management, secrets management, and secure ...

Platform Lead - ML Ops

Location
Greater London, England, United Kingdom
/LLMOps pipelines and CI/CD frameworks for continuous model deployment. GenAI Production Deployment: Deploy Large Language Models (LLMs) into production environments, ensuring high availability, low latency, and optimal performance for GenAI applications. FinOps & Budget Management: Establish FinOps frameworks to track, allocate, and forecast AI infrastructure spend … managing high-cost GPU/CPU cloud budgets. Resource Efficiency & Unit Economics: Implement auto-scaling, spot instances, and down-scaling policies to eliminate waste, while providing full visibility into the unit economics of training and serving LLM models. Observability & Incident Response: Establish 24/7 incident response, telemetry ...

Lead PostgreSQL Database Administrator

Hiring Organisation
CYNET SYSTEMS
Location
Raleigh, North Carolina, United States
Employment Type
Permanent
Salary
USD Hourly
database migration and modernization projects, designing and implementing solutions on AWS platforms such as RDS and Aurora. The role involves driving database performance tuning, high availability, disaster recovery, and backup strategies while providing technical leadership and mentoring to other team members. Responsibilities: Lead PostgreSQL database migration and modernization … across the United States and Canada. We deliver agile, scalable talent solutions across IT, engineering, life sciences, clinical, and professional staffing, powered by a high-performing recruitment engine operating across North America and Asia. As a nationally and locally certified Minority Business Enterprise (MBE), Cynet Systems is committed ...

Infrastructure engineer (UK)

Location
Greater London, England, United Kingdom
infrastructure across AWS (preferred), GCP, and Azure, working fluently across Kubernetes, Helm, Terraform, and the supporting cloud and AI tooling that backs WRITER's high-traffic platform. AI in workflow. Run agents in your daily loop — Claude Code, Droid, Codex, internal skills — to investigate incidents, draft Terraform/Helm … need Technical Track record. 5+ years of experience in infrastructure engineering, DevOps, or a similar role focused on building and operating large-scale, high-availability production systems at a high-growth product company. Breadth. Experience running containerisation in production (a real cluster, not a lab), with experience ...

Senior IBM MQ Integration Engineer

Hiring Organisation
CYNET SYSTEMS
Location
Richmond, Virginia, United States
Employment Type
Permanent
Salary
USD Hourly
SFTP, CloudFormation, CloudTrail, CloudWatch, Lambda, etc. AWS Certification and/or experience utilizing AWS services in a production environment is preferred. Strong knowledge of high-availability, load-balancing, and failover configurations across application, infrastructure, and platform components. Understanding of security design for enterprise software systems. Proven ability … across the United States and Canada. We deliver agile, scalable talent solutions across IT, engineering, life sciences, clinical, and professional staffing, powered by a high-performing recruitment engine operating across North America and Asia. As a nationally and locally certified Minority Business Enterprise (MBE), Cynet Systems is committed ...

Elasticsearch Consultant

Location
Greater London, England, United Kingdom
Develop Kibana dashboards, alerts, detection rules and visualisations. Configure data parsing, enrichment, transformation and indexing within Logstash and Elasticsearch. Integrate Kafka to support reliable, high‐volume event streaming and log ingestion. Deploy and operate Elastic components within Kubernetes environments. Automate infrastructure provisioning, configuration and deployment using Ansible and Argo … Build and maintain GitLab CI/CD pipelines. Develop scripts and automation tools to improve platform administration and operational efficiency. Monitor platform health, performance, availability and storage capacity. Troubleshoot ingestion failures, data‐quality issues and performance bottlenecks. Implement security controls, access management, data‐retention policies and platform hardening. Work ...

Application Architect III

Hiring Organisation
BC Forward
Location
Chandler, Arizona, United States
Employment Type
Permanent
Salary
USD Hourly
Computer Science, Information Technology, or a related field. Proven experience in production support or a similar role, preferably in a CICD platform environment. Availability for on-call support, including weekend rotations, on a round-robin basis. 7+ years of experience in production operations. Excellent communication and collaboration skills … cross-functional work. Familiarity with Oracle or PostgreSQL. Knowledge of F5 load balancers, GTM, high availability architectures, and disaster recovery strategies. Preferred Skills: Certifications in Red Hat Linux, DevOps methodologies, or related fields. Experience in large enterprise financial services environments. Why BCforward? At BCforward, we believe in advancing ...

Infrastructure Engineer

Location
Bracknell, England, United Kingdom
practices Participate in incident response, threat simulation and operational runbooks Collaborate with development teams to resolve technical issues and support deployments Design systems with high availability and disaster recovery in mind Support a global, always-on environment About You Bachelor's degree in Computer Science, a STEM field … automation skills (e.g. PowerShell, Python) Understanding of containerisation and orchestration (e.g. Docker, Kubernetes) Excellent troubleshooting and problem-solving skills Strong written and verbal communication High attention to detail and ability to work under pressure Understanding of security principles and monitoring practices is desirable About The Company Content Guru ...

Head of Cloud

Location
Norwich, England, United Kingdom
self-service for product teams. Drive adoption of best practices, automation, and tooling to improve velocity, reliability, and operational efficiency. Documentation & Continuous Improvement Champion high standards for technical documentation and process optimization across cloud teams. Promote a culture of knowledge sharing, observability, and proactive problem-solving. Success Metrics (KPIs … Team Growth: High-performing cloud engineers recruited, onboarded, and retained. System Evolution: Delivery excellence in cross-team technical initiatives with measurable improvements. Organizational Impact: Clear communication of cloud strategy to executives and stakeholders. Force Multiplier: Platform improvements increase developer velocity and adoption across engineering teams. Our Stack Cloud ...

Cloud Network Engineer

Hiring Organisation
AMS CWS
Location
London, United Kingdom
Employment Type
Contract
across AWS and GCP, including VPC/VNet design, subnetting, routing, IP addressing schemes, and network segmentation Hybrid Cloud Connectivity: Establish and maintain robust, high-performance, and secure connectivity between on-premise data centers and public cloud environments using technologies such as AWS Direct Connect, Google Cloud Interconnect, VPNs … Reliability Engineering (SRE) for Networks: Embrace a 'you build it, you run it' mindset for network services. Take ownership of the reliability, performance, and availability of the cloud network platforms. Implement robust network monitoring, alerting, and logging solutions, and participate in on-call rotations to ensure rapid incident response ...

Product Delivery Owner - Voice Services

Hiring Organisation
London Ambulance Service NHS Trust
Location
LONDON, SE1 8SD, United Kingdom
Salary
£75328.00 to £86114.00
services role delivering voice services to support critical services and applications Experience of designing, developing, implementing and managing voice systems and architectures focusing on high availability and disaster recovery to meet business needs. Extensive experience of working with architecture principles and governance Knowledge and Skills Essential Knowledge … mission critical computer systems and relevant interfaces and the need to provide high levels of service availability in support of key business objectives Ability to design, architect and implement complex Avaya Contact Centre and UC (Unified Communications) solutions and improvements, that meet specific business needs and apply best ...

Cloud DBA

Location
Greater London, England, United Kingdom
optimization of our PostgreSQL databases hosted on the AWS Cloud platform. The successful candidate will collaborate with cross-functional teams to ensure database reliability, availability, and performance, while also contributing to the overall architecture and strategy of our cloud‐based database solution. 1. Database Design and Architecture: Collaborate with … considering factors such as instance sizing, storage, and security. Implement and manage replication, clustering, and backup/recovery strategies to ensure high availability and disaster recovery. 3. Performance Monitoring and Optimization: Monitor database performance, proactively identifying and resolving performance bottlenecks, slow queries, and other issues affecting system responsiveness. ...

DV Cleared Network Engineer (2nd Line)

Location
Cheltenham, England, United Kingdom
compliant, and aligned to strict operational standards Maintain strong security controls, including segmentation, access policies, and device hardening Proactively identify and resolve risks to availability, performance, or security Implement controlled changes in line with rigorous governance and audit requirements Support incident resolution and TAC escalations when needed, with … clearance Experience working with Cisco networking and security technologies Hands-on exposure to Cat Center, ISE, and/or FMC Background in secure or high-availability environments Strong understanding of network security, patching, and lifecycle management CCNA Nice to Have Cisco security or enterprise certifications e.g. CCNP Experience ...

Sr Lead Infrastructure Engineer- Devops/AWS

Hiring Organisation
Hackajob Ltd
Location
Glasgow, Lanarkshire, Scotland, United Kingdom
Employment Type
Permanent
change-management controls Strong communication skills and the ability to set operational standards across a team Formal SRE experience in a regulated or high-availability environment Master's degree in Computer Science, Engineering, or a related technical field (or equivalent applied experience) Preferred qualifications, capabilities, and skills Experience … diverse talents they bring to our global workforce are directly linked to our success. We are an equal opportunity employer and place a high value on diversity and inclusion at our company. We do not discriminate on the basis of any protected attribute, including race, religion, color, national origin ...

Senior Developer (London)

Location
Greater London, England, United Kingdom
platform requirements. Ensure technical solutions meet strategic technical and business vision. Provide a hands-on technical approach to take ownership of designs from a high-level strategic approach down to low-level design, configuration, and installation detail. Provide participation throughout the project lifecycle, defining and refining requirements where required … including JQuery and Bootstrap) Windows Server/Workstation Active Directory Unit and integration testing DevOps tools including Nexus, Bitbucket, Sonarqube, Jenkins and AWX. DNS High Availability services Linux (CentOS) Secure private cloud hosting environments Networking protocols Desirable: Understanding and knowledge of Policing including MPS, National (UK Force wide ...

Oracle OCI Lead Engineer

Location
Leeds, England, United Kingdom
party feeds (e.g. Splunk). Compute, database, and storage knowledge – compute instances, patching, hardening, functions, autonomous databases, and block/object/file storage. High availability and disaster recovery – dual region, multi-AD, backups, replication, and region/service failover planning. Automation and infrastructure as code …/platform performance. More About the Department – DGCIO CS&G Within DGCIO CS&G you will work with people who are passionate about delivering high quality products and services. Unlike many large organisations, we provide both engineering and development in-house and this internal expertise allows us to understand ...

Specialist, Cloud Software Engineer

Hiring Organisation
L3Harris Technologies
Location
Melbourne, Florida, United States
Employment Type
Permanent
Salary
USD Annual
L3Harris is dedicated to recruiting and developing high-performing talent who are passionate about what they do. Our employees are unified in a shared dedication to our customers' mission and quest for professional growth. L3Harris provides an inclusive, engaging environment designed to empower employees and promote work-life success. … build and continuously improve the automated build, test and deployment environments for the applications. Provide operational support for applications and tools to meet high availability and other SLAs. Hands-on experience designing, developing, and deploying microservices to a Kubernetes based environments such as EKS, OpenShift, etc. Knowledge ...

Lead, Cloud Software Engineer

Hiring Organisation
L3Harris Technologies
Location
Melbourne, Florida, United States
Employment Type
Permanent
Salary
USD Annual
L3Harris is dedicated to recruiting and developing high-performing talent who are passionate about what they do. Our employees are unified in a shared dedication to our customers' mission and quest for professional growth. L3Harris provides an inclusive, engaging environment designed to empower employees and promote work-life success. … build and continuously improve the automated build, test and deployment environments for the applications. Provide operational support for applications and tools to meet high availability and other SLAs. Hands-on experience designing, developing, and deploying microservices to a Kubernetes based environments such as EKS, OpenShift, etc. Knowledge ...

Senior Infrastructure & Network Systems Engineer

Hiring Organisation
Galldris Group
Location
Enfield, England, United Kingdom
Infrastructure & Cloud Management Design, implement, and support Linux and Windows Server environments (Windows Server 2022+). Manage and maintain virtualisation platforms, including Proxmox, ensuring high availability and performance. Administer cloud-hosted infrastructure and integrate on-premises systems with cloud-based identity and access management solutions. Support and maintain … enterprise storage and data services, including TrueNAS environments. Manage file services, backup solutions, disaster recovery processes, and business continuity requirements. Ensure data integrity, availability, and recovery capabilities are maintained to a high standard. Identity & Access Management Administer authentication, authorisation, and access control solutions across multiple platforms. Develop ...

Senior VMWare Engineer

Location
Greater London, England, United Kingdom
solid understanding of networking, storage, and automation within enterprise-scale environments. Key Responsibilities Design, implement, and manage VMware-based virtual infrastructure, ensuring high availability, performance, and scalability. Administer VMware vSphere, ESXi hosts, vCenter, and associated tools to maintain a stable and secure environment. Manage and optimize virtual networking ...