101 to 125 of 442 Disaster Recovery Jobs in London

Mandarin speaking Manager of IT - Banking

Location
City Of London, England, United Kingdom
associated budgets. Plan for and execute necessary changes of the Branch’s IT environment and systems. Plan for and execute the Branch’s local Disaster Recovery drills and/or contingency process. Monitor and manage the IT system vulnerabilities, take remediation actions to solve the relevant risks. Provide … support and validation for HO’s IT/system changes, Disaster Recovery drills and/or contingency process. Mandarin speaking Manager of IT - The Skills You'll Need to Succeed: Proficient in English and Mandarin language skills. Proven knowledge of IT infrastructure, network and system related technology. Solid ...

Vice President, Risk and Control - Digital Engineering

Location
Greater London, England, United Kingdom
support to Digital Engineering Services and Solutions for pen test findings* Produce and manage Digital Engineering Solutions and Services owned Key Risk Indicators.* Support disaster recovery exercises, ensuring new services are documented and deployed with BCP/DR in mind* Provide advisory assistance to IT Risk and Control … Experienced in dealing with vendors and third-party suppliers* Strong track record of managing teams and building effective partnerships with peers* Experience with comprehensive disaster recovery architecture and operations, including storage area network and redundant, highly available server and network architectures* Experience with regulatory compliance issues as they ...

Enterprise Platform Lead

Location
City Of London, England, United Kingdom
mapping, error handling, retries, security and resilience Ensure platform and integration solutions meet enterprise architecture and non-functional requirements, including security, performance, scalability, resilience, disaster recovery, recovery objectives, business continuity and operational support Provide technical leadership and assurance throughout discovery, design, delivery and transition to operations, working ...

SRE, London

Location
Greater London, England, United Kingdom
Experience with deploying, supporting and monitoring new and existing services, platforms, and application stacks Excellent troubleshooting and problem solving skills Experience with scale testing, disaster recovery, and capacity planning Familiarity with microservices architecture and container orchestration with Kubernetes Preferred Qualifications Hands-on experience managing large numbers of diverse … Experience with deploying, supporting and monitoring new and existing services, platforms, and application stacks Excellent troubleshooting and problem solving skills Experience with scale testing, disaster recovery, and capacity planning Familiarity with microservices architecture and container orchestration with Kubernetes At Apple, we're not all the same. And that ...

Operations and SRE Manager

Location
Carshalton, England, United Kingdom
actions are owned, tracked and completed. Strengthen operational process adherence, ensuring responsibilities are clear and delegation is effective. Drive SRE practices across observability, automation, disaster recovery, design for reliability, on-call readiness and production support. Protect service levels by ensuring engineering effort is balanced across InfoSec commitments, operational … including coaching, performance management, prioritisation and development of team leads and engineers. Practical experience of SRE, incident management, problem management, post-mortems, RCA quality, disaster recovery and operational resilience. Ability to influence across Ops, engineering, product and wider technology teams, especially where priorities are competing or ownership ...

IT Project Manager

Hiring Organisation
TEKsystems
Location
London, UK
Coordinate infrastructure architecture across Linux, IBM Z/LinuxONE, networking, storage, security, monitoring, and automation teams. Establish clear accountability for provisioning, patching, monitoring, backup, recovery, and operational support activities. Validate workload and application readiness for IBM Z (s390x) environments, ensuring compatibility and supportability. Drive pilot deployments to prove provisioning … processes, performance, resilience, recovery procedures, and support models. Manage production readiness activities and secure operational acceptance from all stakeholders. Deliver service onboarding processes, support documentation, operational runbooks, and service catalogue content. Own programme governance, planning, risks, issues, dependencies, budgets, and stakeholder communications across multiple infrastructure teams. Essential Experience Proven ...

IT Project Manager - Banking, Technology Resilience / BC / DR

Hiring Organisation
Hays Specialist Recruitment Limited
Location
London, South East England, United Kingdom
Employment Type
Full-Time
Salary
£650.00 - £750.00 per day
urgently require an IT Project Manager with proven Investment Banking industry experience and working knowledge of Technology Resilience/Business Continuity (BC)/Disaster Recovery (DR) Projects/Programmes to join a business-critical Operational Resilience Project. Key Requirements: Proven experience as an IT Project Manager with … focused Projects/Programmes with Bank-wide impact across the entire organisation Previous experience of working on Technology Resilience/Business Continuity (BC)/Disaster Recovery (DR) related Projects/Programmes Working knowledge of testing (DR Testing, scenario testing, etc) and ability to prepare, plan and oversee ...

Solutions Architect

Location
Greater London, England, United Kingdom
platform evolution. Maintain comprehensive documentation of design decisions, patterns, standards, and trade-offs. Define non-functional platform architecture standards covering resilience, backup/restore, disaster recovery, observability, service levels, auditability and operational readiness. Define platform security and privacy architecture in partnership with Cyber Security and Data Governance, including … environments transitioning from outsourced to in-house ownership is beneficial. Experience defining non-functional requirements and architecture patterns for enterprise-grade resilience, observability, disaster recovery, data lifecycle management and operational readiness. Experience using architecture decision records, design authorities, exception management and measurable standards adoption to embed architectural governance ...

Head of Operational Resilience

Location
Greater London, England, United Kingdom
exercises across critical business services, identifying vulnerabilities and overseeing remediation activity. Partner with Technology, Information Security, Risk and business leaders to strengthen technology resilience, disaster recovery and operational resilience capabilities. Provide expert advice and constructive challenge to senior leaders, Executive Committees and Board governance forums through clear reporting … service mapping, governance and scenario based testing. Proven experience working closely with Technology and Information Security teams, with a strong understanding of technology resilience, disaster recovery and cyber resilience. Experience engaging and influencing senior executives, Board or Executive Committees, with the ability to communicate complex resilience matters clearly ...

Hosting Platform Engineer (Unix/Linux)

Location
City Of London, England, United Kingdom
platform patching, upgrades, and lifecycle maintenance across UNIX/Linux environments. Perform security hardening, vulnerability remediation, configuration management, cluster patching, rolling upgrades, rollback planning, disaster recovery validation, and post-patch testing while ensuring compliance with enterprise security and change management standards. Skills requireed for the role UNIX/… Linux (RHEL, Solaris, AIX) Red Hat Satellite Ansible Oracle WebLogic OPatch JBoss EAP Java (JDK/JRE) JVM Tuning Security & Vulnerability Management Change Management Disaster Recovery Smoke Testing #J-18808-Ljbffr ...

Enterprise Infrastructure Lead

Location
City Of London, England, United Kingdom
software upgrades, resilience improvements, security remediation, service transitions and platform decommissioning. Ensure changes are appropriately designed, tested, risk‐assessed and governed, with robust implementation, recovery and stakeholder‐communication plans. Act as a senior escalation point during major incidents, bringing together engineering teams, application owners, service management, suppliers and senior … teams responsible for mainframe environments. Developing specialist engineering teams, addressing key‐person dependencies and creating credible succession and skills‐transition plans. Leading major incident recovery, root‐cause investigation, problem management and preventative improvement activity. Operating mission‐critical, high‐availability platforms within financial services or another highly regulated industry. Ownership ...

LEAD ADMINISTRATOR L1(CONTRACT)

Hiring Organisation
Wipro
Location
London, UK
Employment Type
Full-time
load utilities, or ETL tools. Configure and support high-availability solutions such as Sybase Replication Server, Warm Standby, MSA, Failover Clustering, and disaster recovery strategies. Automate database administrative tasks using Shell/Perl/Python scripts. Sybase IQ Data warehouse support reporting, BI workloads, data analytics, and archival. … initiatives to enhance overall system efficiency. Perform system Improvements-Create mitigation measures to evaluate new tools or improvements by collaborating with technical teams. Handling disaster recovery-Perform security assessments and recovery drills, and implement necessary controls to ensure compliance with established security standards. Reporting and stakeholder engagement ...

LEAD ADMINISTRATOR L1(CONTRACT)

Location
Greater London, England, United Kingdom
load utilities, or ETL tools. Configure and support high-availability solutions such as Sybase Replication Server, Warm Standby, MSA, Failover Clustering, and disaster recovery strategies. Automate database administrative tasks using Shell/Perl/Python scripts. Sybase IQ Data warehouse support reporting, BI workloads, data analytics, and archival. … initiatives to enhance overall system efficiency. Perform system Improvements-Create mitigation measures to evaluate new tools or improvements by collaborating with technical teams. Handling disaster recovery-Perform security assessments and recovery drills, and implement necessary controls to ensure compliance with established security standards. Reporting and stakeholder engagement ...

Solution Architect

Hiring Organisation
ECS Resource Group Ltd
Location
London, United Kingdom
Employment Type
Contract
Contract Rate
£750 - £800/day Inside IR35
Financial Services experience is desirable but not essential. Technical Expertise Technical expertise in Nutanix technologies including AHV, Prism Central, lifecycle management, automation, resilience and disaster recovery. Technical expertise in Dell enterprise storage platforms, including PowerStore, PowerScale, PowerMax, ECS and associated storage networking technologies. Strong knowledge of enterprise compute platforms … platform architectures, including Kubernetes, container orchestration, platform operations, workload placement, application hosting, infrastructure integration, resilience, observability and hybrid cloud deployment models. Experience with backup, recovery, cyber resilience and business continuity strategies, including immutable storage, replication, ransomware recovery and disaster recovery architectures. Experience working with Rubrik ...

Cloud Infrastructure Engineer

Location
Greater London, England, United Kingdom
serverless (AWS Lambda) applications. Own infrastructure as code — provisioning resources declaratively so environments are reproducible, version-controlled, and safe to change. Ensure effective backup, recovery, and disaster recovery strategies to minimise downtime. Manage operational and analytical data stores (Aurora MySQL, DynamoDB) Drive cost optimisation across the infrastructure … Linux‐based operating systems and shell scripting. Familiarity with Infrastructure as Code tools (Terraform, CloudFormation). Experience with incident management, troubleshooting, and platform recovery in high‐pressure environments. Strong communication skills with a proven ability to work both independently and collaboratively It's a plus if you have Experience ...

Senior Systems Engineer (Linux & Public Cloud Specialist)

Hiring Organisation
Trayport
Location
London, UK
Employment Type
Full-time
paths of Linux infrastructure and public cloud resources. Identify, deploy, and maintain appropriate hybrid connectivity options across environments. Maintain and test Business Continuity and Disaster Recovery (BC/DR) plans and systems across Private and Public Cloud environments. Contribute to the overall quality of service delivered to internal … compliant provisioning across all environments. Lead architectural design reviews for new workload migrations and deployments, defining criteria for high availability, fault tolerance, and automated disaster recovery. Continuous improvementDrive infrastructure automation (IaC) and design improvements for efficient support of environments, applications, and clients. Actively contribute to increasing service availability through ...

Senior Platform Engineer (12 month FTC)

Hiring Organisation
National Physical Laboratory
Location
London, UK
. Supporting enterprise security initiatives, including privileged access management, monitoring, logging, and data protection. Producing technical documentation, operational procedures, and support runbooks. Contributing to disaster recovery, business continuity, and redundancy planning. Travelling to sites as required to support deployment, implementation, and operational activities. About You Essential Criteria ...

Systems Administrator

Location
Greater London, England, United Kingdom
production: IAM, networking, EC2 and managed services, patch compliance through Systems Manager, tagging, monitoring and cost control. Maintain and improve backup, restore testing and disaster recovery, including documented and rehearsed recovery procedures. Build and maintain automation in PowerShell, Bash or Python to remove manual work from routine … Kandji: baseline configuration, hardening, patching, encryption and application deployment. Maintain joiner, mover and leaver processes end to end, including timely revocation and device recovery or wipe on exit. Keep the asset register accurate and audit ready. Internal support pipeline Action the internal support queue in Jira ...

eTrading Platform Product Owner

Hiring Organisation
ING Banking
Location
London, UK
Employment Type
Full-time
helping others succeed. Useful experience Observability strategies and tooling, including meaningful monitoring, alerting and service-level reporting. Resilient architecture concepts, including active-active designs, disaster recovery scenarios and recovery testing. ITIL practices, incident management, problem management, change management and post-incident improvement. Electronic markets, foreign exchange, trading ...

Senior Desktop Engineer

Location
Greater London, England, United Kingdom
Infrastructure & Cloud Support the wider infrastructure environment alongside senior engineers and technology partners. Work with Azure across areas including compute, networking, storage, monitoring and disaster recovery. Support Okta administration, including SSO, MFA, application integrations and lifecycle management. Assist with Cato SASE, Meraki networking and SecureW2. Support backup and recovery technologies, including monitoring, restores and recovery testing. Contribute to Microsoft 365 administration across Exchange Online, SharePoint, Teams, security and compliance. Participate in infrastructure projects, improvements and disaster recovery testing. Security & Compliance Triage and investigate endpoint security alerts, including CrowdStrike detections. Support the implementation and enforcement ...

Lead Network Engineer

Location
Greater London, England, United Kingdom
principles, including segmentation, encryption, secure remote access, threat prevention, and security policy enforcement. Proven experience supporting 24x7 production environments with demanding availability, incident response, disaster recovery, and change-control requirements. Experience leading technical delivery through the full lifecycle, from requirements and design through implementation, testing, documentation, and operational ...

Software Engineering & Development, Vice President

Location
Greater London, England, United Kingdom
production. Manage application hosting platforms and ensure platform stability, availability, and performance. Monitor platform health and proactively identify opportunities for optimization and improvement. Support disaster recovery, resilience, and business continuity requirements. Infrastructure & Operations Work closely with infrastructure, cloud, and security teams to ensure platforms are secure, scalable ...

platform engineer in legal services

Location
Greater London, England, United Kingdom
host processes, and environment lifecycles Oversee Azure SQL Databases and Managed Instances, including availability, firewalls, private endpoints, Entra ID integration, performance monitoring, backups, and recovery readiness Manage certificates and secrets across Azure workloads using Azure Key Vault, App Service certificates, and SSL/TLS automation, including rotation, renewal … operational metrics using Azure Monitor, Log Analytics, and Application Insights Work with Infrastructure, Applications, and Security Architects to align systems with Business Continuity and Disaster Recovery guidelines and ensure documented plans are regularly reviewed and tested Act as a third-level escalation point for IT Operations, troubleshoot cloud ...

Professional Services Engineer

Hiring Organisation
Atrium Associates Ltd
Location
Greater London, United Kingdom
Employment Type
Permanent
Exchange migrations. SharePoint Online deployment and migration projects. Strong networking knowledge, including firewalls, switching, LAN and multi-VLAN environments. Experience with business continuity and disaster recovery solutions such as Datto, Veeam, StorageCraft, or Altaro. End-user device deployments, including Windows, Apple, and Android devices. Experience supporting both cloud ...

DevOps Engineer

Location
Greater London, England, United Kingdom
processes. Implement and maintain monitoring, alerting, logging, and observability solutions. Participate in technical projects including platform upgrades, migrations, and cloud transformation initiatives. Contribute to disaster recovery, business continuity, and high‐availability strategies. Create and maintain technical documentation, standards, and operational procedures. Participate in an Overnight on‐call rota. ...