Operations Resilience Engineer

Job DescriptionRole Overview:The Operational Resilience Engineer supports Cboe Technology and Operations by owning and evolving the systems, automations, and processes that underpin operational risk control. This role goes beyond managing incidents — it focuses on independent delivery of end-to-end engineering solutions that make Cboe's operational environment faster, smarter, and more resilient.The Operational Resilience Engineer designs and builds automated workflows across incident management, BCP/DR, change management, and compliance evidence collection. They integrate operational systems, develop AI-assisted workflows, and create self-service tooling that reduces manual toil and improves operational metrics across Technology and Operations.In this role you will be responsible for:Maintaining a comprehensive inventory of attributes essential to our trading services, including:Business services and their impact tolerancesSupporting functions and their criticalityProcesses, sub-processes, assets, and controls underpinning these services and functionsRecovery Time Objectives (RTO) and Recovery Point Objectives (RPO) for the aboveRisk identification, operational resilience planning and testing by:Supporting functions with their business impact assessment, along with identification and articulation of technology and operational risksSupporting functions in documenting business continuity plans (BCP) aligned with RTOs, RPOs, and our operational resilience strategyIdentifying vulnerabilities, single points of failure, and interdependencies within business servicesDesigning and executing resilience testing programs, including severe but plausible desktop scenario testsDocumenting test results and implementing recommendations for improvementSupporting governance, monitoring, and reporting of operational resilience by:Preparing regulatory self-assessments for operational resiliencePreparing regular reports for senior management and board committeesMaintaining operational resilience management informationSupporting the analysis of test results and tracking/reporting remedial actionsBuilding and engineering resilience infrastructure, including:Building end-to-end automations that streamline incident lifecycle management, from detection through Learning Review and post-incident action trackingIntegrating operational systems to enable real-time data flow across incident, change, and compliance platformsDeveloping AI-assisted workflows to enrich incidents, surface risk signals, and accelerate decision-makingAutomating change management processes including risk scoring and compliance evidence collectionCreating policy validation pipelines and maintaining operational documentation and proceduresDelivering self-service tooling that empowers Technology and Operations staff to act independentlyDriving continuous improvement in operational metrics through instrumentation and observabilityA successful Operational Resilience Engineer brings knowledge in one or more critical operational risk control processes — such as incident management, BCP/DR, change management, capacity planning, or asset management — combined with strong engineering capabilities including APIs, cloud automation, CI/CD, event-driven architecture, workflow orchestration, and observability tooling.Typical deliverables include automated DR evidence collection systems, incident enrichment pipelines, change automation with risk scoring, and policy validation frameworks.This role requires strong communication, collaboration, and critical thinking skills, with the ability to operate independently and deliver complete solutions in a fast-paced, multi-faceted technical environment.The ideal candidate has:Minimum 3 years' experience in technology risk management, business continuity, or operational resilienceMinimum 2 years of demonstrated computer science, computer networking, and/or computer infrastructure related experienceMinimum Education Requirement: Bachelor's degree in Project Management, Computer Science, Software Engineering, Math, Business, Financial Services, or a related disciplineAptitude to learn our business services and systems, backed by a keen interest in technologyStrong written and verbal communication skills, including demonstrated ability to write in explanatory and procedural styles for multiple audiences, and ability to effectively lead meetings in a technical setting among multiple stakeholders with varied backgrounds and viewpointsStrong stakeholder management and influencing skillsStrong troubleshooting and problem-solving skills, and ability to work well under pressureExperience with data analysis and reporting toolsYou will really stand out with:Knowledge of PRA Policy Statement 21/3 for Building Operational Resilience and the Digital Operational Resilience Act (DORA)Experience in financial market infrastructure or trading environmentsProfessional certifications (CBCI, CISA, FRM, PRM)Experience with Atlassian Suite productsExperience responding to and/or documenting technical incidentsExperience developing AI-assisted workflowsDemonstrated leadership experience, especially in a technical settingAny communication from Cboe regarding this position will only come from a Cboe recruiter who has a @cboe.com email or via LinkedIn Recruiter. Cboe does not use any other third party communication tools for recruiting purposes.SummaryLocation: London, United KingdomType: Full time

Job Details

Company
Appcast
Location
London, UK
Posted