Operational Resilience Engineer

cboe

Greater London

On-site

GBP 90,000 - 120,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

The Cboe team seeks an Operational Resilience Engineer to own and evolve the systems and automations underpinning risk controls. You will design end-to-end resilient solutions across incident management, BCP/DR, change management, and compliance evidence collection, integrating systems and delivering AI-assisted workflows.

Minimum 3+ years in technology risk management or operational resilience, with strong communication and collaboration skills for a fast-paced trading environment.

Qualifications

  • Bachelor’s degree in a relevant field.
  • 3+ years in technology risk management, business continuity, or operational resilience.
  • Strong communication and collaboration skills.
  • Experience in financial market infrastructure or trading environments.

Responsibilities

  • Own and evolve systems, automations, and processes for operational resilience.
  • Design and build automated workflows across incident management, BCP/DR, change management, and compliance evidence collection.
  • Integrate operational systems and develop AI-assisted workflows.
  • Deliver self-service tooling and reduce manual toil.
  • Prepare regulatory self-assessments and reports for senior management.

Skills

APIs
Cloud automation
CI/CD
Event-driven architecture
Workflow orchestration
Observability tooling
Incident management
Data analysis
Reporting tools

Education

Bachelor's degree in Project Management, Computer Science, Software Engineering, Math, Business, Financial Services, or a related discipline

Tools

Atlassian Suite

Job description

  • The Operational Resilience Engineer supports Cboe Technology and Operations by owning and evolving the systems, automations, and processes that underpin operational risk control
  • This role goes beyond managing incidents — it focuses on independent delivery of end-to-end engineering solutions that make Cboe’s operational environment faster, smarter, and more resilient
  • The Operational Resilience Engineer designs and builds automated workflows across incident management, BCP/DR, change management, and compliance evidence collection
  • They integrate operational systems, develop AI-assisted workflows, and create self-service tooling that reduces manual toil and improves operational metrics across Technology and Operations
  • Maintaining a comprehensive inventory of attributes essential to our trading services, including:
  • Business services and their impact tolerances
  • Supporting functions and their criticality
  • Processes, sub-processes, assets, and controls underpinning these services and functions
  • Recovery Time Objectives (RTO) and Recovery Point Objectives (RPO) for the above
  • Risk identification, operational resilience planning and testing by:
  • Supporting functions with their business impact assessment, along with identification and articulation of technology and operational risks
  • Supporting functions in documenting business continuity plans (BCP) aligned with RTOs, RPOs, and our operational resilience strategy
  • Identifying vulnerabilities, single points of failure, and interdependencies within business services
  • Designing and executing resilience testing programs, including severe but plausible desktop scenario tests
  • Documenting test results and implementing recommendations for improvement
  • Supporting governance, monitoring, and reporting of operational resilience by:
  • Preparing regulatory self-assessments for operational resilience
  • Preparing regular reports for senior management and board committees
  • Maintaining operational resilience management information
  • Supporting the analysis of test results and tracking/reporting remedial actions
  • Building and engineering resilience infrastructure, including:
  • Building end-to-end automations that streamline incident lifecycle management, from detection through Learning Review and post-incident action tracking
  • Integrating operational systems to enable real-time data flow across incident, change, and compliance platforms
  • Developing AI-assisted workflows to enrich incidents, surface risk signals, and accelerate decision-making
  • Automating change management processes including risk scoring and compliance evidence collection
  • Creating policy validation pipelines and maintaining operational documentation and procedures
  • Delivering self-service tooling that empowers Technology and Operations staff to act independently
  • Driving continuous improvement in operational metrics through instrumentation and observability
  • A successful Operational Resilience Engineer brings knowledge in one or more critical operational risk control processes — such as incident management, BCP/DR, change management, capacity planning, or asset management — combined with strong engineering capabilities including APIs, cloud automation, CI/CD, event-driven architecture, workflow orchestration, and observability tooling
  • Typical deliverables include automated DR evidence collection systems, incident enrichment pipelines, change automation with risk scoring, and policy validation frameworks

This role requires strong communication, collaboration, and critical thinking skills, with the ability to operate independently and deliver complete solutions in a fast-paced, multi-faceted technical environmentAptitude to learn our business services and systems, backed by a keen interest in technologyStrong written and verbal communication skills, including demonstrated ability to write in explanatory and procedural styles for multiple audiences, and ability to effectively lead meetings in a technical setting among multiple stakeholders with varied backgrounds and viewpointsStrong stakeholder management and influencing skillsExperience with data analysis and reporting toolsMinimum Education Requirement: Bachelor’s degree in Project Management, Computer Science, Software Engineering, Math, Business, Financial Services, or a related disciplineMinimum 2 years of demonstrated computer science, computer networking, and/or computer infrastructure related experienceMinimum 3 years’ experience in technology risk management, business continuity, or operational resilienceStrong troubleshooting and problem-solving skills, and ability to work well under pressureProfessional certifications (CBCI, CISA, FRM, PRM)Knowledge of PRA Policy Statement 21/3 for Building Operational Resilience and the Digital Operational Resilience Act (DORA)Experience in financial market infrastructure or trading environmentsExperience with Atlassian Suite productsExperience responding to and/or documenting technical incidentsExperience developing AI-assisted workflowsDemonstrated leadership experience, especially in a technical setting

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Operational Resilience Engineer: AI-Driven Automation
Operational Resilience Engineer: AI-Driven Automation

cboe • Greater London

On-site
GBP 90,000 - 120,000
Senior Executive Operational Resilience
Senior Executive Operational Resilience

iFAST Global Bank Limited • Greater London

On-site
GBP 90,000 - 130,000
25 days annual leave entitlement plus
Pension scheme, 4% employer contrib.
Private Medical Insurance
+3
Senior Executive Operational Resilience
Senior Executive Operational Resilience

iFAST Global Bank Ltd • Greater London

On-site
GBP 60,000 - 90,000
25 days annual leave
Pension scheme
Private Medical Insurance
+3
Operational Resilience Manager
Operational Resilience Manager

Brown & Brown UK • City Of London

Hybrid
GBP 70,000 - 120,000
Operational Resilience Manager
Operational Resilience Manager

Brown & Brown UK • Greater London

Hybrid
GBP 80,000 - 110,000
Operational Resilience Manager
Operational Resilience Manager

i-confidential Limited • Greater London

On-site
GBP 80,000 - 110,000
Operational Resilience Lead
Operational Resilience Lead

Oliver James • City Of London

On-site
GBP 70,000 - 90,000
Operational Resilience Manager
Operational Resilience Manager

TRIA • Slough

On-site
GBP 77,000 - 94,000
Operational Resilience & BCP Manager
Operational Resilience & BCP Manager

Edenbrook • Greater London

On-site
GBP 70,000 - 90,000
Operational Resilience Manager, Financial Services, Fully remote
Operational Resilience Manager, Financial Services, Fully remote

FDO CONSULTING • England

On-site
GBP 75,000 - 85,000
Bonus
Benefits