×
Register Here to Apply for Jobs or Post Jobs. X

Operations Resilience Engineer

Job in London, Greater London, W1B, England, UK
Listing for: Cboe Exchange
Full Time position
Listed on 2026-08-15
Job specializations:
  • IT/Tech
    SRE/Site Reliability, Cybersecurity, IT Specialist, Disaster Recovery IT
Job Description & How to Apply Below
Job Description

Role Overview :

The Operational Resilience Engineer supports Cboe Technology and Operations by owning and evolving the systems, automations, and processes that underpin operational risk control. This role goes beyond managing incidents — it focuses on independent delivery of end-to-end engineering solutions that make Cboe's operational environment faster, smarter, and more resilient.

The Operational Resilience Engineer designs and builds automated workflows across incident management, BCP/DR, change management, and compliance evidence collection. They integrate operational systems, develop AI-assisted workflows, and create self-service tooling that reduces manual toil and improves operational metrics across Technology and  this role you will be responsible for:

Maintaining a comprehensive inventory of attributes essential to our trading services, including:

Business services and their impact tolerances

Supporting functions and their criticality

Processes, sub-processes, assets, and controls underpinning these services and functions

Recovery Time Objectives (RTO) and Recovery Point Objectives (RPO) for the above

Risk identification, operational resilience planning and testing by:

Supporting functions with their business impact assessment, along with identification and articulation of technology and operational risks

Supporting functions in documenting business continuity plans (BCP) aligned with RTOs, RPOs, and our operational resilience strategy

Identifying vulnerabilities, single points of failure, and interdependencies within business services

Designing and executing resilience testing programs, including severe but plausible desktop scenario tests

Documenting test results and implementing recommendations for improvement

Supporting governance, monitoring, and reporting of operational resilience by:

Preparing regulatory self-assessments for operational resilience

Preparing regular reports for senior management and board committees

Maintaining operational resilience management information

Supporting the analysis of test results and tracking/reporting remedial actions

Building and engineering resilience infrastructure, including:

Building end-to-end automations that streamline incident lifecycle management, from detection through Learning Review and post-incident action tracking

Integrating operational systems to enable real-time data flow across incident, change, and compliance platforms

Developing AI-assisted workflows to enrich incidents, surface risk signals, and accelerate decision-making

Automating change management processes including risk scoring and compliance evidence collection

Creating policy validation pipelines and maintaining operational documentation and procedures

Delivering self-service tooling that empowers Technology and Operations staff to act independently

Driving continuous improvement in operational metrics through instrumentation and observabilityA successful Operational Resilience Engineer brings knowledge in one or more critical operational risk control processes — such as incident management, BCP/DR, change management, capacity planning, or asset management — combined with strong engineering capabilities including APIs, cloud automation, CI/CD, event-driven architecture, workflow orchestration, and observability tooling.

Typical deliverables include automated DR evidence collection systems, incident enrichment pipelines, change automation with risk scoring, and policy validation frameworks.

This role requires strong communication, collaboration, and critical thinking skills, with the ability to operate independently and deliver complete solutions in a fast-paced, multi-faceted technical environment.

The ideal candidate has:

Minimum 3 years' experience in technology risk management, business continuity, or operational resilience

Minimum 2 years of demonstrated computer science, computer networking, and/or computer infrastructure related experience

Minimum Education Requirement:
Bachelor's degree in Project Management, Computer Science, Software Engineering, Math, Business, Financial Services, or a related discipline

Aptitude to learn our business services and systems, backed by a…
Note that applications are not being accepted from your jurisdiction for this job currently via this jobsite. Candidate preferences are the decision of the Employer or Recruiting Agent, and are controlled by them alone.
To Search, View & Apply for jobs on this site that accept applications from your location or country, tap here to make a Search:
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary