×
Register Here to Apply for Jobs or Post Jobs. X

Disaster Recovery Architect

Remote / Online - Candidates ideally in
New York, USA
Listing for: ESR Healthcare
Remote/Work from Home position
Listed on 2026-08-16
Job specializations:
  • IT/Tech
    Cloud Computing: Infrastructure & Operations, Disaster Recovery IT, IT Consultant, Systems Engineer
Job Description & How to Apply Below

Disaster Recovery Architect

Duration:
Expected to be 12 months contract

Location:

100% Remote Work

# of Positions: 1

Skills/

Experience:

Key Responsibilities:

Architectural Leadership

- Define end to end DR and high availability (HA) architectures for enterprise wide workloads, incorporating multi region cloud, hybrid, and on prem solutions.

- Develop architectural blueprints, reference designs, and pattern libraries that align with LM's security, compliance, and cost optimization policies.

Solution Design & Implementation

- Design and implement automated fail over, replication, and fail back mechanisms (e.g., Site Recovery Manager, Kubernetes based HA, database mirroring, storage level replication).

- Evaluate and integrate emerging technologies (e.g., Immutable Infrastructure, Chaos Engineering, Serverless DR) to improve resiliency and reduce mean time to recover (MTTR).

Governance & Compliance

- Ensure all DR solutions meet corporate policies CRX 301, CRX 302, and relevant regulatory requirements (e.g., NIST 800 34, ISO 22301, FedRAMP).

- Create and maintain DR documentation, run books, and test plans; conduct periodic reviews and updates.

Testing & Validation

- Lead full scale DR exercise planning, execution, and post mortem analysis for multi site, multi cloud environments.

- Define success criteria, metrics, and KPIs; report findings to senior leadership and stakeholders.

Stakeholder Collaboration

- Partner with IT Infrastructure, Cloud Engineering, Application Development, Security, and Governance teams to embed DR/HA considerations early in the SDLC.

- Serve as the technical authority for DR during design reviews (SRR, PDR, CDR, TRR) and program risk assessments.

Continuous Improvement

- Conduct risk assessments, threat modeling, and capacity planning to anticipate emerging resiliency challenges.

- Drive adoption of Model Based Systems Engineering (MBSE) and automated documentation tools to keep architecture artefacts current.

DR Plan Modernization & Compliance

- Review existing DR plan architectures across the enterprise, assessing their alignment with current resilience standards, best practices, and organizational Recovery Objectives.

- Collaborate with internal teams (Application Owners, IT Service Managers, Engineering) to update and refine DR plans, ensuring that all applications and IT services meet the latest RTO/RPO targets.

- Develop and implement remediation plans to bring legacy systems and applications up to date with modern resilience standards, ensuring compliance with corporate policies (CRX 301, CRX 302) and regulatory requirements.

- Track progress and report status to senior leadership, providing insights into plan modernization efforts and risk mitigation strategies.

Basic Qualifications :

- 5+years of experience designing and implementing DR/HA solutions for enterprise scale workloads in cloud, hybrid, and on prem environments.

- Bachelor's degree in Computer Science, Information Technology, Engineering, or a related technical discipline (or equivalent experience).

- Proven hands on experience with cloud platforms (AWS, Azure, GCP) and related services (e.g., Disaster Recovery, Site Recovery Manager, cross region replication, networking, IAM).

- Strong understanding of networking, storage, virtualization, container orchestration (Kubernetes), and database technologies as they relate to resiliency.

- Excellent written and verbal communication skills; ability to translate complex technical concepts for both technical and non technical audiences.

- U.S. citizenship required

Desired skills :

Familiarity with automated disaster recovery (DR) solutions, including but not limited to:

Amazon Web Services (AWS) Disaster Recovery Service (DRS):
Experience with configuring and managing replication, fail over, and fail back processes for AWS workloads.

Microsoft Azure Site Recovery (ASR):
Knowledge of setting up and managing site recovery between on premises environments, Azure, and other clouds.

Zerto:
Hands on experience with continuous data protection (CDP) and near zero RPO replication across VMware, Hyper V, and cloud environments.

Veeam Backup & Replication:
Experience with agent less backup, replication,…

To View & Apply for jobs on this site that accept applications from your location or country, tap the button below to make a Search.
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).
 
 
 
Search for further Jobs Here:
(Try combinations for better Results! Or enter less keywords for broader Results)
Location
Increase/decrease your Search Radius (miles)
0
200
Filters
Education Level
Experience Level (years)
Posted in last:
Salary