Cloud Platform Lead - Disaster Recovery

Charlotte, NC, US • Posted 2 days ago • Updated 4 hours ago
Full Time
On-site
Fitment

Dice Job Match Score™

🤯 Applying directly to the forehead...

Job Details

Skills

  • Finance
  • Workflow
  • Provisioning
  • Configuration Management
  • Continuous Improvement
  • Scripting
  • Dashboard
  • Business Continuity Planning
  • Financial Services
  • Computer Science
  • Information Systems
  • Amazon Web Services
  • Recovery
  • Cloud Computing
  • Disaster Recovery
  • Terraform
  • Ansible
  • Progress Chef
  • Puppet
  • Continuous Integration
  • Continuous Delivery
  • CHAOS
  • IT Management
  • Communication
  • Testing
  • Reporting
  • Python
  • JavaScript
  • ISACA
  • CISA

Summary

Cloud Platform Lead - Disaster Recovery

Charlotte, NC - Relocation Assistance Provided

Hybrid Work Schedule - 3 Days Per Week In-Office

The Company

Our client is one of the largest and most trusted financial companies in the world, focusing on improving people's lives through successful investing.

The Job

As a Cloud Platform Lead - Disaster recovery, you will be responsible for the operational side of disaster recovery and resilience. You are going to be developing, implementing, and maintaining resiliency framework and capabilities, that applications teams can consume via automated product offerings or repeatable patterns to attest to validity and viability of their disaster recovery plans in line with business outcomes.

Functions
  • You will partner with infrastructure and application teams to design and implement scripts, templates, and workflows that automate their product's disaster recovery. This includes automation for all relevant resiliency elements including disaster recovery provisioning and scaling, configuration management, monitoring and observability, resyncing and reconciliation, and testing.
  • You will work and partner closely with the project managers, technical leads, and business stakeholders to identify testing scenarios for potential threats, assess impacts, and design testing solutions to ensure business continuity and minimize risks.
  • You will perform detailed evaluations of platform and application resiliency readiness to identify areas of concern.
  • You will conduct regular testing, monitoring, and reporting of the resiliency and disaster recovery plans and activities. You will develop the capability to capture the book of record for all disaster recovery related data. You will identify gaps and continuous improvement opportunities.
  • You can design and implement data collecting scripts, implement and maintain monitoring tools, and develop front-end dashboards to monitor the health, performance, and utilization of the recovery environment to enable prompt response when signs dictate.
  • You will support Global Risk and their requirements to report to regulators on our client's disaster recovery effort.

Qualifications
  • 7+ years of hands-on experience in resiliency, disaster recovery, or business continuity for midsize to large enterprises, with proven technical leadership delivering enterprise-scale DR and resiliency solutions, preferably in regulated or financial services environments. You have a bachelor's degree in computer science, information systems, engineering, or a related field.
  • Strong AWS platform engineering expertise, including hands-on experience with AWS Resiliency Hub and AWS Fault Injector Service, and the ability to design, implement, validate, and operationalize AWS first resiliency and recovery strategies across cloud, hybrid, and on prem environments.
  • Demonstrated ability to design and evolve enterprise disaster recovery and resiliency frameworks, delivering repeatable, automated patterns and scalable solutions consumable by application teams.
  • Deep experience with Infrastructure as Code (IaC) and automation first delivery, using tools such as Terraform, Ansible, Chef, or Puppet, along with strong knowledge of CI/CD principles, resiliency frameworks, DR testing strategies, chaos engineering, and risk based analysis.
  • Proven technical leadership and communication skills, with the ability to lead cross-functional delivery, influence without authority, conduct DR and resiliency testing, support regulatory and risk reporting, and clearly communicate outcomes to both technical and non-technical stakeholders.
  • Coding experience (e.g., Python, JavaScript) and relevant certifications (CBCP, CRISC, CISA) are a plus.
Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: 90986595
  • Position Id: dc4060482609372dc3d9627a232c5225
  • Posted 2 days ago
Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

Charlotte, North Carolina

Today

Easy Apply

Full-time

$60 - $70 per hour

Hybrid in Charlotte, North Carolina

Yesterday

Easy Apply

Contract

50

Charlotte, North Carolina

Today

Easy Apply

Contract

Remote

Today

Full-time

Search all similar jobs