IT 灾难恢复专员
IT Disaster Recovery Specialist
拥有75年的经验,我们的重点是帮助最弱势的儿童克服贫困,体验充实的生活。我们帮助来自各种背景的儿童,即使是在最危险的地方,这源于我们的基督教信仰。
加入我们的31,000多名员工,在近100个国家工作,分享改变弱势儿童人生故事的喜悦!
关键职责:
重要信息:
- 所有简历必须用英文提交。
- 该职位仅对World Vision International在法律上允许运营的国家的候选人开放。
该职位在WVI全球技术环境中提供技术领导力,负责设计、实施并持续改进企业级灾难恢复能力。IT灾难恢复专家确保关键系统、云平台和应用程序具有弹性和可恢复性,符合定义的RTO和RPO目标。该职位推动灾难恢复框架、测试计划和恢复实践的采用,以减少业务中断并增强组织韧性。
资格要求:
- 信息技术、计算机科学、工程或相关领域的学士学位 • ITIL基础(最低要求);ITIL中级或ITIL 4管理专业为优势
- 具有云平台(Azure、AWS)或灾难恢复/业务连续性(例如DRII、CBCI)的相关认证者优先考虑
- 5年以上IT灾难恢复、基础设施运维或业务连续性相关工作经验
- 在云(Azure/AWS)、混合和本地环境中,具备设计和实施企业级灾难恢复策略的实际经验
- 具有灾难恢复技术(如Azure Site Recovery、AWS DR模式、备份/复制工具)的实际操作经验
- 具有定义和实施关键服务的RTO/RPO经验
- 具有领导灾难恢复测试(故障转移、桌面演练、实际恢复演练)和恢复执行的经验
- 对混合环境(云+本地基础设施)有了解
- 具有将灾难恢复与ITSM流程(事件、重大事件、变更、问题管理)集成的经验。
根据雇佣国家的不同,该职位可享受远程或混合办公。需要与全球团队持续协作,跨越多个时区。如需,该职位需要具备并愿意进行国内和国际出差。
技术与功能技能
- 对灾难恢复有深入理解
查看英文原文
With 75 years of experience, our focus is on helping the most vulnerable children overcome poverty and experience fullness of life. We help children of all backgrounds, even in the most dangerous places, inspired by our Christian faith.
Come join our 31,000+ staff working in nearly 100 countries and share the joy of transforming vulnerable children’s life stories!
Key Responsibilities:
IMPORTANT INFORMATION:
- All CVs should be submitted in English.
- This position is open to candidates based in countries where World Vision International is legally registered to operate.
This role provides technical leadership in designing, implementing, and continuously improving enterprise disaster recovery capabilities across WVI’s global technology environment. The IT Disaster Recovery Specialist ensures that critical systems, cloud platforms, and applications are resilient and recoverable, aligned to defined RTO and RPO targets. The role drives adoption of DR frameworks, testing programmes, and recovery practices to minimize business disruption and strengthen organisational resilience.
QUALIFICATIONS:
- Bachelor’s degree in information technology, Computer Science, Engineering, or related field • ITIL Foundation (minimum); ITIL Intermediate or ITIL 4 Managing Professional is an advantage
- Relevant certifications in cloud platforms (Azure, AWS) or disaster recovery / business continuity (e.g., DRII, CBCI) are desirable
- 5+ years’ experience in IT Disaster Recovery, Infrastructure Operations, or Business Continuity roles
- Proven experience designing and implementing enterprise disaster recovery strategies across cloud (Azure/AWS), hybrid, and on-prem environments
- Hands-on experience with DR technologies (e.g., Azure Site Recovery, AWS DR patterns, backup/replication tools)
- Experience defining and operationalizing RTO / RPO for critical services
- Experience leading DR testing (failover, tabletop, live recovery drills) and recovery execution
- Exposure to hybrid environments (cloud + on-prem infrastructure)
- Experience integrating DR with ITSM processes (Incident, Major Incident, Change, Problem Management).
This position is eligible for Remote or Hybrid-Work based dependent on the country of hire. It involves continuous collaboration with global teams across various time zones. The position requires ability and willingness to travel domestically and internationally if needed.
Technical & Functional Skills
- Strong understanding of disaster recovery frameworks and standards (e.g., ITIL, ISO 22301, NIST)
- Experience designing DR architectures including failover, geo-redundancy, and backup strategies
- Knowledge of infrastructure resilience across compute, storage, network, and identity layers
- Experience with automation and infrastructure-as-code (e.g., Terraform) for repeatable recovery environments
- Familiarity with cybersecurity incident response and ransomware recovery integration
- Experience with dashboards, reporting, and DR readiness tracking
Core Competencies:
- Strong analytical thinking with ability to assess recovery risk and design mitigation strategies
- Structured, process-driven mindset with focus on governance, documentation, and audit readiness
- Strong collaboration across infrastructure, cloud, security, and application teams
- Ability to lead calmly and decisively during incident recovery scenarios
- Continuous improvement mindset with focus on resilience, readiness, and operational maturity
- Customer-focused approach, ensuring minimal business disruption and reliable service recovery
CONTINUATION OF MAJOR RESPONSIBILITIES:
DR Strategy, Framework & Governance
- Define and maintain enterprise DR policies, standards, and governance framework
- Establish and refine RTO/RPO targets aligned to business criticality
- Ensure DR documentation is audit-ready and integrated with ITSM processes
- Align DR practices with business continuity strategy
DR Architecture & Implementation
- Design DR solutions across cloud (Azure/AWS), hybrid, and on-prem environments
- Implement failover, geo-redundancy, and backup strategies
- Embed DR requirements into solution architecture and project delivery
- Evaluate DR tools and vendor capabilities.
Testing, Exercising & Validation This includes but not limited to the following:
- Plan and execute DR drills, failover tests, and tabletop exercises
- Validate RTO/RPO compliance for critical systems
- Document outcomes and track remediation actions
- Drive continuous improvement based on test results.
Incident Recovery & Coordination
- Act as technical authority during major incidents and recovery scenarios
- Coordinate recovery across infrastructure, cloud, and application teams
- Maintain DR runbooks and recovery playbooks
- Support PIR and integrate lessons learned.
Stakeholder Enablement & Reporting
- Train teams on DR practices and readiness
- Engage business stakeholders on recovery requirements
- Maintain DR dashboards and reporting
- Collaborate with security, cloud, and platform teams.
Applicant Types Accepted:
Local Applicants OnlyOriginally posted on Himalayas