Cloud Site Reliability Engineer

Ambition Hong Kong
Full time Contract On-site Negotiable

About the job

Our client is seeking a Cloud Site Reliability Engineer (Cloud SRE) to support a critical data center exit and Disaster Recovery (DR) migration initiative. This role will be responsible for ensuring the continuity, resiliency, and recoverability of business-critical applications by planning, migrating, validating, and operationalizing DR and backup environments.

Key Responsibilities

  • Plan and execute DR and backup environment migrations.
  • Support the relocation and re-platforming of DR environments for applications, databases, and related services.
  • Conduct and document DR testing, recovery validation, and failover exercises.
  • Support cloud and hybrid infrastructure environments associated with backup and recovery solutions.
  • Coordinate with infrastructure, application, security teams, and external vendors throughout migration and testing activities.
  • Maintain operational documentation, runbooks, and test outcomes.

Requirements

  • Experience in Site Reliability Engineering (SRE), Cloud Operations, or Infrastructure Reliability roles.
  • Strong hands-on experience with Disaster Recovery, Backup & Recovery, and Infrastructure Resiliency.
  • Solid understanding of DR architecture, business continuity, recovery testing, and failover validation.
  • Experience supporting cloud and hybrid infrastructure environments.
  • Proven ability to deliver complex, time-sensitive infrastructure migration projects within enterprise or regulated environments.
  • Strong stakeholder management, documentation, and vendor coordination skills.

Preferred

  • Data center migration or colocation exit experience.
  • Experience with cloud-based backup and recovery platforms.
  • Financial services or other regulated industry experience.
  • Strong documentation and operational governance skills.