You will work remotely in a hybrid team model (Local Hong Kong coordination + remote execution), contributing to daily ops,incident resolution, upgrades, testing, and 24x7 rotation (with on-call premiums).
Key Responsibilities
Daily System Support & Multi-Cloud Monitoring
Perform daily health checks (post-6pm HKT scripts for non-business hours) across on-prem VMware and cloud environments.Evaluate/optimize monitoring triggers, alerts, and scripts (vSphere alarms, Horizon dashboards, Azure Monitor, AWS CloudWatch,Alibaba Cloud Monitor).
Triage alerts: initial diagnosis, escalation, and resolution support.
Incident & Problem Management
Handle incidents (S1/S2 priority) with impact analysis and root-cause resolution within SLAs.Participate in on-call rotation (Airport 24x7; premium for standby/nights/holidays).
Document and improve processes for different scenarios
Upgrades, Patching & Multi-Cloud Operations
Assess and apply patches/updates for VMware stack + cloud services (Azure VMs, AWS EC2, Alibaba ECS or servers).Execute after-hours upgrades/standby (snapshots, rollbacks, verification).Support hybrid setups if needed (e.g., Azure Arc, AWS Outposts, Alibaba hybrid cloud). (Not a must)
Infrastructure as Code (IaC) & Automation
Implement/maintain IaC for infrastructure provisioning and configuration.Automate repetitive tasks (deployments, backups, Horizon pool management, cloud resource scaling).
Regression Testing & CI/CD Support
Execute and maintain regression test scripts .Create/modify tests for platform/software changes across on-prem and cloud.
Security, Compliance & Platform Awareness
Monitor vendor/security advisories for VMware, Azure, AWS, Alibaba Cloud.Recommend/implement patches, IAM best practices, and security configurations.Assist with account/user management and compliance (PDPO alignment).
Collaboration & Knowledge Sharing
Work with local HK team and software vendors.
Maintain runbooks, knowledge base, and change request impact analysis.
Required Skills & Experience
Must-Have
3+ years hands-on DevOps/infra ops experience.
Strong VMware expertise: vSphere/vCenter/ESXi administration, Horizon VDI (pool provisioning, golden images, troubleshooting for 20–50 users), PowerCLI scripting.
Experience in government/critical infrastructure (transport/aviation preferred).
Tell employers what skills you have
Problem and Incident Management VMware Infrastructure Process Improvement Monitoring servers Verification Patch Management EC2 Incident Handling Server Management Hybrid Cloud Early Diagnosis