1. Bachelor’s Degree in Computer Science, Electrical Engineering, or related fields.
2. At least 5 years of experience in cloud infrastructure operations, server operations, or large-scale infrastructure environments.
3. Strong leadership, ownership, and decision-making capabilities in high-pressure operational environments.
4. Strong communication and cross-functional collaboration skills in English; Mandarin is preferred.
5. Deep understanding of cloud infrastructure operations, server lifecycle management, and large-scale operational ecosystems.
6. Experience working with public cloud platforms such as Oracle, Amazon Web Services, Google, or Microsoft.
7. Strong experience with automated provisioning technologies, bare metal lifecycle management, and infrastructure deployment pipelines.
8. Strong scripting and automation capabilities using Shell, Python, or infrastructure APIs.
9. Familiarity with automation and infrastructure management tools such as Terraform, Ansible, GitLab CI/CD, or cloud SDKs.
10.Strong Linux troubleshooting and infrastructure diagnostic capabilities.
11.Strong understanding of networking concepts including TCP/IP, subnetting, VLANs, DNS, IPv6, routing, and cloud networking architectures.
12.Experience with infrastructure monitoring, incident management, operational governance, and service reliability initiatives.
13.Strong documentation, workflow standardization, and operational process management capabilities.
14.Experience supporting GPU infrastructure, large-scale AI clusters, firmware lifecycle management, or RDMA networking is highly preferred.
15.Experience leading cloud operational programs, vendor management, or infrastructure transformation initiatives is preferred.