DevOps Platform Eng/SWE, Infra (Mandarin Required)
Location: Palo Alto, California
Fluent Mandarin Chinese and professional English are mandatory. Candidates must be able to conduct technical discussions and collaborate effectively in Mandarin with engineering teams in China, while also being comfortable operating professionally in English.
About The Role
Our client is a rapidly scaling global consumer internet platform serving hundreds of millions of users worldwide. The position is for a Platform / DevOps Engineer with strong software engineering and infrastructure fundamentals. The role focuses on building and scaling engineering platforms to support the company’s international expansion. This is not a traditional IT DevOps or pipeline maintenance role; instead, it involves working on large-scale engineering infrastructure, including CI/CD, Kubernetes, cloud infrastructure, Infrastructure as Code, multi-region deployment, High Availability, and Disaster Recovery. You will help extend and evolve engineering platforms originally built at massive consumer-internet scale into global cloud and data-centre environments.
Responsibilities
- Build and scale global DevOps and engineering platforms.
- Develop and improve CI/CD, release, and change-management systems.
- Build Kubernetes and cloud-native infrastructure platforms.
- Develop infrastructure automation and Infrastructure as Code (IaC) capabilities.
- Design deployment solutions across public cloud, private cloud, and hybrid environments.
- Build and improve multi-region deployment infrastructure.
- Design High Availability, Disaster Recovery, and automated failover capabilities.
- Conduct resilience testing and Chaos Engineering.
- Troubleshoot complex production and infrastructure issues and drive Root Cause Analysis (RCA).
- Improve platform observability, reliability, and operational efficiency.
The underlying platform spans networking, storage, Kubernetes, security, and compliance across different infrastructure environments.
Requirements
- 3-7 years of experience in Platform Engineering, DevOps, SRE, or Cloud Infrastructure Engineering.
- Hands-on experience building or operating large-scale production infrastructure.
- Strong Linux, networking, and infrastructure fundamentals.
- Strong hands-on experience with Kubernetes.
- Experience with Terraform, Helm, or other Infrastructure as Code technologies.
- Programming experience in Go, Java, Python, and/or Shell, with the ability to build automation tooling and debug backend services.
- Experience with at least one major cloud platform: AWS, GCP, Azure, or enterprise private cloud.
- Understanding of multi-region deployment, networking, IAM, and security isolation.
- Familiarity with observability technologies such as Prometheus, Grafana, ELK/OpenSearch, and OpenTelemetry.
- Strong troubleshooting and production engineering capabilities.
- Ability to independently own and drive complex infrastructure initiatives.
Particularly Relevant Backgrounds
- Engineers who have built or operated internal engineering platforms at large-scale technology or consumer internet companies, especially in:
- Platform Engineering / Developer Infrastructure
- DevOps Platforms
- CI/CD and Release Engineering
- Kubernetes Platforms
- Cloud Infrastructure
- Infrastructure Automation
- Site Reliability Engineering
- Global Infrastructure Deployment
- High Availability / Disaster Recovery
Nice To Have
- Experience with international infrastructure deployment, Disaster Recovery, Chaos Engineering, capacity planning, large-scale incident management, or regional infrastructure/data-residency requirements.
Why This Role
- Opportunity to work on engineering infrastructure at genuine consumer-internet scale.
- Solve the additional complexity of taking mature internal platforms global.
- Work across software engineering, Kubernetes, cloud infrastructure, distributed deployment, and HA/DR, rather than being limited to traditional DevOps operations.