Remotery

DevOps/SRE Engineer, Bilingual Mandarin

Posted 6 hours ago

This is a fully remote position, open to applicants in California, +4 more states.

📋 Description

• Carry out daily operations for cloud resources and infrastructure as part of the operations and monitoring framework established by the team in China.

• Conduct daily monitoring, alerting, backup, and recovery tasks in accordance with standard operating procedures.

• Support first-line alert triage and coordinate with the China-based team for incident response across time zones.

• Execute initial troubleshooting and first-response actions prior to transferring issues to the China-based team.

• Manage the daily operations of local US data centers and/or cloud resources while ensuring compliance with relevant requirements.

• Implement local regulations such as access controls for privacy data and data localization/storage.

• Act as the first responder for incidents impacting systems and services in North America.

• Independently address common incidents and collaborate with the China-based team on more complex issues.

• Oversee execution and liaison activities related to GDPR, SOC2, CCPA, EO 14117, US local laws, legal compliance, and third-party audit obligations.


⛳️ Requirements

• Proficient in Mandarin Chinese with skills in listening, speaking, reading, and writing.

• Authorized to work for any employer; sponsorship is not currently available.

• Bachelor's degree in Computer Science or a related discipline.

• 3–4 years of practical experience in DevOps, SRE, or Platform Engineering roles.

• Extensive experience with at least one major cloud service provider (AWS, Azure, or GCP), covering VPC, EC2, EKS/Kubernetes, RDS, and IAM.

• In-depth understanding of Linux systems, networking basics, containers (Docker, Kubernetes), load balancing, and service governance.

• Skilled in Infrastructure as Code tools such as Terraform, Ansible, and Helm.

• Experience in developing and maintaining CI/CD pipelines using Jenkins, Argo CD, CodeBuild, or similar tools.

• Practical experience with monitoring, logging, and tracing systems, including Prometheus, Grafana, ELK Stack, OpenTelemetry, or equivalent solutions.

• Proficient in at least one scripting or programming language such as Python, Shell, or Go.

• Strong system design abilities, analytical thinking, and advanced troubleshooting skills.

• Excellent cross-team communication skills; experience in technical knowledge sharing or evangelism is advantageous.


🏝️ Benefits

• 401(k)

• PTO

• Paid Holidays

• Insurance (Medical+Vision)

People also viewed

a378 hours ago

Forward Deployed DevOps Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
GT8 hours ago

Site Reliability Engineer, SRE

PL flagPoland, +2 more statesFreelanceDevOps & Site Reliability Engineer (SRE)
ApplyView job
Sigma Software Group8 hours ago

DevOps Engineer

PL flagPoland OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Applaudo9 hours ago

Google Cloud DevOps Engineer – Temporary Contract

CO flagColombia OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Branch9 hours ago

Cloud Operations Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$135k – $150k/year
ApplyView job
Rimutee10 hours ago

DevOps / SRE, Part Time

Latin AmericaPart-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers