DevOps Engineer, Infra

Posted 1 day ago

This is a fully remote position, open to applicants in Singapore.

📋 Description

• Design, implement, and uphold backend infrastructure to guarantee high availability, scalability, and reliability of production systems.

• Oversee cloud infrastructure on AWS and AliCloud, focusing on performance enhancement, cost management, and operational excellence.

• Deploy, monitor, and sustain Kubernetes clusters, distributed backend systems, and associated infrastructure services.

• Create and manage CI/CD pipelines and infrastructure automation utilizing GitHub Actions, Terraform, and Ansible.

• Work collaboratively with software engineers to facilitate backend service deployment, troubleshooting, and production operations.

• Investigate production incidents, conduct root cause analysis, and implement preventive enhancements.

• Apply infrastructure security best practices to ensure secure and dependable production environments.

• Develop internal DevOps platforms and automation tools aimed at boosting engineering productivity and operational efficiency.

• Establish monitoring, observability, and alerting solutions to improve service reliability.

• Explore and incorporate AI technologies into infrastructure operations, including intelligent alert analysis, ChatOps, and operational automation.


⛳️ Requirements

• Over 5 years of practical experience in Kafka and Redis operations within large-scale production settings, with the ability to collaborate with developers for code optimization.

• Proficient in Python, Go, or Java (at least one language) and SQL programming languages.

• Direct experience with containerization and orchestration technologies (Docker, Kubernetes).

• Significant experience with CI/CD tools such as GitHub Actions, Ansible, Terraform, etc.

• Minimum of 3 years of experience with the AWS cloud platform.

• Experience with GCP, Azure, or Ali Cloud is a plus.

• Exceptional problem-solving and troubleshooting abilities.

• Strong collaboration skills and the ability to develop partnerships with other teams and business units.

• Practical experience in building or managing AIOps systems (anomaly detection, alert correlation, automated healing, or RCA).

• Familiarity with LLM-based DevOps automation (e.g., developing chat-based ops assistants or AI-driven observability workflows).

• Experience in utilizing or integrating tools like Dify, Agno, or LangChain into operational workflows.


🏝️ Benefits

• Competitive salary and comprehensive company benefits.

• Flexible work-from-home arrangements (subject to the nature of the business team's work).

• Opportunities for career advancement and ongoing learning.

• Collaborate with top-tier talent in a user-focused global organization characterized by a flat structure.

• Engage in distinctive, fast-paced projects with autonomy in an innovative environment.

• Results-oriented workplace.

People also viewed

In All Media22 hours ago

DevOps Engineer – Cloud

BR flagBrazil, +5 more countriesFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Verity Group23 hours ago

SRE Engineer

BR flagBrazil OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Fingerprint1 day ago

Senior Site Reliability Engineer

US flagUnited States OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$152k – $205k/year
ApplyView job
Endava1 day ago

Senior DevOps Engineer, Terraform

IN flagIndia OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
CVS Health1 day ago

Staff DevSecOps Engineer, Health

US flagConnecticut, +3 more statesFull-timeDevOps & Site Reliability Engineer (SRE)$130.3k – $260.6k/year
ApplyView job
GoFasti1 day ago

Senior DevOps Engineer

Latin AmericaFull-timeDevOps & Site Reliability Engineer (SRE)$5,000 – $6,000/month
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers