Remotery

Site Reliability Engineer – APAC

Posted Jun 20

This is a fully remote position, open to applicants in Malaysia.

📋 Description

• Oversee the health and performance metrics of the platform.

• Address production incidents and guide them to resolution.

• Analyze failures, pinpoint root causes, and coordinate solutions.

• Ensure that issues are identified, comprehended, and resolved promptly.

• Detect recurring operational challenges and work to eliminate them.

• Enhance software, deployment procedures, and operational workflows.

• Engage in incident reviews and contribute to preventative enhancements.

• Implement reliability-focused modifications directly within production systems.

• Develop and oversee dashboards, metrics, alerting, and monitoring frameworks.

• Enhance signal quality while minimizing alert fatigue.

• Create automation and internal tools that facilitate easier platform operations.

• Assist in establishing reliability best practices throughout the engineering team.


⛳️ Requirements

• Extensive experience with Linux and cloud infrastructure.

• Proven experience in operating and supporting production systems.

• Familiarity with Docker and containerized environments.

• Experience with observability and incident-management tools such as Grafana, Prometheus, PagerDuty, or similar solutions.

• Proficiency in automating workflows using Rust, Python, Bash, or similar programming languages.

• Strong skills in troubleshooting and debugging.

• A high level of ownership and the capacity to make sound independent decisions.

• Nice to Have: Experience with distributed systems, high-availability and low-latency services, CI/CD systems, deployment automation, and designing secure operational workflows and access controls.


🏝️ Benefits

• Competitive salary (~$100k USD/year).

• Meaningful token/equity allocation.

• Genuine ownership and responsibility from day one.

• Flexibility to work from any location within the target timezone range (UTC+7 to UTC+1).

• Opportunities for occasional travel to Europe and beyond for team gatherings.

People also viewed

Ontrac Solutions2 days ago

Site Reliability Engineer

PK flagPakistan OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
CyberSheath2 days ago

Cloud Operations Engineer

US flagVirginia OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$110k – $127k/year
ApplyView job
Ontrac Solutions2 days ago

Site Reliability Engineer

IN flagIndia OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
NVIDIA2 days ago

Service Reliability Engineer

US flagTexas OnlyFull-timeDevOps & Site Reliability Engineer (SRE)$168k – $333.5k/year
ApplyView job
Nagarro2 days ago

Senior Site Reliability Engineer, AWS Cloud

RO flagRomania OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Capgemini2 days ago

Senior DevOps Engineer

UA flagUkraine OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers