
Platform Site Reliability Engineer
Posted Jul 14

Posted Jul 14
This is a fully remote position, open to applicants in Indonesia.
• Oversee the daily management of our production infrastructure.
• Diagnose issues when downtime occurs and ensure resolution.
• Develop shared, self-service tools that benefit all teams.
• Establish technical guidelines: standards, defaults, and platform choices for infrastructure.
• Implement monitoring solutions (VictoriaMetrics, Jaeger, Grafana) and create alerts.
• Administer data stores (PostgreSQL, MySQL, MongoDB, Elasticsearch, Kafka, Redis, Memcache).
• Manage infrastructure as code (Opentofu, Ansible, ArgoCD, Helm) and maintain fast and secure delivery through CI/CD (GitHub Actions, GitLab CI, CircleCI).
• 2 to 4 years of experience in SRE, DevOps, Systems Engineering, or Software Engineering.
• Proficient in developing tools and automation using Bash, Python, or Go.
• Practical experience with containerization and orchestration (Docker, Kubernetes).
• Familiarity with cloud environments (GCP or AWS; GCP is an advantage).
• Experience with relevant technologies in our stack: databases (Postgres, MySQL), caching and queues (Redis, Kafka), CI/CD, and IaC (Terraform/Opentofu, Helm, Ansible, ArgoCD).
• Strong user focus and empathy.
• 100% Remote, Work from Home.
The Codest
IRIUM
Sólides
Resilinc
Get handpicked remote jobs straight to your inbox weekly.