Remotery

Site Reliability Engineer

Posted Jul 17

This is a fully remote position, open to applicants in Singapore.

📋 Description

• Design, construct, and sustain systems that are highly available, scalable, and fault-tolerant.

• Collaborate with software engineering teams to ensure applications are crafted with reliability and performance as priorities.

• Develop and uphold automation procedures to enhance system efficiency, reduce human intervention, and streamline routine tasks.

• Monitor and assess system performance to pinpoint and resolve bottlenecks before they affect users.

• Ensure the infrastructure can accommodate rapid increases in web traffic and machine learning data processing.

• Participate in 24/7 on-call rotations, which include scheduled shifts and holidays.

• Engage in sustainable on-call response practices, perform root-cause analyses, and lead blameless post-mortem reviews to avert future occurrences.

• Implement monitoring tools (SLIs/SLOs/SLAs) and establish automated alerting and metrics to oversee system health and performance.

• Enforce security best practices and guarantee that all systems comply with regulatory standards.


⛳️ Requirements

• Bachelor’s or Master’s degree in Computer Science, Information Technology, Computer Engineering, or a related discipline.

• Over 3 years of experience as a Site Reliability Engineer, Systems Engineer, or Software Engineer.

• Proficient in at least one high-level programming language (e.g., Python, Go, C++, or Java) and shell scripting.

• Strong grasp of data structures and algorithms.

• Solid understanding of Linux operating systems and open-source technologies, along with a comprehensive knowledge of network architecture.

• Competent in relational database systems and database modeling.

• Experience with containers and container orchestration platforms, such as Docker and Kubernetes (preferred).

• Proficiency in or exposure to machine learning frameworks like TensorFlow, PyTorch, MXNet, or PaddlePaddle (preferred).

• Practical experience with monitoring tools and methodologies (e.g., Prometheus, Grafana) (preferred).

• Strategic thinker with exceptional communication skills and the ability to collaborate effectively with cross-functional teams in a dynamic environment.


🏝️ Benefits

• Attractive remuneration and excellent perks.

• Comprehensive medical, insurance, and social security coverage.

• World-class workspaces.

• Engaging activities and recognition programs.

• Strong learning and development plans to foster your career growth.

• Positive work culture that supports your future.

• Convenient location with direct public transport links.

• Flexible working arrangements.

• Coaching and mentoring from industry experts.

People also viewed

The CodestJul 26

DevOps Engineer

PL flagPoland OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
IRIUMJul 26

Ingeniero/a Cloud DevOps

ES flagSpain OnlyFull-timeDevOps & Site Reliability Engineer (SRE)€33k – €40k/year
ApplyView job
SólidesJul 26

Senior DevOps Engineer

BR flagBrazil OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
ResilincJul 25

Junior/Senior Site Reliability Engineer – Night Shift

IN flagIndia OnlyFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
Verity GroupJul 25

Senior SRE / DevOps Engineer

Anywhere in the WorldFull-timeDevOps & Site Reliability Engineer (SRE)
ApplyView job
HOESSLER & HOESSLERJul 25

DevOps Software Engineer – Career Ambitions

DE flagGermany OnlyFull-timeDevOps & Site Reliability Engineer (SRE)€65k – €75k/year
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers